2026-04-19
cs.AR - Architecture
| 标题 | 作者 | 发布日期 | 摘要 | |
|---|---|---|---|---|
| Privatar: Scalable Privacy-preserving Multi-user VR via Secure Offloading | Jianming Tong, Hanshen Xiao, Krishna Kumar Nair, Hao Kang, Ashish Sirasao, Ziqi Zhang, G. Edward Suh, Tushar Krishna | 2026-04-19 | 下载 | Multi-user virtual reality enables immersive interaction. However, rendering avatars for numerous participants on each headset incurs prohibitive computational overhead, limiting scalability. |
| RISC-V Functional Safety for Autonomous Automotive Systems: An Analytical Framework and Research Roadmap for ML-Assisted Certification | Nick Andreasyan, Mikhail Struve, Alexey Popov, Maksim Nikolaev, Vadim Vashkelis | 2026-04-19 | 下载 | RISC-V is emerging as a viable platform for automotive-grade embedded computing, with recent ISO 26262 ASIL-D certifications demonstrating readiness for safety-critical deployment in autonomous drivin... |
| Clover: A Neural-Symbolic Agentic Harness with Stochastic Tree-of-Thoughts for Verified RTL Repair | Zizhang Luo, Yansong Xu, Runlin Guo, Fan Cui, Kexing Zhou, Mile Xia, Hongyuan Hou, Yuhao Luo, Yun Liang | 2026-04-19 | 下载 | RTL program repair remains a critical bottleneck in hardware design and verification. Traditional automatic program repair (APR) methods rely on predefined templates and synthesis, limiting their bug ... |
| Bit-Flip Vulnerability of Shared KV-Cache Blocks in LLM Serving Systems | Yuji Yamamoto, Satoshi Matsuura | 2026-04-19 | 下载 | Rowhammer on GPU DRAM has enabled adversarial bit flips in model weights; shared KV-cache blocks in LLM serving systems present an analogous but previously unexamined target. |
cs.DC - Distributed, Parallel, and Cluster Computing
| 标题 | 作者 | 发布日期 | 摘要 | |
|---|---|---|---|---|
| Towards Energy Efficient Co-Scheduling in HPC | Zhong Zheng, Michael E. Papka, Zhiling Lan | 2026-04-19 | 下载 | Modern multi GPU HPC systems expose substantial computational capacity, yet inefficient GPU allocation often leads to wasted energy and underutilization. |
| EcoShift: Performance-Aware Power Management for Power-Constrained Heterogeneous Systems | Zhong Zheng, Michael E. Papka, Zhiling Lan | 2026-04-19 | 下载 | Power-constrained HPC systems increasingly run heterogeneous CPU--GPU applications under strict cluster-wide power limits. Existing cluster-wide power management policies rely on fair-share or utiliza... |
| SLO-Guard: Crash-Aware, Budget-Consistent Autotuning for SLO-Constrained LLM Serving | Christian Lysenstøen | 2026-04-19 | 下载 | Serving large language models under latency service-level objectives (SLOs) is a configuration-heavy systems problem with an unusually failure-prone search space: many plausible configurations crash o... |
| Flint: Compiler Enabled Cluster-Free Design Space Exploration for Distributed ML | Jinsun Yoo, Meghan Cowan, Zheng Du, Changhai Man, Srinivas Sridharan, Tushar Krishna | 2026-04-19 | 下载 | Design space exploration for future distributed Machine Learning systems suffers from a lack of readily available workload representation that enables flexible exploration across the stack. |
| Active Inference-Based Adaptive Routing for Heterogeneous Edge AI Services | Zihang Wang, Boris Sedlak, Schahram Dustdar | 2026-04-19 | 下载 | Edge computing enables AI inference closer to data sources, reducing latency and bandwidth costs. However, orchestrating AI services across the cloud-edge continuum remains challenging due to dynamic ... |
| Hive: A Multi-Agent Infrastructure for Algorithm- and Task-Level Scaling | Zizhang Luo, Yuhao Luo, Youwei Xiao, Yansong Xu, Runlin Guo, Yun Liang | 2026-04-19 | 下载 | Large language models are increasingly deployed as complex agentic systems that scale with task complexity. While prior work has extensively explored model- and system-level scaling, algorithm- and ta... |
| Cloud-native and Distributed Systems for Efficient and Scalable Large Language Models -- A Research Agenda | Minxian Xu, Jingfeng Wu, Shengye Song, Satish Narayana Srirama, Bahman Javad, Rajiv Ranjan, Devki Nandan Jha, Sa Wang, Wenhong Tian, Huanle Xu, Li Li, Zizhao Mo, Shuo Ren, Thomas Kunz, Petar Kochovski, Vlado Stankovski, Kejiang Ye, Chengzhong Xu, Rajkumar Buyya | 2026-04-19 | 下载 | The rapid rise of Large Language Models (LLMs) has revolutionized various artificial intelligence (AI) applications, from natural language processing to code generation. |
| CCCL: In-GPU Compression-Coupled Collective Communication | Chon Lam Lao, Zhiying Xu, Zhuang Wang, Ziming Mao, Delong Meng, Jia Zhen, Jun Wu, Ion Stoica, Yida Wang, Yang Zhou | 2026-04-19 | 下载 | Collective communication incurs significant overhead in LLM workloads. Although overlapping communication with computation in application-level is a common strategy, it often requires substantial code... |
cs.NI - Networking and Internet Architecture
| 标题 | 作者 | 发布日期 | 摘要 | |
|---|---|---|---|---|
| Scheduling in Multi-Hop Wireless Networks With Deadlines | Nicholas Jones, Eytan Modiano | 2026-04-19 | 下载 | We analyze the problem of scheduling in wireless networks to meet end-to-end service guarantees, defined by instantaneous throughput and hard packet deadlines. |
| Safety-Aware AoI Scheduling for LEO Satellite-Assisted Autonomous Driving | Kangkang Sun, Junyi He, Juntong Liu, Xiuzhen Chen, Jianhua Li, Minyi Guo | 2026-04-19 | 下载 | Autonomous platoons traversing infrastructure gaps increasingly depend on LEO satellite backhaul for safety-critical updates, yet no existing framework jointly addresses compound Doppler from simultan... |
| Decentralised Trust and Security Mechanisms for IoT Networks at the Edge: A Comprehensive Review | Khandoker Ashik Uz Zaman, Mahdi H. Miraz, Mohammed N. M. Ali | 2026-04-19 | 下载 | INTRODUCTION: The proliferation of the amalgamation of IoT and edge computing has increased the demand for decentralised trust and security mechanisms capable of operating across heterogeneous and res... |
cs.PF - Performance
| 标题 | 作者 | 发布日期 | 摘要 | |
|---|---|---|---|---|
| SLO-Guard: Crash-Aware, Budget-Consistent Autotuning for SLO-Constrained LLM Serving | Christian Lysenstøen | 2026-04-19 | 下载 | Serving large language models under latency service-level objectives (SLOs) is a configuration-heavy systems problem with an unusually failure-prone search space: many plausible configurations crash o... |
| Active Inference-Based Adaptive Routing for Heterogeneous Edge AI Services | Zihang Wang, Boris Sedlak, Schahram Dustdar | 2026-04-19 | 下载 | Edge computing enables AI inference closer to data sources, reducing latency and bandwidth costs. However, orchestrating AI services across the cloud-edge continuum remains challenging due to dynamic ... |
| BranchBench: Aligning Database Branching with Agentic Demands | Elaine Ang, Sam Weldon, In Keun Kim, Kevin Durand, Kostis Kaffes, Eugene Wu | 2026-04-19 | 下载 | Branchable databases are evolving from developer tools to infrastructure for agentic workloads characterized by speculative mutations and non-linear state exploration. |