Skip to content

2026-04-19 ​

cs.AR - Architecture ​

标题作者发布日期PDF摘要
Privatar: Scalable Privacy-preserving Multi-user VR via Secure OffloadingJianming Tong, Hanshen Xiao, Krishna Kumar Nair, Hao Kang, Ashish Sirasao, Ziqi Zhang, G. Edward Suh, Tushar Krishna2026-04-19下载Multi-user virtual reality enables immersive interaction. However, rendering avatars for numerous participants on each headset incurs prohibitive computational overhead, limiting scalability.
RISC-V Functional Safety for Autonomous Automotive Systems: An Analytical Framework and Research Roadmap for ML-Assisted CertificationNick Andreasyan, Mikhail Struve, Alexey Popov, Maksim Nikolaev, Vadim Vashkelis2026-04-19下载RISC-V is emerging as a viable platform for automotive-grade embedded computing, with recent ISO 26262 ASIL-D certifications demonstrating readiness for safety-critical deployment in autonomous drivin...
Clover: A Neural-Symbolic Agentic Harness with Stochastic Tree-of-Thoughts for Verified RTL RepairZizhang Luo, Yansong Xu, Runlin Guo, Fan Cui, Kexing Zhou, Mile Xia, Hongyuan Hou, Yuhao Luo, Yun Liang2026-04-19下载RTL program repair remains a critical bottleneck in hardware design and verification. Traditional automatic program repair (APR) methods rely on predefined templates and synthesis, limiting their bug ...
Bit-Flip Vulnerability of Shared KV-Cache Blocks in LLM Serving SystemsYuji Yamamoto, Satoshi Matsuura2026-04-19下载Rowhammer on GPU DRAM has enabled adversarial bit flips in model weights; shared KV-cache blocks in LLM serving systems present an analogous but previously unexamined target.

cs.DC - Distributed, Parallel, and Cluster Computing ​

标题作者发布日期PDF摘要
Towards Energy Efficient Co-Scheduling in HPCZhong Zheng, Michael E. Papka, Zhiling Lan2026-04-19下载Modern multi GPU HPC systems expose substantial computational capacity, yet inefficient GPU allocation often leads to wasted energy and underutilization.
EcoShift: Performance-Aware Power Management for Power-Constrained Heterogeneous SystemsZhong Zheng, Michael E. Papka, Zhiling Lan2026-04-19下载Power-constrained HPC systems increasingly run heterogeneous CPU--GPU applications under strict cluster-wide power limits. Existing cluster-wide power management policies rely on fair-share or utiliza...
SLO-Guard: Crash-Aware, Budget-Consistent Autotuning for SLO-Constrained LLM ServingChristian Lysenstøen2026-04-19下载Serving large language models under latency service-level objectives (SLOs) is a configuration-heavy systems problem with an unusually failure-prone search space: many plausible configurations crash o...
Flint: Compiler Enabled Cluster-Free Design Space Exploration for Distributed MLJinsun Yoo, Meghan Cowan, Zheng Du, Changhai Man, Srinivas Sridharan, Tushar Krishna2026-04-19下载Design space exploration for future distributed Machine Learning systems suffers from a lack of readily available workload representation that enables flexible exploration across the stack.
Active Inference-Based Adaptive Routing for Heterogeneous Edge AI ServicesZihang Wang, Boris Sedlak, Schahram Dustdar2026-04-19下载Edge computing enables AI inference closer to data sources, reducing latency and bandwidth costs. However, orchestrating AI services across the cloud-edge continuum remains challenging due to dynamic ...
Hive: A Multi-Agent Infrastructure for Algorithm- and Task-Level ScalingZizhang Luo, Yuhao Luo, Youwei Xiao, Yansong Xu, Runlin Guo, Yun Liang2026-04-19下载Large language models are increasingly deployed as complex agentic systems that scale with task complexity. While prior work has extensively explored model- and system-level scaling, algorithm- and ta...
Cloud-native and Distributed Systems for Efficient and Scalable Large Language Models -- A Research AgendaMinxian Xu, Jingfeng Wu, Shengye Song, Satish Narayana Srirama, Bahman Javad, Rajiv Ranjan, Devki Nandan Jha, Sa Wang, Wenhong Tian, Huanle Xu, Li Li, Zizhao Mo, Shuo Ren, Thomas Kunz, Petar Kochovski, Vlado Stankovski, Kejiang Ye, Chengzhong Xu, Rajkumar Buyya2026-04-19下载The rapid rise of Large Language Models (LLMs) has revolutionized various artificial intelligence (AI) applications, from natural language processing to code generation.
CCCL: In-GPU Compression-Coupled Collective CommunicationChon Lam Lao, Zhiying Xu, Zhuang Wang, Ziming Mao, Delong Meng, Jia Zhen, Jun Wu, Ion Stoica, Yida Wang, Yang Zhou2026-04-19下载Collective communication incurs significant overhead in LLM workloads. Although overlapping communication with computation in application-level is a common strategy, it often requires substantial code...

cs.NI - Networking and Internet Architecture ​

标题作者发布日期PDF摘要
Scheduling in Multi-Hop Wireless Networks With DeadlinesNicholas Jones, Eytan Modiano2026-04-19下载We analyze the problem of scheduling in wireless networks to meet end-to-end service guarantees, defined by instantaneous throughput and hard packet deadlines.
Safety-Aware AoI Scheduling for LEO Satellite-Assisted Autonomous DrivingKangkang Sun, Junyi He, Juntong Liu, Xiuzhen Chen, Jianhua Li, Minyi Guo2026-04-19下载Autonomous platoons traversing infrastructure gaps increasingly depend on LEO satellite backhaul for safety-critical updates, yet no existing framework jointly addresses compound Doppler from simultan...
Decentralised Trust and Security Mechanisms for IoT Networks at the Edge: A Comprehensive ReviewKhandoker Ashik Uz Zaman, Mahdi H. Miraz, Mohammed N. M. Ali2026-04-19下载INTRODUCTION: The proliferation of the amalgamation of IoT and edge computing has increased the demand for decentralised trust and security mechanisms capable of operating across heterogeneous and res...

cs.PF - Performance ​

标题作者发布日期PDF摘要
SLO-Guard: Crash-Aware, Budget-Consistent Autotuning for SLO-Constrained LLM ServingChristian Lysenstøen2026-04-19下载Serving large language models under latency service-level objectives (SLOs) is a configuration-heavy systems problem with an unusually failure-prone search space: many plausible configurations crash o...
Active Inference-Based Adaptive Routing for Heterogeneous Edge AI ServicesZihang Wang, Boris Sedlak, Schahram Dustdar2026-04-19下载Edge computing enables AI inference closer to data sources, reducing latency and bandwidth costs. However, orchestrating AI services across the cloud-edge continuum remains challenging due to dynamic ...
BranchBench: Aligning Database Branching with Agentic DemandsElaine Ang, Sam Weldon, In Keun Kim, Kevin Durand, Kostis Kaffes, Eugene Wu2026-04-19下载Branchable databases are evolving from developer tools to infrastructure for agentic workloads characterized by speculative mutations and non-linear state exploration.

基于 VitePress 构建 · 使用本地搜索查找论文