2026-05-21
cs.AR - Architecture
| 标题 | 作者 | 发布日期 | 摘要 | |
|---|---|---|---|---|
| ACALSim: A Scalable Parallel Simulation Framework for High-Performance System Design Space Exploration | Wei-Fen Lin, Jen-Chien Chang, Yen-Po Chen, Zi-Yi Tai, Yu-Cheng Chang, Chia-Pao Chiang, Yu-Yang Lee, Yu-Jie Wan | 2026-05-21 | 下载 | Architectural simulation has become the critical bottleneck limiting design space exploration for high-performance computing systems. Modern GPUs and AI accelerators -- with hundreds to thousands of t... |
| ORBIS: Output-Guided Token Reduction with Distribution-Aware Matching for Video Diffusion Acceleration | Hangyeol Lee, Joo-Young Kim | 2026-05-21 | 下载 | Diffusion Transformer (DiT) has emerged as a powerful model architecture for generating high-quality images and videos. In the case of video DiT, 3D Spatio-Temporal Attention increases token length in... |
| NasZip: Software and Hardware Co-Design to Accelerate Approximate Nearest Neighbor Search with DIMM-Based Near-Data Processing | Cheng Zou, Shuo Yang, Chen Nie, Yu Zou, Yu He, Chao Jiang, Limin Xiao, Weifeng Zhang, Zhezhi He | 2026-05-21 | 下载 | As large language models (LLMs) continue to advance, retrieval-augmented generation (RAG) has become the key mechanism for expanding model knowledge and reducing hallucinations. |
| Emerging memory technologies at room/cryogenic temperature | Siddhartha Raman Sundara Raman | 2026-05-21 | 下载 | As conventional technology scaling approaches physical and power limitations, modern computing systems increasingly face performance bottlenecks arising from memory latency, energy consumption, scalab... |
| CompPow: A Case for Component-level GPU Power Management | Shaizeen Aga, Mohamed Assem Ibrahim | 2026-05-21 | 下载 | The ever increasing demand for ML-driven intelligence in a wide spectrum of domains has led to ubiquity of GPUs. At the same time, GPUs are notorious for their power consumption needs and often domina... |
cs.DC - Distributed, Parallel, and Cluster Computing
| 标题 | 作者 | 发布日期 | 摘要 | |
|---|---|---|---|---|
| Orbax: Distributed Checkpointing with JAX | Colin Gaffney, Shutong Li, Daniel Ng, Anastasia Petrushkina, Niket Kumar, Adam Cogdell, Mridul Sahu, Yaning Liang, Nikhil Bansal, Justin Pan, Angel Mau, Abhishek Agrawal, Marco Berlot, Ruoxin Sang, Kiranbir Sodhia, Rakesh Iyer | 2026-05-21 | 下载 | In a landscape of high-performance distributed ML systems, JAX has emerged as a framework of choice. However, JAX's modular design philosophy leaves it without a standardized checkpointing solution. |
| AI-Driven Multi-Region Provisioning for Cloud Services Using Spot Fleets | Javier Fabra, Enrique Molina-Giménez, Pedro García-López | 2026-05-21 | 下载 | Cloud service platforms increasingly rely on elastic infrastructures to support dynamic workloads. Spot instances provide discounted computing resources but introduce uncertainty due to dynamic pricin... |
| A Generalized Nash Equilibrium-Seeking Scheme for Trauma Resuscitation | Promise Ekpo, Angelique Taylor, Lekan Molu | 2026-05-21 | 下载 | Trauma resuscitation is a clinical process for treating life-threatening physiological disorders in safety-critical environments, driven by the experience of healthcare workers (HCWs). |
| Relay-Based Synchronization of Replicated Data Types in Opportunistic Networks | Frédéric Guidec, Yves Mahéo | 2026-05-21 | 下载 | In Opportunistic Networks (OppNets), the dissemination of information can only rely on transient pairwise radio contacts between mobile devices (peers). |
| Exploiting Multicast for Accelerating Collective Communication | Chao Xu, Xu Zhang, Zihang Luo, Yuyan Wu, Guoxin Qian, Yufeng Yao, Chihyung Wang, Jingbin Zhou | 2026-05-21 | 下载 | Reducing collective communication latency is a critical goal for large model training and inference in both academia and industry. Many-to-many communications, such as AllGather and AlltoAll (dispatch... |
| Monotone Erasure Codes | Vivien Bammert, Annalisa Cimatti, Orestis Alpos, Giuliano Losa, Christian Cachin | 2026-05-21 | 下载 | Erasure codes are a critical component in reliable storage systems today, and many blockchain systems use consensus protocols that involve erasure codes to reduce their communication cost. |
| Asymmetric Virtual Memory Paging for Hybrid Mamba-Transformer Inference | An Xuan Nguyen | 2026-05-21 | 下载 | Hybrid language models like Jamba mix attention layers with State Space Models (SSMs), creating two memory cache types with opposite profiles: Key-Value (KV) caches grow linearly with sequence length,... |
| Nf-PEAK: Process-Based Energy Attribution for Nextflow Workflows on Kubernetes Clusters | Philipp Thamm, Somayeh Mohammadi, Kathleen West, Knut Reinert, Lauritz Thamsen, Ulf Leser | 2026-05-21 | 下载 | Scientific workflows are pipelines of interdependent tasks. They are increasingly executed on shared Kubernetes clusters via workflow engines such as Nextflow. |
| SepsisAI Orchestrator: A Containerized and Scalable Platform for Deploying AI Models and Real-Time Monitoring in Early Sepsis Detection | Santiago Ospitia, John Sanabria, John Garcia-Henao | 2026-05-21 | 下载 | Despite strong predictive results in the clinical machine learning literature, the translation of these models into bedside use remains limited by systems-level barriers: heterogeneous data representa... |
| Secure and Parallel Determinant Computation for Large-Scale Matrices in Edge Environments | Prajwal Panth | 2026-05-21 | 下载 | The advent of edge computing has enabled resource-constrained clients to delegate intensive computational tasks to distributed edge servers, especially within Internet of Things (IoT) environments. |
| LiveR: Fine-Grained Elasticity via Live Reconfiguration for Model Training | Haoyuan Liu, Kairui Zhou, Shuyao Qi, Qinwei Yang, Shengkai Lin, Shizhen Zhao, Wei Zhang | 2026-05-21 | 下载 | To reduce user costs and maximize cluster utilization, large model training increasingly leverages volatile but inexpensive GPU capacity, such as spot instances and reclaimable resources in shared clu... |
| NasZip: Software and Hardware Co-Design to Accelerate Approximate Nearest Neighbor Search with DIMM-Based Near-Data Processing | Cheng Zou, Shuo Yang, Chen Nie, Yu Zou, Yu He, Chao Jiang, Limin Xiao, Weifeng Zhang, Zhezhi He | 2026-05-21 | 下载 | As large language models (LLMs) continue to advance, retrieval-augmented generation (RAG) has become the key mechanism for expanding model knowledge and reducing hallucinations. |
| CompPow: A Case for Component-level GPU Power Management | Shaizeen Aga, Mohamed Assem Ibrahim | 2026-05-21 | 下载 | The ever increasing demand for ML-driven intelligence in a wide spectrum of domains has led to ubiquity of GPUs. At the same time, GPUs are notorious for their power consumption needs and often domina... |
cs.NI - Networking and Internet Architecture
| 标题 | 作者 | 发布日期 | 摘要 | |
|---|---|---|---|---|
| DRL-Driven Edge-Aware Utility Optimization for Multi-Slice 6G Networks | Khaled M. Naguib, Soumaya Cherkaoui, Mahmoud M. Elmessalawy, Ahmed M. Abd El-Haleem, Ibrahim I. Ibrahim | 2026-05-21 | 下载 | Virtual Reality (VR) services delivered over 6G networks demand ultra-low latency and high bandwidth to ensure seamless user experiences. This paper presents an intelligent resource allocation and edg... |
| An intensive vRAN deployment with OpenAirInterface | Romain Beurdouche, Raymond Knopp | 2026-05-21 | 下载 | The advent of 5G virtualized Radio Access Networks (vRANs) brings a new challenge with regards to computer architectures. It requires to select or design computing technologies that provide a sufficie... |
| UNAD+: An Explainable Hybrid Framework for Unknown Network Attack Detection | Saif Alzubi, Frederic Stahl | 2026-05-21 | 下载 | The detection of previously unseen network attacks remains a major challenge for intrusion detection systems. Although supervised learning methods often perform well on known attack classes, they are ... |
| SCALE: Sensitivity-Aware Federated Unlearning with Information Freshness Optimization for Mobile Edge Computing | Zihao Ding, Beining Wu, Jun Huang | 2026-05-21 | 下载 | Federated Unlearning (FU) is emerging as a powerful tool that enables the selective removal of client data to effectively address data contamination and meet strict privacy regulations in mobile edge ... |
| EnCoR: An end-to-end architecture for simplifying cellular networks | Wesley Woo, Zhuowei Wen, Monniiesh Velmurugan, Richard Raad, Sylvia Ratnasamy, Scott Shenker, Shaddi Hasan | 2026-05-21 | 下载 | Since their creation, cellular networks have made in-network mobility support a key feature of their service model. While this approach provides seamless connectivity for legacy traffic, it has the si... |
| Relay-Based Synchronization of Replicated Data Types in Opportunistic Networks | Frédéric Guidec, Yves Mahéo | 2026-05-21 | 下载 | In Opportunistic Networks (OppNets), the dissemination of information can only rely on transient pairwise radio contacts between mobile devices (peers). |
| Eliminating Premature Termination in Multihop Rendezvous for Cognitive Radio-based Emergency Response Network | Zahid Ali, Saritha Unnikrishnan, Eoghan Furey, Ian McLoughlin, Saim Ghafoor | 2026-05-21 | 下载 | In post-disaster environments, damaged communication infrastructure severely limits coordination among emergency response teams. Cognitive radio networks (CRNs) enable rapidly deployable communication... |
| Throughput and Delay Performance of Slotted Aloha in SmartBANs under Saturation Conditions | Anastasios C. Politis, Constantinos S. Hilas | 2026-05-21 | 下载 | This letter evaluates the performance of the slotted Aloha protocol defined by the European Telecommunication Standard Institute (ETSI) SmartBAN specification, under saturation conditions. |
| Impact of Atmospheric Turbulence and Pointing Error on Earth Observation | Celia Sánchez-de-Miguel, Antonio M. Mercado-Martínez, Beatriz Soret, Antonio Jurado-Navas, Miguel Castillo-Vázquez | 2026-05-21 | 下载 | Earth Observation (EO) imagery is often degraded by atmospheric turbulence and pointing jitter; yet, these effects are rarely considered in datasets used to train AI-based detection models. |
| Latency in Real-Time 3D Volumetric Streaming: A Comprehensive Study | Seungwoo Hong, Hosun Yoon, Seong Moon, Inayat Ali | 2026-05-21 | 下载 | Real-time 3D volumetric streaming is a transformative technology that enables the seamless transmission and rendering of high-fidelity 3D models, enhancing applications in virtual reality (VR), augmen... |
| Astragalus: Automatic Configuration Repair for Production Networks | Zhenrong Gu, Peng Zhang, Xing Feng, Xu Liu | 2026-05-21 | 下载 | Network configurations are prone to errors, which can lead to catastrophic service outages. A tool that can achieve automatic configuration repair (ACR) is highly desired by operators. |
| Toward Realistic Wi-Fi Fault Diagnosis: A Multi-Modal Benchmark | Junjian Zhang, Haobo Deng, Xinxin Li, Ming Zhao, Fengxiao Tang, Nei Kato | 2026-05-21 | 下载 | Intelligent network operation and maintenance systems in modern networks continuously generate large volumes of multi-modal operational data. However, Wi-Fi fault diagnosis under heterogeneous operati... |
| Lost in the Prefix: Revisiting IP Geolocation Accuracy Across Networks and Geographies | Syed Tauhidun Nabi, Jocelyn Bliton, Tijay Chung, Shaddi Hasan | 2026-05-21 | 下载 | IP geolocation databases are widely used in research, policy, and industry, yet their accuracy across network types and geographies remains poorly characterized. |
| Resilience Characterization of AI-Native Wireless Receivers via Persistent Homology | Christo Kurisummoottil Thomas, Emilio Calvanese Strinati | 2026-05-21 | 下载 | AI-native wireless receivers based on deep learning exhibit remarkable performance under stationary channel conditions, yet their resilience to distributional shifts remains poorly characterized by co... |
| AdaPTwin: Adaptive Multi-Fidelity Predictive Digital Twin for Proactive Radio Resource Management in Vehicular Networks | Armin Makvandi, Md. Zoheb Hassan, Md. Jahangir Hossain | 2026-05-21 | 下载 | The highly dynamic nature of vehicular networks necessitates proactive and site-specific radio resource management (RRM) to achieve ultra-reliable low-latency communications. |
cs.OS - Operating Systems
| 标题 | 作者 | 发布日期 | 摘要 | |
|---|---|---|---|---|
| DeltaBox: Scaling Stateful AI Agents with Millisecond-Level Sandbox Checkpoint/Rollback | Yunpeng Dong, Jingkai He, Yuze Hou, Dong Du, Zhonghu Xu, Si Yu, Yubin Xia, Haibo Chen | 2026-05-21 | 下载 | LLM-powered AI agents require high-frequency state exploration (e.g., test-time tree search and reinforcement learning), relying on rapid checkpoint and rollback (C/R) of the complete sandbox state, i... |
cs.PF - Performance
| 标题 | 作者 | 发布日期 | 摘要 | |
|---|---|---|---|---|
| ModeSwitch-LLM: A Lightweight Phase-Aware Controller for Cross-Mode LLM Inference on a Single GPU | Aman Sunesh, Ali Alshehhi, Hivansh Dhakne | 2026-05-21 | 下载 | ModeSwitch-LLM is a lightweight request-boundary controller for improving single-GPU large language model inference efficiency by routing each request to an appropriate fixed inference mode. |
| ACALSim: A Scalable Parallel Simulation Framework for High-Performance System Design Space Exploration | Wei-Fen Lin, Jen-Chien Chang, Yen-Po Chen, Zi-Yi Tai, Yu-Cheng Chang, Chia-Pao Chiang, Yu-Yang Lee, Yu-Jie Wan | 2026-05-21 | 下载 | Architectural simulation has become the critical bottleneck limiting design space exploration for high-performance computing systems. Modern GPUs and AI accelerators -- with hundreds to thousands of t... |
| Asymmetric Virtual Memory Paging for Hybrid Mamba-Transformer Inference | An Xuan Nguyen | 2026-05-21 | 下载 | Hybrid language models like Jamba mix attention layers with State Space Models (SSMs), creating two memory cache types with opposite profiles: Key-Value (KV) caches grow linearly with sequence length,... |