Skip to content

2026-05-21 ​

cs.AR - Architecture ​

标题作者发布日期PDF摘要
ACALSim: A Scalable Parallel Simulation Framework for High-Performance System Design Space ExplorationWei-Fen Lin, Jen-Chien Chang, Yen-Po Chen, Zi-Yi Tai, Yu-Cheng Chang, Chia-Pao Chiang, Yu-Yang Lee, Yu-Jie Wan2026-05-21下载Architectural simulation has become the critical bottleneck limiting design space exploration for high-performance computing systems. Modern GPUs and AI accelerators -- with hundreds to thousands of t...
ORBIS: Output-Guided Token Reduction with Distribution-Aware Matching for Video Diffusion AccelerationHangyeol Lee, Joo-Young Kim2026-05-21下载Diffusion Transformer (DiT) has emerged as a powerful model architecture for generating high-quality images and videos. In the case of video DiT, 3D Spatio-Temporal Attention increases token length in...
NasZip: Software and Hardware Co-Design to Accelerate Approximate Nearest Neighbor Search with DIMM-Based Near-Data ProcessingCheng Zou, Shuo Yang, Chen Nie, Yu Zou, Yu He, Chao Jiang, Limin Xiao, Weifeng Zhang, Zhezhi He2026-05-21下载As large language models (LLMs) continue to advance, retrieval-augmented generation (RAG) has become the key mechanism for expanding model knowledge and reducing hallucinations.
Emerging memory technologies at room/cryogenic temperatureSiddhartha Raman Sundara Raman2026-05-21下载As conventional technology scaling approaches physical and power limitations, modern computing systems increasingly face performance bottlenecks arising from memory latency, energy consumption, scalab...
CompPow: A Case for Component-level GPU Power ManagementShaizeen Aga, Mohamed Assem Ibrahim2026-05-21下载The ever increasing demand for ML-driven intelligence in a wide spectrum of domains has led to ubiquity of GPUs. At the same time, GPUs are notorious for their power consumption needs and often domina...

cs.DC - Distributed, Parallel, and Cluster Computing ​

标题作者发布日期PDF摘要
Orbax: Distributed Checkpointing with JAXColin Gaffney, Shutong Li, Daniel Ng, Anastasia Petrushkina, Niket Kumar, Adam Cogdell, Mridul Sahu, Yaning Liang, Nikhil Bansal, Justin Pan, Angel Mau, Abhishek Agrawal, Marco Berlot, Ruoxin Sang, Kiranbir Sodhia, Rakesh Iyer2026-05-21下载In a landscape of high-performance distributed ML systems, JAX has emerged as a framework of choice. However, JAX's modular design philosophy leaves it without a standardized checkpointing solution.
AI-Driven Multi-Region Provisioning for Cloud Services Using Spot FleetsJavier Fabra, Enrique Molina-Giménez, Pedro García-López2026-05-21下载Cloud service platforms increasingly rely on elastic infrastructures to support dynamic workloads. Spot instances provide discounted computing resources but introduce uncertainty due to dynamic pricin...
A Generalized Nash Equilibrium-Seeking Scheme for Trauma ResuscitationPromise Ekpo, Angelique Taylor, Lekan Molu2026-05-21下载Trauma resuscitation is a clinical process for treating life-threatening physiological disorders in safety-critical environments, driven by the experience of healthcare workers (HCWs).
Relay-Based Synchronization of Replicated Data Types in Opportunistic NetworksFrédéric Guidec, Yves Mahéo2026-05-21下载In Opportunistic Networks (OppNets), the dissemination of information can only rely on transient pairwise radio contacts between mobile devices (peers).
Exploiting Multicast for Accelerating Collective CommunicationChao Xu, Xu Zhang, Zihang Luo, Yuyan Wu, Guoxin Qian, Yufeng Yao, Chihyung Wang, Jingbin Zhou2026-05-21下载Reducing collective communication latency is a critical goal for large model training and inference in both academia and industry. Many-to-many communications, such as AllGather and AlltoAll (dispatch...
Monotone Erasure CodesVivien Bammert, Annalisa Cimatti, Orestis Alpos, Giuliano Losa, Christian Cachin2026-05-21下载Erasure codes are a critical component in reliable storage systems today, and many blockchain systems use consensus protocols that involve erasure codes to reduce their communication cost.
Asymmetric Virtual Memory Paging for Hybrid Mamba-Transformer InferenceAn Xuan Nguyen2026-05-21下载Hybrid language models like Jamba mix attention layers with State Space Models (SSMs), creating two memory cache types with opposite profiles: Key-Value (KV) caches grow linearly with sequence length,...
Nf-PEAK: Process-Based Energy Attribution for Nextflow Workflows on Kubernetes ClustersPhilipp Thamm, Somayeh Mohammadi, Kathleen West, Knut Reinert, Lauritz Thamsen, Ulf Leser2026-05-21下载Scientific workflows are pipelines of interdependent tasks. They are increasingly executed on shared Kubernetes clusters via workflow engines such as Nextflow.
SepsisAI Orchestrator: A Containerized and Scalable Platform for Deploying AI Models and Real-Time Monitoring in Early Sepsis DetectionSantiago Ospitia, John Sanabria, John Garcia-Henao2026-05-21下载Despite strong predictive results in the clinical machine learning literature, the translation of these models into bedside use remains limited by systems-level barriers: heterogeneous data representa...
Secure and Parallel Determinant Computation for Large-Scale Matrices in Edge EnvironmentsPrajwal Panth2026-05-21下载The advent of edge computing has enabled resource-constrained clients to delegate intensive computational tasks to distributed edge servers, especially within Internet of Things (IoT) environments.
LiveR: Fine-Grained Elasticity via Live Reconfiguration for Model TrainingHaoyuan Liu, Kairui Zhou, Shuyao Qi, Qinwei Yang, Shengkai Lin, Shizhen Zhao, Wei Zhang2026-05-21下载To reduce user costs and maximize cluster utilization, large model training increasingly leverages volatile but inexpensive GPU capacity, such as spot instances and reclaimable resources in shared clu...
NasZip: Software and Hardware Co-Design to Accelerate Approximate Nearest Neighbor Search with DIMM-Based Near-Data ProcessingCheng Zou, Shuo Yang, Chen Nie, Yu Zou, Yu He, Chao Jiang, Limin Xiao, Weifeng Zhang, Zhezhi He2026-05-21下载As large language models (LLMs) continue to advance, retrieval-augmented generation (RAG) has become the key mechanism for expanding model knowledge and reducing hallucinations.
CompPow: A Case for Component-level GPU Power ManagementShaizeen Aga, Mohamed Assem Ibrahim2026-05-21下载The ever increasing demand for ML-driven intelligence in a wide spectrum of domains has led to ubiquity of GPUs. At the same time, GPUs are notorious for their power consumption needs and often domina...

cs.NI - Networking and Internet Architecture ​

标题作者发布日期PDF摘要
DRL-Driven Edge-Aware Utility Optimization for Multi-Slice 6G NetworksKhaled M. Naguib, Soumaya Cherkaoui, Mahmoud M. Elmessalawy, Ahmed M. Abd El-Haleem, Ibrahim I. Ibrahim2026-05-21下载Virtual Reality (VR) services delivered over 6G networks demand ultra-low latency and high bandwidth to ensure seamless user experiences. This paper presents an intelligent resource allocation and edg...
An intensive vRAN deployment with OpenAirInterfaceRomain Beurdouche, Raymond Knopp2026-05-21下载The advent of 5G virtualized Radio Access Networks (vRANs) brings a new challenge with regards to computer architectures. It requires to select or design computing technologies that provide a sufficie...
UNAD+: An Explainable Hybrid Framework for Unknown Network Attack DetectionSaif Alzubi, Frederic Stahl2026-05-21下载The detection of previously unseen network attacks remains a major challenge for intrusion detection systems. Although supervised learning methods often perform well on known attack classes, they are ...
SCALE: Sensitivity-Aware Federated Unlearning with Information Freshness Optimization for Mobile Edge ComputingZihao Ding, Beining Wu, Jun Huang2026-05-21下载Federated Unlearning (FU) is emerging as a powerful tool that enables the selective removal of client data to effectively address data contamination and meet strict privacy regulations in mobile edge ...
EnCoR: An end-to-end architecture for simplifying cellular networksWesley Woo, Zhuowei Wen, Monniiesh Velmurugan, Richard Raad, Sylvia Ratnasamy, Scott Shenker, Shaddi Hasan2026-05-21下载Since their creation, cellular networks have made in-network mobility support a key feature of their service model. While this approach provides seamless connectivity for legacy traffic, it has the si...
Relay-Based Synchronization of Replicated Data Types in Opportunistic NetworksFrédéric Guidec, Yves Mahéo2026-05-21下载In Opportunistic Networks (OppNets), the dissemination of information can only rely on transient pairwise radio contacts between mobile devices (peers).
Eliminating Premature Termination in Multihop Rendezvous for Cognitive Radio-based Emergency Response NetworkZahid Ali, Saritha Unnikrishnan, Eoghan Furey, Ian McLoughlin, Saim Ghafoor2026-05-21下载In post-disaster environments, damaged communication infrastructure severely limits coordination among emergency response teams. Cognitive radio networks (CRNs) enable rapidly deployable communication...
Throughput and Delay Performance of Slotted Aloha in SmartBANs under Saturation ConditionsAnastasios C. Politis, Constantinos S. Hilas2026-05-21下载This letter evaluates the performance of the slotted Aloha protocol defined by the European Telecommunication Standard Institute (ETSI) SmartBAN specification, under saturation conditions.
Impact of Atmospheric Turbulence and Pointing Error on Earth ObservationCelia Sánchez-de-Miguel, Antonio M. Mercado-Martínez, Beatriz Soret, Antonio Jurado-Navas, Miguel Castillo-Vázquez2026-05-21下载Earth Observation (EO) imagery is often degraded by atmospheric turbulence and pointing jitter; yet, these effects are rarely considered in datasets used to train AI-based detection models.
Latency in Real-Time 3D Volumetric Streaming: A Comprehensive StudySeungwoo Hong, Hosun Yoon, Seong Moon, Inayat Ali2026-05-21下载Real-time 3D volumetric streaming is a transformative technology that enables the seamless transmission and rendering of high-fidelity 3D models, enhancing applications in virtual reality (VR), augmen...
Astragalus: Automatic Configuration Repair for Production NetworksZhenrong Gu, Peng Zhang, Xing Feng, Xu Liu2026-05-21下载Network configurations are prone to errors, which can lead to catastrophic service outages. A tool that can achieve automatic configuration repair (ACR) is highly desired by operators.
Toward Realistic Wi-Fi Fault Diagnosis: A Multi-Modal BenchmarkJunjian Zhang, Haobo Deng, Xinxin Li, Ming Zhao, Fengxiao Tang, Nei Kato2026-05-21下载Intelligent network operation and maintenance systems in modern networks continuously generate large volumes of multi-modal operational data. However, Wi-Fi fault diagnosis under heterogeneous operati...
Lost in the Prefix: Revisiting IP Geolocation Accuracy Across Networks and GeographiesSyed Tauhidun Nabi, Jocelyn Bliton, Tijay Chung, Shaddi Hasan2026-05-21下载IP geolocation databases are widely used in research, policy, and industry, yet their accuracy across network types and geographies remains poorly characterized.
Resilience Characterization of AI-Native Wireless Receivers via Persistent HomologyChristo Kurisummoottil Thomas, Emilio Calvanese Strinati2026-05-21下载AI-native wireless receivers based on deep learning exhibit remarkable performance under stationary channel conditions, yet their resilience to distributional shifts remains poorly characterized by co...
AdaPTwin: Adaptive Multi-Fidelity Predictive Digital Twin for Proactive Radio Resource Management in Vehicular NetworksArmin Makvandi, Md. Zoheb Hassan, Md. Jahangir Hossain2026-05-21下载The highly dynamic nature of vehicular networks necessitates proactive and site-specific radio resource management (RRM) to achieve ultra-reliable low-latency communications.

cs.OS - Operating Systems ​

标题作者发布日期PDF摘要
DeltaBox: Scaling Stateful AI Agents with Millisecond-Level Sandbox Checkpoint/RollbackYunpeng Dong, Jingkai He, Yuze Hou, Dong Du, Zhonghu Xu, Si Yu, Yubin Xia, Haibo Chen2026-05-21下载LLM-powered AI agents require high-frequency state exploration (e.g., test-time tree search and reinforcement learning), relying on rapid checkpoint and rollback (C/R) of the complete sandbox state, i...

cs.PF - Performance ​

标题作者发布日期PDF摘要
ModeSwitch-LLM: A Lightweight Phase-Aware Controller for Cross-Mode LLM Inference on a Single GPUAman Sunesh, Ali Alshehhi, Hivansh Dhakne2026-05-21下载ModeSwitch-LLM is a lightweight request-boundary controller for improving single-GPU large language model inference efficiency by routing each request to an appropriate fixed inference mode.
ACALSim: A Scalable Parallel Simulation Framework for High-Performance System Design Space ExplorationWei-Fen Lin, Jen-Chien Chang, Yen-Po Chen, Zi-Yi Tai, Yu-Cheng Chang, Chia-Pao Chiang, Yu-Yang Lee, Yu-Jie Wan2026-05-21下载Architectural simulation has become the critical bottleneck limiting design space exploration for high-performance computing systems. Modern GPUs and AI accelerators -- with hundreds to thousands of t...
Asymmetric Virtual Memory Paging for Hybrid Mamba-Transformer InferenceAn Xuan Nguyen2026-05-21下载Hybrid language models like Jamba mix attention layers with State Space Models (SSMs), creating two memory cache types with opposite profiles: Key-Value (KV) caches grow linearly with sequence length,...

基于 VitePress 构建 · 使用本地搜索查找论文