Skip to content

2026-10-01 ​

cs.DC - Distributed, Parallel, and Cluster Computing ​

标题作者发布日期PDF摘要
MoE-CORE: Coordinated Expert Offloading and Residency for Memory-Constrained MoE InferenceKe Yang, Yongji Gao, Xushi Li, Kui Luo, Sicheng Zhang, Tianming Zhou, Keyi Liu, Shufang Lu, Aoxuan Chen, Jie Meng, Jingchun Gao, Dan Li, Xinkai You, Dan Li, Zhixiang Xia, Yan Shi, Yang Liu, Yanjia Zeng, Liangjun Feng2026-10-01下载Sparse expert activation reduces MoE models' computation, yet expert weights can exceed limited device memory. Offloading makes inference feasible on a compact AI appliance but exposes host-to-device ...
ePACT: Energy-Performance-Aware Commitment Tracking for LLM ServingYou Peng, Youhe Jiang, Chen Wang, Binhang Yuan2026-10-01下载Reducing LLM serving energy does not by itself guarantee lower deployment cost when electricity procurement exposes operators to unfavorable deviations from preset commitments.
Towards a Cloud Fog Edge System for Smart BuildingChristophe Cérin, Mamadou Sow, Frédéric Andrès2026-10-01下载In this article, we present our vision and recent advancements toward creating a decentralized system capable of learning from real-time data within buildings to support sustainable and privacy-preser...
Exploiting the Interplay of Compute- and Memory-Bound kernels in MPI ApplicationsAyesha Afzal, Krishna Manda, Georg Hager2026-10-01下载Parallel applications are often designed for synchronous, lock-step execution, treating communication stalls as performance hazards. Yet, in a communication-light application without frequent synchron...
GridSMR: Causal Compression for Sharded BlockchainsShir Cohen, Adam Alon, Raz Omessi, Amir Sarid, Dana Shamir, Ofir Zohar2026-10-01下载We present GridSMR, a sharded blockchain that scales execution horizontally while allowing dependent cross-shard operations to progress within a single block.
GPU-Initiated Communication: Dissecting Down to the BoneJavid Baydamirli, Ismayil Ismayilov, Kaan Oktay, Didem Unat2026-10-01下载GPU-initiated communication lets GPU threads post RDMA operations directly to the NIC. It underpins NVSHMEM, NCCL GIN, and DeepEP, which serve the fine-grained, latency-critical communication of Mixtu...
Blockchain Lifecycle Prediction - Dead CoinsUwe A. Kuehn, Syed Muhammad Adnan2026-10-01下载The cryptocurrency ecosystem has experienced extraordinary growth alongside an equally remarkable rate of failure, with over 52 percent of all tokens launched since 2021 ceasing to trade by early 2025...
RapidMoE: Exploiting Cross-Asymmetry via Adaptive Residual Offloading for Large-Scale MoE InferenceWenxun Wang, Likai Ma, Zongle Huang, Chen Tang, Yongpan Liu2026-10-01下载The widespread adoption of Mixture-of-Experts (MoE) has created a growing need for deployment on heterogeneous platforms. However, it exposes a fundamental mismatch between the algorithmic demands of ...
Serving a Revisable World: Versioned Execution for Interruptible AgentsYanxin Zhang, Rahul Sharma, Nitin Vegesna, Zheyu Fu, Chang Liu, Trivikram Krishnamurthy2026-10-01下载LLM agents revise running tasks when users change instructions, tools fail, or new information changes a plan. Today's servers express a revision as aborting old requests and submitting replacements.
The Other Half of Workflow Portability: Evidence-Backed HPC Site Profiles with Agentic DiscoveryMd Saiful Islam, Douglas Thain2026-10-01下载Moving a workflow developed and tested at one HPC site to another rarely succeeds without some amount of trial and error. Package managers rebuild software environments, containers ship whole filesyst...
Seamless Reconfiguration for DAG-BFTMichael Yiqing Hu, Leander Jehl, Paul Franke-Bergmann, Jialin Li, Christian Berger2026-10-01下载Byzantine Atomic Broadcast in Asynchronous networks has been studied extensively for decades. The FLP impossibility result rules out deterministic consensus in a fully asynchronous setting, motivating...
HakiCC: LLM-Driven Multi-Agent Design and Optimization of Concurrency Control ProtocolsFarzad Habibi, Juncheng Fang, Faisal Nawab2026-10-01下载Large language models (LLMs) have recently been applied in systems research as a tool to reduce human-intensive engineering effort through cost-efficient automation.

cs.NI - Networking and Internet Architecture ​

标题作者发布日期PDF摘要
From Network Intrusion Detection to Blockchain-Backed Endpoint Detection and Response: Mapping the Landscape of Decentralized Detection-and-Response ArchitecturesYahya Shahsavari, Sara Rouhani, Kaiwen Zhang2026-10-01下载While the literature on blockchain-assisted intrusion detection and prevention systems (IDS/IPS) for Internet of Things (IoT) and Industrial Internet of Things (IIoT) networks is mature, existing syst...
SL-RFSIM: Enabling Scalable Multi-Hop 5G NR Sidelink Mesh Networking in OpenAirInterfaceSimone Pio Candido, Jin Yan, Jérôme Härri2026-10-01下载Recent 3GPP releases have extended 5G New Radio (NR) Sidelink (SL) to support device-to-device (D2D) relay and multi-hop capabilities. This paper presents SL-RFSIM, a component-based experimentation f...
Protocol Integration of Physical Layer Deception into EAP-TEAP Wi-Fi AuthenticationMoustafa Ibrahim, Bin Han, Hans D. Schotten2026-10-01下载Credential-based Extensible Authentication Protocol (EAP) authentication cannot distinguish a legitimate credential holder from an adversary using compromised credentials.
Managing Context and Communication in Distributed Agentic UAV SwarmsAndrea Iannoli, Ivan Zyrianoff, Angelo Trotta, Lorenzo Gigli, Marco Di Felice2026-10-01下载Unmanned aerial vehicle (UAV) swarms increasingly rely on language-model agents to provide adaptive mission-level reasoning in uncertain environments.
DRL-driven RAN Slicing Management: A V2X-oriented Approach In Multi-service ScenariosDaniel E. Garcia-Fernandez, Pablo Vera-Soto, Sergio Fortes, M. Martinez, I. de-la-Bandera, M. L. Luque, A. Mendo, J. Ramiro, Raquel Barco2026-10-01下载The integration of Vehicle-to-Everything (V2X) communications is driving a profound transformation in vehicular connectivity, expected to significantly enhance traffic efficiency and safety.
GPU-Initiated Communication: Dissecting Down to the BoneJavid Baydamirli, Ismayil Ismayilov, Kaan Oktay, Didem Unat2026-10-01下载GPU-initiated communication lets GPU threads post RDMA operations directly to the NIC. It underpins NVSHMEM, NCCL GIN, and DeepEP, which serve the fine-grained, latency-critical communication of Mixtu...
Federated Learning for LLMs over Mobile Networks: Issues and Solutions in the RAN TransportEmilio Paolini, Andrea Pinto, Flavio Esposito, Luca Valcarenghi2026-10-01下载Federated LLM fine-tuning enables large models to be adapted using private and geographically distributed data at the network edge, creating recurring and deadline-sensitive communication workloads ac...
Physics-Guided Bayesian Optimization for High-Dimensional Mixed-Variable MIMO Base Station DesignKoki Kanzaki, Koya Sato2026-10-01下载This paper proposes a physics-guided Bayesian optimization for high-dimensional mixed-variable multiple-input and multiple-output (MIMO) base station (BS) design.

cs.PF - Performance ​

标题作者发布日期PDF摘要
Catscan: Visualizing Pipelines of CPU Performance SimulationAaron Lindsay, Nicholas Kelly, Scott Witscher, Mahesh Madhav2026-10-01下载Processor pipeline visualization tools are routine inside industry CPU teams, but few of them are described or released publicly. As a result, students, researchers, and other practitioners rarely see...
Exploiting the Interplay of Compute- and Memory-Bound kernels in MPI ApplicationsAyesha Afzal, Krishna Manda, Georg Hager2026-10-01下载Parallel applications are often designed for synchronous, lock-step execution, treating communication stalls as performance hazards. Yet, in a communication-light application without frequent synchron...
GPU-Initiated Communication: Dissecting Down to the BoneJavid Baydamirli, Ismayil Ismayilov, Kaan Oktay, Didem Unat2026-10-01下载GPU-initiated communication lets GPU threads post RDMA operations directly to the NIC. It underpins NVSHMEM, NCCL GIN, and DeepEP, which serve the fine-grained, latency-critical communication of Mixtu...
Is Your AI Fast Enough to Run a Fusion Reactor?Nathaniel Chen, Andrew Rothstein, Ricardo Shousha, Hiro Farre-Kaga, Peter Steiner, Azarakhsh Jalalvand, Egemen Kolemen2026-10-01下载Machine learning models are increasingly used in feedback control loops for nuclear fusion, where inference speed and predictable timing are critical.

基于 VitePress 构建 · 使用本地搜索查找论文