Skip to content

2026-06-23 ​

cs.AR - Architecture ​

标题作者发布日期PDF摘要
PDS Joint: A Parametric Double-Spiral Joint Tailored for Dexterous HandsHaoyang Li, Yibo Wen, Yixiang Fan, Yiheng Xu, Yufeng Yue2026-06-23下载Compliant joints can embed safety and adaptability into dexterous hands, but achieving large-stroke anthropomorphic motion while maintaining joint-specific, directiondependent stiffness and reliable p...

cs.DC - Distributed, Parallel, and Cluster Computing ​

标题作者发布日期PDF摘要
Ambulance: saving BFT through racingNeil Giridharan, Shubham Mishra, Lorenzo Alvisi, Natacha Crooks, Benjamin Marsh, Hein Meling, Kartik Nayak, Grzegorz Prusak2026-06-23下载Today's practical Byzantine Fault Tolerant (BFT) state machine replication deployments are vulnerable to slowdowns. The main culprit is timeouts.
Power-Flexible AI Data Centers: A New Paradigm for Grid-Responsive ComputeChris Williams, Philip Colangelo, Ayse Coskun, Ethan Levine, Andy Neale, Ciaran Roberts, Shayan Sengupta, Nikhil Shirolkar, Varun Sivaram, Sarah Soares, Ethan Tiao, Scott Underwood, Daniel Wilson, Frank Sharp, Luke Wainwright, Harry Petty, Scott Wallace, Brandon Records2026-06-23下载The rapid expansion of artificial intelligence (AI) infrastructure is driving unprecedented growth in electricity demand from data centers. Traditional power-system planning treats large computing fac...
Speculation at a Distance: Where Edge-Cloud Speculative Decoding Actually Pays OffYuan Lyu, Bharath Irukulapati, Jaya Prakash Champati2026-06-23下载Speculative decoding (SD) accelerates LLM inference by 1.51.5-33 times when the draft and target models are co-located. This has motivated a distributed variant (DSD) that places the draft model on an...
Energy Efficient Scheduling of AI/ML Workloads on Multi Instance GPUs with Dynamic RepartitioningEllie Lipe, Neel Karia, Connor Espenshade, Clifford Stein, Asser Tantawi, Olivier Tardieu2026-06-23下载Increasing demand from AI/ML workloads is exacerbating the rising energy consumption of data centers. Recent advances in hardware such as NVIDIA's Multi Instance GPUs (MIGs) offer improvements in flex...
A Natively Blocked, Device-Resident Algebraic Multigrid GPU Path in PETScMark F. Adams2026-06-23下载Smoothed-aggregation algebraic multigrid (AMG) is widely used for the linear systems arising from finite-element discretizations of vector PDEs such as elasticity, but its GPU implementations have use...
Solvability of Approximate Agreement on Graphs and Simplicial ComplexesJoel Rybicki, Yaroslav Verbitsky2026-06-23下载Approximate agreement tasks on graphs are discrete relaxations of consensus, where each process in a distributed system is given as input a vertex on a graph GG, and processes have to output vertices...
Unified Position-Invariant Random Access Through Two Compression Layers via Absolute-Offset Coordinates: A Bit-Perfect Device-Resident ProofYakiv Shavidze2026-06-23下载Random access into compressed data is normally confined to a single layer. Entropy-layer methods (Recoil) seek within rANS by storing intermediate decoder states; dictionary/match-layer methods seek w...
CrossPool: Efficient Multi-LLM Serving for Cold MoE Models through KV-Cache and Weight DisaggregationZhuoren Ye, Tianyu Wo, Dinghao Xue, Mingming Zhang, Yuchen Teng, Chunming Hu, Renyu Yang2026-06-23下载Emerging LLM services increasingly host many sparse MoE models, yet most models receive sparse requests and remain cold. This creates a GPU memory problem: model weights are stable and model-determine...
cuSBF: A Minimizer-Aware Bloom Filter for Genomic Sequence Data on Modern GPUsTim Dortmann, Markus Vieth, Bertil Schmidt2026-06-23下载Efficient genomic k-mer indexing depends on approximate membership query (AMQ) structures that must deliver high throughput, low false-positive rates (FPR), and modest memory footprints.
BiJuTy: An Interactive HPC-Aware Big Data Cluster Lifecycle Manager and Performance Assessment Utility for JupyterHubApurv Deepak Kulkarni, Jan Frenzel, Siavash Ghiasvand2026-06-23下载The increasing demand for data processing has created a pressing need for access to high-performance computing (HPC) systems. Nevertheless, leveraging these systems to execute complex big data process...
Accelerating Disaggregated RL for Visual Generative LLMs with Diffusion-Based Parallelism and Trainer-Assisted GenerationSijie Wang, Zhengyu Qing, Zhiqiang Tan, Yiming Yin, Yeqing Zhang, Yaoyuan Wang, Qiang Wang, Xiaowen Chu, Shaohuai Shi2026-06-23下载Reinforcement learning (RL) has become a dominant post-training paradigm, driving the emergence of high-performance RL systems such as veRL for autoregressive large language models (LLMs).
OpenMP GPU Acceleration and Portability of TRIMEG-C1 for Electromagnetic Gyrokinetic Simulations in Tokamak PlasmasGiorgio Daneri, Zhixin Lu, Matthias Hoelzl, Luca Venerando Greco, Edoardo Carrà2026-06-23下载The Triangular mesh-based gyrokinetic code TRIMEG-C1 solves the gyrokinetic equations using the particle-in-cell scheme to simulate electromagnetic instabilities in tokamak plasmas.
Semantic Lock: Synchronization Based on the Analysis of the Operation Conflict GraphDenis Korotchenko, Vitaly Aksenov2026-06-23下载This paper presents a new lock, SemanticLock, based on the conflict graph between operations. We can consider it a generalization of a read-write lock where conflicts exist between write operations an...
Semi-asynchronous Federated Learning in Flower: Framework Extension and Performance AssessmentVíctor Hidalgo-Izquierdo, Carmen Carrión, Blanca Caminero2026-06-23下载This paper presents an extension of the Flower federated learning framework to support Semi-Asynchronous Federated Learning. The proposed approach adapts the traditional synchronous paradigm to better...
SkyChain Intelligence: A Blockchain-Secured Multi-Agent DRL Framework for Low-Altitude Embodied Artificial IntelligenceHaoxiang Luo, Tianqi Jiang, Ruichen Zhang, Yinqiu Liu, Gang Sun, Hongfang Yu, Abbas Jamalipour, Dong In Kim2026-06-23下载With the rapid development of the Low-Altitude Economy (LAE) ecosystem, Low-Altitude Embodied Artificial Intelligence (LAEAI) agents have become the core carriers of autonomous aerial services, thereb...
Aquifer: Hierarchical Memory Pooling with CXL and RDMA for MicroVM SnapshotsJunliang Hu, Huaicheng Li, Ming-Chang Yang2026-06-23下载Memory stranding wastes 25-35% of installed DRAM in production cloud clusters. Memory pooling over CXL and RDMA offers a remedy, but neither technology alone suffices: CXL provides low-latency, load/s...
Hash Table Design for RDMA:Challenges and OpportunitiesShuchen She, Haipeng Dai2026-06-23下载Hash tables complete the insertion, lookup, and deletion of a single key in constant time on average, and they are widely used in databases, key-value stores, and network systems.

cs.NI - Networking and Internet Architecture ​

标题作者发布日期PDF摘要
Forget to Improve: On-Device LLM-Agent Continual Learning via Budget-Curated MemoryBeining Wu, Zihao Ding, Jun Huang, Yanxiao Zhao2026-06-23下载On-device language-model agents improve by accumulating experience in retrieved memory rather than by updating weights. This memory is hard-bounded and exposed: it consumes RAM and energy, reaches pee...
Building a Low-cost Network Digital Twin for the IoT-Edge-Cloud Continuum Using Open-Source ToolingJosevany do Amaral, Rute C. Sofia2026-06-23下载Validating network configurations and testing failure scenarios in IoT-edge-cloud environments without disrupting live infrastructure remains an open operational challenge.
BRAVR: An AP-Assisted Online DRL Mechanism for Interactive VR Bitrate Adaptation over Wi-FiMiguel Casasnovas, Francesc Wilhelmi, Boris Bellalta2026-06-23下载Interactive virtual reality (VR) streaming over Wi-Fi requires stringent latency and reliability guarantees, which become increasingly difficult to achieve under dynamic channel conditions and shared ...
Accelerating Disaggregated RL for Visual Generative LLMs with Diffusion-Based Parallelism and Trainer-Assisted GenerationSijie Wang, Zhengyu Qing, Zhiqiang Tan, Yiming Yin, Yeqing Zhang, Yaoyuan Wang, Qiang Wang, Xiaowen Chu, Shaohuai Shi2026-06-23下载Reinforcement learning (RL) has become a dominant post-training paradigm, driving the emergence of high-performance RL systems such as veRL for autoregressive large language models (LLMs).
Importance of Intent-Sharing for V2X-based Maneuver CoordinationRafael Molina-Masegosa, Sergei S. Avedisov, Miguel Sepulcre, Javier Gozalvez, Yashar Z. Farid, Onur Altintas2026-06-23下载This paper examines the critical role of intent-sharing in enabling effective maneuver coordination for connected and automated vehicles (CAVs).
FORESEE: A Cooperative Lane Change Model for Connected and Automated DrivingRafael Molina-Masegosa, Sergei S. Avedisov, Miguel Sepulcre, Javier Gozalvez, Yashar Z. Farid, Onur Altintas2026-06-23下载This paper presents FORESEE, a novel cooperative lane change model for connected and automated driving. FORESEE leverages Vehicle-to-Everything (V2X) data to anticipate traffic conditions and effectiv...
SkyChain Intelligence: A Blockchain-Secured Multi-Agent DRL Framework for Low-Altitude Embodied Artificial IntelligenceHaoxiang Luo, Tianqi Jiang, Ruichen Zhang, Yinqiu Liu, Gang Sun, Hongfang Yu, Abbas Jamalipour, Dong In Kim2026-06-23下载With the rapid development of the Low-Altitude Economy (LAE) ecosystem, Low-Altitude Embodied Artificial Intelligence (LAEAI) agents have become the core carriers of autonomous aerial services, thereb...
Overconfident Coordinates: Quantifying Confidence in Traceroute GeolocationSantiago Klein, Caleb J. Wang, Fabián E. Bustamante2026-06-23下载Studies of Internet paths often attach router locations to traceroute hops using commercial geolocation databases, rDNS labels, Geofeeds, and IXP metadata.

cs.OS - Operating Systems ​

标题作者发布日期PDF摘要
ActPlane: Programmable OS-Level Policy Enforcement for Agent HarnessesYusheng Zheng, Tianyuan Wu, Quanzhi Fu, Tong Yu, Wenan Mao, Wei Wang, Dan Williams, Andi Quinn2026-06-23下载AI agents increasingly run in production through harnesses, the software around the LLM, including an engine that enforces safety and effectiveness policies, e.g., 'run tests before committing.
Kops: Safely Extending the eBPF Compilation Pipeline with Native OperationsYusheng Zheng, Zhengjie Ji, Weichen Tao, Hao Sun, Wei Zhang, Dan Williams, Andi Quinn2026-06-23下载eBPF safely extends OS kernels in domains such as networking, observability, and security. The safety comes from an in-kernel compilation pipeline where a verifier checks every program, and a kernel j...
Aquifer: Hierarchical Memory Pooling with CXL and RDMA for MicroVM SnapshotsJunliang Hu, Huaicheng Li, Ming-Chang Yang2026-06-23下载Memory stranding wastes 25-35% of installed DRAM in production cloud clusters. Memory pooling over CXL and RDMA offers a remedy, but neither technology alone suffices: CXL provides low-latency, load/s...

cs.PF - Performance ​

标题作者发布日期PDF摘要
Power-Flexible AI Data Centers: A New Paradigm for Grid-Responsive ComputeChris Williams, Philip Colangelo, Ayse Coskun, Ethan Levine, Andy Neale, Ciaran Roberts, Shayan Sengupta, Nikhil Shirolkar, Varun Sivaram, Sarah Soares, Ethan Tiao, Scott Underwood, Daniel Wilson, Frank Sharp, Luke Wainwright, Harry Petty, Scott Wallace, Brandon Records2026-06-23下载The rapid expansion of artificial intelligence (AI) infrastructure is driving unprecedented growth in electricity demand from data centers. Traditional power-system planning treats large computing fac...
Scheduling jobs with unknown size distribution in a M/G/1 queue: the shifted empirical GittinsNicolas Gast, Bruno Gaujal, Adrien Obrecht2026-06-23下载In this paper we consider a M/G/1 queue for which we want to minimize the expected response time. We show how to compute indices from nn samples of the job size distribution such that the correspondi...
AI-PAVE-Br: Leveraging Large Language Models for Enhanced Product Attribute Value Extraction through a Golden Set ApproachMurilo Gazzola, Hugo Gobato Souto, Samuel Silva, Júlia Schubert Peixoto, Felipe Siqueira, André Luis Pedroso de Morais, Caio Gomes2026-06-23下载The explosive growth and complexity of product data within the dynamic Brazilian e-commerce landscape demand robust and specialized methods for structured information extraction.
CrossPool: Efficient Multi-LLM Serving for Cold MoE Models through KV-Cache and Weight DisaggregationZhuoren Ye, Tianyu Wo, Dinghao Xue, Mingming Zhang, Yuchen Teng, Chunming Hu, Renyu Yang2026-06-23下载Emerging LLM services increasingly host many sparse MoE models, yet most models receive sparse requests and remain cold. This creates a GPU memory problem: model weights are stable and model-determine...
Accelerating Disaggregated RL for Visual Generative LLMs with Diffusion-Based Parallelism and Trainer-Assisted GenerationSijie Wang, Zhengyu Qing, Zhiqiang Tan, Yiming Yin, Yeqing Zhang, Yaoyuan Wang, Qiang Wang, Xiaowen Chu, Shaohuai Shi2026-06-23下载Reinforcement learning (RL) has become a dominant post-training paradigm, driving the emergence of high-performance RL systems such as veRL for autoregressive large language models (LLMs).

基于 VitePress 构建 · 使用本地搜索查找论文