2026-05-31
cs.AR - Architecture
| 标题 | 作者 | 发布日期 | 摘要 | |
|---|---|---|---|---|
| OpenEye: A Scalable Open-Source Hardware Accelerator for DNNs | Denis Lebold, Hendrik Wöhrle | 2026-05-31 | 下载 | The increasing computational complexity of deep neural network inference poses significant challenges for efficient hardware acceleration on embedded platforms, particularly with respect to resource c... |
| Formal Verification of Secure Encrypted Virtualization | Hansika Weerasena, Amitabh Das, Prabhat Mishra | 2026-05-31 | 下载 | Trusted execution environments (TEEs) provide a secure environment for data and code in use, ensuring that they are protected with respect to confidentiality and integrity. |
| Can AI Review Improve Paper Drafting? An Empirical Study on 20 Computer Architecture Submissions | Di Wu | 2026-05-31 | 下载 | Research is advancing faster than ever with artificial intelligence (AI); and so are the corresponding research papers. The exploding volume of AI-generated papers have put a strain to peer review, le... |
| Linear Complexity Fermionic Simulation on Quantum Devices with Hardware Connectivity Constraints | Xiangyu Gao, Winston Li, Jiakang Li, Zirui Li, Yipeng Huang, Costin Iancu, Eddy Z. Zhang | 2026-05-31 | 下载 | Simulating fermionic systems on quantum hardware requires compiling fermionic Hamiltonians into executable quantum circuits. Existing approaches treat each compilation stage independently, applying he... |
cs.DC - Distributed, Parallel, and Cluster Computing
| 标题 | 作者 | 发布日期 | 摘要 | |
|---|---|---|---|---|
| Move the Query, Not the Cache: Characterizing Cross-Instance Latent Attention Redistribution Across GPU Fabrics | Bole Ma, Jan Eitzinger, Harald Köstler, Gerhard Wellein | 2026-05-31 | 下载 | Frontier LLMs increasingly decide what a query attends to with a sparse-attention indexer that picks a few KV-cache blocks per query: attention's unit is now a small, reusable chunk. |
| Hierarchical Online Prompt Mutation with Dual-Loop Feedback for Guardrailed Evidence Document Generation: A Production-Evaluation Case Study | Nataraj Agaram Sundar Tejas Morabia | 2026-05-31 | 下载 | High-stakes production document-generation systems require language models to be adaptive, evidence-grounded, and auditable. We present HOPM, a hierarchical online prompt mutation framework evaluated ... |
| Understanding Cross-Cloud Interconnects: Hands-On Measurements and Cost Optimization | Eitan Eliav, Isaac Keslassy, David Breitgand, Dean H. Lorenz, Avi Weit | 2026-05-31 | 下载 | New services such as Google Cross-Cloud Interconnect (CCI) address the rise in fast and large-scale cross-cloud data transfers. CCI offers dedicated high-throughput links with low per-GB transfer cost... |
| Fail-Closed Lowering of Resident KV Claims onto LLM Serving Runtimes | Lukas Stepanek | 2026-05-31 | 下载 | LLM serving runtimes increasingly expose KV-cache primitives that resemble future-reuse controls: retention priority, TTL-like duration, host or storage offload, block events, active no-evict scheduli... |
| GuidaPA: Privacy-Preserving Chatbot for Public Administration via Federated Learning | Daniel M. Jimenez-Gutierrez, Albenzio Cirillo, Raffaele Nicolussi, Alessio Beltrame, Andrea Vitaletti | 2026-05-31 | 下载 | We present GuidaPA, a privacy-preserving chatbot for the Italian Public Administration (PA) trained via Federated Learning (FL) on documentation from two national PA platforms, SIGESON and SIDFORS. |
| Residual-Weighted Randomized Jacobi: Sharpened Bounds via Residual Concentration and Asynchronous Extension | Evan Coleman | 2026-05-31 | 下载 | We study randomized stationary methods for symmetric positive definite linear systems in which component is selected with probability proportional to . |
| GPU Acceleration of Learning With Errors KEMs Using OpenACC for Post-Quantum Cryptography | Tiziana Liberati, Nitin Shukla, Matteo Barbieri, Gabriella Bettonte, Elisabetta Boella, Simone Rizzo, Daniele Gregori, Marco Pedicini | 2026-05-31 | 下载 | Shor's algorithm proved that asymmetric cryptographic protocols based on the integer factorization and discrete logarithm problems are no longer safe in a world with large-scale quantum computers. |
| The World's Fastest Matching Engine Algorithm | Jake Yoon | 2026-05-31 | 下载 | Every electronic exchange relies on an order book whose storage layer determines matching latency. The dominant implementation -- linked lists chained through a balanced tree -- imposes two costs on e... |
| AcOrch: Accelerating Sampling-based GNN Training under CPU-NPU Heterogeneous Environments | Kefu Chen, Xin Ai, Qiange Wang, Yanfeng Zhang, Ge Yu | 2026-05-31 | 下载 | Graph Neural Networks (GNNs) have achieved remarkable success in various applications. Sampling-based GNN training, which conducts mini-batch training on sampled subgraphs, has become a promising solu... |
| Schedule-Level Shared-Prefix Reuse for LLM RL Training | Pengbo Li, Feiyuan Zhang, Guangming Sheng, Guangxin He, Di Chai, Ziniu Li, Taiqiang Wu, Binhang Yuan, Kai Chen | 2026-05-31 | 下载 | GRPO- and PPO-style LLM post-training commonly sample multiple trajectories from the same prompt and then train on the resulting group. In long-context RL workloads, this shared prompt-side prefix can... |
| AMP: A Vendor-Neutral Wire Format for Agent Memory Operations | Thamilvendhan Munirathinam | 2026-05-31 | 下载 | Agent-memory frameworks - mem0, Letta/MemGPT, Cognee, Zep/Graphiti, MemoryOS, MemTensor - each ship their own SDK, storage layout, and operational vocabulary. |
| Magnum.np.distributed: Accelerating Finite Difference Micromagnetic Simulations with Multiple GPUs | Tsz Chung Cheng, Yuichiro Kurokawa, Hiromi Yuasa | 2026-05-31 | 下载 | Micromagnetic simulations are essential tools in nanomagnetism and spintronics research. Although widely adopted solvers like Mumax3 and the Python-native magnum. |
| Leyline: KV Cache Directives for Agentic Inference | Bole Ma, Jan Eitzinger, Harald Koestler | 2026-05-31 | 下载 | Modern KV cache management assumes the chatbot workload: prompts arrive once and the cache grows append-only, so prefix caching and forward-only eviction are correct by construction. |
| Lodestar: An Online-Learning LLM Inference Router | Gangmuk Lim, Wanyu Zhao, Brighten Godfrey, Jiaxin Shan, Le Xu, Liguang Xie | 2026-05-31 | 下载 | Efficiently serving large language model (LLM) inference tasks is crucial both for user-perceived latency such as time-to-first-token (TTFT) and for GPU utilization. |
| Characterizing Metastable Faults and Failures | Ali Farahbakhsh, Qingjie Lu, Lorenzo Alvisi, Andreas Haeberlen, Robbert Van Renesse | 2026-05-31 | 下载 | Metastable failures are hard to detect, prevent, and mitigate. During a metastable failure, a system exhibits self-sustaining bad behavior even in the absence of adversarial conditions. |
cs.NI - Networking and Internet Architecture
| 标题 | 作者 | 发布日期 | 摘要 | |
|---|---|---|---|---|
| Move the Query, Not the Cache: Characterizing Cross-Instance Latent Attention Redistribution Across GPU Fabrics | Bole Ma, Jan Eitzinger, Harald Köstler, Gerhard Wellein | 2026-05-31 | 下载 | Frontier LLMs increasingly decide what a query attends to with a sparse-attention indexer that picks a few KV-cache blocks per query: attention's unit is now a small, reusable chunk. |
| A Reproducible UAV-Assisted VANET Dataset Generator for Fragmentation Risk Analysis in Intelligent Transportation Systems | Bappa Muktar, Justin Moskolaï Ngossaha, Adama Nouboukpo | 2026-05-31 | 下载 | Vehicular Ad Hoc Networks (VANETs) are a key component of Intelligent Transportation Systems, enabling cooperative communication among vehicles and between vehicles and roadside infrastructure. |
| FlexLink: Decoupling Control and Data Beams for Next-Generation Wideband Networks | Ish Kumar Jain, Rohith Reddy Vennam, Dinesh Bharadia | 2026-05-31 | 下载 | The next generation of 6G networks aims to utilize ultra-wideband spectrum and massive antenna arrays to serve multiple users with both control and data channels at low latency and high efficiency. |
| Understanding Cross-Cloud Interconnects: Hands-On Measurements and Cost Optimization | Eitan Eliav, Isaac Keslassy, David Breitgand, Dean H. Lorenz, Avi Weit | 2026-05-31 | 下载 | New services such as Google Cross-Cloud Interconnect (CCI) address the rise in fast and large-scale cross-cloud data transfers. CCI offers dedicated high-throughput links with low per-GB transfer cost... |
| SEArch: Optimistic Policy Selection Between Scene Noise and Drift for UAV Radar Search | Noor Khial, Naram Mhaisen, Loay Ismail, Amr Mohamed | 2026-05-31 | 下载 | Unmanned Aerial Vehicles (UAVs) equipped with radar sensors are deployed for target search missions in diverse environments, where targets exhibit characteristic signatures (e.g. |
| A Communication-Centric 6G-LLM Architecture for Scalable Tactical Autonomous Defense Vehicle Networks | Kiran Khurshid, Shumaila Javaid, Nasir Saeed | 2026-05-31 | 下载 | The integration of Artificial Intelligence (AI) and emerging 6G networks introduces new opportunities for scalable coordination in tactical autonomous vehicle systems. |
| AI-IoT-Robotics Integration: Survey of Frameworks, Emerging Trends, and the Path Toward Connected Robotics | Ranulfo Bezerra, Satoshi Tadokoro, Kazunori Ohno | 2026-05-31 | 下载 | The convergence of Artificial Intelligence, the Internet of Things, and Robotics is no longer a futuristic vision; it is rapidly becoming the foundation of real-time, intelligent, and context-aware sy... |
cs.OS - Operating Systems
| 标题 | 作者 | 发布日期 | 摘要 | |
|---|---|---|---|---|
| Characterizing Metastable Faults and Failures | Ali Farahbakhsh, Qingjie Lu, Lorenzo Alvisi, Andreas Haeberlen, Robbert Van Renesse | 2026-05-31 | 下载 | Metastable failures are hard to detect, prevent, and mitigate. During a metastable failure, a system exhibits self-sustaining bad behavior even in the absence of adversarial conditions. |
cs.PF - Performance
| 标题 | 作者 | 发布日期 | 摘要 | |
|---|---|---|---|---|
| The World's Fastest Matching Engine Algorithm | Jake Yoon | 2026-05-31 | 下载 | Every electronic exchange relies on an order book whose storage layer determines matching latency. The dominant implementation -- linked lists chained through a balanced tree -- imposes two costs on e... |