Skip to content

2026-05-31 ​

cs.AR - Architecture ​

标题作者发布日期PDF摘要
OpenEye: A Scalable Open-Source Hardware Accelerator for DNNsDenis Lebold, Hendrik Wöhrle2026-05-31下载The increasing computational complexity of deep neural network inference poses significant challenges for efficient hardware acceleration on embedded platforms, particularly with respect to resource c...
Formal Verification of Secure Encrypted VirtualizationHansika Weerasena, Amitabh Das, Prabhat Mishra2026-05-31下载Trusted execution environments (TEEs) provide a secure environment for data and code in use, ensuring that they are protected with respect to confidentiality and integrity.
Can AI Review Improve Paper Drafting? An Empirical Study on 20 Computer Architecture SubmissionsDi Wu2026-05-31下载Research is advancing faster than ever with artificial intelligence (AI); and so are the corresponding research papers. The exploding volume of AI-generated papers have put a strain to peer review, le...
Linear Complexity Fermionic Simulation on Quantum Devices with Hardware Connectivity ConstraintsXiangyu Gao, Winston Li, Jiakang Li, Zirui Li, Yipeng Huang, Costin Iancu, Eddy Z. Zhang2026-05-31下载Simulating fermionic systems on quantum hardware requires compiling fermionic Hamiltonians into executable quantum circuits. Existing approaches treat each compilation stage independently, applying he...

cs.DC - Distributed, Parallel, and Cluster Computing ​

标题作者发布日期PDF摘要
Move the Query, Not the Cache: Characterizing Cross-Instance Latent Attention Redistribution Across GPU FabricsBole Ma, Jan Eitzinger, Harald Köstler, Gerhard Wellein2026-05-31下载Frontier LLMs increasingly decide what a query attends to with a sparse-attention indexer that picks a few KV-cache blocks per query: attention's unit is now a small, reusable chunk.
Hierarchical Online Prompt Mutation with Dual-Loop Feedback for Guardrailed Evidence Document Generation: A Production-Evaluation Case StudyNataraj Agaram Sundar Tejas Morabia2026-05-31下载High-stakes production document-generation systems require language models to be adaptive, evidence-grounded, and auditable. We present HOPM, a hierarchical online prompt mutation framework evaluated ...
Understanding Cross-Cloud Interconnects: Hands-On Measurements and Cost OptimizationEitan Eliav, Isaac Keslassy, David Breitgand, Dean H. Lorenz, Avi Weit2026-05-31下载New services such as Google Cross-Cloud Interconnect (CCI) address the rise in fast and large-scale cross-cloud data transfers. CCI offers dedicated high-throughput links with low per-GB transfer cost...
Fail-Closed Lowering of Resident KV Claims onto LLM Serving RuntimesLukas Stepanek2026-05-31下载LLM serving runtimes increasingly expose KV-cache primitives that resemble future-reuse controls: retention priority, TTL-like duration, host or storage offload, block events, active no-evict scheduli...
GuidaPA: Privacy-Preserving Chatbot for Public Administration via Federated LearningDaniel M. Jimenez-Gutierrez, Albenzio Cirillo, Raffaele Nicolussi, Alessio Beltrame, Andrea Vitaletti2026-05-31下载We present GuidaPA, a privacy-preserving chatbot for the Italian Public Administration (PA) trained via Federated Learning (FL) on documentation from two national PA platforms, SIGESON and SIDFORS.
Residual-Weighted Randomized Jacobi: Sharpened Bounds via Residual Concentration and Asynchronous ExtensionEvan Coleman2026-05-31下载We study randomized stationary methods for symmetric positive definite linear systems in which component jj is selected with probability proportional to ∣rj∣ℓ|r_j|^\ell.
GPU Acceleration of Learning With Errors KEMs Using OpenACC for Post-Quantum CryptographyTiziana Liberati, Nitin Shukla, Matteo Barbieri, Gabriella Bettonte, Elisabetta Boella, Simone Rizzo, Daniele Gregori, Marco Pedicini2026-05-31下载Shor's algorithm proved that asymmetric cryptographic protocols based on the integer factorization and discrete logarithm problems are no longer safe in a world with large-scale quantum computers.
The World's Fastest Matching Engine AlgorithmJake Yoon2026-05-31下载Every electronic exchange relies on an order book whose storage layer determines matching latency. The dominant implementation -- linked lists chained through a balanced tree -- imposes two costs on e...
AcOrch: Accelerating Sampling-based GNN Training under CPU-NPU Heterogeneous EnvironmentsKefu Chen, Xin Ai, Qiange Wang, Yanfeng Zhang, Ge Yu2026-05-31下载Graph Neural Networks (GNNs) have achieved remarkable success in various applications. Sampling-based GNN training, which conducts mini-batch training on sampled subgraphs, has become a promising solu...
Schedule-Level Shared-Prefix Reuse for LLM RL TrainingPengbo Li, Feiyuan Zhang, Guangming Sheng, Guangxin He, Di Chai, Ziniu Li, Taiqiang Wu, Binhang Yuan, Kai Chen2026-05-31下载GRPO- and PPO-style LLM post-training commonly sample multiple trajectories from the same prompt and then train on the resulting group. In long-context RL workloads, this shared prompt-side prefix can...
AMP: A Vendor-Neutral Wire Format for Agent Memory OperationsThamilvendhan Munirathinam2026-05-31下载Agent-memory frameworks - mem0, Letta/MemGPT, Cognee, Zep/Graphiti, MemoryOS, MemTensor - each ship their own SDK, storage layout, and operational vocabulary.
Magnum.np.distributed: Accelerating Finite Difference Micromagnetic Simulations with Multiple GPUsTsz Chung Cheng, Yuichiro Kurokawa, Hiromi Yuasa2026-05-31下载Micromagnetic simulations are essential tools in nanomagnetism and spintronics research. Although widely adopted solvers like Mumax3 and the Python-native magnum.
Leyline: KV Cache Directives for Agentic InferenceBole Ma, Jan Eitzinger, Harald Koestler2026-05-31下载Modern KV cache management assumes the chatbot workload: prompts arrive once and the cache grows append-only, so prefix caching and forward-only eviction are correct by construction.
Lodestar: An Online-Learning LLM Inference RouterGangmuk Lim, Wanyu Zhao, Brighten Godfrey, Jiaxin Shan, Le Xu, Liguang Xie2026-05-31下载Efficiently serving large language model (LLM) inference tasks is crucial both for user-perceived latency such as time-to-first-token (TTFT) and for GPU utilization.
Characterizing Metastable Faults and FailuresAli Farahbakhsh, Qingjie Lu, Lorenzo Alvisi, Andreas Haeberlen, Robbert Van Renesse2026-05-31下载Metastable failures are hard to detect, prevent, and mitigate. During a metastable failure, a system exhibits self-sustaining bad behavior even in the absence of adversarial conditions.

cs.NI - Networking and Internet Architecture ​

标题作者发布日期PDF摘要
Move the Query, Not the Cache: Characterizing Cross-Instance Latent Attention Redistribution Across GPU FabricsBole Ma, Jan Eitzinger, Harald Köstler, Gerhard Wellein2026-05-31下载Frontier LLMs increasingly decide what a query attends to with a sparse-attention indexer that picks a few KV-cache blocks per query: attention's unit is now a small, reusable chunk.
A Reproducible UAV-Assisted VANET Dataset Generator for Fragmentation Risk Analysis in Intelligent Transportation SystemsBappa Muktar, Justin Moskolaï Ngossaha, Adama Nouboukpo2026-05-31下载Vehicular Ad Hoc Networks (VANETs) are a key component of Intelligent Transportation Systems, enabling cooperative communication among vehicles and between vehicles and roadside infrastructure.
FlexLink: Decoupling Control and Data Beams for Next-Generation Wideband NetworksIsh Kumar Jain, Rohith Reddy Vennam, Dinesh Bharadia2026-05-31下载The next generation of 6G networks aims to utilize ultra-wideband spectrum and massive antenna arrays to serve multiple users with both control and data channels at low latency and high efficiency.
Understanding Cross-Cloud Interconnects: Hands-On Measurements and Cost OptimizationEitan Eliav, Isaac Keslassy, David Breitgand, Dean H. Lorenz, Avi Weit2026-05-31下载New services such as Google Cross-Cloud Interconnect (CCI) address the rise in fast and large-scale cross-cloud data transfers. CCI offers dedicated high-throughput links with low per-GB transfer cost...
SEArch: Optimistic Policy Selection Between Scene Noise and Drift for UAV Radar SearchNoor Khial, Naram Mhaisen, Loay Ismail, Amr Mohamed2026-05-31下载Unmanned Aerial Vehicles (UAVs) equipped with radar sensors are deployed for target search missions in diverse environments, where targets exhibit characteristic signatures (e.g.
A Communication-Centric 6G-LLM Architecture for Scalable Tactical Autonomous Defense Vehicle NetworksKiran Khurshid, Shumaila Javaid, Nasir Saeed2026-05-31下载The integration of Artificial Intelligence (AI) and emerging 6G networks introduces new opportunities for scalable coordination in tactical autonomous vehicle systems.
AI-IoT-Robotics Integration: Survey of Frameworks, Emerging Trends, and the Path Toward Connected RoboticsRanulfo Bezerra, Satoshi Tadokoro, Kazunori Ohno2026-05-31下载The convergence of Artificial Intelligence, the Internet of Things, and Robotics is no longer a futuristic vision; it is rapidly becoming the foundation of real-time, intelligent, and context-aware sy...

cs.OS - Operating Systems ​

标题作者发布日期PDF摘要
Characterizing Metastable Faults and FailuresAli Farahbakhsh, Qingjie Lu, Lorenzo Alvisi, Andreas Haeberlen, Robbert Van Renesse2026-05-31下载Metastable failures are hard to detect, prevent, and mitigate. During a metastable failure, a system exhibits self-sustaining bad behavior even in the absence of adversarial conditions.

cs.PF - Performance ​

标题作者发布日期PDF摘要
The World's Fastest Matching Engine AlgorithmJake Yoon2026-05-31下载Every electronic exchange relies on an order book whose storage layer determines matching latency. The dominant implementation -- linked lists chained through a balanced tree -- imposes two costs on e...

基于 VitePress 构建 · 使用本地搜索查找论文