2026-06-17
cs.AR - Architecture
| 标题 | 作者 | 发布日期 | 摘要 | |
|---|---|---|---|---|
| A Tool for the Synthesis of Adaptive Probabilistic Processors Based on the Ising Model | Jonathan Juracy Carneiro da Silva, Leonardo R. Gobatto, Jose Rodrigo Azambuja | 2026-06-17 | 下载 | This work presents a tool for the synthesis and simulation of probabilistic architectures for solving combinatorial optimization problems by mapping them to the Ising model. |
| SPINE: A Fault Injection Profiler for Quantized Neural Networks under Accumulated Faults | Nathan Guimarães, Ian Kersz, Leonardo R. Gobatto, Fabio Benevenuti, Michael G. Jordan, Antonio Carlos S. Beck, Fernanda L. Kastensmidt, Jose Rodrigo Azambuja | 2026-06-17 | 下载 | Deploying deep neural networks at the edge demands efficient inference under strict cost and power constraints. Quantized neural networks address these demands by replacing floating-point parameters w... |
| PuDGhost: Experimental Analysis of Computation Result Corruption in Processing-using-DRAM Operations on Real DRAM Chips and Implications for Future Systems | Daichi Tokuda, İsmail Emir Yüksel, Tatsuya Kubo, Ataberk Olgun, Haocong Luo, Nisa Bostanci, Jikun Wang, A. Giray Yağlıkçı, Shinya Takamaeda-Yamazaki, Onur Mutlu | 2026-06-17 | 下载 | Processing-using-DRAM (PuD) is a promising computation paradigm that alleviates frequent data movement between main memory and processing units by using each DRAM column as a computation engine via si... |
| CHERI-D: Secure and efficient inline object ID for CHERI temporal memory safety | Yuecheng Wang, Jonathan Woodruff, Alfredo Mazzinghi, Peter Rugg, Samuel W. Stark, Alexandre Joannou, Robert N. M. Watson, Simon W. Moore | 2026-06-17 | 下载 | We propose CHERI-D, an architectural extension to CHERI that supports efficient temporal memory safety. Efficient memory safety is an increasing priority for programming languages, operating systems, ... |
| Nanoscale memristive devices: Threats and solutions | Amir M. Hajisadeghi, Javad Talafy, Hamid R. Zarandi | 2026-06-17 | 下载 | Due to their incentivizing features, memristors are a promising candidate for replacing CMOS-based memories, which are faced with various functional challenges in deep submicron process technologies. |
cs.DC - Distributed, Parallel, and Cluster Computing
| 标题 | 作者 | 发布日期 | 摘要 | |
|---|---|---|---|---|
| REMOP: REmote-Memory-aware OPerator Optimization | Shiquan Zhang, Yunhao Mao, Yuqiu Zhang, Gengrui Zhang, Jeyhun Karimov, Hans-Arno Jacobsen | 2026-06-17 | 下载 | Remote and disaggregated memory tiers expand the effective memory capacity of analytical database engines, but they also reshape the cost structure of out-of-memory query processing. |
| Mesh Inference: A Formal Model of Collective Intelligence Without a Center | Hongwei Xu | 2026-06-17 | 下载 | We present a formal model of mesh inference: how a population of independent agents, each holding private state and exchanging only admitted, typed observations, derives a conclusion none of them hold... |
| The Sheaf Laplacian: A Topological Framework for Data Fusion and Consensus in Distributed Sensing Networks | Manuel Hernández, Eduardo Sánchez-Soto | 2026-06-17 | 下载 | We argue here that traditional network models, which are overwhelmingly based on the mathematical construct of a simple graph, are fundamentally insufficient for capturing the complexity of modern dis... |
| A Topos-Theoretic Interpretation of Blockchain Systems: Sheaves of Consensus and the Logic of Decentralized Truth | Manuel Hernández, Eduardo Sánchez-Soto | 2026-06-17 | 下载 | The predominant formal models for blockchain systems, particularly smart contracts, have largely been drawn from the classical theory of computation, with the finite state machine (FSM) or labeled tra... |
| TurboServe: Serving Streaming Video Generation Efficiently and Economically | Youhe Jiang, Haoxu Wang, Haotong Bao, Kai Jiang, Jianfei Chen, Jun Zhu, Fangcheng Fu, Jintao Zhang | 2026-06-17 | 下载 | Streaming video generation is emerging as a new serving workload in which users interact with long-lived sessions that generate video progressively, chunk by chunk. |
| Pulse: Training Acceleration for Large Diffusion Models with Automatic Pipeline Parallelism | Boran Sun, Guoyong Jiang, Lin Zhang, Chen Chen, Yuechen Tao, Zhishu Che, Jieling Yu, Shan Chang, Huaxi Gu, Fangming Liu, Bo Li | 2026-06-17 | 下载 | Diffusion models are now a dominant approach for high-fidelity image and video generation, yet scaling their training across GPU clusters remains challenging. |
| PuDGhost: Experimental Analysis of Computation Result Corruption in Processing-using-DRAM Operations on Real DRAM Chips and Implications for Future Systems | Daichi Tokuda, İsmail Emir Yüksel, Tatsuya Kubo, Ataberk Olgun, Haocong Luo, Nisa Bostanci, Jikun Wang, A. Giray Yağlıkçı, Shinya Takamaeda-Yamazaki, Onur Mutlu | 2026-06-17 | 下载 | Processing-using-DRAM (PuD) is a promising computation paradigm that alleviates frequent data movement between main memory and processing units by using each DRAM column as a computation engine via si... |
| A performance portable fast Ewald summation for Stokes flow | Gabriel Kosmacher, Ziyu Du, Joar Bagge, George Biros | 2026-06-17 | 下载 | We present GPU algorithms for Ewald summation methods for accelerating N-body Stokes flow problems in periodic domains. Like most N-body codes, Ewald sums use a near-field/far-field decomposition. |
| FoMoE: Breaking the Full-Replica Barrier with a Federation of MoEs | Lorenzo Sani, Zeyu Cao, Meghdad Kurmanji, Alex Iacob, Andrej Jovanovic, Yan Gao, Wanru Zhao, Nicholas D. Lane | 2026-06-17 | 下载 | Pre-training Large Language Models (LLMs) typically demands large-scale infrastructure with tightly coupled hardware accelerators. While increasing model and dataset scale remains the dominant driver ... |
| Spotlight: Synergizing Seed Exploration and Spot GPUs for DiT RL Post-Training | Ruiqi Lai, Dakai An, Wei Gao, Ju Huang, Siran Yang, Jiamang Wang, Lin Qu, Dmitrii Ustiugov, Wei Wang | 2026-06-17 | 下载 | Reinforcement learning (RL) post-training of Diffusion Transformers (DiTs) is prohibitively expensive, requiring thousands of high-end GPUs. Existing works explore two directions to reduce cost: seed ... |
| On the Notions of Bounded Bypass, and How to Make any Deadlock-Free MUTEX Protocol Satisfy One of Them | Rob van Glabbeek, Daniele Gorla, Myrthe Spronck | 2026-06-17 | 下载 | In the literature on mutual exclusion, bounded bypass has been used for a long time as a strengthening of starvation-freedom, but, to the best of our knowledge, it still lacks a satisfying definition ... |
| A Composable CRDT Layer for Byzantine-Resilient Deterministic Reconstruction | Amos Brocco | 2026-06-17 | 下载 | Conflict-free Replicated Data Types (CRDTs) ensure Strong Eventual Consistency without coordination, but typically assume benign participants and rely on validation or exclusion to handle Byzantine be... |
| LiveStack: OS Support for Cluster-Scale Full-Stack Live Simulation | Yiliang Wan, Haifeng Sun, Yihan Yang, Jonas Kaufmann, Antoine Kaufmann, Jialin Li | 2026-06-17 | 下载 | Cluster-scale full-stack simulation is essential for evaluating distributed software stacks and emerging hardware components before deployment. |
| Urban Limits as Design Constraints: Identifying Suitable Locations for Distributed, Photovoltaic-Powered Servers | Justin Chikhaoui, Thomas Leduc, Daniel Siret, Abdoulaye Gamatie | 2026-06-17 | 下载 | Urban territories face growing tensions between increasing digital demand, limited resources, and socially constrained built environments. Although distributed computing paradigms such as edge and fog... |
| Compressed-Resident Genomics: Full-Pipeline Device-Resident GPU LZ77 Decode with Position-Invariant Random Access | Yakiv Shavidze | 2026-06-17 | 下载 | Genomic archives grow faster than decompression keeps up: the European Nucleotide Archive holds tens of petabytes of fastq.gz, and gzip is fundamentally sequential. |
| ReMP: Low-Downtime Runtime Model-Parallelism Reconfiguration for LLM Serving | Haipeng Yuan, Kaining Zheng, Yongshu Bai, Yuchen Zhang, Yunquan Zhang, Baodong Wu, Xiang Gao, Daning Cheng | 2026-06-17 | 下载 | Current large language model (LLM) inference systems universally deploy ultra-large-scale models using a combination of Tensor Parallelism (TP) and Pipeline Parallelism (PP). |
| Closed-Form and Constant-Time New-Source Selection for Fault-Tolerant Broadcasting in Dense Gaussian Networks | Bader Albader | 2026-06-17 | 下载 | Fault-tolerant broadcasting in dense Gaussian networks is recovered by re-rooting the broadcast at a new source at maximum graph distance from the faulty nodes. |
| Closed-Form and Constant-Time New-Source Selection for Fault-Tolerant Broadcasting in Dense Eisenstein--Jacobi Networks | Bader Albader | 2026-06-17 | 下载 | Fault-tolerant broadcasting in dense Eisenstein--Jacobi networks requires efficient recovery when faulty nodes disrupt the original broadcast structure. |
| Re-Rooting-Based Fault-Tolerant One-to-All Broadcasting in Dense Eisenstein--Jacobi Networks | Bader Albader | 2026-06-17 | 下载 | Dense Eisenstein--Jacobi networks are degree-six algebraic interconnection topologies with regular structure, vertex symmetry, small diameter, and efficient communication algorithms. |
| HI-HCQC: A Tightly-Coupled Hardware Interface with High-Efficiency Communication for Hybrid Classical-Quantum Computing | Shibo Liang, Junchao Wang, Zeyuan Wang, Feng Wang, Xiaoyu Li, Lei Li, FuDong Liu, Zheng Shan | 2026-06-17 | 下载 | Hybrid classical-quantum computing requires frequent data exchange between classical processors and quantum control hardware. However, existing superconducting quantum control systems are commonly con... |
| ShuntServe: Cost-Efficient LLM Serving on Heterogeneous Spot GPU Clusters | Seungwoo Jeong, Moohyun Song, Juhyun Park, Kyungyong Lee | 2026-06-17 | 下载 | As large language model (LLM) services become widely adopted, the cost of GPU resources for serving these models in cloud environments has emerged as a critical concern. |
| Splaxel: Efficient Distributed Training of 3D Gaussian Splatting for Large-scale Scene Reconstruction via Pixel-level Communication | Wenqi Jia, Zhewen Hu, Ying Huang, Yu Gong, Stavros Kalafatis, Yuke Wang, Wei Niu, Chengming Zhang, Ang Li, Sheng Di, Yuede Ji, Bo Fang, Miao Yin | 2026-06-17 | 下载 | 3D Gaussian Splatting (3DGS) enables high-fidelity and real-time 3D scene reconstruction, but scaling training to large-scale scenes requires optimizing hundreds of millions of Gaussians across multip... |
cs.NI - Networking and Internet Architecture
| 标题 | 作者 | 发布日期 | 摘要 | |
|---|---|---|---|---|
| A Technical Taxonomy of LLM Agent Communication Protocols | Linus Sander, Habtom Kahsay Gidey, Alexander Lenz, Alois Knoll | 2026-06-17 | 下载 | As large language models (LLMs) advance and multi-agent systems aim to overcome the limits of standalone agents, robust communication protocols are becoming essential infrastructure for distributed ag... |
| Atomic Handover for 6G Nomadic Non-Public Networks Using Edge-Based Spectrum Brokering | Daniel Lindenschmitt, Hans D. Schotten | 2026-06-17 | 下载 | Nomadic Non-Public Networks (NNPN) are expected to play an important role in future 6G systems by enabling mobile and rapidly deployable network infrastructures for scenarios such as emergency respons... |
| Direct-V2X Support with 5G Network-based Communications: Performance, Challenges and Solutions | M. C. Lucas-Estañ, B. Coll-Perales, T. Shimizu, J. Gozálvez, T. Higuchi, S. Avedisov, O. Altintas, M. Sepulcre | 2026-06-17 | 下载 | This study analyzes the feasibility of supporting critical V2X services using 5G network-based Vehicle-to-Network-to-Vehicle (V2N2V) communications. |
| An open-source implementation and validation of 5G NR Configured Grant for URLLC in ns-3 5G LENA: a scheduling case study in Industry 4.0 scenarios | Ana Larrañaga, M. Carmen Lucas-Estañ, Sandra Lagén, Zoraze Ali, Imanol Martinez, Javier Gozálvez | 2026-06-17 | 下载 | Factories are undergoing a digital transformation towards cost-efficient, zero-defect manufacturing, creating the need for communication networks capable of meeting stringent latency and reliability r... |
| 5G UE and Network Asset Administration Shells for the Integration of 5G and Industry 4.0 Systems | Jorge Gómez-Jerez, Jorge Cañete-Martín, M. Carmen Lucas-Estañ, Javier Gozalvez | 2026-06-17 | 下载 | 5G is a fundamental technology for the full digitalization of smart manufacturing. The use of Asset Administration Shells (AAS) can facilitate the integration of 5G with Industry 4. |
| Robustness Analysis of Australia's Internet Using a Multilayer Network Model | Benjamin Lang, Matthew Roughan, Mengbin Ye | 2026-06-17 | 下载 | Australia depends on an Internet built from multiple networks of long-haul links. We study the interactions of these independent provider networks to investigate how the peering between these networks... |
| Closed-Form and Constant-Time New-Source Selection for Fault-Tolerant Broadcasting in Dense Gaussian Networks | Bader Albader | 2026-06-17 | 下载 | Fault-tolerant broadcasting in dense Gaussian networks is recovered by re-rooting the broadcast at a new source at maximum graph distance from the faulty nodes. |
| Closed-Form and Constant-Time New-Source Selection for Fault-Tolerant Broadcasting in Dense Eisenstein--Jacobi Networks | Bader Albader | 2026-06-17 | 下载 | Fault-tolerant broadcasting in dense Eisenstein--Jacobi networks requires efficient recovery when faulty nodes disrupt the original broadcast structure. |
| Re-Rooting-Based Fault-Tolerant One-to-All Broadcasting in Dense Eisenstein--Jacobi Networks | Bader Albader | 2026-06-17 | 下载 | Dense Eisenstein--Jacobi networks are degree-six algebraic interconnection topologies with regular structure, vertex symmetry, small diameter, and efficient communication algorithms. |
cs.OS - Operating Systems
| 标题 | 作者 | 发布日期 | 摘要 | |
|---|---|---|---|---|
| LiveStack: OS Support for Cluster-Scale Full-Stack Live Simulation | Yiliang Wan, Haifeng Sun, Yihan Yang, Jonas Kaufmann, Antoine Kaufmann, Jialin Li | 2026-06-17 | 下载 | Cluster-scale full-stack simulation is essential for evaluating distributed software stacks and emerging hardware components before deployment. |