2026-07-30
cs.AR - Architecture
| 标题 | 作者 | 发布日期 | 摘要 | |
|---|---|---|---|---|
| Open-Source LLM-Driven Formal Verification: A Multi-Agent Pipeline for RTL Repair | Ha Trung Tran | 2026-07-30 | 下载 | Verification consumes the majority of modern chip design effort, yet the formal verification tools that provide mathematical guarantees of correctness remain expensive and restrictively licensed. |
| Characterizing LLM Kernel Access and Memory Interaction in Multi-Partition NUMA GPUs | Donghyeon Joo, Sooraj Puthoor, Nuwan Jayasena, Bahar Asgari | 2026-07-30 | 下载 | Large language model (LLM) workloads motivate multi-partition GPUs as a path to scaling compute and memory capacity, but their non-uniform memory access characteristics and inter-partition communicati... |
| Demystifying DRAM Read Disturbance: Bridging the Gap Between Experimental Characterization and Device-Level Modeling of RowHammer and RowPress Phenomena | Haocong Luo, Longda Zhou, Ataberk Olgun, İsmail Emir Yüksel, Nisa Bostanci, Zhigang Ji, Xing Wu, Onur Mutlu | 2026-07-30 | 下载 | DRAM read disturbance, like RowHammer and RowPress, is a critical robustness issue where accessing DRAM can cause unintended bitflips in other unaccessed DRAM locations. |
| WitCert: Sound Runtime Risk Observability and Gating for KV-Cache Quantization | Fanzhe Wei, Li Liu | 2026-07-30 | 下载 | KV-cache quantization is validated today by offline benchmark averages; a deployed system cannot tell whether compression is damaging the request it is serving right now. |
| ARES: Adaptive Reasoning-Effort Steering for PPA- and Cost-Aware RTL Optimization with LLM Agents | Stef Cuyckens, Mihaela Jivanescu, Jun Yin, Chao Fang, Marian Verhelst | 2026-07-30 | 下载 | Large language model (LLM) agents optimize the power, performance, and area (PPA) of register-transfer-level (RTL) designs by iterating over edits, synthesis, and PPA analysis, paying a dollar cost fo... |
| Nanoparticle Networks for Neuromorphic Computing | Jonas Mensing, Wilfred G. van der Wiel, Andreas Heuer | 2026-07-30 | 下载 | Physical computing leverages complex dynamical systems for energy-efficient data processing. In this work, we present a neuromorphic architecture based on metallic nanoparticles interconnected by mole... |
| LightRot: A Light-Weighted Rotation Scheme and Architecture for Accurate Low-Bit Large Language Model Inference | Sangjin Kim, Yuseon Choi, Jungjun Oh, Byeongcheol Kim, Hoi-Jun Yoo | 2026-07-30 | 下载 | As large language models (LLMs) continue to demonstrate exceptional capabilities across various domains, the challenge of achieving energy-efficient and accurate inference becomes increasingly critica... |
| GyRot: Leveraging Hidden Synergy between Rotation and Fine-grained Group Quantization for Low-bit LLM Inference | Sangjin Kim, Yuseon Choi, Byeongcheol Kim, Jungjun Oh, Hoi-jun Yoo | 2026-07-30 | 下载 | Low-bit quantization is essential for efficient LLM inference, and both rotation and fine-grained group quantization have shown individual promise. |
| Optical Flow Sensor: A Direction-Selective Bionic Retina Design | Juchen Zhou, Bonan Yan, Yuchao Yang | 2026-07-30 | 下载 | Optical flow characterizes motion in the visual field and is fundamental to motion perception and tracking in biological and artificial vision systems. |
| Analog Courant Numbers and their Role in Analog Computing | Arash Ghasemi | 2026-07-30 | 下载 | This paper identifies a dynamical constraint on analog-computing approaches in which a row of the matrix is represented by an impedance network. |
cs.DC - Distributed, Parallel, and Cluster Computing
| 标题 | 作者 | 发布日期 | 摘要 | |
|---|---|---|---|---|
| High-Level Big Integer Arithmetic in Futhark for GPUs | Cosmin E. Oancea, Stephen M. Watt | 2026-07-30 | 下载 | We report on GPU implementations of block-level addition, subtraction, multiplication and division for midsize integers, with operands of to bits using the high-level functional lang... |
| LayoutBench: Performance Benchmarking of Cloud Storage Layouts for Multimedia Data | Debopam Sanyal, Hongjie Chen, Alexey Tumanov, Joshua Kimball | 2026-07-30 | 下载 | Modern multimedia machine learning workloads increasingly store large-scale datasets in cloud object storage services such as AWS S3. How these samples are physically organized in storage (i.e. |
| DeltaServe: Host-Agnostic Co-Serving of Inference and Fine-Tuning for LLMs | Jiaxuan Chen, Jianshu She, Ye Yuan, Rajat Ghosh, Karan Gupta, Qirong Ho, Xue Liu, Oana Balmau | 2026-07-30 | 下载 | LLM serving systems are provisioned for peak load to meet strict latency targets, leaving substantial GPU compute idle whenever traffic falls below peak. |
| Characterizing LLM Kernel Access and Memory Interaction in Multi-Partition NUMA GPUs | Donghyeon Joo, Sooraj Puthoor, Nuwan Jayasena, Bahar Asgari | 2026-07-30 | 下载 | Large language model (LLM) workloads motivate multi-partition GPUs as a path to scaling compute and memory capacity, but their non-uniform memory access characteristics and inter-partition communicati... |
| Safe Quotes for Retroactive Liquidity Pools | Peter Bro Miltersen | 2026-07-30 | 下载 | Automated market makers exchange assets through liquidity pools whose quoted prices depend on their reserves, with constant product pools being the most common. |
| A Taxonomy of Performance Metrics for the Distributed Computing Continuum | Praveen Kumar Donta, Boris Sedlak, Alfreds Lapkovskis, Alaa Saleh, Ying Li, Victor Casamayor Pujol, Ilir Murturi, Manuel Otero Barbasan, Schahram Dustdar | 2026-07-30 | 下载 | Performance evaluation is essential for understanding, comparing, and improving computing systems, including Distributed Computing Continuum Systems (DCCS). |
| Anonymous sharing is pairwise phase-blind | Brieuc Le roux tardif | 2026-07-30 | 下载 | Independent training jobs sharing a storage system write their checkpoints through the same finite bandwidth, and the resulting bursts of correlated I/O are commonly described as a self-reinforcing "c... |
| Encryption-Compatible Clustered Federated Learning via Distributed Expectation-Maximization over Metadata | Michael Ben Ali, Imen Megdiche, André Péninou, Olivier Teste | 2026-07-30 | 下载 | Clustered Federated Learning (CFL) addresses data heterogeneity in federated settings by grouping clients with similar data distributions to enable effective training. |
| FAIR-Compute: A Roadmap for Fair and Efficient Allocation of Federated Digital Research Infrastructure | Konstantinos, E. Zachariadis, Ahmed Sayed, Dimitris Fotakis, Angeliki Mathioudaki, Wan Shuen Siaw | 2026-07-30 | 下载 | As demand for high-performance computing (HPC), high-throughput computing and data storage grows, the way scarce compute is allocated -- not just how much exists -- has become a decisive factor in the... |
| Queue-Theoretic Admission Control for Multi-Tenant GPU Clusters | Sohan Kunkerkar | 2026-07-30 | 下载 | GPU cluster operators cannot predict how long pending workloads will wait for admission. Existing systems use greedy heuristics with no formal wait time guarantees. |
| A Cloud Continuum Research Infrastructure for Distributed CPS Experimentation | Fabio Orazio Mirto, Giuseppe Tricomi, Luca D'Agati, Andrea Sabbioni, Stefano Silvestri, Francesco Longo, Giovanni Merlino, Armir Bujari, Paolo Bellavista, Antonio Puliafito | 2026-07-30 | 下载 | Cloud Continuum applications require experimental environments capable of combining heterogeneous Edge, Fog, Cloud, and high-performance computing resources while preserving reproducibility, observabi... |
| Secure Aggregation for Privacy-Preserving Federated Learning on Clinical EEG Data | Pouya Rajabi, Mohsen Toorani | 2026-07-30 | 下载 | Federated learning enables multiple institutions to train shared models without exchanging raw clinical EEG data, but it does not fully prevent privacy leakage from individual model updates. |
| SmartGen: Seamless Disaggregated LLM Inference with Selective KV Cache Transfer | Xuchuan Luo, Jiacheng Shen, Xin Wang, Yangfan Zhou | 2026-07-30 | 下载 | Disaggregating the prefill and decoding stages of large language model (LLM) inference into two separate sets of nodes is widely adopted in today's LLM serving systems. |
| ESBT: A Scalable and Deterministic Sequence CRDT for Distributed Collaborative Editing | Moulay Driss Mechaoui, Abdessamad Imine | 2026-07-30 | 下载 | Modern collaborative editing systems require efficient mechanisms for managing concurrent updates across distributed replicas. Sequence Conflict-free Replicated Data Types (CRDTs) have become the de f... |
| A CPU+DCU Heterogeneous Parallel Framework for Post-Processing Reconstruction in Quantum Circuit Cutting | Qingqing Jiang, Weidong Liu, Yufu Liu, Ruiqing He, Jiandong Shang, Hengliang Guo, Qiang Chen | 2026-07-30 | 下载 | In the NISQ era, limited qubit resources make it difficult to execute large quantum circuits directly on real hardware. Quantum circuit cutting mitigates this limitation by decomposing a large circuit... |
| Argonaut: Interactive Visual Exploration for Distributed Optimization | Srijoni Majumdar, Chuhao Qin, Evangelos Pournaras | 2026-07-30 | 下载 | Distributed discrete-choice optimization in decentralized settings is often hard to explore and navigate: disentangling what other agents choose, how their choices are interdependent, and how they col... |
| Learning Color Grading, No Photo Sharing: Federated Aesthetic Preference Learning for Personalized Image Enhancement | Chuanzhi Xu, Ziyuan Tao, Jean Julien KNell, Yanrong Chen, Haolan Guo, Xuanhua Yin, Adnan Mahmood, Weidong Cai | 2026-07-30 | 下载 | Personalized image enhancement should reflect individual aesthetic taste, yet learning such preferences commonly depends on private photos and ratings that are unsuitable for centralized collection. |
cs.NI - Networking and Internet Architecture
| 标题 | 作者 | 发布日期 | 摘要 | |
|---|---|---|---|---|
| The AnyLog Edge Data Fabric | Roy Shadmon, Mark Davidson, Eric Aquaronne, Massimiliano Pinto, Ori Shadmon, Moshe Shadmon | 2026-07-30 | 下载 | Industrial and autonomous systems increasingly depend on AI, automation, and real-time coordination to act on operational data as it is generated. |
| When Unlearning Fails: Reliable Data Deletion under Post-Training in Agent Networks | Zihao Ding, Jun Huang, Liang Dong | 2026-07-30 | 下载 | Self-improving federated agent networks keep training after deployment by collecting new trajectories with the current policy and feeding them back into later rounds. |
| Sovereign Cognitive Digital Twins: Fusing 6G ISAC, AI-RAN, and Zero-Trust Edge Grids for National Resilience in the Global South | Zoe Aiyanna M. Cayetano, George M. Gichuru, Taijuo T. Morris | 2026-07-30 | 下载 | Small-island developing states face accelerating sea-level rise, intensifying cyclones, and storm surge, while suffering the sparse ground instrumentation that makes timely hazard perception difficult... |
| A Taxonomy of Performance Metrics for the Distributed Computing Continuum | Praveen Kumar Donta, Boris Sedlak, Alfreds Lapkovskis, Alaa Saleh, Ying Li, Victor Casamayor Pujol, Ilir Murturi, Manuel Otero Barbasan, Schahram Dustdar | 2026-07-30 | 下载 | Performance evaluation is essential for understanding, comparing, and improving computing systems, including Distributed Computing Continuum Systems (DCCS). |
| Observing the Relationship between QoS Unpredictability, Prediction Error, and User Activity in a Remote Desktop Service | Keisuke Ishibashi, Xuliang Deng, Yoshiaki Kitaguchi, Kenichi Nagami, Ichiro Mizukoshi, Akira Sato, Daiyu Nobori | 2026-07-30 | 下载 | With the increasing need for remote work, especially since the COVID-19 era, Remote Desktop Services (RDS) have become widely used. Because interactive RDS usage depends heavily on communication quali... |
| Coexistence of 5G NR and Wi Fi 6E/7 at 6 GHz: Experimental Interference Measurements | Rafik Zitouni, Demos Serghiou, Ali Dagdeviren, Tajinder Randhawa, Edwards Udean, Hanli Dong, Riccardo Pozza, Rahim Tafazolli | 2026-07-30 | 下载 | This paper presents the first conducted-interference measurements of a commercial Very Low Power (VLP) Wi-Fi 6E/7 device into both the gNB uplink and UE downlink receiver chains of a live 5G New Radio... |
| Powering Net-Zero 6G: Packetized Energy Management for Grid-Interactive Telecom Infrastructure | Adnan Aijaz, Xinyi Lin | 2026-07-30 | 下载 | The transition to net-zero 6G requires energy-management approaches that go beyond conventional RAN efficiency mechanisms. As future networks integrate AI-native operation, edge intelligence, dense de... |
| PCAP-LM: An LLM-Native Text Representation for TLS Bulk Traffic Analysis | Xavier Marjou, Lucas Tamic, Ilan Jaffeux-Cheniout | 2026-07-30 | 下载 | Large language models (LLMs) offer powerful reasoning capabilities for network traffic analysis, but standard capture formats and their textual equivalents are prohibitively verbose, overflowing LLM c... |
| Layered Architecture for Mobile Intelligence | Qingwen Liu, Mingqing Liu | 2026-07-30 | 下载 | Artificial intelligence (AI) is rapidly evolving from a centralized computing capability into a pervasive infrastructure that interacts directly with the physical world. |
| Geometric View on Integrated Cascaded Channel of IRS-Aided Communications | Yunli Li, Young Jin Chun | 2026-07-30 | 下载 | The hybrid intelligent reflecting surface (IRS) architecture is a novel technology that leverages the advantages of both passive and active IRS; the passive IRS offers a large aperture, while the acti... |
| Localization and Pursuit of a Mobile Target using Distance-only Measurements | Nabarupa Das, Suvadip Batabyal | 2026-07-30 | 下载 | This paper investigates localization and pursuit of a linearly moving target in a two-dimensional plane using distance measurements only. The distance from the target is estimated from pathloss measur... |
cs.OS - Operating Systems
| 标题 | 作者 | 发布日期 | 摘要 | |
|---|---|---|---|---|
| From C to Idiomatic Rust: A Ship-of-Theseus Agentic Translation | Vasily A. Sartakov | 2026-07-30 | 下载 | C underpins operating systems, embedded platforms, and network infrastructure because its abstractions map directly to machine behaviour. Its explicit memory model, predictable data representations, a... |
cs.PF - Performance
| 标题 | 作者 | 发布日期 | 摘要 | |
|---|---|---|---|---|
| Characterizing LLM Kernel Access and Memory Interaction in Multi-Partition NUMA GPUs | Donghyeon Joo, Sooraj Puthoor, Nuwan Jayasena, Bahar Asgari | 2026-07-30 | 下载 | Large language model (LLM) workloads motivate multi-partition GPUs as a path to scaling compute and memory capacity, but their non-uniform memory access characteristics and inter-partition communicati... |
| Quantum Fidelity-per-Cost: A Metric for Evaluation of Quantum Computing Systems | Siddarth Shinde, Jakub Szefer | 2026-07-30 | 下载 | Cloud-accessible quantum computing has made hardware comparison not only a physics benchmark but also a practical purchasing decision. Cost-aware comparison of quantum computers remains underexplored ... |
| A Taxonomy of Performance Metrics for the Distributed Computing Continuum | Praveen Kumar Donta, Boris Sedlak, Alfreds Lapkovskis, Alaa Saleh, Ying Li, Victor Casamayor Pujol, Ilir Murturi, Manuel Otero Barbasan, Schahram Dustdar | 2026-07-30 | 下载 | Performance evaluation is essential for understanding, comparing, and improving computing systems, including Distributed Computing Continuum Systems (DCCS). |
| Extended Depth-First Representations of -trees | Gabriel Carmona, Paolo Ferragina, Giovanni Manzini, Francesco Tosoni | 2026-07-30 | 下载 | In this paper, we study static, computation-friendly, lossless compression formats for graphs, focusing on memory locality and operational efficiency of -trees. |
| Load balancing in parallel infinite-server queues with action delay via phase representation | Kazuma Abe, Tuan Phung-Duc | 2026-07-30 | 下载 | Spatially distributed service systems rely on state-dependent routing to allocate users, tasks, or requests to less-loaded service nodes. In practice, a routing decision does not take effect immediate... |
| Reflected UAS: Corrected Deterministic Stability and Direct CTMC Drift Calculation | Krishna Subedi | 2026-07-30 | 下载 | We analyze Reflected UAS routing for heterogeneous multi-server queues at fixed parameters under subcritical load. The deterministic surrogate is a reflected ODE on the nonnegative orthant, not the un... |