Skip to content

2026-07-30 ​

cs.AR - Architecture ​

标题作者发布日期PDF摘要
Open-Source LLM-Driven Formal Verification: A Multi-Agent Pipeline for RTL RepairHa Trung Tran2026-07-30下载Verification consumes the majority of modern chip design effort, yet the formal verification tools that provide mathematical guarantees of correctness remain expensive and restrictively licensed.
Characterizing LLM Kernel Access and Memory Interaction in Multi-Partition NUMA GPUsDonghyeon Joo, Sooraj Puthoor, Nuwan Jayasena, Bahar Asgari2026-07-30下载Large language model (LLM) workloads motivate multi-partition GPUs as a path to scaling compute and memory capacity, but their non-uniform memory access characteristics and inter-partition communicati...
Demystifying DRAM Read Disturbance: Bridging the Gap Between Experimental Characterization and Device-Level Modeling of RowHammer and RowPress PhenomenaHaocong Luo, Longda Zhou, Ataberk Olgun, İsmail Emir Yüksel, Nisa Bostanci, Zhigang Ji, Xing Wu, Onur Mutlu2026-07-30下载DRAM read disturbance, like RowHammer and RowPress, is a critical robustness issue where accessing DRAM can cause unintended bitflips in other unaccessed DRAM locations.
WitCert: Sound Runtime Risk Observability and Gating for KV-Cache QuantizationFanzhe Wei, Li Liu2026-07-30下载KV-cache quantization is validated today by offline benchmark averages; a deployed system cannot tell whether compression is damaging the request it is serving right now.
ARES: Adaptive Reasoning-Effort Steering for PPA- and Cost-Aware RTL Optimization with LLM AgentsStef Cuyckens, Mihaela Jivanescu, Jun Yin, Chao Fang, Marian Verhelst2026-07-30下载Large language model (LLM) agents optimize the power, performance, and area (PPA) of register-transfer-level (RTL) designs by iterating over edits, synthesis, and PPA analysis, paying a dollar cost fo...
Nanoparticle Networks for Neuromorphic ComputingJonas Mensing, Wilfred G. van der Wiel, Andreas Heuer2026-07-30下载Physical computing leverages complex dynamical systems for energy-efficient data processing. In this work, we present a neuromorphic architecture based on metallic nanoparticles interconnected by mole...
LightRot: A Light-Weighted Rotation Scheme and Architecture for Accurate Low-Bit Large Language Model InferenceSangjin Kim, Yuseon Choi, Jungjun Oh, Byeongcheol Kim, Hoi-Jun Yoo2026-07-30下载As large language models (LLMs) continue to demonstrate exceptional capabilities across various domains, the challenge of achieving energy-efficient and accurate inference becomes increasingly critica...
GyRot: Leveraging Hidden Synergy between Rotation and Fine-grained Group Quantization for Low-bit LLM InferenceSangjin Kim, Yuseon Choi, Byeongcheol Kim, Jungjun Oh, Hoi-jun Yoo2026-07-30下载Low-bit quantization is essential for efficient LLM inference, and both rotation and fine-grained group quantization have shown individual promise.
Optical Flow Sensor: A Direction-Selective Bionic Retina DesignJuchen Zhou, Bonan Yan, Yuchao Yang2026-07-30下载Optical flow characterizes motion in the visual field and is fundamental to motion perception and tracking in biological and artificial vision systems.
Analog Courant Numbers and their Role in Analog ComputingArash Ghasemi2026-07-30下载This paper identifies a dynamical constraint on analog-computing approaches in which a row of the matrix is represented by an impedance network.

cs.DC - Distributed, Parallel, and Cluster Computing ​

标题作者发布日期PDF摘要
High-Level Big Integer Arithmetic in Futhark for GPUsCosmin E. Oancea, Stephen M. Watt2026-07-30下载We report on GPU implementations of block-level addition, subtraction, multiplication and division for midsize integers, with operands of 2152^{15} to 2192^{19} bits using the high-level functional lang...
LayoutBench: Performance Benchmarking of Cloud Storage Layouts for Multimedia DataDebopam Sanyal, Hongjie Chen, Alexey Tumanov, Joshua Kimball2026-07-30下载Modern multimedia machine learning workloads increasingly store large-scale datasets in cloud object storage services such as AWS S3. How these samples are physically organized in storage (i.e.
DeltaServe: Host-Agnostic Co-Serving of Inference and Fine-Tuning for LLMsJiaxuan Chen, Jianshu She, Ye Yuan, Rajat Ghosh, Karan Gupta, Qirong Ho, Xue Liu, Oana Balmau2026-07-30下载LLM serving systems are provisioned for peak load to meet strict latency targets, leaving substantial GPU compute idle whenever traffic falls below peak.
Characterizing LLM Kernel Access and Memory Interaction in Multi-Partition NUMA GPUsDonghyeon Joo, Sooraj Puthoor, Nuwan Jayasena, Bahar Asgari2026-07-30下载Large language model (LLM) workloads motivate multi-partition GPUs as a path to scaling compute and memory capacity, but their non-uniform memory access characteristics and inter-partition communicati...
Safe Quotes for Retroactive Liquidity PoolsPeter Bro Miltersen2026-07-30下载Automated market makers exchange assets through liquidity pools whose quoted prices depend on their reserves, with constant product pools being the most common.
A Taxonomy of Performance Metrics for the Distributed Computing ContinuumPraveen Kumar Donta, Boris Sedlak, Alfreds Lapkovskis, Alaa Saleh, Ying Li, Victor Casamayor Pujol, Ilir Murturi, Manuel Otero Barbasan, Schahram Dustdar2026-07-30下载Performance evaluation is essential for understanding, comparing, and improving computing systems, including Distributed Computing Continuum Systems (DCCS).
Anonymous sharing is pairwise phase-blindBrieuc Le roux tardif2026-07-30下载Independent training jobs sharing a storage system write their checkpoints through the same finite bandwidth, and the resulting bursts of correlated I/O are commonly described as a self-reinforcing "c...
Encryption-Compatible Clustered Federated Learning via Distributed Expectation-Maximization over MetadataMichael Ben Ali, Imen Megdiche, André Péninou, Olivier Teste2026-07-30下载Clustered Federated Learning (CFL) addresses data heterogeneity in federated settings by grouping clients with similar data distributions to enable effective training.
FAIR-Compute: A Roadmap for Fair and Efficient Allocation of Federated Digital Research InfrastructureKonstantinos, E. Zachariadis, Ahmed Sayed, Dimitris Fotakis, Angeliki Mathioudaki, Wan Shuen Siaw2026-07-30下载As demand for high-performance computing (HPC), high-throughput computing and data storage grows, the way scarce compute is allocated -- not just how much exists -- has become a decisive factor in the...
Queue-Theoretic Admission Control for Multi-Tenant GPU ClustersSohan Kunkerkar2026-07-30下载GPU cluster operators cannot predict how long pending workloads will wait for admission. Existing systems use greedy heuristics with no formal wait time guarantees.
A Cloud Continuum Research Infrastructure for Distributed CPS ExperimentationFabio Orazio Mirto, Giuseppe Tricomi, Luca D'Agati, Andrea Sabbioni, Stefano Silvestri, Francesco Longo, Giovanni Merlino, Armir Bujari, Paolo Bellavista, Antonio Puliafito2026-07-30下载Cloud Continuum applications require experimental environments capable of combining heterogeneous Edge, Fog, Cloud, and high-performance computing resources while preserving reproducibility, observabi...
Secure Aggregation for Privacy-Preserving Federated Learning on Clinical EEG DataPouya Rajabi, Mohsen Toorani2026-07-30下载Federated learning enables multiple institutions to train shared models without exchanging raw clinical EEG data, but it does not fully prevent privacy leakage from individual model updates.
SmartGen: Seamless Disaggregated LLM Inference with Selective KV Cache TransferXuchuan Luo, Jiacheng Shen, Xin Wang, Yangfan Zhou2026-07-30下载Disaggregating the prefill and decoding stages of large language model (LLM) inference into two separate sets of nodes is widely adopted in today's LLM serving systems.
ESBT: A Scalable and Deterministic Sequence CRDT for Distributed Collaborative EditingMoulay Driss Mechaoui, Abdessamad Imine2026-07-30下载Modern collaborative editing systems require efficient mechanisms for managing concurrent updates across distributed replicas. Sequence Conflict-free Replicated Data Types (CRDTs) have become the de f...
A CPU+DCU Heterogeneous Parallel Framework for Post-Processing Reconstruction in Quantum Circuit CuttingQingqing Jiang, Weidong Liu, Yufu Liu, Ruiqing He, Jiandong Shang, Hengliang Guo, Qiang Chen2026-07-30下载In the NISQ era, limited qubit resources make it difficult to execute large quantum circuits directly on real hardware. Quantum circuit cutting mitigates this limitation by decomposing a large circuit...
Argonaut: Interactive Visual Exploration for Distributed OptimizationSrijoni Majumdar, Chuhao Qin, Evangelos Pournaras2026-07-30下载Distributed discrete-choice optimization in decentralized settings is often hard to explore and navigate: disentangling what other agents choose, how their choices are interdependent, and how they col...
Learning Color Grading, No Photo Sharing: Federated Aesthetic Preference Learning for Personalized Image EnhancementChuanzhi Xu, Ziyuan Tao, Jean Julien KNell, Yanrong Chen, Haolan Guo, Xuanhua Yin, Adnan Mahmood, Weidong Cai2026-07-30下载Personalized image enhancement should reflect individual aesthetic taste, yet learning such preferences commonly depends on private photos and ratings that are unsuitable for centralized collection.

cs.NI - Networking and Internet Architecture ​

标题作者发布日期PDF摘要
The AnyLog Edge Data FabricRoy Shadmon, Mark Davidson, Eric Aquaronne, Massimiliano Pinto, Ori Shadmon, Moshe Shadmon2026-07-30下载Industrial and autonomous systems increasingly depend on AI, automation, and real-time coordination to act on operational data as it is generated.
When Unlearning Fails: Reliable Data Deletion under Post-Training in Agent NetworksZihao Ding, Jun Huang, Liang Dong2026-07-30下载Self-improving federated agent networks keep training after deployment by collecting new trajectories with the current policy and feeding them back into later rounds.
Sovereign Cognitive Digital Twins: Fusing 6G ISAC, AI-RAN, and Zero-Trust Edge Grids for National Resilience in the Global SouthZoe Aiyanna M. Cayetano, George M. Gichuru, Taijuo T. Morris2026-07-30下载Small-island developing states face accelerating sea-level rise, intensifying cyclones, and storm surge, while suffering the sparse ground instrumentation that makes timely hazard perception difficult...
A Taxonomy of Performance Metrics for the Distributed Computing ContinuumPraveen Kumar Donta, Boris Sedlak, Alfreds Lapkovskis, Alaa Saleh, Ying Li, Victor Casamayor Pujol, Ilir Murturi, Manuel Otero Barbasan, Schahram Dustdar2026-07-30下载Performance evaluation is essential for understanding, comparing, and improving computing systems, including Distributed Computing Continuum Systems (DCCS).
Observing the Relationship between QoS Unpredictability, Prediction Error, and User Activity in a Remote Desktop ServiceKeisuke Ishibashi, Xuliang Deng, Yoshiaki Kitaguchi, Kenichi Nagami, Ichiro Mizukoshi, Akira Sato, Daiyu Nobori2026-07-30下载With the increasing need for remote work, especially since the COVID-19 era, Remote Desktop Services (RDS) have become widely used. Because interactive RDS usage depends heavily on communication quali...
Coexistence of 5G NR and Wi Fi 6E/7 at 6 GHz: Experimental Interference MeasurementsRafik Zitouni, Demos Serghiou, Ali Dagdeviren, Tajinder Randhawa, Edwards Udean, Hanli Dong, Riccardo Pozza, Rahim Tafazolli2026-07-30下载This paper presents the first conducted-interference measurements of a commercial Very Low Power (VLP) Wi-Fi 6E/7 device into both the gNB uplink and UE downlink receiver chains of a live 5G New Radio...
Powering Net-Zero 6G: Packetized Energy Management for Grid-Interactive Telecom InfrastructureAdnan Aijaz, Xinyi Lin2026-07-30下载The transition to net-zero 6G requires energy-management approaches that go beyond conventional RAN efficiency mechanisms. As future networks integrate AI-native operation, edge intelligence, dense de...
PCAP-LM: An LLM-Native Text Representation for TLS Bulk Traffic AnalysisXavier Marjou, Lucas Tamic, Ilan Jaffeux-Cheniout2026-07-30下载Large language models (LLMs) offer powerful reasoning capabilities for network traffic analysis, but standard capture formats and their textual equivalents are prohibitively verbose, overflowing LLM c...
Layered Architecture for Mobile IntelligenceQingwen Liu, Mingqing Liu2026-07-30下载Artificial intelligence (AI) is rapidly evolving from a centralized computing capability into a pervasive infrastructure that interacts directly with the physical world.
Geometric View on Integrated Cascaded Channel of IRS-Aided CommunicationsYunli Li, Young Jin Chun2026-07-30下载The hybrid intelligent reflecting surface (IRS) architecture is a novel technology that leverages the advantages of both passive and active IRS; the passive IRS offers a large aperture, while the acti...
Localization and Pursuit of a Mobile Target using Distance-only MeasurementsNabarupa Das, Suvadip Batabyal2026-07-30下载This paper investigates localization and pursuit of a linearly moving target in a two-dimensional plane using distance measurements only. The distance from the target is estimated from pathloss measur...

cs.OS - Operating Systems ​

标题作者发布日期PDF摘要
From C to Idiomatic Rust: A Ship-of-Theseus Agentic TranslationVasily A. Sartakov2026-07-30下载C underpins operating systems, embedded platforms, and network infrastructure because its abstractions map directly to machine behaviour. Its explicit memory model, predictable data representations, a...

cs.PF - Performance ​

标题作者发布日期PDF摘要
Characterizing LLM Kernel Access and Memory Interaction in Multi-Partition NUMA GPUsDonghyeon Joo, Sooraj Puthoor, Nuwan Jayasena, Bahar Asgari2026-07-30下载Large language model (LLM) workloads motivate multi-partition GPUs as a path to scaling compute and memory capacity, but their non-uniform memory access characteristics and inter-partition communicati...
Quantum Fidelity-per-Cost: A Metric for Evaluation of Quantum Computing SystemsSiddarth Shinde, Jakub Szefer2026-07-30下载Cloud-accessible quantum computing has made hardware comparison not only a physics benchmark but also a practical purchasing decision. Cost-aware comparison of quantum computers remains underexplored ...
A Taxonomy of Performance Metrics for the Distributed Computing ContinuumPraveen Kumar Donta, Boris Sedlak, Alfreds Lapkovskis, Alaa Saleh, Ying Li, Victor Casamayor Pujol, Ilir Murturi, Manuel Otero Barbasan, Schahram Dustdar2026-07-30下载Performance evaluation is essential for understanding, comparing, and improving computing systems, including Distributed Computing Continuum Systems (DCCS).
Extended Depth-First Representations of k2k^2-treesGabriel Carmona, Paolo Ferragina, Giovanni Manzini, Francesco Tosoni2026-07-30下载In this paper, we study static, computation-friendly, lossless compression formats for graphs, focusing on memory locality and operational efficiency of k2k^2-trees.
Load balancing in parallel infinite-server queues with action delay via phase representationKazuma Abe, Tuan Phung-Duc2026-07-30下载Spatially distributed service systems rely on state-dependent routing to allocate users, tasks, or requests to less-loaded service nodes. In practice, a routing decision does not take effect immediate...
Reflected UAS: Corrected Deterministic Stability and Direct CTMC Drift CalculationKrishna Subedi2026-07-30下载We analyze Reflected UAS routing for heterogeneous multi-server queues at fixed parameters under subcritical load. The deterministic surrogate is a reflected ODE on the nonnegative orthant, not the un...

基于 VitePress 构建 · 使用本地搜索查找论文