2026-09-22
cs.AR - Architecture
| 标题 | 作者 | 发布日期 | 摘要 | |
|---|---|---|---|---|
| Agentic-IC3: Enabling Semantic Proof Search in IC3 Model Checking | Yu-Wei Fan, SooHyuk Cho, Aarti Gupta, Sharad Malik | 2026-09-22 | 下载 | IC3 is a state-of-the-art algorithm for hardware model checking that proves safety properties by incrementally constructing an inductive invariant consisting of a set of lemmas. |
| Bringing Chip Tapeout Into University Education | Luca Pezzarossa, Martin Schoeberl, Matti Käyrä, Nooshin Nosrati, Matthias Bo Stuart, Timo D. Hämäläinen, Jean-Max Dutertre, Michael Pehl | 2026-09-22 | 下载 | Providing students with experience from chip specification to fabricated silicon can strengthen chip-design education, but integrating tapeout into regular teaching is difficult to scale, especially a... |
| Dynamic Slack-Aware Clocking for Near-Threshold Tensor Processing Units (TPUs) | Muhammad Usman Nadeem, Sanghamitra Roy, Koushik Chakraborty | 2026-09-22 | 下载 | Operating Tensor Processing Units (TPUs) in the near-threshold computing (NTC) region significantly reduces energy consumption but introduces high delay sensitivity to process variation and data activ... |
| Toki: Profiling HBM Performance on FPGA Systems with RISC-V Soft Cores and PCIe Host DMA Traffic | Andrea Galimberti, Andrea Motta, Gianni Antichi, Davide Zoni | 2026-09-22 | 下载 | Programmable RISC-V soft cores are becoming more widespread in data-center scenarios, making it crucial to design efficient systems that deploy them on FPGA chips with HBM memory. |
| ESupNNet: An Error Supervising Neural Network architecture for error detection against soft errors in parameters | Jorge Cano-Paez, Luis Entrena, Almudena Lindoso | 2026-09-22 | 下载 | This work presents a novel approach to detect misclassification errors in CNNs caused by soft errors in their parameters. We propose an architecture that uses inter-class relations induced by the CNN ... |
| AgenticSizing: A Large Language Model-based Multi-Agent Framework for Analog Circuit Sizing | Yijia Hao, Pratibha Verma, Dongxu Guo, Cristian Sestito, Michael O'Boyle, Christos-Savvas Bouganis, Themis Prodromakis | 2026-09-22 | 下载 | Analog circuit sizing remains a challenging and time-consuming task due to the large design space, strong performance trade-offs, and increasing circuit complexity in scaled technologies. |
| Decoupling Logical Masks from GPU Execution for Dynamic Block-Sparse Attention | Shanghao Liu, Xiaoyun Yu, Wanting Li, Wenqi Jiang | 2026-09-22 | 下载 | Attention computation makes inference expensive in video diffusion transformers (vDiTs), which generate videos through iterative denoising. Block-sparse attention (BSA) reduces this cost by computin... |
| Hot-Cold Tiering of HBM and High Bandwidth Flash for Agentic LLM Serving | Jongjin Baek, Won Ji, Seungjae Yoo, Joo-Young Kim | 2026-09-22 | 下载 | Large language model (LLM) serving is increasingly agentic, with multi-turn sessions that idle between actions yet must retain their full context. |
| SLED-IFV: Solver-Validated LLM-Guided Decomposition for Scalable Hardware Information-Flow Verification | Liangtao Dai, Yimin Gao, Melika Morsali, Mircea Stan | 2026-09-22 | 下载 | Formal hardware information-flow verification (IFV) provides strong guarantees against secret-dependent timing and control behavior, but often scales poorly on realistic RTL. |
| Accelerating the Mitigation of LLM Inference Nondeterminism Across GPU Architectures | Liam Cooper, Shinnung Jeong, Hyeran Jeon, Jeffrey Young, Hyesoon Kim | 2026-09-22 | 下载 | Large language model (LLM) outputs are expected to be reproducible under greedy decoding, yet in practice the same model, prompt, and software stack produce different outputs on different GPUs. |
cs.DC - Distributed, Parallel, and Cluster Computing
| 标题 | 作者 | 发布日期 | 摘要 | |
|---|---|---|---|---|
| Crossflow: Prefill-Decode Elasticity for Agentic LLM Serving | Yi Xu, Ehsan K. Ardestani, Wenyin Fu, Martin Schatz, Krishna Malladi, Zhan Shu, Adnan Aziz, Shobhit Kanaujia, Ajit Mathews, Chunqiang Tang | 2026-09-22 | 下载 | As serving capacity demand surpasses that of training, serving efficiency becomes increasingly important. Prefill-decode (P/D) disaggregation improves serving efficiency through specialization and iso... |
| EMA: Elastic and Performance Transparent Memory Across GPUs | Yi Xu, Tian Xia, Ion Stoica | 2026-09-22 | 下载 | Multi-GPU servers have become the standard building block of modern data centers, providing aggregated capacity through high-bandwidth interconnects. |
| Quantum Advantage for Distributed Symmetry Breaking | Maxime Flin, Longcheng Li, Jukka Suomela | 2026-09-22 | 下载 | We present a distributed quantum algorithm that -colors cycles in rounds, with high probability. It follows that all locally checkable labeling problems (LCLs) that have round complexity $O(... |
| SARA: SLO-Aware Resource Allocation for Disaggregated Agentic LLM Services | Shicong Liu, Xianghao Yu, Zhen Gao, Jun Zhang | 2026-09-22 | 下载 | Recent advances in large language models (LLMs) are driving the emergence of multi-modal and agentic services for mobile users through cloud and edge infrastructures, where long-context workloads pose... |
| Seeking Cost-Optimal Infrastructure Size for Distributed Filesystems: A Ceph Case Study | Niccolo Tosato, Isac Pasianotto, Ruggero Lot, Stefano Cozzini | 2026-09-22 | 下载 | Distributed Filesystems (DFS) are a crucial component of modern computing environments, and their performance is critical to the success of all the facilities that rely on them. |
| A Configurable Heuristic for the MLCS Problem | Farhana Akter Tumpa, Rebin Silva Valan Arasu, Rajiv Gupta | 2026-09-22 | 下载 | Motivation: The Multiple Longest Common Subsequence (MLCS) problem for an arbitrary number of sequences is an NP-hard problem in sequence analysis. |
| Don't let your Memory defy you: Fragmentation-Aware Serverless Allocation with Elastic Memory Locality | Achilleas Tzenetopoulos, Dimosthenis Masouros, Sotirios Xydis, Francky Catthoor, Dimitrios Soudris | 2026-09-22 | 下载 | Serverless platforms commonly rely on bundled, memory-centric configurations, where CPU capacity follows the specified memory size. Resource decoupling reduces this waste, but can create external frag... |
| DHSched: Stateless Control for Stateful Real-Time Avatar Serving | Xin Wang, Haitong Zhang, Xianghong Li, Shumin Lin, Xianzheng Song, Lin Wang, Zhenyu Xu | 2026-09-22 | 下载 | Real-time avatar services run long-lived, GPU-backed sessions that transform a continuous text stream into speech, facial motion, rendered video, and RTC media. |
| Flux: Optimal Scheduling of Optical Circuit Switches for LLM Training | Arno Troch, Seyyidahmed Lahmer, Abubakr Nada, Jeroen Famaey, Michael Peeters | 2026-09-22 | 下载 | Optical Circuit Switching (OCS) offers high bandwidth density and energy efficiency for LLM training, but incurs a non-negligible reconfiguration delay. |
| Hydrozoan: Latency-Adaptive DAG Consensus under Mixed Byzantine and Crash Faults | Qianyu Yu, Lefteris Kokoris-Kogias, Alberto Sonnino | 2026-09-22 | 下载 | DAG-based consensus protocols can achieve great throughput and the optimal three-message-delay limit for n = 3f+1 consensus. While two-delay protocols exist, they pay with reduced resilience (requirin... |
| Lizard: Bandwidth-Adaptive Real-Time Video Analytics through Content-Aware Packet Discarding at Last-Mile Edge Routers | Shan Yu, Yu Chen, Yifan Qiao, Sheng Zhang, Ravi Netravali, Harry Xu | 2026-09-22 | 下载 | The timeliness and accuracy of edge-based video analytics can be hindered by drastic reductions in available bandwidth (ABW) at last-mile edge routers, causing prolonged queuing delays. |
| Co-Fabric: Breaking Host-Domain Boundaries for Unified xPU Interconnection | Zhen Peng, Jiaming Huang, Chaofan Chen, Zhao Zhang, An Wu, Baoyang Liu, Xinglong Wang, Tanlong Ci, Jinfeng Li, Xueke Duan, Hao Wang, Xi Chen, Shunshun Zhang, Zhiyuan Su, Zhu Cao, Zhichong Dou, Shaohua Wu, Lu Jing, Yue Yuan | 2026-09-22 | 下载 | Large-model parameters have grown beyond the capacity of a single xPU, dispersing across multiple xPUs spanning distinct host domains, where xPU-to-xPU communication dominates overall system efficienc... |
cs.NI - Networking and Internet Architecture
| 标题 | 作者 | 发布日期 | 摘要 | |
|---|---|---|---|---|
| RIS-Enabled Integrated Access and Relay: Empowering Collaboration Among BSs | Hao Lin, Mustafa A. Kishk, Mohamed-Slim Alouini | 2026-09-22 | 下载 | The increasing number of Internet of Things (IoT) devices and applications leads to severe access congestion in conventional base station (BS) networks. |
| Reading the Sky to Forecast the Ground: Physics-Informed Link-State Forecasting for LEO Networks at Any Location | Yunxiang Chi, Zhenlin An, Longfei Shangguan, Kyle Jamieson | 2026-09-22 | 下载 | In this paper, we introduce Gnomon, a physics-informed system that forecasts user-perceived low-Earth-orbit (LEO) downlink throughput, uplink throughput, and round-trip time (RTT) under different leve... |
| Frugal Collective Perception: Context-Aware Adaptive Reporting for Safety-Critical C-ITS | Romain Tessier, Bruno Monsuez, Oyunchimeg Shagdar, Adriana Tapus | 2026-09-22 | 下载 | Ensuring safety and scalability in Collective Perception Service (CPS) remains a key challenge for Cooperative Intelligent Transport Systems (C-ITS). |
| Overload-Robust Latency in 5G-TSN: A HoL-Enhanced Hybrid Lyapunov Approach for 3GPP Indoor Factory Environments | Kouros Zanbouri, Md Noor-A-Rahim, Cormac J Sreenan, Dirk Pesch | 2026-09-22 | 下载 | Private 5G networks are a key enabler for flexible industrial automation, especially when used in conjunction with Time-Sensitive Networking (TSN) technology. |
| Flux: Optimal Scheduling of Optical Circuit Switches for LLM Training | Arno Troch, Seyyidahmed Lahmer, Abubakr Nada, Jeroen Famaey, Michael Peeters | 2026-09-22 | 下载 | Optical Circuit Switching (OCS) offers high bandwidth density and energy efficiency for LLM training, but incurs a non-negligible reconfiguration delay. |
| Modulating Retroreflector-Aided UAV-Based FSO/QKD Systems | Duy N. Luong, Duy-Tuan Dao, Cuong T. Nguyen, Hoang D. Le, Anh T. Pham | 2026-09-22 | 下载 | Unmanned aerial vehicles (UAVs)-based free-space optics (FSO)/quantum key distribution (QKD) systems require high-precision pointing mechanisms. |
| Rethinking Web Application Firewalls | Laurin Brandner, Laurent Vanbever | 2026-09-22 | 下载 | In recent years, the threat of application-layer (L7) distributed denial-of-service (DDoS) attacks is ever increasing. To defend against them, network operators deploy web application firewalls (WAFs)... |
| Lizard: Bandwidth-Adaptive Real-Time Video Analytics through Content-Aware Packet Discarding at Last-Mile Edge Routers | Shan Yu, Yu Chen, Yifan Qiao, Sheng Zhang, Ravi Netravali, Harry Xu | 2026-09-22 | 下载 | The timeliness and accuracy of edge-based video analytics can be hindered by drastic reductions in available bandwidth (ABW) at last-mile edge routers, causing prolonged queuing delays. |
| Adaptive Traffic Camouflage: Causal and Resource-Aware Defense Against IoT Fingerprinting | Daniel Adu Worae, Spyridon Mastorakis, Nuno Moniz, Nitesh V. Chawla | 2026-09-22 | 下载 | Encryption hides IoT payloads, but traffic shape can still reveal device identity through packet sizes, timing, direction, and packetization. We present Adaptive Traffic Camouflage, a causal, leakage-... |
| Key Reconciliation with RC-LDPC/Error Estimation for Satellite-based FSO/QKD Systems | Cuong T. Nguyen, Hoang D. Le, Anshul Jaiswal, Swaminathan R., Anh T. Pham | 2026-09-22 | 下载 | Satellite-based free-space optics (FSO) quantum key distribution (QKD) systems have recently attracted significant research interest due to their potential to enable globally secured applications. |
| Semord: Learned Semantic-Preserving Placement and Low-Fanout Routing for Distributed Vector Search | Shengze Wang, Yi Liu, Yifan Hua, Xiaoxue Zhang, Chen Qian | 2026-09-22 | 下载 | Vector databases are increasingly deployed in distributed settings where different users, sites, or domains maintain vector data. Existing vector databases rely on a coordinator to record which shards... |
cs.PF - Performance
| 标题 | 作者 | 发布日期 | 摘要 | |
|---|---|---|---|---|
| SARA: SLO-Aware Resource Allocation for Disaggregated Agentic LLM Services | Shicong Liu, Xianghao Yu, Zhen Gao, Jun Zhang | 2026-09-22 | 下载 | Recent advances in large language models (LLMs) are driving the emergence of multi-modal and agentic services for mobile users through cloud and edge infrastructures, where long-context workloads pose... |
| Seeking Cost-Optimal Infrastructure Size for Distributed Filesystems: A Ceph Case Study | Niccolo Tosato, Isac Pasianotto, Ruggero Lot, Stefano Cozzini | 2026-09-22 | 下载 | Distributed Filesystems (DFS) are a crucial component of modern computing environments, and their performance is critical to the success of all the facilities that rely on them. |