2026-09-20
cs.AR - Architecture
| 标题 | 作者 | 发布日期 | 摘要 | |
|---|---|---|---|---|
| SPLASH: Co-Designing Sparse Attention with High-Bandwidth Flash for Efficient Long-Context Inference | Aditya Anirudh Jonnalagadda, Agasthi Haputhanthri, Pranav Dangi, Rohan Juneja, Wenshuo Yue, Aritra Bagchi, Bin Gao, Tulika Mitra | 2026-09-20 | 下载 | The key-value (KV) cache has become the dominant consumer of memory in large language model (LLM) serving systems as context lengths, concurrency, and request lifetimes grow. |
| VSpector: Specification-Driven Bug Detection for RISC-V CPUs | Tianyu Jia, Zhaoyang Yu, Yuanliang Chen, Wei You, Jianjun Huang, Bin Liang | 2026-09-20 | 下载 | Detecting RTL design bugs in open-source RISC-V CPU implementations is critical for ensuring system reliability. Traditional detection approaches inherently rely on predefined artifacts. |
| WaveletECO: A Closed-Loop Physical ECO Platform and a Specialized Local Language Model | Guoxiang Xu, Guozhen Ji, Zijian Luo, Zhengrui Chen, Qi Sun, Cheng Zhuo | 2026-09-20 | 下载 | Engineering change order (ECO) is an important step in repairing timing and electrical violations during the late stages of chip design. Existing Agentic EDA methods primarily focus on tool invocation... |
cs.DC - Distributed, Parallel, and Cluster Computing
| 标题 | 作者 | 发布日期 | 摘要 | |
|---|---|---|---|---|
| Adaptive Determinantal Client Scheduling in Federated Learning | Wen Xu, Ben Liang, Gary Boudreau, Hamza Sokun | 2026-09-20 | 下载 | Scheduling clients for model training is critical in federated learning due to both data and system heterogeneity. Most previous works focus on the quality of the scheduled clients to achieve faster c... |
| SAGE: Optimal-Stopping Peer Selection for Decentralised Federated Learning | Ke Xiao, Qiyuan Wang, Christos Anagnostopoulos | 2026-09-20 | 下载 | Decentralised federated learning replaces server aggregation with peer-to-peer model exchange, making collaborator selection a local decision under uncertainty. |
| TriFleetRCA: On-Premise LLM Root Cause Analysis for Kubernetes | Rohit Patel, Susil Kumar Mohanty, Jeenal Chaudhary | 2026-09-20 | 下载 | Root cause analysis at a remote site is slow: evidence is scattered across pod logs, Kubernetes events and cluster-level objects, and many operators cannot send production logs to a hosted model at al... |
| Explicit State and Resource Contracts for Low-Precision Pipeline Parallel Training under Captured Graphs | Genlang Chen, Junyi Zhu | 2026-09-20 | 下载 | CUDA Graphs eliminate launch overheads by replaying tensor operations over static virtual addresses. However, FP8 pipeline training continuously alters the scaling states, microbatches, and deferred b... |
| Conflicting Pattern Formation by Teams of Anonymous, Fully Disoriented Robots | Animesh Maiti, Prakhar Shukla, Subhash Bhagat | 2026-09-20 | 下载 | Two groups of autonomous, anonymous, and oblivious mobile robots are deployed in the two-dimensional Euclidean plane, each assigned a distinct task. |
| Economical and efficient big data sharing with i-Cloud | Thepparit Banditwattanawong, Masawee Masdisornchote, Putchong Uthayopas | 2026-09-20 | 下载 | Big data can be hosted on cloud and being shared distributedly through cloud services in an unprecedented volume, variety and velocity. This causes not only cloud network congestions and delayed cloud... |
| Co-occurrence Patterns of LoRA Adapters in Production Diffusion Model Inference Services | Tao Zhang, Bin Liao, Tao Zhou, Yanping Liu | 2026-09-20 | 下载 | Low-rank adaptation (LoRA) has become a key technology for serving large-scale personalized large language models and diffusion models in the cloud. |
| Accurate Distributed Tracing for Large-Scale AI Infrastructure: Time Synchronization as a Foundation for Reliable Observability | Hesham Elbakoury, Ankur Sharma | 2026-09-20 | 下载 | Distributed tracing in large-scale AI infrastructure fails silently when clock accuracy is insufficient: causal events are misordered, fault attribution is corrupted, and performance diagnoses are unr... |
| Accurate Simulation of Distributed Training Jobs with Network Contention Modeling | Yeonho Yoo, Hyunho Lee, Hyunmok Choi, Chuck Yoo, Gyeongsik Yang | 2026-09-20 | 下载 | Trace-driven simulation is widely used to evaluate distributed training (DT) jobs in GPU clusters, but existing simulators either ignore network contention or approximate it with a fixed penalty. |
cs.NI - Networking and Internet Architecture
| 标题 | 作者 | 发布日期 | 摘要 | |
|---|---|---|---|---|
| Remote IoT Source Monitoring with Delayed Feedback | Andrea Munari | 2026-09-20 | 下载 | Remote source monitoring is a key use case for Internet of things (IoT) applications, calling on efficient communications protocols to ensure timely data delivery. |
| SemDHT: Certified Semantic Discovery for Peer-to-Peer Agent Networks over Exact-Key DHTs | Taotao Wang, Chonghe Zhao, Shengli Zhang, Soung Chang Liew | 2026-09-20 | 下载 | Agents may need capabilities exposed through external agent endpoints or service APIs. When a requester is not already bound to a provider, it must discover advertised capabilities matching its task a... |
| Feature Suppression and Differential Privacy for Residential Traffic Classification: A Two-Home Federated Study | Márton Pál Lipcsey-Magyar, Adrian Pekar | 2026-09-20 | 下载 | Residential traffic classification supports service management, but learning across homes must account for heterogeneous traffic and privacy constraints. |
| Protocol-Flexible Custom NFC for Wire-Free Wearable Sensor Networks | Riku Maeda, Akihito Noda | 2026-09-20 | 下载 | This study presents a custom near-field communication (NFC) system with software-defined protocol implementation for wearable sensors distributed across the body. |
| Accurate Simulation of Distributed Training Jobs with Network Contention Modeling | Yeonho Yoo, Hyunho Lee, Hyunmok Choi, Chuck Yoo, Gyeongsik Yang | 2026-09-20 | 下载 | Trace-driven simulation is widely used to evaluate distributed training (DT) jobs in GPU clusters, but existing simulators either ignore network contention or approximate it with a fixed penalty. |
cs.OS - Operating Systems
| 标题 | 作者 | 发布日期 | 摘要 | |
|---|---|---|---|---|
| When the Agent Becomes the Kernel: A Systematization of Security on the Path to AI-Native Operating Systems | Li Zhang, Yang Sun, Jie Shi | 2026-09-20 | 下载 | Large language model agents are now privileged principals that take consequential actions: editing code repositories, operating inboxes, completing purchases. |
cs.PF - Performance
| 标题 | 作者 | 发布日期 | 摘要 | |
|---|---|---|---|---|
| Total Cost of Agency: Exact Attribution of Memory Injection Cost in Multi-Agent LLM Workflows | Vivek Kumar Singh, Preeti Priyam, Gautam Bhowmick | 2026-09-20 | 下载 | Every node in a multi-agent large language model (LLM) workflow retrieves context from memory and injects it into its prompt, where those injected tokens are billed as input tokens at the same per-tok... |