Skip to content

2026-07-31 ​

cs.AR - Architecture ​

标题作者发布日期PDF摘要
Analyzing RV32/RV64 Trade-offs for FreeRTOS Latency on 8-Stage RISC-V Soft ProcessorsHyunwoo Kang, Geonwoo Yu, Jongwon Kim, Seungwoo You, Minchan Gil2026-07-31下载Although many commercial RISC-V platforms provide real-time operating system support, practical examples that explain how to enable a preemptive RTOS on a custom bare-metal RISC-V soft processor remai...
Themis: Software-Defined Hardware PrefetchingKeisuke Kamahori, Neil Adit, Kan Zhu, Yuqi Mai, Victor Lee, Heiner Litz, Chris Kennelly, Snehasish Kumar, Hanna Alam, Milad Hashemi, David Li, Adrian Sampson, Baris Kasikci, Tipp Moseley, Parthasarathy Ranganathan, Akanksha Jain2026-07-31下载Data cache misses represent a significant portion of stall cycles in datacenter workloads. Hardware prefetchers that reduce such stalls by fetching data ahead of time have become increasingly sophisti...
Low-Power PLL-Based Clock Stabilization for Flexible IGZO AMS SystemsPaula Carolina Lozano Duarte, Georgios Zervakis, Mehdi Tahoori2026-07-31下载Flexible electronics (FE) platforms rely on analog and mixed-signal (AMS) circuits - biosensors, readout front-ends, and analog-to-digital converters - that dominate both functionality and energy cons...
RTLCurator: Label-Efficient Data Curation for RTL GenerationSiyang Cai, Cangyuan Li, Wenjing Chang, Kun Wang, Haoyu Gao, Yinhe Han, Ying Wang2026-07-31下载Training large language models (LLMs) to write register-transfer level (RTL) requires large corpora of paired specifications and code, and such data is scarce enough that most public corpora are now s...
Selective KV Cache Protection for Noise-Resilient LLM Inference on Analog Compute-In-Memory SystemsYuannuo Feng, Wenyong Zhou, Yuang Ma, Yizhe Chen, Wenshuai Yao, Yuxin Xie, Ngai Wong, Wang Kang2026-07-31下载Analog compute-in-memory (CIM) arrays have emerged as a promising substrate for energy-efficient LLM inference, particularly for weight-stationary computations in linear layers.
Adaptivity via a Parallel Architecture for Stochastic Gradient MethodsBin Fu2026-07-31下载We develop a parallel framework that assembles static gradient methods to achieve better adaptivity. A static gradient method, denoted by GD(x0,T)\mathrm{GD}(x_0,T), takes as input an initial point $x_0\in...

cs.DC - Distributed, Parallel, and Cluster Computing ​

标题作者发布日期PDF摘要
A Portable and Versatile Limited-Memory BFGS Implementation in PETSc/TAOHansol Suh, Tobin Isaac, Alp Dener, Todd Munson, Hong Zhang, Richard Tran Mills2026-07-31下载The limited-memory BFGS (L-BFGS) Hessian update scheme is the critical kernel in many quasi-Newton optimization algorithms. The most common approach to implementing L-BFGS uses 2m2m sequential rank-1 ...
TokTier: Exact Stateful CPU+GPU Tokenization for Agentic LLM ServingZhenyu Zhang, Zhichao Cao2026-07-31下载LLM serving stacks cache prompt KV state, yet the front end still re-tokenizes the full request text on every call. Coding agents pay the most: each call resubmits a long transcript after a small appe...
GQ-FSL: Green Quantized Federated Split LearningIdan Roth, Lutz Lampe2026-07-31下载Deploying state-of-the-art deep neural networks (DNNs) at the wireless edge is severely bottlenecked by the strict energy and resource constraints of mobile devices.
SLIM: Saturation-Aware Lightweight Performance Modeling for LLM ServingPol G. Recasens, Ferran Agullo, Yue Zhu, Chen Wang, Jordi Torres, Josep Ll. Berral2026-07-31下载Large language model (LLM) serving commonly increases batch size to improve throughput, but performance eventually reaches a deployment-dependent plateau beyond which larger batches provide marginal g...
System-Wide Termination in Distributed Betweenness Centrality ComputationSiamak Abdi, Lucia Cavallaro, Giuseppe Di Fatta2026-07-31下载Computing betweenness centrality on large networks is inherently expensive, as it requires aggregating shortest-path dependencies across all pairs of vertices and becomes increasingly difficult to sca...
CARA: Exact Local Repair with Fresh One-Action Certification for Cloud ConsolidationXiyang Zhang, Yuanhe Tian, Hongzhi Wang2026-07-31下载Simulator-based placement pipelines may inspect many repairs but deploy only when several reliability criteria improve together. Reusing search scenes to test the selected action invalidates nominal e...
Exploring Block Anomaly Detection In HDFS Log Data AnalysisWenYang Zhong, Tutut Herawan2026-07-31下载In recent years, with the development of big data technology, increasingly more companies use HDFS for data processing and storage. As a result, the maintenance of distributed file systems has become ...
Allocation Tracking and Parameter Checking for Parallel Programming Models using ContractsYussur Mustafa Oraji, Christian Bischof2026-07-31下载Correctness checking tools for High-Performance Computing programs are typically limited to specific parallel programming models such as MPI or OpenSHMEM.
Knox: Fortifying Smart Spaces With Safety GuaranteesRishabh Menezes, Jadon T. Schuler, Kaimeng Zhu, Oliver Rogalski, Indranil Gupta2026-07-31下载Internet of Things (IoT) devices in smart spaces and buildings are an emerging class of distributed systems with critical safety requirements.
Rethinking AI Cloud Infrastructure for Agentic Serving Systems with the Aries Experimentation FrameworkLeonid Kondrashov, Hongrui Liu, JooYoung Park, Boxi Zhou, Zonghao Liu, Chengzhi Lu, Riccardo Mancini, Esha Choukse, Haris Javaid, German Sviridov, Tao Peng, Chen Zhao, Anastasia Avdeeva, Aleksei Gusev, Marios Kogias, Luo Mai, Dmitrii Ustiugov2026-07-31下载Autonomous agents challenge conventional LLM serving by coupling repeated inference with persistent context and sandboxed tool execution. We present Aries, a full-stack experimentation framework that ...
Beyond Byzantine: An Organizational Consensus Algorithm for Self-Interested Agents Under Information AsymmetryJiawei Zhang, Jianbo Liu2026-07-31下载Traditional distributed consensus protocols classify nodes as either honest-but-faulty or actively malicious (Byzantine). However, in organizational structures, departmental agents rarely fit this bin...

cs.NI - Networking and Internet Architecture ​

标题作者发布日期PDF摘要
LightCal: Lightweight Optical-Pulse Bootstrap Calibration for Crystal-Free BLE RadiosCheng Wang, Titan Yuan, David Burnett, Filip Maksimovic, Kristofer S. J. Pister, Tengfei Chang2026-07-31下载Crystal-free Bluetooth Low Energy (BLE) radios remove the off-chip high-frequency crystal oscillator and can therefore reduce the cost, size, and integration complexity of Internet of Things (IoT) nod...
Learning the LoS Skyline from LEO Satellite Observations for Proactive HandoverMarius Corici, Manar Zaboub, Fabian Eichhorn, Hauke Buhr2026-07-31下载In non-terrestrial network deployments, local obstructions may block line-of-sight satellite links before the satellite reaches the geometric elevation mask, causing abrupt and unplanned handovers.
RIGEL: Real-time Optical Anomaly Diagnosis with Stateful In-Network Inference based on Distributed On-switch GNNsZhen Wei, Yidong Wang, Yufan Zhu, Xuefeng Yan, Binjun Tang, Xiaoliang Chen, Zuqing Zhu2026-07-31下载The recent booming of data-intensive applications has complicated optical network management, making real-time optical anomaly diagnosis a must-have feature.
METIS: A Declarative Slice Orchestrator for Application-Centric 5G/6G NetworksArman Divband, Ali Yaghoubian, Navid Nikaein2026-07-31下载Network slicing is the cornerstone of application-aware 5G and 6G networks, yet dynamic lifecycle management of network slice instances with coordinated quality-of-service enforcement across the radio...
Mind the Gap: Policy vs Reality in Post-Quantum TLS DeploymentNimesha Wickramasinghe, Frank Li, Sanjay Jha, Arash Shaghaghi2026-07-31下载Post-quantum cryptography (PQC) has evolved from a long-term planning concern into an operational priority. Following NIST's standardization of PQC, governments and standard bodies published transitio...
Point2Radio: A Foundation Model for Cross-Scene Radio Fields from Material-Aware Point CloudsChaozheng Wen, Chenghong Bian, Hongze Chen, Jun Zhang2026-07-31下载High-fidelity radio fields are typically simulated for every scene--transmitter configuration or fitted separately to each scene, failing to exploit propagation structures shared across environments.
Skillsets on the Chain: A Blockchain-based Zero-Trust Framework for Agentic AI NetworkingYayu Gao, Yong Xiao, Hao Hu, Xubo Li, Zhiwei Liu, Yingyu Li, Guangming Shi, Ping Zhang2026-07-31下载Agentic AI networking (AgentNet) systems rely heavily on third-party skillset implementations and distributed multi-agent collaboration, yet they face major claim-to-capability inconsistencies and sec...

cs.OS - Operating Systems ​

标题作者发布日期PDF摘要
Benchmarking LLMs on File System Design and Implementation: The Good, The Bad, and The UglyYuqi Xue, Daixuan Li, Jian Huang2026-07-31下载Large Language Models (LLMs) are fundamentally transforming computer system research and development. As we employ LLMs in file system (fs) development, it is essential to understand their capabilitie...
Themis: Software-Defined Hardware PrefetchingKeisuke Kamahori, Neil Adit, Kan Zhu, Yuqi Mai, Victor Lee, Heiner Litz, Chris Kennelly, Snehasish Kumar, Hanna Alam, Milad Hashemi, David Li, Adrian Sampson, Baris Kasikci, Tipp Moseley, Parthasarathy Ranganathan, Akanksha Jain2026-07-31下载Data cache misses represent a significant portion of stall cycles in datacenter workloads. Hardware prefetchers that reduce such stalls by fetching data ahead of time have become increasingly sophisti...

cs.PF - Performance ​

标题作者发布日期PDF摘要
Lossless Compression Performance for PETRA III DatasetsMalte Buschmann, Yannis Schumann, Christian Voss, Tigran Mkrtchyan, Philipp Neumann2026-07-31下载Large-scale research facilities increasingly face the challenge of managing rapidly growing data volumes while maintaining sustainable archival infrastructures.
TokTier: Exact Stateful CPU+GPU Tokenization for Agentic LLM ServingZhenyu Zhang, Zhichao Cao2026-07-31下载LLM serving stacks cache prompt KV state, yet the front end still re-tokenizes the full request text on every call. Coding agents pay the most: each call resubmits a long transcript after a small appe...
SLIM: Saturation-Aware Lightweight Performance Modeling for LLM ServingPol G. Recasens, Ferran Agullo, Yue Zhu, Chen Wang, Jordi Torres, Josep Ll. Berral2026-07-31下载Large language model (LLM) serving commonly increases batch size to improve throughput, but performance eventually reaches a deployment-dependent plateau beyond which larger batches provide marginal g...
Lightweight Neural Networks for Affordance Segmentation: Enhancement of the Decoder ModuleSimone Lugani, Edoardo Ragusa, Rodolfo Zunino, Paolo Gastaldo2026-07-31下载The deployment of deep neural networks for visual affordance segmentation on wearable robots poses may prove critical, due to some conflicting aspects of the problem.
Studying quantization trade-offs for efficient inference deployment in machine translationJim Zhao, Sohir Maskey, Koen Oostermeijer, Douglas Orr, Teryn Jones2026-07-31下载Deploying large language models in realistic server environments poses challenges, as the system needs to provide high-quality responses with low latency.

基于 VitePress 构建 · 使用本地搜索查找论文