Skip to content

2026-08-15 ​

cs.AR - Architecture ​

标题作者发布日期PDF摘要
TIDE: An FPGA quantum-control processor for deterministic adaptive execution with guarded runtime program revisionXiaoqin Luo, Jiayun Song, Xiaolu Su2026-08-15下载Measurement-responsive quantum experiments require control programs that can revise future operations after execution has begun without disturbing events already committed to precise timing.
NPU Offloading of a Frozen Visual Encoder for Robot Policy TrainingHyojun Yun, Seungjae Won, Hyungpil Moon2026-08-15下载When a robot policy is trained for a new task or dataset, its visual encoder can be frozen and only its action generation module trained, reducing training cost.

cs.DC - Distributed, Parallel, and Cluster Computing ​

标题作者发布日期PDF摘要
MM-BEV: Enhancing Timeliness by Computing Where and When it MattersLiangkai Liu, Kang G. Shin2026-08-15下载Multimodal bird's-eye-view (BEV) perception combines LiDAR depth accuracy with dense camera semantics, but its high computational cost and imperfect sensing conditions make real-time deployment challe...
FloodReasonBench: Benchmarking VLM Reasoning Segmentation for Embodied Flood Response at the EdgeRajat Bhattacharjya, Yoomee Jung, Minwoo Kim, Sing-Yao Wu, Eli Bozorgzadeh, Nalini Venkatasubramanian, Nikil Dutt2026-08-15下载Reasoning segmentation enables vision-language models (VLMs) to translate mission-relevant language requests into pixel-level visual grounding, offering a natural perception interface for embodied age...
LOCAL: Enabling Learning On-device Contiguously for Agent LLMsXinxin Liu, Jiaxin Li, Zibo Wang, Yun Ji, Zhangqi Zhu, Qing Hu, Zhibin Wang, Rong Gu, Sheng Zhong, Chen Tian2026-08-15下载On-device LLM agents interact repeatedly with users on local hardware, producing private traces that are valuable for adaptation but should not be sent to a remote trainer.
TERRA: A Hierarchical Parallel Training and Memory Orchestration Framework for High-Resolution AI-based Earth ModelingRuohan Wu, Ziqi Zhu, Yang Zhao, Jiarui Tang, Yingzhe Cui, Junshi Chen, Zhao Jing, Jun Shi, Hong An2026-08-15下载Training high-resolution AI-based Earth forecasting models is memory-intensive. Window-based Swin Transformers reduce the quadratic cost of global attention, but existing distributed systems such as A...
P-PAS: Prefill-Pressure Adaptive Scheduling for Long-Context LLM ServingTimo Sämann2026-08-15下载Long-context LLM applications such as retrieval-augmented generation (RAG) and agentic systems often process tens of thousands of input tokens to produce short outputs, making end-to-end request laten...
From LLM Inference to Agentic Workloads: Characterization and Implications for Serving SystemsChaokun Chang, Yukun Zhou, Kaihua Fu, Dakai An, Tianyu Feng, Hanfeng Lu, Sheng Yao, Pu Guo, Yinghao Yu, Yizhou Shan, Bo Li, Binhang Yuan, Wei Wang2026-08-15下载Agentic applications are shifting AI serving from isolated model inference to long-running workloads in which LLMs coordinate tools, environments, and persistent state.
Collective Communication for Distributed LLM Systems: Planning, Runtime Adaptation, and Computation CoordinationXuebin Song, Menghao Zhang, Yuezheng Liu, Jinyi Xia, Shucan Yang, Xiaohe Hu, Chunming Hu, Mingwei Xu2026-08-15下载Distributed large language model (LLM) systems increasingly rely on collective communication primitives such as AllReduce (AR), ReduceScatter (RS), AllGather (AG), and AlltoAll (A2A).
Anatomy of a Quantized Agent: VRAM Stability and Forecasting in Code-Synthesis Agentic WorkloadsAnubhab Banerjee2026-08-15下载Analytical models of peak VRAM consumption for LLM inference decompose memory into weight-storage, KV-cache, and activation terms parameterized by step count, tool invocations, and context expansion.
PAS-QFL: Personalized Ansatz Selection for Quantum Federated Learning under Client Data HeterogeneityJindi Wu, Qun Li2026-08-15下载Quantum federated learning (QFL) lets multiple quantum clients collaboratively train quantum neural networks (QNNs) without sharing private local data.
When Does Distributed AI Inference Need More Wide-Area Bandwidth? A Co-Design Evaluation of Optical, Packet, and Software LeversPrasanna C2026-08-15下载Wide-area bandwidth per unit of GPU compute falls every hardware generation: in compute-intensity-ratio terms (CIR, bytes per FLOP), the gap between on-package memory and the conventional WAN is four ...

cs.NI - Networking and Internet Architecture ​

标题作者发布日期PDF摘要
OTel: Building Domain-Specialized Telecom LLM Foundations for Intelligent NetworksFarbod Tavakkoli, Roderic Paulk, Jorden Terrazas, Kenneth Church, Mark Austin, Louis Powell, Gregory Diamos, Lina Bariah, Syed Ali Raza Zaidi, Maryam Hafeez, Ali Maatouk, Imtiaz Karim2026-08-15下载Frontier AI models have advanced rapidly, but they still struggle with telecom-specific tasks. We present Open Telco (OTel), an open telecom AI resource with derived datasets for retrieval, reranking,...
Adaptive Bridge: A Proxy-Based Decoupling Layer for Mitigating DDS Backpressure in ROS 2Kaushalraj Puwar2026-08-15下载In ROS 2 systems using DDS, a single slow subscriber on a RELIABLE topic can cause backpressure that degrades throughput and latency for all subscribers sharing the same publisher, including safety-cr...
Exploring the Suitability of QUIC for the Internet of ThingsCarles Gomez, Nika Soltani-Tehrani, Jon Crowcroft2026-08-15下载QUIC is an emerging transport-layer protocol that provides reliability and security. QUIC was designed to overcome issues from other protocol stacks used in the Internet, such as TCP/TLS, especially f...
ISAC in 3GPP: Evolution Toward 6GNeeraj Varshney2026-08-15下载Integrated sensing and communication (ISAC) is emerging as an important direction in the Third Generation Partnership Project (3GPP) evolution toward 6G because it allows cellular networks to provide ...
Diffused-Beam Laser-Diode LiFi Under Realizable Receiver, Noise, and Safety Constraints: Design-Space Analysis and an Open Cross-Verified Simulation FrameworkHussain Ahmad, Saleem Aslam, Ammara Nasim, Syed Muhammad Talha Gillani, Toheed Omer2026-08-15下载Link-budget studies of indoor optical wireless systems frequently assume receiver parameter sets--large photodetector area, large transimpedance, and wide bandwidth simultaneously--that violate basic ...
ICL-SEC: Iterative Cross-Layer Semantic Error CorrectionYirun Wang, Soung Chang Liew, Yuyang Du2026-08-15下载Iterative decoding has been central to the success of modern channel coding, where reliability information is repeatedly exchanged across decoding components to approach fundamental performance limits...
Agentic AI-Enabled Solar-Powered High-Altitude Platforms for Sustainable SAGINsHaoxiang Luo, Bang Huang, Mohamed-Slim Alouini2026-08-15下载Space-Air-Ground Integrated Networks (SAGINs) can extend connectivity, but their communication, computing, and platform operations create tightly coupled energy demands.
When Does Distributed AI Inference Need More Wide-Area Bandwidth? A Co-Design Evaluation of Optical, Packet, and Software LeversPrasanna C2026-08-15下载Wide-area bandwidth per unit of GPU compute falls every hardware generation: in compute-intensity-ratio terms (CIR, bytes per FLOP), the gap between on-package memory and the conventional WAN is four ...

cs.OS - Operating Systems ​

标题作者发布日期PDF摘要
From LLM Inference to Agentic Workloads: Characterization and Implications for Serving SystemsChaokun Chang, Yukun Zhou, Kaihua Fu, Dakai An, Tianyu Feng, Hanfeng Lu, Sheng Yao, Pu Guo, Yinghao Yu, Yizhou Shan, Bo Li, Binhang Yuan, Wei Wang2026-08-15下载Agentic applications are shifting AI serving from isolated model inference to long-running workloads in which LLMs coordinate tools, environments, and persistent state.

cs.PF - Performance ​

标题作者发布日期PDF摘要
Diffused-Beam Laser-Diode LiFi Under Realizable Receiver, Noise, and Safety Constraints: Design-Space Analysis and an Open Cross-Verified Simulation FrameworkHussain Ahmad, Saleem Aslam, Ammara Nasim, Syed Muhammad Talha Gillani, Toheed Omer2026-08-15下载Link-budget studies of indoor optical wireless systems frequently assume receiver parameter sets--large photodetector area, large transimpedance, and wide bandwidth simultaneously--that violate basic ...
T-LLM Compiler: Trusted LLM-based Code Optimization and Verification FrameworkZahra Fazel, Sunanda Gamage, Shayan Shirahmad Gale Bagi, Amir H. Ashouri, Tomasz S. Czajkowski, Bryan Chan, Reza Azimi, Yaoqing Gao2026-08-15下载Recent advances in Large Language Models (LLMs) have opened opportunities to apply high-level code transformations to the field of code optimization, and it has since emerged as one of the most fundam...

基于 VitePress 构建 · 使用本地搜索查找论文