Skip to content

2026-06-14 ​

cs.AR - Architecture ​

标题作者发布日期PDF摘要
Google's Training Supercomputers from TPU v2 to Ironwood: Architectural Stability, Scale, Resilience, Power Efficiency, and Sustainability Across Five GenerationsNorman P. Jouppi, Sridhar Lakshmanamurthy, Cliff Young, David Patterson2026-06-14下载This paper (to appear in the July/August 2026 issue of IEEE Micro magazine) summarizes five generations of Google s TPUs, from TPU v2 to Ironwood, highlighting their evolution as scalable, resilient, ...
EPIC: A System Framework for Efficient Egocentric Perception on Embodied AR GlassesTianhua Xia, Haiyu Wang, Jiajing Zheng, Su Chen, Sai Qian Zhang2026-06-14下载Modern smart AR glasses are evolving into intelligent systems that support foundation model-based assistance through continuous perception of the user and surrounding environment.
Approaching Shannon Bound with Lossless LLM Weight CompressionHongshi Tan, Yao Chen, Gustavo Alonso, Weng-Fai Wong, Bingsheng He2026-06-14下载Large language models (LLMs) now scale to trillions of parameters, driving weight storage into the terabyte regime and creating an acute mismatch with GPU memory capacity.

cs.DC - Distributed, Parallel, and Cluster Computing ​

标题作者发布日期PDF摘要
Raiders of the Lost Log: Synchronous Parallel In-Place Models and AlgorithmsMichael T. Goodrich, Vinesh Sridhar2026-06-14下载Embedded systems and Internet of Things (IoT) applications motivate in-place parallel algorithms, which avoid allocating additional shared memory past the input.
PreLort: Prefix-Nested LoRA for Federated Fine-Tuning under Rank HeterogeneityMuhammad Waseem, Nurbek Tastan, Andrej Jovanovic, Nicholas D. Lane, Nils Lukas, Karthik Nandakumar, Samuel Horvath2026-06-14下载Federated fine-tuning of large language models using parameter-efficient methods such as LoRA enables privacy-preserving adaptation of foundation models.
Quantifying the Impact of Lossy Compression on Neural Generative Surrogate ModelingZhimin Li, Harshitha Menon, Charles Jekel, Valerio Pascucci, Peter Lindstrom2026-06-14下载Neural networks are used as generative surrogate models for scientific discovery, which are trainable approximations of scientific simulations.
Green SARC: Predictive Cost and Carbon Governance for Agentic AI SystemsGaston Besanson2026-06-14下载Agentic AI systems act through tools and sub-agents, yet the controls meant to bound their financial and environmental cost still sit on dashboards evaluated beside or after execution.
SDVDiag: Multimodal Causal Discovery for Online Diagnosis in Software-defined VehiclesMatthias Weiß, Athreya Hosahalli Prakash, Falk Dettinger, Nasser Jazdi, Michael Weyrich2026-06-14下载The transition toward software-defined vehicles concentrates an increasing share of vehicle functionality into distributed software services, where failures propagate through service dependencies and ...

cs.NI - Networking and Internet Architecture ​

标题作者发布日期PDF摘要
Hidden Degradation Costs in Energy-Cost-Only HEMS Optimisation: Study on Battery and PV SensitivityDawood Butt, Nandor Verba2026-06-14下载Residential battery energy storage systems (BESS) are increasingly deployed alongside photovoltaic (PV) generation to reduce household energy costs under volatile time-of-use (TOU) tariffs.
SINR-Aware Base Station Deployment in Wide Area IoT Sensor NetworksSachin Kadam2026-06-14下载The rapid expansion of Internet of Things (IoT) applications necessitates the effective deployment of base stations (BSs) to enable consistent connectivity across large geographic areas under interfer...
Conflict-Aware Federated Fine-Tuning of Large Language Models with Mixture-of-ExpertsYijun Lu, Zihan Fang, Pengpeng Qiao, Zheng Lin, Jing Yang, Yuxin Zhang, Por Lip Yee, Zhe Chen, Jun Luo2026-06-14下载The continuous scaling of large language models (LLMs) incurs prohibitive computational costs, making Mixture-of-Experts (MoE) a scalable alternative for efficient fine-tuning via sparse activation.

cs.PF - Performance ​

标题作者发布日期PDF摘要
MADAR: An Address-Free ProcessorMohamed Amine Bergach2026-06-14下载In a modern processor, computing is the cheap part. Most of its area and energy go to \emph{addressing} -- moving operands to and from a register file and cache, and running the tags, ports, miss queu...

基于 VitePress 构建 · 使用本地搜索查找论文