2026-06-29
cs.AR - Architecture
| 标题 | 作者 | 发布日期 | 摘要 | |
|---|---|---|---|---|
| AgRefactor: Self-Evolving Agentic Workflow for HLS Compatibility and Performance | Yang Zou, Zijian Ding, Yizhou Sun, Jason Cong | 2026-06-29 | 下载 | High-Level Synthesis (HLS) provides a fast path from concepts to silicon, but converting real-world software into synthesizable HLS code remains challenging due to restrictive language support and the... |
| SpikON: A Dual-Parallel and Efficient Accelerator for Online Spiking Neural Networks Learning | Peilin Chen, Xiaoxuan Yang | 2026-06-29 | 下载 | Spiking neural networks (SNNs) have emerged as a promising paradigm for energy-efficient brain-inspired computing. However, existing online unsupervised SNN learning suffers from low training accuracy... |
| CryoZip: An Efficient Cryogenic Compressor for Quantum Error Correction Syndromes | Guanchen Tao, Alexander Knapen, Jacob Mack, Gokul Subramanian Ravi, Qirui Zhang, Mehdi Saligane, Dennis Sylvester | 2026-06-29 | 下载 | Scaling fault tolerant quantum computing is increasingly constrained by the limited bandwidth and power budget across the 4 K to room temperature (RT) interface. |
| COSM: A Cooperative Scheduling Framework for Concurrent PIM and CPU Execution on Mobile Devices | Yilong Zhao, Fangxin Liu, Onur Mutlu, Mingyu Gao, Jian Liu, Haibing Guan, Li Jiang | 2026-06-29 | 下载 | The development of on-device large language models (LLMs) is driven by the need for privacy and fast response times. Energy-intensive data transfer on mobile devices makes Processing-in-Memory (PIM) a... |
| Model Predictive Current Control with Harmonic Correction for Single-Phase AC-DC EV Charging | Changhong Li, Bharathkumar Hegde, Biswajit Basu, Shreejith Shanker | 2026-06-29 | 下载 | The increasing integration of Electric Vehicles (EVs) has imposed a growing harmonic challenge on the power grid. For AC/DC Power Factor Correction (PFC) in single-phase On-Board Chargers (OBCs), Mode... |
| RQP: Resource-Oriented Quantiser Pruning for Neural Networks on FPGAs | Changhong Li, Biswajit Basu, Shreejith Shanker | 2026-06-29 | 下载 | High granularity quantisation (HGQ) exploits weight-level quantisation and pruning to design resource-efficient neural network accelerators, achieving an attractive trade-off between accuracy and hard... |
| Mega: A 22 nm Convolutional Spiking Neural Network Accelerator Achieving 0.375 pJ/SOP for Efficient Edge Vision | Rick Luiken, Manil Dev Gomony, Sander Stuijk | 2026-06-29 | 下载 | Convolutional Spiking Neural Networks (SNN) offer the potential for highly energy-efficient vision processing by exploiting sparse, event-driven computation. |
| HBM Is Not All You Need: Efficient Disaggregated LLM Serving across Memory-heterogeneous Accelerators | Zhixiang Wei, Yun Wang, James Yen, Mingyuan Xia, Zhengwei Qi | 2026-06-29 | 下载 | LLM inference comprises a compute-bound prefill phase and a memory-bound decode phase, and recent systems disaggregate them onto separate hardware. |
cs.DC - Distributed, Parallel, and Cluster Computing
| 标题 | 作者 | 发布日期 | 摘要 | |
|---|---|---|---|---|
| Towards Transparent Checkpointing with AI-driven Code Generation | Hai Duc Nguyen, Tekin Bicer, Kyle Chard, Ian Foster, Bogdan Nicolae | 2026-06-29 | 下载 | Adding reliable checkpoint/restart support to an MPI scientific application is a time-consuming expert effort that requires deep knowledge of both the application and resilience. |
| Budget-Adaptive Routing: Skipping the Weak When the Strong Answers Anyway | Wei Geng, Nitinder Mohan, Jörg Ott | 2026-06-29 | 下载 | Edge-cloud inference collaborations are often designed with a routing estimator that decides whether to offload each frame from weak models at the edge to stronger models in the cloud. |
| StreamGuard: Low-Overhead Resilience for Real-time HPC Data Streams | Hai Duc Nguyen, Bogdan Nicolae, Tekin Bicer, Amal Gueroudji, Matthieu Dorier, Kyle Chard, Ian Foster | 2026-06-29 | 下载 | Real-time scientific workflows operate on continuous data streams and must produce timely, high-quality results despite executing on complex, failure-prone infrastructure. |
| Protecting Futures against Silent Data Corruption -- Efficient Task Replication for Dynamic Data Dependencies | Rüdiger Nather, Claudia Fohry, Mia Reitz | 2026-06-29 | 下载 | As the size of computational problems grows, so does the likelihood of Silent Data Corruptions (SDCs). A common defense is replication, where the computation is repeated and correct results are determ... |
| Data Replication Meets Function Scheduling in the Edge-Cloud Continuum | Matteo Cenzato, Dario d'Abate, Arianna Dragoni, Matteo Briscini, Alessandro Margara | 2026-06-29 | 下载 | Serverless computing is an appealing model for the edge-cloud continuum, but its stateless assumption breaks down once functions need persistent data: fetching state from a distant cloud store erases ... |
| SubEdge: A Subscriber-Centric Edge Computing Subsystem in 6G Networks for AI | Abdirazak Ali Asir Rage, Riccardo Pozza, Rahim Tafazolli | 2026-06-29 | 下载 | Beyond traditional connectivity, 6G is envisioned to transform mobile networks into a distributed fabric that provides native integrated communication, computing, and intelligence services. |
| COSM: A Cooperative Scheduling Framework for Concurrent PIM and CPU Execution on Mobile Devices | Yilong Zhao, Fangxin Liu, Onur Mutlu, Mingyu Gao, Jian Liu, Haibing Guan, Li Jiang | 2026-06-29 | 下载 | The development of on-device large language models (LLMs) is driven by the need for privacy and fast response times. Energy-intensive data transfer on mobile devices makes Processing-in-Memory (PIM) a... |
| Spandana: Reconciling Strict SLOs with Low Cost under Fine-Grained Load Fluctuations | Dilina Dehigama, Shyam Jesalpura, Zeyu Xu, Marton Nemeth, Shengda Zhu, Marios Kogias, Boris Grot | 2026-06-29 | 下载 | Cloud-based online services face significant sub-second load fluctuations while needing to meet strict Service Level Objectives (SLOs). Cluster operators often over-provision resources to protect SLOs... |
| GPU Parallelization Strategies for Forward and Backward Propagation in Shallow Neural Networks: A CUDA-Based Comparative Study | Rania Zitouni, Nadine Bousdjira, Sarah Hasnaoui, Amel Sadoun, Fatma Salhi | 2026-06-29 | 下载 | We present a comparative study of CUDA optimization strategies applied to forward and backward propagation in a shallow neural network. Three stacked optimizations are evaluated: (1) tiled shared memo... |
| HSAP: A Hierarchical Sequence-aware Parallelism for Hybrid-Context Generative Models | Songxin Zhang, Zejian Xie, Zhuoyang Song, Cong lin, Junyu Lu, Jiaxing Zhang, Bingyi Jing | 2026-06-29 | 下载 | In this paper, we aim to combine the advantages of existing sequence parallelism paradigms and overcomes their drawbacks, the most serious of which is the incapability to correctly compute causal atte... |
| Analyzing Linearizability in Relativistic Distributed Systems | Kahbod Aeini, Wojciech Golab | 2026-06-29 | 下载 | Einstein's theory of relativity correctly predicted that time is relative, and subject to both kinematic and gravitational dilation. Therefore, executions of distributed systems cannot always be model... |
| Energy-Aware Scheduling for Serverless LLM Serving on Shared GPUs | Tianyu Wang, Gourav Rattihalli, Aditya Dhakal, Longfei Shangguan, Dejan Milojicic | 2026-06-29 | 下载 | As LLM inference becomes a major cloud workload, its growing energy footprint makes cluster-wide energy optimization increasingly important. Serverless LLM serving helps platforms absorb traffic volat... |
| FBench: A Flexible Benchmark for CFG-Based What-If Exploration of HPC I/O Patterns | Zhaobin Zhu, Chen Wang, Kathryn Mohror, Sarah Neuwirth | 2026-06-29 | 下载 | The I/O performance of large-scale HPC applications depends on a complex interplay of access patterns, middleware optimizations, and file system configurations. |
| HBM Is Not All You Need: Efficient Disaggregated LLM Serving across Memory-heterogeneous Accelerators | Zhixiang Wei, Yun Wang, James Yen, Mingyuan Xia, Zhengwei Qi | 2026-06-29 | 下载 | LLM inference comprises a compute-bound prefill phase and a memory-bound decode phase, and recent systems disaggregate them onto separate hardware. |
| Beyond Uniform Experts: Cost-Aware Expert Execution for Efficient Multi-Device MoE Inference | Hui Zang, Pengfei Xia, Hong Liu, Jiajia Chu, Tuo Hao, Minghao Chen, Rui Zhang, Ziyang Zhang | 2026-06-29 | 下载 | Mixture-of-Experts (MoE) architectures enable language models to achieve unprecedented scale via sparse activation. However, their inference performance is often limited by data movement bottlenecks. |
| Rethinking Collaborative Trust for Verifiably Decentralized Blockchain Systems | Yunqi Zhang, Shaileshh Bojja Venkatakrishnan | 2026-06-29 | 下载 | Despite the promise of decentralization, measurement studies have identified a conspicuous lack of decentralization in blockchains. Centralization has been observed in almost all layers of the blockch... |
| SMART-MIG: A Learning Framework for Scalable and Energy-Efficient GPU Scheduling | Wenqing Yu, Neel Karia, Tanvi Hisaria, Clifford Stein, Olivier Tardieu, Asser Tantawi | 2026-06-29 | 下载 | The emergence of Multi-Instance GPU (MIG) technology enables us to run smaller machine learning models on partitions of a GPU rather than the entire device, thus improving utilization and reducing ene... |
| Demystifying the Design Space and Best Practices for Heterogeneous LLM Inference and Serving | Zhixin Wang, Zhengbo Wang, Fangcheng Fu, Yinhui Lu, Jinlong Hou, Yijie Chen, Xiaowei Shen, He Liu, Xiangbin Li, Jun Chen, Ruya Gu, Dian Wang, Zhou Tan, Yuan Cheng, Hongzhou Zhang, Xiangjun Huang, Ping Zhang, Xiaohe Hu | 2026-06-29 | 下载 | Heterogeneous prefill-decode (PD) inference is now in production: prefill on cost-efficient or supply-available accelerators, decode on bandwidth-strong ones, and KV state crossing mixed interconnects... |
cs.NI - Networking and Internet Architecture
| 标题 | 作者 | 发布日期 | 摘要 | |
|---|---|---|---|---|
| Budget-Adaptive Routing: Skipping the Weak When the Strong Answers Anyway | Wei Geng, Nitinder Mohan, Jörg Ott | 2026-06-29 | 下载 | Edge-cloud inference collaborations are often designed with a routing estimator that decides whether to offload each frame from weak models at the edge to stronger models in the cloud. |
| A Practical Implementation of Day-3 Cooperative Intersection with Automated Connected Mini-Cars | Lorenzo Farina, Vittorio Todisco, Federico Gavioli, Salvatore Iandolo, Francesco Moretti, Giuseppe Perrone, Matteo Piccoli, Francesco Raviglione, Marco Rapelli, Antonio Solida, Claudio Casetti, Paolo Burgio, Carlo Augusto Grazia, Alessandro Bazzi | 2026-06-29 | 下载 | Cooperative driving enabled by connected and automated vehicles is expected to improve traffic efficiency and safety, particularly at intersections where traditional control mechanisms such as traffic... |
| CALO: Constraint-Aware Learning Optimization for Joint Resource Allocation in Double-Active RIS-Assisted Wireless Networks | Alaa S. Arabiyat, Mohammad J. Abdel-Rahman | 2026-06-29 | 下载 | Double-active reconfigurable intelligent surface (RIS)-assisted wireless systems can improve coverage and achievable rate in blockage-dominated environments. |
| When and Which Sensor to Observe? Timely Tracking of a Joint Markov Source | Ismail Cosandal, Sennur Ulukus, Nail Akar | 2026-06-29 | 下载 | We investigate the problem of remote estimation (at a monitor) of a discrete-time joint Markov process with individual components which can be observed with dedicated sensors. |
| Wireless Backdoor Attack and Defense for Semantic Communications over Multiple Access Channel | Yalin E. Sagduyu, Tugba Erpek, Aylin Yener, Sennur Ulukus | 2026-06-29 | 下载 | Semantic communication (SemCom) aims to preserve semantic meaning and task-oriented information beyond conventional message recovery over wireless channels. |
| SubEdge: A Subscriber-Centric Edge Computing Subsystem in 6G Networks for AI | Abdirazak Ali Asir Rage, Riccardo Pozza, Rahim Tafazolli | 2026-06-29 | 下载 | Beyond traditional connectivity, 6G is envisioned to transform mobile networks into a distributed fabric that provides native integrated communication, computing, and intelligence services. |
| COHORT: Collaborative Orchestration for Hardening via Offensive Replay on Emulated Topologies | Chen Frydman, Aviram Zilberman, Rubin Krief, Abed Showgan, Andres Murillo, Sekiya Motoyoshi, Asaf Shabtai, Yuval Elovici, Rami Puzis | 2026-06-29 | 下载 | Mitigating an observed adversary in an enterprise network typically takes weeks of expert work: an analyst derives a mitigation tailored to that adversary, validates it without breaking production, an... |
| LLMs and Optical Networks: A Symbiotic Relationship | Mëmëdhe Ibrahimi, Qiaolun Zhang, Giovanni S. Sticca, Jiaheng Xiong, Francesco Musumeci, Massimo Tornatore | 2026-06-29 | 下载 | This paper explores the emerging symbiosis between LLMs and optical networks. Massive LLMs require geo-distributed training, which demands advanced optical transport capabilities that require new key ... |
| Selective Deployment of Bidirectional Hollow-Core Fibers in Hybrid SMF/HCF Optical Networks | Mëmëdhe Ibrahimi, Giovanni S. Sticca, Angelo Ferrara, Massimo Tornatore | 2026-06-29 | 下载 | We investigate selectively deploying bidirectional transmission in hybrid Hollow-Core Fiber (HCF) networks. Upgrading 50% of links to bidirectional HCF yields at least a 40% throughput increase compar... |
| Scalable Intention Sharing for ETSI VAMs | Felipe E. Valle, Oscar Amador, Johan Thunberg, Elena Haller, Alexey Vinel | 2026-06-29 | 下载 | Efficient maneuver coordination in dense V2X environments requires accurate short-term prediction while maintaining low communication and computational overhead. |
| LEOSTP: A Spatio-Temporal Traffic Prediction Framework for LEO Satellite Networks | Shaoyou Ao, Yong Niu, Zhu Han, Cheng Li, Bo Ai | 2026-06-29 | 下载 | With the evolution of next-generation mobile communication networks and the commercial boom of Low Earth Orbit (LEO) satellites, globally covered satellite networks are gradually becoming a crucial in... |
cs.OS - Operating Systems
| 标题 | 作者 | 发布日期 | 摘要 | |
|---|---|---|---|---|
| LUMOS: A Semantic Operating-System Layer for Accessibility-Grounded AI Agents | Yogeswar Reddy Thota | 2026-06-29 | 下载 | Current operating systems expose interfaces optimized for human users but not for AI agents. Humans benefit from pixels, icons, windows, visual grouping, mouse movement, and keyboard shortcuts; AI age... |
cs.PF - Performance
| 标题 | 作者 | 发布日期 | 摘要 | |
|---|---|---|---|---|
| The Fourth-Root Complexity of Data Movement | Chen Ding | 2026-06-29 | 下载 | Time complexity typically assumes cost per data access. This paper presents an analysis based on an abstract memory hierarchy. For a common class of applications, it shows that the data-access ... |
| TraceLab: Characterizing Coding Agent Workloads for LLM Serving | Kan Zhu, Mathew Jacob, Chenxi Ma, Yi Pan, Stephanie Wang, Arvind Krishnamurthy, Baris Kasikci | 2026-06-29 | 下载 | Coding agents are rapidly becoming a major application of agentic LLMs, but serving them efficiently remains challenging. Progress on this challenge requires understanding real workload patterns, yet ... |
| FBench: A Flexible Benchmark for CFG-Based What-If Exploration of HPC I/O Patterns | Zhaobin Zhu, Chen Wang, Kathryn Mohror, Sarah Neuwirth | 2026-06-29 | 下载 | The I/O performance of large-scale HPC applications depends on a complex interplay of access patterns, middleware optimizations, and file system configurations. |