2026-04-30
cs.AR - Architecture
| 标题 | 作者 | 发布日期 | 摘要 | |
|---|---|---|---|---|
| DPU or GPU for Accelerating Neural Networks Inference -- Why not both? Split CNN Inference | Ali Emre Oztas, Mahir Demir, James Garside, Mikel Luj'an | 2026-04-30 | 下载 | Video and image streaming on edge devices requires low latency. To address this, Neural Networks (NNs) are widely used, and prior work mainly focuses on accelerating them with single hardware units su... |
| I hope we don't do to trust what advertising has done to love | Jade Alglave | 2026-04-30 | 下载 | Advertising uses love to sell stuff, like nylons. It also uses the word "love" in trivialising ways -- do you "love" your oven? When I hear about trust in the context of AI, especially agentic, I hope... |
| NeuroRing: Scaling Spiking Neural Networks via Multi-FPGA Bidirectional Ring Topologies and Stream-Dataflow Architectures | Muhammad Ihsan Al Hafiz, Artur Podobas | 2026-04-30 | 下载 | Spiking neural networks (SNNs) are a promising paradigm for energy-efficient event-driven computation, but large-scale SNN execution remains challenging because sparse spike communication and synchron... |
| Affinity Tailor: Dynamic Locality-Aware Scheduling at Scale | Jin Xin Ng, Ori Livneh, Richard O'Grady, Josh Don, Peng Ding, Samuel Grossman, Luis Otero, Chris Kennelly, David Lo, Carlos Villavieja | 2026-04-30 | 下载 | Modern large multicore systems often run multiple workloads that share CPUs under schedulers such as Linux CFS. To keep CPUs busy, these schedulers load-balance runnable work, causing each workload to... |
| AME-PIM: Can Memory be Your Next Tensor Accelerator? | Emanuele Venieri, Simone Manoni, Alberto Florian, Jaehyun Park, Kyomin Sohn, Andrea Bartolini | 2026-04-30 | 下载 | High Bandwidth Memory with Processing-in-Memory (HBM-PIM) offers an opportunity to reduce data movement by executing computation directly inside memory, but current commercial platforms expose limited... |
| RuC: HDL-Agnostic Rule Completion Benchmark Generation | Arnau Ayguadé Domingo, Miquel Alberti-Binimelis, Cristian Gutierrez-Gomez, Emanuele Parisi, Razine Moundir Ghorab, Miquel Moreto, Gokcen Kestor, Dario Garcia-Gasulla | 2026-04-30 | 下载 | Large Language Models (LLMs) have rapidly improved in performance across code-related tasks, making their integration into Register Transfer Level (RTL) development increasingly attractive. |
| HAVEN: Hybrid Automated Verification ENgine for UVM Testbench Synthesis with LLMs | Chang-Chih Meng, Yu-Ren Lu, Guan-Yu Lin, Tsung Tai Yeh, Kai-Chiang Wu, I-Chen Wu | 2026-04-30 | 下载 | Integrated Circuit (IC) verification consumes nearly 70% of the IC development cycle, and recent research leverages Large Language Models (LLMs) to automatically generate testbenches and reduce verifi... |
| CuLifter: Lifting GPU Binaries to Typed IR | Jisheng Zhao, Huanzhi Pu, Shinnung Jeong, Chihyo Ahn, Hyesoon Kim | 2026-04-30 | 下载 | GPU compilers merge all data types into a single unified register file, erasing the type information that binary-analysis tools rely on. We show that type recovery from this untyped register file is t... |
| VitaLLM: A Versatile, Ultra-Compact Ternary LLM Accelerator with Dependency-Aware Scheduling | Zi-Wei Lin, Tian-Sheuan Chang | 2026-04-30 | 下载 | Deploying Large Language Models (LLMs) on resource-constrained edge devices faces critical bottlenecks in memory bandwidth and power consumption. While ternary quantization (e.g., BitNet b1. |
| RCW-CIM: A Digital CIM-based LLM Accelerator with Read-Compute/Write | Yan-Cheng Guo, Tian-Sheuan Chang, Jian-Wei Su | 2026-04-30 | 下载 | Digital computing-in-memory (DCIM) has emerged as a promising solution for large language model (LLM) acceleration by minimizing data transfers between external DRAM and on-chip accelerators while mai... |
| Autoformalizing Memory Specifications with Agents | Jan Ole Ernst, Dmitri Michelangelo Saberi, Derek Christ, Thomas Zimmermann, Rajath Salegame, Suhaas M. Bhat, Stanislav Levental, Thomas Dybdahl Ahle, Matthias Jung | 2026-04-30 | 下载 | The primary goal of Design Verification (DV) is to ensure that a proposed chip design implementation (either in code, or physical form) exactly matches its specification and is free of functional erro... |
cs.DC - Distributed, Parallel, and Cluster Computing
| 标题 | 作者 | 发布日期 | 摘要 | |
|---|---|---|---|---|
| or Not : A Tale of Three Algorithms for Streaming: Covariance Estimation after Welford and Chan-Golub-LeVeque | Felix Reichel | 2026-04-30 | 下载 | We place three algorithms for computing the unbiased sample covariance matrix in streaming and distributed settings on a common algebraic, numerical, and statistical foundation. |
| Replication in Graph Partitioning and Scheduling Problems | Pál András Papp, Toni Böhnlein, A. N. Yzelman | 2026-04-30 | 下载 | The efficient parallel execution of complex computations requires balancing the workload across processors while minimizing the communication between them. |
| Network Digital Untwinning: Towards Backward Optimization of Digital Twins | Zifan Zhang, Dianwei Chen, Anjun Gao, Manhua Wang, Mingzhe Chen, Minghong Fang, Xianfeng Yang, Yuchen Liu | 2026-04-30 | 下载 | Network digital twins (NDTs) are transforming network management by offering precise virtual replicas of physical network systems. However, their reliance on diverse and sensitive data introduces sign... |
| Akita: A High Usability Simulation Framework for Computer Architecture | Sabila Al Jannat, Ying Li, Mengyang He, Xuzhong Wang, Huizhi Zhao, Jingxiang Sun, Daoxuan Xu, Enze Xu, Yifan Sun | 2026-04-30 | 下载 | Computer architecture simulation is essential for evaluating new designs without the need for costly tapeout. The community has developed dozens of valuable simulators that have enabled significant ar... |
| NeuroRing: Scaling Spiking Neural Networks via Multi-FPGA Bidirectional Ring Topologies and Stream-Dataflow Architectures | Muhammad Ihsan Al Hafiz, Artur Podobas | 2026-04-30 | 下载 | Spiking neural networks (SNNs) are a promising paradigm for energy-efficient event-driven computation, but large-scale SNN execution remains challenging because sparse spike communication and synchron... |
| Characterizing Path-Independent Fees: A Route to Zero Impermanent Loss in CPMMs | Andrey Voronin, Roman Vlasov, Vladimir Gorgadze, Andrey Seoev, Yury Yanovich | 2026-04-30 | 下载 | Constant Product Market Makers use fees that are typically fixed proportions of trade size. When these fees are automatically reinvested into the pool, as in Uniswap~V2 and some designs of Uniswap V4,... |
| From Impermanent Loss to Sustainable Gain: Quantifying Profitability Zones for Liquidity Providers on DEX | Ignat Melnikov, Roman Vlasov, Vladimir Gorgadze, Andrey Seoev, Yury Yanovich | 2026-04-30 | 下载 | Decentralized Finance (DeFi) is a rapidly evolving segment of blockchain technology that enables a transformative approach to financial services through Web3 applications. |
| Exploring Sparse Matrix Multiplication Kernels on the Cerebras CS-3 | Milan Shah, Sheng Di, Michela Becchi | 2026-04-30 | 下载 | In recent years, novel AI accelerators have emerged as promising alternatives to GPU for AI model training and inference tasks. One such accelerator, the Cerebras CS-3, achieves strong performance on ... |
| Distributed Santa Claus via Global Rounding | Tijn de Vos, Leo Wennmann, Malte Baumecker, Yannic Maus, Florian Schager | 2026-04-30 | 下载 | In this paper, we consider the Santa Claus problem in the CONGEST model. This NP-hard problem can be modeled as a bipartite graph of children and gifts where an edge indicates that a child desires a g... |
| The Origins of MEV: Systematic Attribution of Arbitrage Opportunity Creation at Scale | Andrei Seoev, Dmitry Belousov, Anastasiia Smirnova, Ksenia Kurinova, Aleksei Smirnov, Denis Fedyanin, Yury Yanovich | 2026-04-30 | 下载 | Maximal Extractable Value (MEV) represents billions of dollars in extracted value that fundamentally shapes blockchain network dynamics and participant incentives. |
| Affinity Tailor: Dynamic Locality-Aware Scheduling at Scale | Jin Xin Ng, Ori Livneh, Richard O'Grady, Josh Don, Peng Ding, Samuel Grossman, Luis Otero, Chris Kennelly, David Lo, Carlos Villavieja | 2026-04-30 | 下载 | Modern large multicore systems often run multiple workloads that share CPUs under schedulers such as Linux CFS. To keep CPUs busy, these schedulers load-balance runnable work, causing each workload to... |
| AnTi-MiCS: Analytical Framework for Bounding Time in Embedded Mixed-Criticality Systems | Behnaz Ranjbar, Akash Kumar | 2026-04-30 | 下载 | In Mixed-Criticality (MC) systems, although the high Worst-Case Execution Time (WCET) serves as a conservative upper bound representing the task's maximum execution time under all conditions, obtainin... |
| AI Inference as Relocatable Electricity Demand: A Latency-Constrained Energy-Geography Framework | Xubin Luo, Cheng Yang | 2026-04-30 | 下载 | AI inference is becoming a persistent and geographically distributed source of electricity demand. Unlike many traditional electrical loads, inference workloads can sometimes be executed away from the... |
| ZipCCL: Efficient Lossless Data Compression of Communication Collectives for Accelerating LLM Training | Wenxiang Lin, Xinglin Pan, Ruibo Fan, Shaohuai Shi, Xiaowen Chu | 2026-04-30 | 下载 | Communication has emerged as a critical bottleneck in the distributed training of large language models (LLMs). While numerous approaches have been proposed to reduce communication overhead, the poten... |
| Autonomous Systems Dependability in the era of AI: Design Challenges in Safety, Security, Reliability and Certification | Behnaz Ranjbar, Kirankumar Raveendiran, Sudeep Pasricha, Samarjit Chakraborty, Cecilia Carbonelli, Akash Kumar | 2026-04-30 | 下载 | The design of embedded safety-critical systems such as those used in next-generation automotive and autonomous platforms, is increasingly challenged by escalating system complexity, hardware-software ... |
| Monadic Presburger Predicates have Robust Population Protocols | Philipp Czerner, Javier Esparza, Vincent Fischer, Roland Guttenberg, Julian Pins, Simon Reilich | 2026-04-30 | 下载 | Population protocols are a model of distributed computation in which a collection of indistinguishable finite-state agents interact randomly in pairs to decide a predicate of their initial configurati... |
| Back to the Future: Rethinking Endorsement in Order-Execute Blockchains | Rongji Huang, Yifeng Ye, Gerui Wang, Mingchao Wan, Yuxing Duan, Jingjing Zhang, Guangtao Xue, Shengyun Liu | 2026-04-30 | 下载 | Due to regulatory compliance and governance management, modern (permissioned) blockchains require flexible endorsement, which allows the endorsement policy for each contract or state object to be indi... |
| Lightweight Tamper-Evident Log Integrity Verification for IoT Edge Environments: A Merkle Tree Pipeline with Adaptive Chunking | Muhammet Anil Yagiz, Fahrettin Horasan, Ahmet Hasim Yurttakal | 2026-04-30 | 下载 | Integrity of audit logs produced by Internet of Things (IoT) devices is a prerequisite for post-incident forensics, regulatory compliance, and operational accountability. |
| A Study on the Performance of Distributed Training of Data-driven CFD Simulations | Sergio Iserte, Alejandro González-Barberá, Paloma Barreda, Krzysztof Rojek | 2026-04-30 | 下载 | Data-driven methods for computer simulations are blooming in many scientific areas. The traditional approach to simulating physical behaviors relies on solving partial differential equations (PDE). |
| Towards the Democratization and Standardization of Dynamic Resources with MPI Spawning | Sergio Iserte, Iker Martín-Alvarez, Krzystof Rojek, José I. Aliaga, Maribel Castillo, Antonio J. Peña | 2026-04-30 | 下载 | This paper presents an efficient tool for managing dynamic resources in production high-performance computing (HPC) settings, focusing on flexibility, adaptability, and user-friendliness. |
cs.NI - Networking and Internet Architecture
| 标题 | 作者 | 发布日期 | 摘要 | |
|---|---|---|---|---|
| Rethinking Network Topologies for Cost-Effective Mixture-of-Experts LLM Serving | Junsun Choi, Sam Son, Sunjin Choi, Hansung Kim, Yakun Sophia Shao, Scott Shenker, Sylvia Ratnasamy, Borivoje Nikolic | 2026-04-30 | 下载 | Mixture-of-experts (MoE) architectures have turned LLM serving into a cluster-scale workload in which communication consumes a considerable portion of LLM serving runtime. |
| Fidelity-Guaranteed Entanglement Routing with Distributed Purification Planning | Anthony Gatti, Anoosha Fayyaz, Prashant Krishnamurthy, Kaushik P. Seshadreesan, Amy Babay | 2026-04-30 | 下载 | Many quantum-network applications require end-to-end Bell pairs whose fidelity exceeds a request-specific threshold, but existing entanglement routing algorithms either optimize only throughput withou... |
| A Multi-Perspective Study of the Internet Shutdown in Iran | Ali Sadeghi Jahromi, Jason Jaskolka | 2026-04-30 | 下载 | Iran conducted two nationwide Internet shutdowns in January and March 2026, the latter ongoing at the time of writing and the longest documented Iranian disruption. |
| RouteProfile: Elucidating the Design Space of LLM Profiles for Routing | Jingjun Xu, Hongji Pu, Tao Feng, Haozhen Zhang, Jiaxuan You, Ge Liu | 2026-04-30 | 下载 | As the large language model (LLM) ecosystem expands, individual models exhibit varying capabilities across queries, benchmarks, and domains, motivating the development of LLM routing. |
| Network Digital Untwinning: Towards Backward Optimization of Digital Twins | Zifan Zhang, Dianwei Chen, Anjun Gao, Manhua Wang, Mingzhe Chen, Minghong Fang, Xianfeng Yang, Yuchen Liu | 2026-04-30 | 下载 | Network digital twins (NDTs) are transforming network management by offering precise virtual replicas of physical network systems. However, their reliance on diverse and sensitive data introduces sign... |
| DeGenTWeb: A First Look at LLM-dominant Websites | Sichang Steven He, Calvin Ardi, Ramesh Govindan, Harsha V. Madhyastha | 2026-04-30 | 下载 | Many recent news reports have claimed that content generated by large language models (LLMs) is taking over the web. However, these claims are typically not based on a representative sample of the web... |
| A MEC-Based Optimization Framework for Dynamic Inductive Charging | Emre Akıskalıoğlu, Mustafa Atmaca, Lorenzo Ghiro, Giovanni Perin, Renato Lo Cigno | 2026-04-30 | 下载 | Range anxiety and long recharging times remain critical barriers to electric vehicle adoption. Dynamic Inductive Charging (DIC) offers a compelling solution by enabling wireless power transfer while d... |
| NetSatBench: A Distributed LEO Constellation Emulator with an SRv6 Case Study | Andrea Detti, Shahram Dadras, Giuseppe Tropea | 2026-04-30 | 下载 | NetSatBench is a distributed emulation platform for evaluating communication protocols and application workloads over large-scale LEO satellite systems. |
| Libra: Accelerating Socket I/O via Programmable Selective Data Copying | Kairui Zhou, Shengkai Lin, Wei Zhang, Shizhen Zhao | 2026-04-30 | 下载 | Layer-7 (L7) proxies are critical to modern cloud-native systems, yet their performance is increasingly bottlenecked by copying entire payloads across the kernel-user boundary. |
| LZn : Robust LoRa Frame Synchronization Under Frame Collisions and Ultra-Low SNR Conditions | José Álamos, Thomas C. Schmidt, Matthias Wählisch | 2026-04-30 | 下载 | LoRa has become a widely adopted wireless modulation scheme in LPWANs due to its low cost, long range, and minimal transmission power. However, collisions between frames of the same spreading factor -... |
| Multi-Connectivity for UAVs: A Measurement Study of Integrating Cellular, Aerial Mesh, and LEO Satellite Links | Aygun Baltaci, Irshad A. Meer, Mustafa Ozger, Cicek Cavdar, Dominic Schupke | 2026-04-30 | 下载 | Future uncrewed aerial vehicle (UAV) systems increasingly combine heterogeneous communication technologies, such as low-latency aerial mesh, terrestrial cellular, and satellite links, to improve robus... |
| Unified 5G-IoT Framework with CAMARA Gateways and SDN Federation | Zihan Jia, Ze Wang, Chen Chen, Ziren Xiao, Fung Po Tso | 2026-04-30 | 下载 | The convergence of 5G and IoT enables fully connected, intelligent environments, but it faces challenges from the fragmentation of public/private 5G networks and the heterogeneity of IoT networks. |
| ReVo: A Cross-Layer Reliable Volumetric Videoconferencing System | Ankur Aditya, Diptyaroop Maji, Lingdong Wang, Bhavya Ramakrishna, Ramesh Sitaraman, Prashant Shenoy | 2026-04-30 | 下载 | Volumetric videoconferencing enables immersive six Degrees of Freedom interactions by jointly transmitting visual appearance and 3D geometry. However, delivering volumetric video over today's networks... |
cs.OS - Operating Systems
| 标题 | 作者 | 发布日期 | 摘要 | |
|---|---|---|---|---|
| Crab: A Semantics-Aware Checkpoint/Restore Runtime for Agent Sandboxes | Tianyuan Wu, Chaokun Chang, Lunxi Cao, Wei Gao, Wei Wang | 2026-04-30 | 下载 | Autonomous agents act through sandboxed containers and microVMs whose state spans filesystems, processes, and runtime artifacts. Checkpoint and restore (C/R) of this state is needed for fault toleranc... |
| Affinity Tailor: Dynamic Locality-Aware Scheduling at Scale | Jin Xin Ng, Ori Livneh, Richard O'Grady, Josh Don, Peng Ding, Samuel Grossman, Luis Otero, Chris Kennelly, David Lo, Carlos Villavieja | 2026-04-30 | 下载 | Modern large multicore systems often run multiple workloads that share CPUs under schedulers such as Linux CFS. To keep CPUs busy, these schedulers load-balance runnable work, causing each workload to... |
| treVM: Tiny Rust Embedded Virtual Machines with WASM on Variable Resource-Constrained Hardware | Antoine Lavandier, Bastien Buil, Chrystel Gaber, Emmanuel Baccelli | 2026-04-30 | 下载 | Software stacks embedded on microcontroller-based hardware typically provide rudimentary APIs programmed in C/C++, basic connectivity and, sometimes, a firmware update mechanism. |