2026-08-28
cs.AR - Architecture
| 标题 | 作者 | 发布日期 | 摘要 | |
|---|---|---|---|---|
| Neuromorphic architectures as numerical solvers for computational neuroscience | Jakob Jordan, Ole Richter, Congyang Li, Mihai A. Petrovici, Rajit Manohar | 2026-08-28 | 下载 | Neuromorphic computing is closely associated with spiking neuronal networks. However, an alternative class of so-called "rate-based" models arising from computational neuroscience and machine learning... |
| Beyond Flat Netlist: Hierarchical Graph Representation Learning for Scalable Analysis of Sequential Circuits | Jingyi Zhou, Zhengyuan Shi, Jiaying Zhu, Ziyang Zheng, Qiang Xu | 2026-08-28 | 下载 | Circuit Representation Learning (CRL) offers a powerful paradigm to guide and optimize core Electronic Design Automation (EDA) tasks, but its practical adoption is hindered by the immense scale of ind... |
| Gen-TAS: A Generative AI-Aided Hardware-Software Task Allocation Framework for FPGA-GPP Heterogeneous Systems | Mary Kong, Yuqin Zhao, Semih Vazgecen, Cristian Sestito, Themis Prodromakis | 2026-08-28 | 下载 | FPGA-GPP heterogeneous systems combine software flexibility with the performance and energy efficiency of reconfigurable hardware. However, determining which application tasks should execute on the GP... |
| Great Expectations: Benchmarking the Real-World Performance of RVV 1.0 in HPC | Stepan Nassyr, Prateek Chawla, Daniel Seibel, Jayesh Badwaik, Kaveh Haghighi Mood, Andreas Herten | 2026-08-28 | 下载 | Following the ratification of the RISC-V Vector Extension (RVV 1.0), new commercially available silicon has been adopting the extension. This paper revisits the question of RISC-V viability for High-P... |
| AI Hardware Accelerators for Large Language Models: Architectures and the Memory Wall | Siddharth Patel, Rohit Singh | 2026-08-28 | 下载 | Large language models (LLMs) place unprecedented and still-growing demands on the hardware that trains and serves them. This review surveys the full landscape of AI hardware accelerators for LLMs, inc... |
cs.DC - Distributed, Parallel, and Cluster Computing
| 标题 | 作者 | 发布日期 | 摘要 | |
|---|---|---|---|---|
| Algorithmic Simplification for Million-Vertex Diffusion History Reconstruction | Gökhan Göktürk | 2026-08-28 | 下载 | Diffusion history reconstruction infers latent node states between sparse observations of SI or SIR processes. HERMES combines parameter fitting, a learned graph-neural proposal, and feasibility-aware... |
| Relaxed Sender Anonymity for CBDC Interbank Settlement: A Zero-Knowledge Approach on Permissioned EVM | Pietro Tiberi, Gabriele Marcelli, Vitangelo Lasorella | 2026-08-28 | 下载 | Central Bank Digital Currency (CBDC) interbank settlement systems operating on Distributed Ledger Technology (DLT) face a fundamental trade-off: blockchain transparency enables trustless verification ... |
| No Silver Bullet: Boosting GaussDB Performance on the 30TB TPC-H Workload | Tim Zeyl, Jason Lam, Shu Lin, Reza Pournaghi, Qi Cheng, Calvin Wong, Kaixiang Du, Yuliang He, Yang Sun, Weicheng Wang, Paul Lee, Chen Ruo, Yang Xinyi, Li Qunan, Wang Junjie, Hu Dongxing, Chong Chen, Per-Ake Larson | 2026-08-28 | 下载 | GaussDB is Huawei's premier database system, designed for large-scale deployments and the most demanding workloads. It is a distributed shared-nothing system, capable of handling all types of workload... |
| Memory-efficient GPU pipelines for real-time non-line-of-sight reconstruction | Alfonso López-Ruiz, Diego Royo | 2026-08-28 | 下载 | Non-line-of-sight (NLOS) imaging reconstructs scenes hidden around a corner from indirect light recorded by a single-photon avalanche diode (SPAD). |
| HARTS: Efficient Agentic Reinforcement Learning for Hybrid-Attention Models over Arbitrary Rollout Trees | Boyuan Meng, Peihua Bao, Hong Liu, Xiaowei Zhu, Chao Wang, Gen Li, Zhenxuan Pan | 2026-08-28 | 下载 | Agentic reinforcement learning (RL) often produces irregular rollout trees with shared histories. Training root-to-leaf trajectories independently recomputes these shared prefixes. |
| Great Expectations: Benchmarking the Real-World Performance of RVV 1.0 in HPC | Stepan Nassyr, Prateek Chawla, Daniel Seibel, Jayesh Badwaik, Kaveh Haghighi Mood, Andreas Herten | 2026-08-28 | 下载 | Following the ratification of the RISC-V Vector Extension (RVV 1.0), new commercially available silicon has been adopting the extension. This paper revisits the question of RISC-V viability for High-P... |
| Performance Evaluation of Fast Fourier Transforms on Emerging RISC-V Hardware with Vector Extension Support | Daniel Seibel, Kaveh Haghighi Mood, Jayesh Badwaik, Prateek Chawla, Stepan Nassyr, Andreas Herten | 2026-08-28 | 下载 | This manuscript presents a performance evaluation of Fast Fourier Transform (FFT) implementations on emerging processors supporting the RISC-V Vector Extension (RVV 1.0). |
| AI Hardware Accelerators for Large Language Models: Architectures and the Memory Wall | Siddharth Patel, Rohit Singh | 2026-08-28 | 下载 | Large language models (LLMs) place unprecedented and still-growing demands on the hardware that trains and serves them. This review surveys the full landscape of AI hardware accelerators for LLMs, inc... |
| Characterization of Request and Token Energy Costs for LLM Inference Workloads on GPU Platforms | Prabhu Vellaisamy, Vanessa Lam, Shawn Blanton, John Paul Shen | 2026-08-28 | 下载 | Large language model (LLM) inference serving is priced by tokens, but GPU energy is consumed over inference windows. This accounting mismatch makes token-normalized metrics incomplete, since average o... |
| Learning-Augmented Heuristics: Simple, yet Smart, Robust and Interpretable Cache Eviction | Haocheng Xia, William Nixon, Bintang Dwi Marthen, Pranav Bhandari, Juncheng Yang | 2026-08-28 | 下载 | Caching is widely used across the system stack to improve performance and efficiency, with eviction algorithms at its core. Existing cache eviction policies fall into two broad categories: static heur... |
| TerraceMoE: A Cost Model for Hierarchical MoE All-to-All Communication | Weicheng Xue, Bingqiang Wang, Li Yuan, Huihui Zhou, Yonghong Tian | 2026-08-28 | 下载 | Hierarchical two-hop dispatch can reduce slow-fabric traffic in expert-parallel Mixture-of-Experts training, but it adds a second collective and an arrival-side operator chain. |
| FISGuard: Defending Against Membership Inference via Fixed Input Subspaces | Haocheng Jiang, Hua Shen | 2026-08-28 | 下载 | As large language models are increasingly adopted in federated learning, protecting user privacy while performing parameter-efficient fine-tuning on distributed private data has become an important ch... |
cs.NI - Networking and Internet Architecture
| 标题 | 作者 | 发布日期 | 摘要 | |
|---|---|---|---|---|
| Adaptive RIS-aided Communications through ML-based Generation of Phase Masks | Corwin Carpenter, Thomas Daltzis, George C. Trichopoulos, Jacek Kibilda, Joao F. Santos | 2026-08-28 | 下载 | Reconfigurable Intelligent Surfaces (RISs) are an attractive technology for Millimeter Wave (mmWave) communications due to their ability to passively reflect incident signals. |
| Uncertainty-Aware Multi-Task Learning for Joint Modulation Recognition and SINR Estimation | Kosar Nourolahi, Vahid Ghasemi | 2026-08-28 | 下载 | Joint modulation recognition and signal-to-interference-plus-noise ratio (SINR) estimation can reduce duplicated processing in intelligent receivers, but the two tasks have different uncertainty chara... |
| xTRUCE: A Provably Safe Arbiter for Multi-xApp Conflict Mitigation in Agentic O-RAN | Le Xia, Rose Qingyang Hu, Paul S. Kudyba, Zhenlin An, Haijian Sun | 2026-08-28 | 下载 | The open radio access network (O-RAN) is evolving toward agentic operation, where large language model (LLM)-driven xApps/rApps generate control proposals under operator intents. |
| Network Topologies for QKD Networks | Ori Rottenstreich, Ran Hasson Ruso, Eliahu Cohen | 2026-08-28 | 下载 | Quantum key distribution (QKD) is a method for distributing cryptographic keys between remote endpoints, enjoying security based on quantum physics. |
cs.OS - Operating Systems
| 标题 | 作者 | 发布日期 | 摘要 | |
|---|---|---|---|---|
| CrabOS: An Operating System for Human-AI Co-inhabitation | Qi Yang, Yun Ma | 2026-08-28 | 下载 | AI agents are evolving into long-running computational entities that can invoke tools, maintain memory, and complete complex tasks across applications. |
cs.PF - Performance
| 标题 | 作者 | 发布日期 | 摘要 | |
|---|---|---|---|---|
| Characterization of Request and Token Energy Costs for LLM Inference Workloads on GPU Platforms | Prabhu Vellaisamy, Vanessa Lam, Shawn Blanton, John Paul Shen | 2026-08-28 | 下载 | Large language model (LLM) inference serving is priced by tokens, but GPU energy is consumed over inference windows. This accounting mismatch makes token-normalized metrics incomplete, since average o... |
| FFSlim: An Efficient and Lightweight Format for Multi-modal Data Storage and Retrieval | Long Yang, Yu Mao, Yuchen Shao, Yumiao Zhao, Yaqi Li, Xuan Liu, Xiaolong Shen, Tao Yu, Gezi Li, Jing Wang, Chengcheng Wan, Liang Shi | 2026-08-28 | 下载 | With the rapid expansion of large-scale media-text corpora, multi-modal datasets increasingly require efficient storage and retrieval. Existing formats such as Files, TDP, and FFRecord work adequately... |