Skip to content

2026-08-28 ​

cs.AR - Architecture ​

标题作者发布日期PDF摘要
Neuromorphic architectures as numerical solvers for computational neuroscienceJakob Jordan, Ole Richter, Congyang Li, Mihai A. Petrovici, Rajit Manohar2026-08-28下载Neuromorphic computing is closely associated with spiking neuronal networks. However, an alternative class of so-called "rate-based" models arising from computational neuroscience and machine learning...
Beyond Flat Netlist: Hierarchical Graph Representation Learning for Scalable Analysis of Sequential CircuitsJingyi Zhou, Zhengyuan Shi, Jiaying Zhu, Ziyang Zheng, Qiang Xu2026-08-28下载Circuit Representation Learning (CRL) offers a powerful paradigm to guide and optimize core Electronic Design Automation (EDA) tasks, but its practical adoption is hindered by the immense scale of ind...
Gen-TAS: A Generative AI-Aided Hardware-Software Task Allocation Framework for FPGA-GPP Heterogeneous SystemsMary Kong, Yuqin Zhao, Semih Vazgecen, Cristian Sestito, Themis Prodromakis2026-08-28下载FPGA-GPP heterogeneous systems combine software flexibility with the performance and energy efficiency of reconfigurable hardware. However, determining which application tasks should execute on the GP...
Great Expectations: Benchmarking the Real-World Performance of RVV 1.0 in HPCStepan Nassyr, Prateek Chawla, Daniel Seibel, Jayesh Badwaik, Kaveh Haghighi Mood, Andreas Herten2026-08-28下载Following the ratification of the RISC-V Vector Extension (RVV 1.0), new commercially available silicon has been adopting the extension. This paper revisits the question of RISC-V viability for High-P...
AI Hardware Accelerators for Large Language Models: Architectures and the Memory WallSiddharth Patel, Rohit Singh2026-08-28下载Large language models (LLMs) place unprecedented and still-growing demands on the hardware that trains and serves them. This review surveys the full landscape of AI hardware accelerators for LLMs, inc...

cs.DC - Distributed, Parallel, and Cluster Computing ​

标题作者发布日期PDF摘要
Algorithmic Simplification for Million-Vertex Diffusion History ReconstructionGökhan Göktürk2026-08-28下载Diffusion history reconstruction infers latent node states between sparse observations of SI or SIR processes. HERMES combines parameter fitting, a learned graph-neural proposal, and feasibility-aware...
Relaxed Sender Anonymity for CBDC Interbank Settlement: A Zero-Knowledge Approach on Permissioned EVMPietro Tiberi, Gabriele Marcelli, Vitangelo Lasorella2026-08-28下载Central Bank Digital Currency (CBDC) interbank settlement systems operating on Distributed Ledger Technology (DLT) face a fundamental trade-off: blockchain transparency enables trustless verification ...
No Silver Bullet: Boosting GaussDB Performance on the 30TB TPC-H WorkloadTim Zeyl, Jason Lam, Shu Lin, Reza Pournaghi, Qi Cheng, Calvin Wong, Kaixiang Du, Yuliang He, Yang Sun, Weicheng Wang, Paul Lee, Chen Ruo, Yang Xinyi, Li Qunan, Wang Junjie, Hu Dongxing, Chong Chen, Per-Ake Larson2026-08-28下载GaussDB is Huawei's premier database system, designed for large-scale deployments and the most demanding workloads. It is a distributed shared-nothing system, capable of handling all types of workload...
Memory-efficient GPU pipelines for real-time non-line-of-sight reconstructionAlfonso López-Ruiz, Diego Royo2026-08-28下载Non-line-of-sight (NLOS) imaging reconstructs scenes hidden around a corner from indirect light recorded by a single-photon avalanche diode (SPAD).
HARTS: Efficient Agentic Reinforcement Learning for Hybrid-Attention Models over Arbitrary Rollout TreesBoyuan Meng, Peihua Bao, Hong Liu, Xiaowei Zhu, Chao Wang, Gen Li, Zhenxuan Pan2026-08-28下载Agentic reinforcement learning (RL) often produces irregular rollout trees with shared histories. Training root-to-leaf trajectories independently recomputes these shared prefixes.
Great Expectations: Benchmarking the Real-World Performance of RVV 1.0 in HPCStepan Nassyr, Prateek Chawla, Daniel Seibel, Jayesh Badwaik, Kaveh Haghighi Mood, Andreas Herten2026-08-28下载Following the ratification of the RISC-V Vector Extension (RVV 1.0), new commercially available silicon has been adopting the extension. This paper revisits the question of RISC-V viability for High-P...
Performance Evaluation of Fast Fourier Transforms on Emerging RISC-V Hardware with Vector Extension SupportDaniel Seibel, Kaveh Haghighi Mood, Jayesh Badwaik, Prateek Chawla, Stepan Nassyr, Andreas Herten2026-08-28下载This manuscript presents a performance evaluation of Fast Fourier Transform (FFT) implementations on emerging processors supporting the RISC-V Vector Extension (RVV 1.0).
AI Hardware Accelerators for Large Language Models: Architectures and the Memory WallSiddharth Patel, Rohit Singh2026-08-28下载Large language models (LLMs) place unprecedented and still-growing demands on the hardware that trains and serves them. This review surveys the full landscape of AI hardware accelerators for LLMs, inc...
Characterization of Request and Token Energy Costs for LLM Inference Workloads on GPU PlatformsPrabhu Vellaisamy, Vanessa Lam, Shawn Blanton, John Paul Shen2026-08-28下载Large language model (LLM) inference serving is priced by tokens, but GPU energy is consumed over inference windows. This accounting mismatch makes token-normalized metrics incomplete, since average o...
Learning-Augmented Heuristics: Simple, yet Smart, Robust and Interpretable Cache EvictionHaocheng Xia, William Nixon, Bintang Dwi Marthen, Pranav Bhandari, Juncheng Yang2026-08-28下载Caching is widely used across the system stack to improve performance and efficiency, with eviction algorithms at its core. Existing cache eviction policies fall into two broad categories: static heur...
TerraceMoE: A Cost Model for Hierarchical MoE All-to-All CommunicationWeicheng Xue, Bingqiang Wang, Li Yuan, Huihui Zhou, Yonghong Tian2026-08-28下载Hierarchical two-hop dispatch can reduce slow-fabric traffic in expert-parallel Mixture-of-Experts training, but it adds a second collective and an arrival-side operator chain.
FISGuard: Defending Against Membership Inference via Fixed Input SubspacesHaocheng Jiang, Hua Shen2026-08-28下载As large language models are increasingly adopted in federated learning, protecting user privacy while performing parameter-efficient fine-tuning on distributed private data has become an important ch...

cs.NI - Networking and Internet Architecture ​

标题作者发布日期PDF摘要
Adaptive RIS-aided Communications through ML-based Generation of Phase MasksCorwin Carpenter, Thomas Daltzis, George C. Trichopoulos, Jacek Kibilda, Joao F. Santos2026-08-28下载Reconfigurable Intelligent Surfaces (RISs) are an attractive technology for Millimeter Wave (mmWave) communications due to their ability to passively reflect incident signals.
Uncertainty-Aware Multi-Task Learning for Joint Modulation Recognition and SINR EstimationKosar Nourolahi, Vahid Ghasemi2026-08-28下载Joint modulation recognition and signal-to-interference-plus-noise ratio (SINR) estimation can reduce duplicated processing in intelligent receivers, but the two tasks have different uncertainty chara...
xTRUCE: A Provably Safe Arbiter for Multi-xApp Conflict Mitigation in Agentic O-RANLe Xia, Rose Qingyang Hu, Paul S. Kudyba, Zhenlin An, Haijian Sun2026-08-28下载The open radio access network (O-RAN) is evolving toward agentic operation, where large language model (LLM)-driven xApps/rApps generate control proposals under operator intents.
Network Topologies for QKD NetworksOri Rottenstreich, Ran Hasson Ruso, Eliahu Cohen2026-08-28下载Quantum key distribution (QKD) is a method for distributing cryptographic keys between remote endpoints, enjoying security based on quantum physics.

cs.OS - Operating Systems ​

标题作者发布日期PDF摘要
CrabOS: An Operating System for Human-AI Co-inhabitationQi Yang, Yun Ma2026-08-28下载AI agents are evolving into long-running computational entities that can invoke tools, maintain memory, and complete complex tasks across applications.

cs.PF - Performance ​

标题作者发布日期PDF摘要
Characterization of Request and Token Energy Costs for LLM Inference Workloads on GPU PlatformsPrabhu Vellaisamy, Vanessa Lam, Shawn Blanton, John Paul Shen2026-08-28下载Large language model (LLM) inference serving is priced by tokens, but GPU energy is consumed over inference windows. This accounting mismatch makes token-normalized metrics incomplete, since average o...
FFSlim: An Efficient and Lightweight Format for Multi-modal Data Storage and RetrievalLong Yang, Yu Mao, Yuchen Shao, Yumiao Zhao, Yaqi Li, Xuan Liu, Xiaolong Shen, Tao Yu, Gezi Li, Jing Wang, Chengcheng Wan, Liang Shi2026-08-28下载With the rapid expansion of large-scale media-text corpora, multi-modal datasets increasingly require efficient storage and retrieval. Existing formats such as Files, TDP, and FFRecord work adequately...

基于 VitePress 构建 · 使用本地搜索查找论文