2026-08-14
cs.AR - Architecture
| 标题 | 作者 | 发布日期 | 摘要 | |
|---|---|---|---|---|
| Beyond Capacity: Scalable MoE LLM Inference via High-Bandwidth Flash with Direct GPU and HBM Paths | Seeyeon Kim, Juhyeong Jin, Joo-Young Kim | 2026-08-14 | 下载 | Modern mixture-of-experts (MoE) language models increasingly strain the capacity and cost efficiency of high-bandwidth memory (HBM), as rapidly growing expert weights must be provisioned close to GPUs... |
| Qu-Trefoil: Large-Scale Quantum Circuit Simulator Working on FPGA With SATA Storages | Kaijie Wei, Hideharu Amano, Ryohei Niwase, Yoshiki Yamaguchi, Takefumi Miyoshi | 2026-08-14 | 下载 | Quantum circuits are fundamental components of quantum computing, and state-vector-based quantum circuit simulation is a widely used technique for tracking qubit behavior throughout circuit evolution. |
| The Quartic Hessian Conjecture in Dimension Four | Zixiang Ni | 2026-08-14 | 下载 | The Hessian conjecture asks whether a polynomial with nonzero constant Hessian determinant has a polynomial gradient inverse. It is known in dimensions at most three, false in dimensions at least five... |
| Experimental Study on System-Level Performance Impact of Read Disturbance in Modern SSDs | Yonggon Park, Hyunuk Cho, Onur Mutlu, Sungjin Lee, Jisung Park | 2026-08-14 | 下载 | This work investigates the system-level performance impact of read disturbance in modern NAND flash-based SSDs, aiming to provide new insights that can help develop better storage architectures and op... |
| MoE Expert Execution in Disaggregated LLM Serving with a High-Bandwidth ReRAM Near-Memory Architecture | Kunming Shao, Ming Zeng, Xin Yuan, Binbin Liao, Yangming Zhang, Wei Wang, Tim Kwang-Ting Cheng, Chi-Ying Tsui | 2026-08-14 | 下载 | Attention-FFN disaggregation maps LLM modules to specialized pools, creating an opening to keep Mixture-of-Experts (MoE) weights resident in a high-bandwidth FFN pool. |
| Characterizing the Variance Envelope: A Multi-Dimensional Analysis of Spectre Telemetry Across Architectures and Workloads | Jaya Keshava Chandra Kotha, Jean-Luc Gaudiot | 2026-08-14 | 下载 | Hardware attacks like Spectre exploit built-in processor vulnerabilities, leaving anomalous footprints in Hardware Performance Counter (HPC) metrics. |
| Exploring High-Bandwidth Flash for Modern LLM Inference: Opportunities and Challenges | Dowon Son, Yonggon Park, Hyunuk Cho, Hyungkyu Ham, Onur Mutlu, Sungjin Lee, Gwangsun Kim, Jisung Park | 2026-08-14 | 下载 | This work investigates the potential benefits and technical challenges of using high-bandwidth flash (HBF) for large language model (LLM) inference. |
cs.DC - Distributed, Parallel, and Cluster Computing
| 标题 | 作者 | 发布日期 | 摘要 | |
|---|---|---|---|---|
| Evaluating Agentic Code Repair Capabilities in Distributed Systems | Yibo Yan, Huijuan Wang, Junzhou He, Yizhuo Liang, Shaoyu Wang, Huanchen Sun, Seo Jin Park | 2026-08-14 | 下载 | LLM-based coding agents have advanced rapidly on single-process SWE tasks, with frontier models now clustering in the high-70s on SWE-bench Verified. |
| Enabling Hybrid HPCQC Workflows with a Heterogeneous Software Stack | Muhammad Nufail Farooqi, Minh Chung, Burak Mete, Eric Mansfield, Bernd Hoffmann, Teemu Mattsson, Laura Schulz, Jorge Echavarria | 2026-08-14 | 下载 | In this work, we demonstrate hybrid High Performance Computing-Quantum Computing (HPCQC) workflows on a production petascale system. The demonstration combines three components: the SuperMUC-NG superc... |
| Porting and Benchmarking Chapel on Emerging RISC-V Hardware: an HPC Viability Study | Ian Henriksen, Chris Taylor, Patrick Diehl, Jade Abraham, Palmer Cox, Bradford L. Chamberlain, Stephen L. Olivier | 2026-08-14 | 下载 | The Chapel programming language recently added support for the RISC-V architecture. Here we discuss what changes were needed for Chapel to work on RISC-V as well as lessons learned from the porting pr... |
| Validating LLM-Modernized Scientific Software Through Differential Fault Injection | Evan Coleman, Yuzhong Shen, Masha Sosonkina, Peng Xu | 2026-08-14 | 下载 | Large language model (LLM) agents are increasingly used to modernize the legacy Fortran underlying production scientific software, but validation of these transformations emphasizes nominal executions... |
| Rollplex: Cross-Phase GPU Spatial Sharing for Vision Language Model Post-Training | Hanfeng Lu, Tianyu Feng, Suyi Li, Yuheng Zhao, Wei Gao, Shaopan Xiong, Ju Huang, Siran Yang, Jiamang Wang, Lin Qu, Wei Wang | 2026-08-14 | 下载 | Vision-language models (VLMs) enable embodied agents to reason and act from visual observations and language instructions. Reinforcement learning (RL) post-training enhances these capabilities using t... |
| Large-scale workflow placement in serverless computing using integer nonlinear programming | Joshua Adamek, Natalie Carl, Trever Schirmer, Moritz Heinlein, David Bermbach, Sergio Lucia | 2026-08-14 | 下载 | Serverless edge computing has become a powerful cloud framework that enables the execution of large workflows without the need for the user to manage the underlying servers and edge devices. |
| Could Model Partitioning Make Federated Learning More Sustainable? | Tobias Frohlich, Tiffany Vlaar, Lauritz Thamsen | 2026-08-14 | 下载 | As federated learning (FL) extends from distributed machine learning between low-power devices to cross-silo scenarios involving edge servers and data centres, its carbon footprint has become a growin... |
| Hybrid Quantum-inspired Kolmogorov-Arnold Networks for Privacy-Aware Federated Biosignal Learning | Chun-Hua Lin, Samuel Yen-Chi Chen, Yu-Chao Hsu, Kuo-Chung Peng, Jiun-Cheng Jiang, Chi-Sheng Chen, Tai-Yue Li, Nan-Yow Chen, En-Jui Kuo, Hsi-Sheng Goan | 2026-08-14 | 下载 | Electrocardiogram (ECG) recordings are sensitive biomedical data, limiting the ability of hospitals and wearable devices to share raw signals for centralized model training. |
| Federated Prompt Learning: A Unified Framework, Empirical Analysis, and Future Directions | Qinglin Yang, Chen Qiu, Hongyuan Zhang, Pengdeng Li, Yuan Liu, Zhihong Tian | 2026-08-14 | 下载 | Large language models (LLMs) have become core components of cloud-based intelligent services in academia and industry, yet their training and deployment are hindered by high computational costs, data ... |
cs.NI - Networking and Internet Architecture
| 标题 | 作者 | 发布日期 | 摘要 | |
|---|---|---|---|---|
| Deep Reinforcement Learning for 6G AI-RAN: A Comprehensive Survey | Jie Lu, Peihao Yan, Qijun Wang, Ruxin Lin, Huacheng Zeng | 2026-08-14 | 下载 | The evolution toward sixth-generation (6G) networks is transforming the radio access network (RAN) into a programmable and intelligent control platform that must continuously adapt to heterogeneous se... |
| Scaling 5G-TSN Bridges: Operating Regimes, Scheduling, and Time Synchronisation Under Heterogeneous Industrial Traffic | Mohamed Seliem, Utz Roedig, Cormac Sreenan, Dirk Pesch | 2026-08-14 | 下载 | 3GPP Release 16 enables a 5G system to operate as a transparent IEEE 802.1 TSN bridge, but its scalability under heterogeneous industrial workloads remains insufficiently characterised. |
| Robust Constraint-Aware Bayesian Tuning of BBRv2 for QUIC under Tactile Internet Constraints | Muhammad Hanif Lashari, Shakil Ahmed, Wafa Batayneh, Ashfaq Khokhar | 2026-08-14 | 下载 | Tactile Internet applications place strict require- ments on latency, jitter, loss, and responsiveness, which makes transport configuration a critical design factor. |
| CipherSight: Robust Website Fingerprinting via Record-Resource Semantic Supervision under Distribution Shifts | Runhan Song, Qiqi Liu, Chuanzhou Pan, Zhenquan Ding, Youquan Xian, Chongru Fan, Lei Cui, Wei Wang, Zhiyu Hao | 2026-08-14 | 下载 | HTTPS website fingerprinting (WF) aims to identify visited websites from metadata observable in encrypted traffic. However, real-world deployments introduce a significant out-of-distribution (OOD) pro... |
cs.OS - Operating Systems
| 标题 | 作者 | 发布日期 | 摘要 | |
|---|---|---|---|---|
| CoRun: Padding is Simple and Efficient for Deterministic LLM Inference | Shiju Zhao, Jiacheng Yang, Qihang Chen, Junhao Hu, Jiaqi Zheng, Guihai Chen, Xusheng Chen | 2026-08-14 | 下载 | Despite fixed sampling parameters and random seeds, Large Language Model (LLM) inference exhibits output inconsistency, which undermines downstream tasks such as model evaluation and reinforcement lea... |
cs.PF - Performance
| 标题 | 作者 | 发布日期 | 摘要 | |
|---|---|---|---|---|
| CoRun: Padding is Simple and Efficient for Deterministic LLM Inference | Shiju Zhao, Jiacheng Yang, Qihang Chen, Junhao Hu, Jiaqi Zheng, Guihai Chen, Xusheng Chen | 2026-08-14 | 下载 | Despite fixed sampling parameters and random seeds, Large Language Model (LLM) inference exhibits output inconsistency, which undermines downstream tasks such as model evaluation and reinforcement lea... |