2026-09-11
cs.NI - Networking and Internet Architecture
| 标题 | 作者 | 发布日期 | 摘要 | |
|---|---|---|---|---|
| mKernel: Fast Multi-GPU, Multi-Node Fused Kernels | Ziming Mao, Yihan Zhang, Shawn Wei Chew, Shuang Ma, Costin Raiciu, Yang Zhou, Scott Shenker, Ion Stoica | 2026-09-11 | 下载 | Communication has become a bottleneck in distributed training and inference of large models. Overlapping communication with computation at the granularity of kernels, on separate streams, reduces only... |
| Virtualized 5G Tesbed using OpenAirInterface: Tutorial and Benchmarking Tests | M. Dória, V. Sousa, A. Campos, N. Oliveira, P. Eduardo, P. Filho, C. Lima, J. Guilherme, D. Luna, I. Rego, M. Fernandes, A. Neto | 2026-09-11 | 下载 | The development of 5G and its evolutionary path to 6G brings virtualization as close as possible to the antennas. Native 3GPP systems are now software running at servers boosted by accelerator cards t... |
| Support-Aware Telemetry Compression for 5G Positioning via Conditional Conflict Graphs | Mohammad Reza Deylam Salehi, Hakima Chaouchi | 2026-09-11 | 下载 | Geographically separated transmission/reception points (TRPs) report quantized measurements to a Location Management Function (LMF), even when the application requires only a coarse location region. |
| Vertical Assessment of RF-EMF Exposure in a Building Adjacent to a Multi-Operator Shared Base Station | Ricardo Q. de F. H. Silva, Marcio E. C. Rodrigues, Fred S. R. Pinheiro, Gutembergue S. da Silva, Halysson B. Mendonça, Vicente A. de Sousa | 2026-09-11 | 下载 | This paper extends a previously published proposed approach for estimating worst-case exposure scenarios in buildings using publicly available Base Station (BS) parameters, applying it to a critical c... |
| Hybrid Monitoring for Early Fault Detection in Cloud-Native 5G Systems | Anton Andersson, Sai Akshara Naineni, Mats Jansborg, Yixing Zhang, Romaric Duvignau | 2026-09-11 | 下载 | This paper presents the design implementation and evaluation of NetMon a hybrid network monitoring system designed for Kubernetes-based 5G packet core deployments specifically evaluated on Ericssons A... |
| A Feature-Rich Embedded NIDS with eBPF/XDP: Detector and Architecture Trade-offs | Shiqi Wu, Oleksii Koshovyi, Georgios Pseiridis Pseiras, Victor Morel, Romaric Duvignau | 2026-09-11 | 下载 | Distributed Denial-of-Service (DDoS) attacks remain a serious threat to transport networks, with recent attack volumes exceeding 30 Tbps, and the telecommunications industry being the main target. |
cs.OS - Operating Systems
| 标题 | 作者 | 发布日期 | 摘要 | |
|---|---|---|---|---|
| SeqMoE: Toward Full-Load Performance via Predictive and Graph-Compatible MoE Offloading | Zihan Wang, Yuqi Wang, Lei Gong, Cheng Tang, Wenqi Lou, Teng Wang, Chao Wang, Xuehai Zhou | 2026-09-11 | 下载 | Mixture-of-Experts (MoE) creates a structural advantage for offloading: only a small fraction of activated experts need to reside in device memory, and if they can be loaded in time for computation, o... |
| Invisible Yet Dominant: Big Stalls of Kernel I/O Mechanisms in Cloud OLTP Databases | Mitsumasa Kondo | 2026-09-11 | 下载 | Most databases, including PostgreSQL, RocksDB, and recent AI KV-cache middleware, rely on buffered I/O, delegating write-back to the Linux kernel. |
cs.PF - Performance
| 标题 | 作者 | 发布日期 | 摘要 | |
|---|---|---|---|---|
| Transpilation-Aware Runtime Prediction for Noisy Quantum Circuit Simulation | Davud Azizov, Javier Vela-Tambo, Tian Guo | 2026-09-11 | 下载 | Predicting the runtime of noisy quantum circuit simulations is important for scheduling, resource allocation, and performance optimization. However, accurate prediction is challenging because backend-... |
| Dissecting GPU Utilization for LLM Inference on Nvidia Hopper | Mohammad Siavashi, Gerald Q. Maguire, Dejan Kostic, Marco Chiesa | 2026-09-11 | 下载 | A single SM utilization percentage can make an LLM inference workload look compute-saturated while hiding how much useful work is being done. The problem is not that the counter is wrong, but that it ... |
| Input Resolution Matters: Real-Time Object Detection Latency | Qingyang Zhang, Fumio Machida, Laura Carnevali | 2026-09-11 | 下载 | We model total latency as the convolution of preprocessing, inference, and postprocessing distributions under a simplifying independence approximation, with selected stage parameters expressed as func... |
| 4D Parallelism Unlocks Exascale Bayesian Neural Networks for High-Fidelity Atmospheric Modeling | Deifilia Kieckhefen, Juan Pedro Gutiérrez Hermosillo Muriedas, Lars Helge Heyen, Mathis Bode, Iida Hakulinen, Andreas Herten, Chelsea Maria John, Thorsten Kurth, Anni Moisala, Asena Karolin Özdemir, Kaleb Phipps, Oskar Taubert, Arvid Weyrauch, Markus Götz, Charlotte Debus | 2026-09-11 | 下载 | We present BEAST, the first-ever Bayesian Swin Transformer for atmospheric forecasting on 0.25 global resolution able to accurately quantify both aleatoric and epistemic uncertainty. |
| HeatCache: Thermal-aware Energy-efficient LLM Inference Scheduling for Chassis-level Liquid Cooling in Sustainable Edge Server Rooms | Rui Lu, Huanghuang Liang, Kaiqi Guan, Dan Wang | 2026-09-11 | 下载 | LLM inference is increasingly deployed at institution-scale edges to meet service requirements. However, multi-GPU inference consumes a large amount of electricity and produces substantial heat. |
| HoliBench: A Cross-Platform Benchmarking and Deployment Toolkit for Foundation Models in CPS-IoT Applications | Inesh Chakrabarti, Zejun Xiong, Pragya Sharma, Mani Srivastava | 2026-09-11 | 下载 | Foundation models, including large language models, vision-language models, and time-series foundation models, are increasingly deployed on embedded and edge platforms for CPS and IoT applications, wh... |
| Argus: Orchestrating Cross-Layer GPU Performance Measurements around Semantic Regions | Jianzhu Yao, Yue Guan, Srivatsan Ramesh, Yuanwei Fang, Jian Jiao, Boda Li, Yueming Hao, Xinwei Qiang, Pramod Viswanath, Yufei Ding, Bill Yoshimi, Alexey Loginov, Shane Nay, Adnan Aziz | 2026-09-11 | 下载 | GPU developers and automated optimizers need performance evidence for semantic code regions--such as neural-network operator implementations and pipeline stages--but this evidence is fragmented across... |