Skip to content

2026-09-11 ​

cs.NI - Networking and Internet Architecture ​

标题作者发布日期PDF摘要
mKernel: Fast Multi-GPU, Multi-Node Fused KernelsZiming Mao, Yihan Zhang, Shawn Wei Chew, Shuang Ma, Costin Raiciu, Yang Zhou, Scott Shenker, Ion Stoica2026-09-11下载Communication has become a bottleneck in distributed training and inference of large models. Overlapping communication with computation at the granularity of kernels, on separate streams, reduces only...
Virtualized 5G Tesbed using OpenAirInterface: Tutorial and Benchmarking TestsM. Dória, V. Sousa, A. Campos, N. Oliveira, P. Eduardo, P. Filho, C. Lima, J. Guilherme, D. Luna, I. Rego, M. Fernandes, A. Neto2026-09-11下载The development of 5G and its evolutionary path to 6G brings virtualization as close as possible to the antennas. Native 3GPP systems are now software running at servers boosted by accelerator cards t...
Support-Aware Telemetry Compression for 5G Positioning via Conditional Conflict GraphsMohammad Reza Deylam Salehi, Hakima Chaouchi2026-09-11下载Geographically separated transmission/reception points (TRPs) report quantized measurements to a Location Management Function (LMF), even when the application requires only a coarse location region.
Vertical Assessment of RF-EMF Exposure in a Building Adjacent to a Multi-Operator Shared Base StationRicardo Q. de F. H. Silva, Marcio E. C. Rodrigues, Fred S. R. Pinheiro, Gutembergue S. da Silva, Halysson B. Mendonça, Vicente A. de Sousa2026-09-11下载This paper extends a previously published proposed approach for estimating worst-case exposure scenarios in buildings using publicly available Base Station (BS) parameters, applying it to a critical c...
Hybrid Monitoring for Early Fault Detection in Cloud-Native 5G SystemsAnton Andersson, Sai Akshara Naineni, Mats Jansborg, Yixing Zhang, Romaric Duvignau2026-09-11下载This paper presents the design implementation and evaluation of NetMon a hybrid network monitoring system designed for Kubernetes-based 5G packet core deployments specifically evaluated on Ericssons A...
A Feature-Rich Embedded NIDS with eBPF/XDP: Detector and Architecture Trade-offsShiqi Wu, Oleksii Koshovyi, Georgios Pseiridis Pseiras, Victor Morel, Romaric Duvignau2026-09-11下载Distributed Denial-of-Service (DDoS) attacks remain a serious threat to transport networks, with recent attack volumes exceeding 30 Tbps, and the telecommunications industry being the main target.

cs.OS - Operating Systems ​

标题作者发布日期PDF摘要
SeqMoE: Toward Full-Load Performance via Predictive and Graph-Compatible MoE OffloadingZihan Wang, Yuqi Wang, Lei Gong, Cheng Tang, Wenqi Lou, Teng Wang, Chao Wang, Xuehai Zhou2026-09-11下载Mixture-of-Experts (MoE) creates a structural advantage for offloading: only a small fraction of activated experts need to reside in device memory, and if they can be loaded in time for computation, o...
Invisible Yet Dominant: Big Stalls of Kernel I/O Mechanisms in Cloud OLTP DatabasesMitsumasa Kondo2026-09-11下载Most databases, including PostgreSQL, RocksDB, and recent AI KV-cache middleware, rely on buffered I/O, delegating write-back to the Linux kernel.

cs.PF - Performance ​

标题作者发布日期PDF摘要
Transpilation-Aware Runtime Prediction for Noisy Quantum Circuit SimulationDavud Azizov, Javier Vela-Tambo, Tian Guo2026-09-11下载Predicting the runtime of noisy quantum circuit simulations is important for scheduling, resource allocation, and performance optimization. However, accurate prediction is challenging because backend-...
Dissecting GPU Utilization for LLM Inference on Nvidia HopperMohammad Siavashi, Gerald Q. Maguire, Dejan Kostic, Marco Chiesa2026-09-11下载A single SM utilization percentage can make an LLM inference workload look compute-saturated while hiding how much useful work is being done. The problem is not that the counter is wrong, but that it ...
Input Resolution Matters: Real-Time Object Detection LatencyQingyang Zhang, Fumio Machida, Laura Carnevali2026-09-11下载We model total latency as the convolution of preprocessing, inference, and postprocessing distributions under a simplifying independence approximation, with selected stage parameters expressed as func...
4D Parallelism Unlocks Exascale Bayesian Neural Networks for High-Fidelity Atmospheric ModelingDeifilia Kieckhefen, Juan Pedro Gutiérrez Hermosillo Muriedas, Lars Helge Heyen, Mathis Bode, Iida Hakulinen, Andreas Herten, Chelsea Maria John, Thorsten Kurth, Anni Moisala, Asena Karolin Özdemir, Kaleb Phipps, Oskar Taubert, Arvid Weyrauch, Markus Götz, Charlotte Debus2026-09-11下载We present BEAST, the first-ever Bayesian Swin Transformer for atmospheric forecasting on 0.25∘^\circ global resolution able to accurately quantify both aleatoric and epistemic uncertainty.
HeatCache: Thermal-aware Energy-efficient LLM Inference Scheduling for Chassis-level Liquid Cooling in Sustainable Edge Server RoomsRui Lu, Huanghuang Liang, Kaiqi Guan, Dan Wang2026-09-11下载LLM inference is increasingly deployed at institution-scale edges to meet service requirements. However, multi-GPU inference consumes a large amount of electricity and produces substantial heat.
HoliBench: A Cross-Platform Benchmarking and Deployment Toolkit for Foundation Models in CPS-IoT ApplicationsInesh Chakrabarti, Zejun Xiong, Pragya Sharma, Mani Srivastava2026-09-11下载Foundation models, including large language models, vision-language models, and time-series foundation models, are increasingly deployed on embedded and edge platforms for CPS and IoT applications, wh...
Argus: Orchestrating Cross-Layer GPU Performance Measurements around Semantic RegionsJianzhu Yao, Yue Guan, Srivatsan Ramesh, Yuanwei Fang, Jian Jiao, Boda Li, Yueming Hao, Xinwei Qiang, Pramod Viswanath, Yufei Ding, Bill Yoshimi, Alexey Loginov, Shane Nay, Adnan Aziz2026-09-11下载GPU developers and automated optimizers need performance evidence for semantic code regions--such as neural-network operator implementations and pipeline stages--but this evidence is fragmented across...

基于 VitePress 构建 · 使用本地搜索查找论文