2026-05-15
cs.AR - Architecture
| 标题 | 作者 | 发布日期 | 摘要 | |
|---|---|---|---|---|
| TTP: A Hardware-Efficient Design for Precise Prefetching in Ray Tracing | Yavuz Selim Tozlu, Anshul Naithani, Huiyang Zhou | 2026-05-15 | 下载 | Ray tracing (RT) is a 3D graphics technique that offers highly realistic visuals. It is becoming prominent and accessible as GPU vendors have integrated dedicated ray tracing acceleration hardware. |
| ADS-IMC: Accelerating Data Sorting with In-Memory Computation | Narendra Singh Dhakad, Santosh Kumar Vishvakarma | 2026-05-15 | 下载 | Sorting is a fundamental operation across numerous computational domains. Traditionally, this process involves transferring data from main memory to a processing unit for sorting, followed by writing ... |
| SRAM Based Digital Custom Compute Engine for Improved Area Efficiency of AI Hardware | Narendra Singh Dhakad, Santosh Kumar Vishvakarma | 2026-05-15 | 下载 | This paper presents a novel architecture utilizing a 10T SRAM cell for XNOR-based in-memory computing, aimed at mitigating the extensive routing challenges typically encountered in conventional in-mem... |
| Certificate-Aware Property-Directed Reachability | Arman Ferdowsi, Laura Kovacs | 2026-05-15 | 下载 | Property-Directed Reachability (PDR/IC3) is a standard workhorse for hardware safety verification, but most implementations are tuned primarily for time-to-answer and treat the produced invariant or c... |
| ICP: Exploiting Instruction Correlation for Prefetching Irregular Memory Accesses | Mengming Li, Chenlu Miao, Buqing Xu, Qijun Zhang, Xiangfeng Sun, Ceyu Xu, Yuan Xie, Wenkai Li, Shang Liu, Zhiyao Xie | 2026-05-15 | 下载 | Irregular memory accesses pose challenges for effective and efficient data prefetching. While temporal prefetchers have recently shown promise for irregular memory access patterns, their effectiveness... |
| ITHICA: Intra-Thread Instruction Checking Approach for Defect-Induced Silent Data Corruptions | Ioanna Vavelidou, Subho S. Banerjee, Eric X. Liu, Mike Fuller, Subhasish Mitra, Caroline Trippel | 2026-05-15 | 下载 | Hyperscaler reports of silent data corruptions (SDCs), presumed to be caused by silicon manufacturing defects, have motivated the development of functional tests for detecting defective CPUs. |
cs.DC - Distributed, Parallel, and Cluster Computing
| 标题 | 作者 | 发布日期 | 摘要 | |
|---|---|---|---|---|
| HexAGenT: Efficient Agentic LLM Serving via Workflow- and Heterogeneity-Aware Scheduling | You Peng, Youhe Jiang, Wenshuang Li, Xu Xu, Ke Zhou, Jiawei Jiang, Chen Wang, Binhang Yuan | 2026-05-15 | 下载 | Agentic LLM applications increasingly execute user requests as multi-step workflows involving planning, tool use, branching, refinement, and synthesis. |
| Exceeding the Numerical and Performance Characteristics of IEEE-754 SGEMM with BFloat16 Tensor Cores on GPUs for Scientific Computing | Harun Bayraktar, Cole Brower, John Gunnels, Greg Henry, Cherin Joseph, Jack Kosaian, Dmitry Lyakh, Lukas Mosimann, Victor Podlozhnyuk, Addison Richards, Paul Springer, Haicheng Wu | 2026-05-15 | 下载 | Largely due to their increased native capacity for numerical intensity and power efficiency, reduced-precision floating-point computing resources, primarily used in artificial intelligence (AI) applic... |
| Designing Datacenter Power Delivery Hierarchies for the AI Era | Grant Wilkins, Fiodar Kazhamiaka, Alok Gautam Kumbhare, Chaojie Zhang, Ricardo Bianchini | 2026-05-15 | 下载 | Demand for AI accelerators is rapidly increasing rack power density, with projections approaching 1MW per deployment by 2027. This poses a major challenge for datacenter power delivery designers. |
| Runtime-Orchestrated Second-Order Optimization for Scalable LLM Training | Yishun Lu, Junhao Zhang, Zeyu Yang, Wes Armour | 2026-05-15 | 下载 | Second-order methods offer an attractive path toward more sample-efficient LLM training, but their practical use is often blocked by the systems cost of maintaining and updating large matrix-based opt... |
| A GPU Accelerated Temporal Window-Based Random Walk Sampler | Md Ashfaq Salehin, George Parisis, Luc Berthouze | 2026-05-15 | 下载 | Temporal random walks, which sample causality-preserving paths, are widely used to analyze time-stamped interactions in domains such as microservices, finance, and online platforms. |
| From Backup Restoration to Minimum Viable Factory Recovery: A Systematization of Ransomware Recovery in Manufacturing Systems | Chun Yin Chiu | 2026-05-15 | 下载 | Ransomware recovery in critical manufacturing infrastructure is not only a backup-restoration problem. Production capability depends on coupled information-technology, operational-technology, physical... |
| PCDM: A Diffusion-Based Data Poisoning Attack Against Federated Learning Systems | Wei Sun, Yijun Chen, Bo Gao, Ke Xiong, Yuwei Wang, Pingyi Fan, Khaled Ben Letaief | 2026-05-15 | 下载 | Federated learning (FL) is vulnerable to data poisoning attacks due to its distributed nature. Although recent GAN-based data poisoning methods have indicated the potential of using generative AI to g... |
| An efficient multi-GPU implementation for the Discontinuous Galerkin ocean model SLIM | Miguel De Le Court, Vincent Legat, Ange P. Ishimwe, Colin Scherpereel, Emmanuel Hanert, Jonathan Lambrechts | 2026-05-15 | 下载 | Unstructured-mesh ocean models are increasingly used for coastal applications due to their ability to represent complex geometries and apply local grid refinement where needed. |
| High-Performance Star-M SVD for Big Data Compression | Md Taufique Hussain, Grey Ballard, Aditya Devarakonda, Srinivas Eswar, Naman Pesricha, Vishwas Rao | 2026-05-15 | 下载 | In the era of big data, effectively compressing large datasets while performing complex mathematical operations is crucial. Tensor-based decomposition methods have shown superior compression capabilit... |
| Evaluating Container Orchestration for Neuromorphic Workloads in Virtual Edge Environments | Huyen Pham, Bilhanan Silverajan | 2026-05-15 | 下载 | The growing adoption of edge computing has created an increasing need for workloads capable of operating under strict resource and energy constraints. |
| ADAPT: A Self-Calibrating Proactive Autoscaler for Container Orchestration | Himanshu Singh Baghel | 2026-05-15 | 下载 | Proactive autoscaling for containerized workloads depends on knowing the provisioning delay, i.e., the time between a scaling decision and the moment new capacity is ready to serve traffic. |
| Scale: Deep Reinforcement Learning for Container Scheduling in Serverless Edge Computing | Chen Chen, Zihan Jia, Andrea Sabbioni, Reza Farahani, Lei Jiao | 2026-05-15 | 下载 | Serverless computing has emerged as a promising computing paradigm for edge computing. However, adopting the event driven model in highly dynamic, heterogeneous, and distributed edge systems poses sig... |
| ParamSpMM: Adaptive and Efficient Sparse Matrix-Matrix Multiplication on GPUs for GNNs | Lixing Zhang, Guanhua Ye, Hongzheng Li, Shigang Li, Yingxia Shao | 2026-05-15 | 下载 | Fueled by the ability to mine real-world graph data, GNN applications have experienced phenomenal growth. Sparse Matrix-Matrix Multiplication (SpMM) is a critical operator in GNNs. |
| A Few GPUs, A Whole Lotta Scale: Faithful LLM Training Emulation with PrismLLM | Shaoke Xi, ChonLam Lao, Boyi Jia, Jiaqi Gao, Zhipeng Zhang, Jiamin Cao, Brian Sutioso, Erci Xu, Minlan Yu, Kui Ren, Yong Li, Zhengping Qian, Ennan Zhai, Jingren Zhou | 2026-05-15 | 下载 | Large language model (LLM) training today runs on clusters spanning thousands of GPUs. While this scale enables rapid model advances, developing, debugging, and performance-tuning the training framewo... |
| On the Fragility of Data Attribution When Learning Is Distributed | Xian Gao, Bo Hui, Min-Te Sun, Wei-Shinn Ku | 2026-05-15 | 下载 | Data attribution has become an important component of pricing, auditing, and governance in machine learning pipelines, yet most attribution methods implicitly assume that attribution values faithfully... |
cs.NI - Networking and Internet Architecture
| 标题 | 作者 | 发布日期 | 摘要 | |
|---|---|---|---|---|
| Against the Monolithic Wireless World Model: Why NextG Needs Composable and Agentic Intelligence | Aladin Djuhera, Farhan Ahmed, Vlad C. Andrei, Swanand Ravindra Kadhe, Alecio Binotto, Haris Gacanin, Holger Boche | 2026-05-15 | 下载 | AI-native 6G visions increasingly invoke wireless foundation models, large multimodal models, and wireless world models as the natural endpoint of AI-native networking, drawing an analogy to recent de... |
| Re/Imagining Smart Home Automation Framework in the Era of 6G-Enabled Smart Cities | Byungkwan Jung, Suman Kumar, Adityasinh Manthansinh Chauhan | 2026-05-15 | 下载 | Smart home automation systems represent a seamless integration of Internet of Things technologies, facilitating the monitoring, management, and regulation of various aspects of our daily life. |
| Seasonal Statistics of Shannon Capacity in a Dynamical Poisson-Voronoi Cellular Network | Sanjoy Kumar Jhawar, François Baccelli | 2026-05-15 | 下载 | In this work we consider a dynamical cellular communication network in which mobile BSs are modeled as a homogeneous Poisson point process on . |
| End-to-End Simulation of 5G NR Integrated Access and Backhaul Networks for Remote Maritime Connectivity | Alessandro Traspadini, Matteo Pagin, Raphaël Ihamouine, Rupert Lucas, Andrew Noren, Michele Zorzi, Marco Giordani | 2026-05-15 | 下载 | Millimeter wave (mmWave) 5th generation (5G) networks offer high data rates but face coverage challenges due to severe path loss and blockage. |
| Preemption Revisited: Multi-Threshold Preemption Policies for AoI Minimization | Sahan Liyanaarachchi, Sennur Ulukus, Nail Akar | 2026-05-15 | 下载 | The study of optimal preemption policies for status update systems has been a recurring topic in the age of information (AoI) literature, where threshold-based structures have been shown to be optimal... |
| Near-optimal Online Traffic Engineering | Arvin Ghavidel, Pooria Namyar, Nikolai Matni, Walter Willinger, Ramesh Govindan | 2026-05-15 | 下载 | Most deployed WAN Traffic Engineering (TE) systems use a logically centralized controller that periodically gathers traffic demands, runs a TE optimization or heuristic, and then programs the network. |
| How Far Back in Time a Digital Twin Reflects the State of the Physical Object: Age of Staleness | Ismail Cosandal, Sennur Ulukus | 2026-05-15 | 下载 | The groundbreaking metric age of information (AoI) has been introduced to measure information freshness in communication networks. As transformational as it is, AoI metric falls short in some applicat... |
| Restoring CFAR Validity for Single-Channel IoT Sensor Streams: A Monte Carlo Comparison of Five Detectors under Cortex-M0+ Constraints | Sergii Makovetskyi, Lars Thomsen | 2026-05-15 | 下载 | Real-time event detection in IoT mesh sensor networks must balance sensitivity against false-positive load on a constrained mesh radio. We present a Monte Carlo comparison of the Temporal Spectral Noi... |
| MAxLM: Multi-Agent Language Model-Based Scheduling and Resource Allocation in MU-MIMO-OFDMA-Enabled Wireless Networks | Adnan Quadri, Hongxiang Li | 2026-05-15 | 下载 | Wireless networks support multi-user (MU) communication with multiple-input multiple-output (MIMO) and orthogonal frequency-division multiple access (OFDMA) technologies. |
| IoT and Massive Connectivity: Massive MIMO Optimization for IoT Connectivity in 5G and Beyond Networks | Praveen Hegde, Robin Joseph Varughese | 2026-05-15 | 下载 | The IoT's explosive growth has led to a massive number of connected devices, which demand high-speed and pervasive connectivity, posing significant challenges for current-generation wireless communica... |
| Sustainability in Telecom: Energy-Efficient Networks and Circular Economy Models to Reduce Carbon Footprints and Increase Efficiency | Praveen Hegde, Robin Joseph Varughese | 2026-05-15 | 下载 | The increasing environmental impact of the telecom industry has heightened the need for sustainable telecommunications networks. With skyrocketing data traffic and 5G gaining a foothold, telecom opera... |
| HOPPER: A Hop-by-hop Entanglement Distribution Protocol for Asynchronous Quantum Networks | Claudio Cicconetti | 2026-05-15 | 下载 | The quantum Internet relies on the ability to distribute entangled quantum bits (ebits) between quantum memories at the end nodes, to perform applications like blind or distributed quantum computing t... |
| Joint Mobile User Positioning and Passive Target Sensing using Optimized Sequential Beamforming | Aymen Hamrouni, Sofie Pollin, Hazem Sallouha | 2026-05-15 | 下载 | Integrated sensing and communication (ISAC) relies on monostatic sensing (MS) and bistatic positioning (BP) to enable comprehensive environmental awareness and user localization. |
| The Shared Prosperity Internet | Juan A. Cabrera, Pit Hofmann, Jonas Schulz, Frederic Benken, Hrjehor Mark, Giang T. Nguyen, Holger Boche, Frank H. P. Fitzek | 2026-05-15 | 下载 | The Shared Prosperity Internet (SPI) is a network-computing architecture that makes the benefits of automation and Artificial Intelligence (AI) broadly accessible to the society. |
| The Internet Runs on Names | Geoff Huston, Lixia Zhang | 2026-05-15 | 下载 | The Internet's TCP/IP architecture was designed for resilient packet delivery between hosts identified by IP addresses. Over time, however, the consolidation of applications and services into large-sc... |
| Operator-Controlled 6G: From Connectivity Infrastructure to Guaranteed Digital Services | David Soldani | 2026-05-15 | 下载 | Sixth-generation mobile networks (6G) are approaching a structural inflection point. Five generations of vendor-led architectures have left operators procuring and operating networks they do not own, ... |
| TG-DIN: Theory-Guided Demand Inference Network for Generalizable QoS Measurement and Prediction | Fuliang Yang, Feng Ye | 2026-05-15 | 下载 | In this paper, we introduce TG-DIN, a theory-guided demand inference network that infers latent user demand from observable network quality-of-service (QoS) measurements. |
cs.OS - Operating Systems
| 标题 | 作者 | 发布日期 | 摘要 | |
|---|---|---|---|---|
| Skim: Speculative Execution for Fast and Efficient Web Agents | Mike Wong, Kevin Hsieh, Suman Nath, Ravi Netravali | 2026-05-15 | 下载 | Skim is a speculative execution framework for web agents that exploits the predictable structure of purpose-built websites. Today's web-agent expense is not intrinsic to the tasks but a property of ho... |
cs.PF - Performance
| 标题 | 作者 | 发布日期 | 摘要 | |
|---|---|---|---|---|
| Heuristic-Based Merging of HPC Traces to Extend Hardware Counter Coverage | Júlia Orteu Aubach, Fabio Banchelli, Marc Clascà Ramírez, Marta Garcia-Gasulla | 2026-05-15 | 下载 | This work extends a framework for predicting the performance of High-Performance Computing (HPC) workloads using Machine Learning (ML). A common limitation in performance modeling is the restricted nu... |
| Ghosted Layers: Unconstrained Activation Alignment for Recovering Layer-Pruned LLMs | Vincent-Daniel Yun, Junhyuk Jo, Sai Praneeth Karimireddy, Sunwoo Lee | 2026-05-15 | 下载 | Layer pruning removes entire Transformer decoder blocks from large language models, but introduces a mismatch between the hidden state received by the next surviving layer and the distribution it was ... |