2026-06-16
cs.AR - Architecture
| 标题 | 作者 | 发布日期 | 摘要 | |
|---|---|---|---|---|
| Deep-Learning-Based Pixelated Microwave Filter Design and Characterization using Electro-Optical Electric-Field Measurements | Han Zhou, Richard Bannister, Caspar Pierce, Haojie Chang, David Widen, Ludvig Fornstedt, Gabriel Melin, Alexander Bohlin, Pontus Lindeberg Fredriksson, Dilbagh Singh, Christian Fager, Koen Buisman | 2026-06-16 | 下载 | Traditional microwave filter design typically relies on iterative parameter tuning and predefined topologies, which limits design space and increases development time. |
| Deep Learning-Driven Inverse Design of Doherty Power Amplifiers Using Pixelated Combiners and Dual-State Impedance Synthesis | Han Zhou, Haojie Chang, David Widen, Christian Fager | 2026-06-16 | 下载 | The output combiner of a Doherty power amplifier (PA) integrates load modulation, impedance matching, and phase compensation within a single network, making its design and synthesis highly challenging... |
| ComPart: Community-Guided Post-Coarsening for High-Quality Hypergraph Partitioning | Yugao Zhu, Zhicheng Guo, Yuchao Wu, Mengming Li, Jing Wang, Zhiyao Xie | 2026-06-16 | 下载 | Hypergraph partitioning is a critical step in the design of complex embedded systems, essential for optimizing task mapping on heterogeneous MPSoCs and enabling multi-FPGA prototyping. |
| Embedded Machine Learning for Microcontroller-Class Edge Devices: Data, Feature, Evaluation, and Deployment Pipelines | Mostafa Darvishi | 2026-06-16 | 下载 | Embedded machine learning moves inference from cloud services to resource-constrained devices that must acquire data, preprocess signals, run a model, and act within tight limits on memory, energy, an... |
| IMPart: Integration of Memetic Operations into Multi-Level Framework for Large-k-Way Hypergraph Partitioning | Yugao Zhu, Zhicheng Guo, Shang Liu, Mengming Li, Jing Wang, Zhiyao Xie | 2026-06-16 | 下载 | The problem of k-way hypergraph partitioning is fundamental with significant applications in various fields, including VLSI design and scientific computing. |
| CUTh-Solver: GPU-Accelerated Sparse Matrix Solver for High-Resolution Thermal Simulation of 3D ICs | Chenghan Wang, Zhen Zhuang, Shui Jiang, Siyuan Liang, Xiaoman Yang, Kai Zhu, Darong Huang, Luis Costero, Rongmei Chen, Tsung-Wei Huang, David Atienza, Tsung-Yi Ho | 2026-06-16 | 下载 | Coarse-grained thermal simulation tends to underestimate localized thermal issues, potentially missing critical hotspots. Accurate analysis, therefore, demands fine-grained information, which dramatic... |
| MIVE: A Minimalist Integer Vector Engine for Softmax LayerNorm and RMSNorm Acceleration | Kosmas Alexandridis, Giorgos Dimitrakopoulos | 2026-06-16 | 下载 | The rapid growth of Large Language Models (LLMs) has intensified the need for specialized hardware accelerators that can satisfy stringent inference latency and power constraints. |
| Reconfigurable Computing Challenge: Transformer for Jet Tagging on Versal AI Engines | Gram Koski, Sean Lipps, Zhenghua Ma, G. Abarajithan, Ryan Kastner | 2026-06-16 | 下载 | Transformer-based models achieve strong performance for jet tagging at the CERN LHC, but deploying them in low-latency, resource-constrained trigger systems is challenging. |
| AUTOGATE: Automated Clock Gating via Toggling-Aware LLM-based RTL Rewriting | Yiting Wang, Chenhui Deng, Chia-Tung Ho, Yanqing Zhang, Zhuo Feng, Cunxi Yu, Ang Li, Gang Qu, Brucek Khailany | 2026-06-16 | 下载 | Fine-grain clock gating (FGCG) is among the most effective techniques for reducing dynamic power, yet current FGCG optimization flows remain largely manual. |
cs.DC - Distributed, Parallel, and Cluster Computing
| 标题 | 作者 | 发布日期 | 摘要 | |
|---|---|---|---|---|
| Flexible Distributed Particle Filtering for the Internet of Things via Aggregate Computing | Angela Cortecchia, Davide Domini, Giovanni Ciatto, Roberto Casadei, Danilo Pianini, Mirko Viroli | 2026-06-16 | 下载 | State estimation from uncertain, distributed observations is central in many cyber-physical applications. While Distributed Particle Filtering (DPF) algorithms address nonlinear and non-Gaussian estim... |
| Mixed-Precision Communication-Avoiding SGD for Generalized Linear Models on GPUs | Aditya Devarakonda, Irene Simó Muñoz, Giulia Guidi | 2026-06-16 | 下载 | Distributed stochastic gradient descent (SGD) is limited by communication rather than computation, since each iteration requires an AllReduce across processes. |
| Beyond Prediction: Tail-Aware Scheduling for LLM Inference | Yueying Li, Yuanfan Chen, Jiayang Chen, Esha Choukse, Haoran Qiu, G. Edward Suh, Rodrigo Fonseca, Ziv Scully, Udit Gupta | 2026-06-16 | 下载 | LLM serving exhibits extreme length variability, making size-based scheduling difficult in practice. Recent LLM schedulers approximate SJF/SRPT using predicted decode lengths or ranks and primarily re... |
| From Specification to Execution: AI Assisted Scientific Workflow Management | Komal Thareja, Hamza Safri, Rajiv Mayani, Anirban Mandal, Ewa Deelman | 2026-06-16 | 下载 | Scientific workflow management systems (WMS) support scalable and reproducible execution of complex pipelines, but workflow design, implementation, and debugging remain largely manual and require sign... |
| SCOPE-FL: A Strategy-proof Chain-based Optimal pareto efficient Federated Learning System | Seyed Salar Ghazi, Kaiwen Zhang, Mehdi feizi, Hans-Arno Jacobsen | 2026-06-16 | 下载 | Hierarchical Federated Learning (HFL) enables scalable collaborative model training across distributed devices while preserving data privacy. However, existing HFL client selection mechanisms suffer f... |
| Gatling: Rapid-Fire Consensus from Parallel Composition | Giulia Scaffino, Max Resnick, Joachim Neu | 2026-06-16 | 下载 | Consensus protocols form the core of blockchains and other replicated state machines, ensuring that all correct nodes process the same totally ordered log of input transactions. |
| S4oP: Operator-level Pruning of Structured State Space Models for Resource-Constrained Devices | Marco Deano, Filippo Ziche, Nicola Bombieri | 2026-06-16 | 下载 | Structured State Space Models (SSMs), including the S4 and S4D architectures, have recently emerged as powerful alternatives to attention-based models for capturing long-range dependencies in sequenti... |
| Latency Prediction for LLM Inference on NPU Systems | Juhyun Park, Seungwoo Jeong, Jingyu Lee, Kyungyong Lee | 2026-06-16 | 下载 | Deploying Large Language Models (LLMs) requires exploring a large configuration space spanning parallelization strategies, batching techniques, and scheduling policies. |
| RouteBalance: Fused Model Routing and Load Balancing for Heterogeneous LLM Serving | Wei Da, Evangelia Kalyvianaki | 2026-06-16 | 下载 | Heterogeneous LLM serving stacks split scheduling into two layers that optimize in isolation: model routers pick a model from quality and cost signals while ignoring instance load, and serving load ba... |
| An Epistemic Analysis of Random Coordinated Attack | Sophia Knight, David Lehnherr, Sergio Rajsbaum | 2026-06-16 | 下载 | The coordinated attack problem models the challenge of coordinating a joint action within a bounded time by communicating over unreliable links. |
| LUMEN: Coordinated Failure Recovery for Distributed LLM Serving | Zhang Cao, Shujie Han, Juncheng Zhang, Yuanming Ren, Yongkun Li, Patrick P. C. Lee | 2026-06-16 | 下载 | Modern large language model (LLM) serving clusters distribute inference requests across multiple worker processes on different GPUs, but failures are prevalent at scale. |
| TIGER: Inverting Transformer Gradients via Embedding-Subspace Distance Optimization | William Kalikman, Ivo Petrov, Dimitar I. Dimitrov, Martin Vechev | 2026-06-16 | 下载 | Federated learning allows multiple clients to jointly train a shared model by sending gradient updates to a central server while keeping raw inputs local. |
| From GPU to Microcontroller: Online Ridge Regression for Edge-Deployable Traffic Prediction | Suresh Purini, Archit Narwadkar, Deepak Gangadharan | 2026-06-16 | 下载 | State-of-the-art traffic flow forecasting models, including Graph Convolutional Networks and graph-less MLPs, require centralized GPU training across all sensors, making them impractical for resource-... |
| AoiZora: Topology-Aware Auto-Parallel Optimization for Inference of Diffusion Transformers | Kaijian Wang, Yuanyuan Xu, Fanjiang Ye, Ye Cao, Jingwei Zuo, T. S. Eugene Ng, Yarong Mu, Yuke Wang | 2026-06-16 | 下载 | Video diffusion has quickly grown into a key generative serving workload, yet producing each clip demands many denoising iterations over large spatio-temporal latents, which puts low-latency inference... |
| Multi-Orientation Edge-Minimum Repair for Non-Redundant Fault-Tolerant Broadcasting in Dense Gaussian Networks | Bader Albader | 2026-06-16 | 下载 | Dense Gaussian networks are degree-four algebraic interconnection networks with compact diameter and simple modular routing. This paper studies non-redundant one-to-all broadcast repair in the dense G... |
| Local Fault Repair of Perfect Resource Placements in Dense Gaussian Networks | Bader Albader | 2026-06-16 | 下载 | Perfect resource placement in dense Gaussian networks partitions the network into Lee balls centered at resource nodes. The fault-free placement problem is already classified; this paper studies the c... |
| SpecGen: Accelerating Agentic Kernel Optimization with Speculative Generation | Jihu Guo, Sitian Lu, Tenghui Ma, Wei Gao, Zhisheng Ye, Xingcheng Zhang, Dahua Lin | 2026-06-16 | 下载 | Agentic kernel optimization automates manual GPU kernel tuning via iterative generation, validation, and profiling with reasoning LLMs, casting the optimization task as feedback-guided search. |
| When the Next Step Is Not One Step: Distribution-Aware Execution Modeling for Concurrent Go Programs | Kaviru Hapuarachchi | 2026-06-16 | 下载 | Training a model to predict the next step in a concurrent program is harder than it looks: two runs of the same program from the same trace prefix can produce different next events, both valid, becaus... |
| RISE: Relay Inference and Online Scheduling for Efficient Edge-Device Collaborative Diffusion Model Services | Zilan Huang, Zhiqing Tang, Hanshuai Cui, Tian Wang, Yuan Wu, Weijia Jia, Wei Zhao | 2026-06-16 | 下载 | Text-to-image diffusion models are increasingly deployed at the network edge to serve heterogeneous workloads with diverse quality and latency requirements. |
cs.NI - Networking and Internet Architecture
| 标题 | 作者 | 发布日期 | 摘要 | |
|---|---|---|---|---|
| Understanding the "Airport" Censorship Circumvention Ecosystem in China | Rumaisa Habib, Mingshi Wu, Shiva Shahandeh, Min Ni, Eric Wustrow, Zakir Durumeric | 2026-06-16 | 下载 | In China, a burgeoning underground market sells citizens subscription-based censorship circumvention proxies known as ''airports''. We present the first systematic study of this ecosystem, combining u... |
| The Multipath Reliable Connection (MRC) Transport | Rip Sohan, Eric Spada, Eric Davis, Mark Handley, Idan Burstein, Tony Hurson, Jithin Jose, Vivek Kashyap, Rong Pan, Sayantan Sur, Sreevatsa Anantharamu, Aviv Barnea, Adrian Caulfield, Elazar Cohen, Elliot Edmunds, Yamin Friedman, Mahdieh Ghazi, Murali Guramali, Torsten Hoefler, Vipin Jain, Abdul Kabbani, Noam Katz, Yanfang Le, Charlie Mbariky, Guglielmo Morandin, Masoud Moshref, Shane O'Neil, Michael Papamichael, Jonas Pfefferle, Siva Santosh Pyla, Costin Raiciu, David Riddoch, Karen Schramm, Yuval Shpigelman, Shahaf Shuler, Shy Shyman, Raghava Sivaramu, Amin Tootoonchian, Yang Wang | 2026-06-16 | 下载 | MRC is an open, production-grade transport designed for large-scale AI/ML training over best-effort Ethernet. It extends RoCEv2 with explicit, composable primitives for per-packet multipath and sender... |
| OmniPlan: An Adaptive Framework for Timely and Near-Optimal Network Planning Optimization | Longlong Zhu, Jiashuo Yu, Zedi Chen, Yuhan Wu, Zhifan Jiang, Yuchen Xian, Yimeng Liu, Jiajie Su, Shaopeng Zhou, Xingyuan Li, Hongyan Liu, Xuan Liu, Dong Zhang, Chunming Wu, Xiang Chen | 2026-06-16 | 下载 | Network planning optimization is a fundamental problem across diverse domains, including transportation systems, communication networks, and power grids. |
| Energy-Efficient FSO Reconfiguration under User Mobility in Hybrid Fiber-IAB Backhaul | Piotr Lechowicz, Charitha Madapatha, Carlos Natalino, Tommy Svensson, Paolo Monti | 2026-06-16 | 下载 | User mobility creates stochastic, time-varying backhaul demand that static capacity provisioning cannot match. We propose a closed-loop, load-aware hysteresis controller for hybrid fiber-IAB-FSO backh... |
| User-Mobility-Aware Optimization of Fiber Placement in Hybrid Fiber-IAB Networks | Piotr Lechowicz, Charitha Madapatha, Carlos Natalino, Tommy Svensson, Paolo Monti | 2026-06-16 | 下载 | Metaheuristic optimization of hybrid fiber-IAB networks demonstrates that integrating user dynamics into topology design enables more adaptive and cost-efficient backhaul architectures, contributing t... |
| A T-API-Compliant ReAct Agentic Loop for Optical Networks: Generic vs. Domain-Specific Tool Abstractions | Seyed Morteza Ahmadian, Paolo Monti, Carlos Natalino | 2026-06-16 | 下载 | Optical networks need intent-driven, closed-loop agentic management, a key enabler for higher autonomy levels. We present the first T-API-compliant reasoning and act (ReAct) loop. |
| Security-Induced Braess Paradoxes in Service Function Chain Orchestration | Daniel Commey, Bin Mai | 2026-06-16 | 下载 | NFV/SDN orchestration lets operators instantiate and steer traffic through virtual firewalls, IDS/IPS replicas, WAF clusters, zero-trust gateways, backup inspection paths, and migration targets on dem... |
| UAV-CAS: A Calibrated Digital-Twin Dataset for Intrusion Detection in UAV Swarm Networks | Sripath Mishra, Bharat Bhargava, Zizheng Liu, Shafkat Islam | 2026-06-16 | 下载 | Intrusion detection systems (IDS) trained on wired-network benchmarks degrade sharply in real-world unmanned aerial vehicle (UAV) swarms, where mobility, fluctuating link quality, and decentralized ro... |
| FlowCLIP: Contrastive Pretraining Using Domain Names for Encrypted Traffic Classification | Eun Hun Choi | 2026-06-16 | 下载 | Network traffic classification enables website fingerprinting, intrusion detection, and Quality of Service management. However, developing methods that capture stable and generalizable traffic pattern... |
| DPDS: A DPDK-Based Packet Delayer and Spacer | Etienne Zink, Fabian Ihle, Michael Menth | 2026-06-16 | 下载 | In this paper we tackle the problem of adding varying delay to packets for link emulation. Naive approaches either add more delay than desired or cause packet reordering, both of which are undesirable... |
| Integration of 5G and Industrial Digital Models: A Case Study with AGVs | J. Cañete-Martín, J. Gómez-Jerez, M. C. Lucas-Estañ, J. Gozálvez | 2026-06-16 | 下载 | 5G is a fundamental technology for the digitalization of smart manufacturing. Smart manufacturing relies on the use of digital models to optimize industrial processes before implementation on the manu... |
| 5G Network Architecture and Configuration Choices to Support Teleoperated Driving at Scale | M. C. Lucas-Estañ, B. Coll-Perales, M. I. Khan, J. Gozálvez, S. S. Avedisov, O. Altintas, M. Sepulcre | 2026-06-16 | 下载 | Teleoperated driving (ToD) enables the remote driving or control of vehicles. For this purpose, vehicles must transmit video feeds to the ToD control center so that the remote operator is fully aware ... |
| Predictive Configured Grant Scheduling for Deterministic Wireless Communications | Syed Morsleen Riaz, M. Carmen Lucas-Estañ, Baldomero Coll-Perales, Javier Gozalvez | 2026-06-16 | 下载 | Future wireless networks must enhance their capacity to sustain deterministic service levels and support emerging time-sensitive services in key verticals. |
| Multi-Orientation Edge-Minimum Repair for Non-Redundant Fault-Tolerant Broadcasting in Dense Gaussian Networks | Bader Albader | 2026-06-16 | 下载 | Dense Gaussian networks are degree-four algebraic interconnection networks with compact diameter and simple modular routing. This paper studies non-redundant one-to-all broadcast repair in the dense G... |
| Local Fault Repair of Perfect Resource Placements in Dense Gaussian Networks | Bader Albader | 2026-06-16 | 下载 | Perfect resource placement in dense Gaussian networks partitions the network into Lee balls centered at resource nodes. The fault-free placement problem is already classified; this paper studies the c... |
| RATIO: Redundancy-Controlled Stochastic Routing for Reliable Vehicular Multi-Hop Networking | Lei Lei, Xudong Wang | 2026-06-16 | 下载 | Reliable, low-latency multi-hop data delivery in vehicular networks is increasingly demanded, yet remains challenging due to frequent route failures caused by high mobility and intermittent blockage. |
| ResAware: Cross-Environment Website Fingerprinting via Resource-Privileged Distillation | Chongru Fan, Wei Wang, Wentao Huang, Zhenquan Ding, Jinqiao Shi, Lei Cui, Zhiyu Hao, Xiaochun Yun | 2026-06-16 | 下载 | While Website Fingerprinting (WF) attacks achieve high accuracy in controlled laboratory settings, they often degrade substantially in real-world environments due to spatio-temporal drift, browser het... |
cs.OS - Operating Systems
| 标题 | 作者 | 发布日期 | 摘要 | |
|---|---|---|---|---|
| CloakLM: Obfuscating GPU Memory Layout to Mitigate Model Ex-filtration for Serving | Kunal Jain, Seokjin Go, Divya Mahajan | 2026-06-16 | 下载 | Large foundation models deployed on third-party and shared accelerator infrastructure face a practical risk of model exfiltration that existing defenses do not fully address. |
| Cordon: Semantic Transactions for Tool-Using LLM Agents | Zheng Chen, Hanqing Liu, Duling Xu, Dong Dong, Jialin Li, Bangzheng Pu, Jidong Zhai | 2026-06-16 | 下载 | Tool-using LLM agents are shifting the unit of computation from explicit human-issued commands to model-driven tasks with stateful consequences. |
cs.PF - Performance
| 标题 | 作者 | 发布日期 | 摘要 | |
|---|---|---|---|---|
| Group Commit Self-Clocks: Why Tuning Is Unnecessary Above a Device-Set Load Threshold | Madhulatha Mandarapu, Sandeep Kunkunuru | 2026-06-16 | 下载 | Group commit amortizes the fixed cost of a durable log flush across many committing transactions; the release rule - a timer, a batch size, or an adaptive policy - is a classic tuning knob. |
| Optimal Calibration of Quantum Network Links | Vinay Kumar, Claudio Cicconetti, Marco Conti, Andrea Passarella | 2026-06-16 | 下载 | The reliable distribution of entanglement is essential for the effective operation of quantum networks. Due to fundamental differences between quantum and classical communication systems, it is necess... |