Skip to content

2026-06-16 ​

cs.AR - Architecture ​

标题作者发布日期PDF摘要
Deep-Learning-Based Pixelated Microwave Filter Design and Characterization using Electro-Optical Electric-Field MeasurementsHan Zhou, Richard Bannister, Caspar Pierce, Haojie Chang, David Widen, Ludvig Fornstedt, Gabriel Melin, Alexander Bohlin, Pontus Lindeberg Fredriksson, Dilbagh Singh, Christian Fager, Koen Buisman2026-06-16下载Traditional microwave filter design typically relies on iterative parameter tuning and predefined topologies, which limits design space and increases development time.
Deep Learning-Driven Inverse Design of Doherty Power Amplifiers Using Pixelated Combiners and Dual-State Impedance SynthesisHan Zhou, Haojie Chang, David Widen, Christian Fager2026-06-16下载The output combiner of a Doherty power amplifier (PA) integrates load modulation, impedance matching, and phase compensation within a single network, making its design and synthesis highly challenging...
ComPart: Community-Guided Post-Coarsening for High-Quality Hypergraph PartitioningYugao Zhu, Zhicheng Guo, Yuchao Wu, Mengming Li, Jing Wang, Zhiyao Xie2026-06-16下载Hypergraph partitioning is a critical step in the design of complex embedded systems, essential for optimizing task mapping on heterogeneous MPSoCs and enabling multi-FPGA prototyping.
Embedded Machine Learning for Microcontroller-Class Edge Devices: Data, Feature, Evaluation, and Deployment PipelinesMostafa Darvishi2026-06-16下载Embedded machine learning moves inference from cloud services to resource-constrained devices that must acquire data, preprocess signals, run a model, and act within tight limits on memory, energy, an...
IMPart: Integration of Memetic Operations into Multi-Level Framework for Large-k-Way Hypergraph PartitioningYugao Zhu, Zhicheng Guo, Shang Liu, Mengming Li, Jing Wang, Zhiyao Xie2026-06-16下载The problem of k-way hypergraph partitioning is fundamental with significant applications in various fields, including VLSI design and scientific computing.
CUTh-Solver: GPU-Accelerated Sparse Matrix Solver for High-Resolution Thermal Simulation of 3D ICsChenghan Wang, Zhen Zhuang, Shui Jiang, Siyuan Liang, Xiaoman Yang, Kai Zhu, Darong Huang, Luis Costero, Rongmei Chen, Tsung-Wei Huang, David Atienza, Tsung-Yi Ho2026-06-16下载Coarse-grained thermal simulation tends to underestimate localized thermal issues, potentially missing critical hotspots. Accurate analysis, therefore, demands fine-grained information, which dramatic...
MIVE: A Minimalist Integer Vector Engine for Softmax LayerNorm and RMSNorm AccelerationKosmas Alexandridis, Giorgos Dimitrakopoulos2026-06-16下载The rapid growth of Large Language Models (LLMs) has intensified the need for specialized hardware accelerators that can satisfy stringent inference latency and power constraints.
Reconfigurable Computing Challenge: Transformer for Jet Tagging on Versal AI EnginesGram Koski, Sean Lipps, Zhenghua Ma, G. Abarajithan, Ryan Kastner2026-06-16下载Transformer-based models achieve strong performance for jet tagging at the CERN LHC, but deploying them in low-latency, resource-constrained trigger systems is challenging.
AUTOGATE: Automated Clock Gating via Toggling-Aware LLM-based RTL RewritingYiting Wang, Chenhui Deng, Chia-Tung Ho, Yanqing Zhang, Zhuo Feng, Cunxi Yu, Ang Li, Gang Qu, Brucek Khailany2026-06-16下载Fine-grain clock gating (FGCG) is among the most effective techniques for reducing dynamic power, yet current FGCG optimization flows remain largely manual.

cs.DC - Distributed, Parallel, and Cluster Computing ​

标题作者发布日期PDF摘要
Flexible Distributed Particle Filtering for the Internet of Things via Aggregate ComputingAngela Cortecchia, Davide Domini, Giovanni Ciatto, Roberto Casadei, Danilo Pianini, Mirko Viroli2026-06-16下载State estimation from uncertain, distributed observations is central in many cyber-physical applications. While Distributed Particle Filtering (DPF) algorithms address nonlinear and non-Gaussian estim...
Mixed-Precision Communication-Avoiding SGD for Generalized Linear Models on GPUsAditya Devarakonda, Irene Simó Muñoz, Giulia Guidi2026-06-16下载Distributed stochastic gradient descent (SGD) is limited by communication rather than computation, since each iteration requires an AllReduce across processes.
Beyond Prediction: Tail-Aware Scheduling for LLM InferenceYueying Li, Yuanfan Chen, Jiayang Chen, Esha Choukse, Haoran Qiu, G. Edward Suh, Rodrigo Fonseca, Ziv Scully, Udit Gupta2026-06-16下载LLM serving exhibits extreme length variability, making size-based scheduling difficult in practice. Recent LLM schedulers approximate SJF/SRPT using predicted decode lengths or ranks and primarily re...
From Specification to Execution: AI Assisted Scientific Workflow ManagementKomal Thareja, Hamza Safri, Rajiv Mayani, Anirban Mandal, Ewa Deelman2026-06-16下载Scientific workflow management systems (WMS) support scalable and reproducible execution of complex pipelines, but workflow design, implementation, and debugging remain largely manual and require sign...
SCOPE-FL: A Strategy-proof Chain-based Optimal pareto efficient Federated Learning SystemSeyed Salar Ghazi, Kaiwen Zhang, Mehdi feizi, Hans-Arno Jacobsen2026-06-16下载Hierarchical Federated Learning (HFL) enables scalable collaborative model training across distributed devices while preserving data privacy. However, existing HFL client selection mechanisms suffer f...
Gatling: Rapid-Fire Consensus from Parallel CompositionGiulia Scaffino, Max Resnick, Joachim Neu2026-06-16下载Consensus protocols form the core of blockchains and other replicated state machines, ensuring that all correct nodes process the same totally ordered log of input transactions.
S4oP: Operator-level Pruning of Structured State Space Models for Resource-Constrained DevicesMarco Deano, Filippo Ziche, Nicola Bombieri2026-06-16下载Structured State Space Models (SSMs), including the S4 and S4D architectures, have recently emerged as powerful alternatives to attention-based models for capturing long-range dependencies in sequenti...
Latency Prediction for LLM Inference on NPU SystemsJuhyun Park, Seungwoo Jeong, Jingyu Lee, Kyungyong Lee2026-06-16下载Deploying Large Language Models (LLMs) requires exploring a large configuration space spanning parallelization strategies, batching techniques, and scheduling policies.
RouteBalance: Fused Model Routing and Load Balancing for Heterogeneous LLM ServingWei Da, Evangelia Kalyvianaki2026-06-16下载Heterogeneous LLM serving stacks split scheduling into two layers that optimize in isolation: model routers pick a model from quality and cost signals while ignoring instance load, and serving load ba...
An Epistemic Analysis of Random Coordinated AttackSophia Knight, David Lehnherr, Sergio Rajsbaum2026-06-16下载The coordinated attack problem models the challenge of coordinating a joint action within a bounded time by communicating over unreliable links.
LUMEN: Coordinated Failure Recovery for Distributed LLM ServingZhang Cao, Shujie Han, Juncheng Zhang, Yuanming Ren, Yongkun Li, Patrick P. C. Lee2026-06-16下载Modern large language model (LLM) serving clusters distribute inference requests across multiple worker processes on different GPUs, but failures are prevalent at scale.
TIGER: Inverting Transformer Gradients via Embedding-Subspace Distance OptimizationWilliam Kalikman, Ivo Petrov, Dimitar I. Dimitrov, Martin Vechev2026-06-16下载Federated learning allows multiple clients to jointly train a shared model by sending gradient updates to a central server while keeping raw inputs local.
From GPU to Microcontroller: Online Ridge Regression for Edge-Deployable Traffic PredictionSuresh Purini, Archit Narwadkar, Deepak Gangadharan2026-06-16下载State-of-the-art traffic flow forecasting models, including Graph Convolutional Networks and graph-less MLPs, require centralized GPU training across all sensors, making them impractical for resource-...
AoiZora: Topology-Aware Auto-Parallel Optimization for Inference of Diffusion TransformersKaijian Wang, Yuanyuan Xu, Fanjiang Ye, Ye Cao, Jingwei Zuo, T. S. Eugene Ng, Yarong Mu, Yuke Wang2026-06-16下载Video diffusion has quickly grown into a key generative serving workload, yet producing each clip demands many denoising iterations over large spatio-temporal latents, which puts low-latency inference...
Multi-Orientation Edge-Minimum Repair for Non-Redundant Fault-Tolerant Broadcasting in Dense Gaussian NetworksBader Albader2026-06-16下载Dense Gaussian networks are degree-four algebraic interconnection networks with compact diameter and simple modular routing. This paper studies non-redundant one-to-all broadcast repair in the dense G...
Local Fault Repair of Perfect Resource Placements in Dense Gaussian NetworksBader Albader2026-06-16下载Perfect resource placement in dense Gaussian networks partitions the network into Lee balls centered at resource nodes. The fault-free placement problem is already classified; this paper studies the c...
SpecGen: Accelerating Agentic Kernel Optimization with Speculative GenerationJihu Guo, Sitian Lu, Tenghui Ma, Wei Gao, Zhisheng Ye, Xingcheng Zhang, Dahua Lin2026-06-16下载Agentic kernel optimization automates manual GPU kernel tuning via iterative generation, validation, and profiling with reasoning LLMs, casting the optimization task as feedback-guided search.
When the Next Step Is Not One Step: Distribution-Aware Execution Modeling for Concurrent Go ProgramsKaviru Hapuarachchi2026-06-16下载Training a model to predict the next step in a concurrent program is harder than it looks: two runs of the same program from the same trace prefix can produce different next events, both valid, becaus...
RISE: Relay Inference and Online Scheduling for Efficient Edge-Device Collaborative Diffusion Model ServicesZilan Huang, Zhiqing Tang, Hanshuai Cui, Tian Wang, Yuan Wu, Weijia Jia, Wei Zhao2026-06-16下载Text-to-image diffusion models are increasingly deployed at the network edge to serve heterogeneous workloads with diverse quality and latency requirements.

cs.NI - Networking and Internet Architecture ​

标题作者发布日期PDF摘要
Understanding the "Airport" Censorship Circumvention Ecosystem in ChinaRumaisa Habib, Mingshi Wu, Shiva Shahandeh, Min Ni, Eric Wustrow, Zakir Durumeric2026-06-16下载In China, a burgeoning underground market sells citizens subscription-based censorship circumvention proxies known as ''airports''. We present the first systematic study of this ecosystem, combining u...
The Multipath Reliable Connection (MRC) TransportRip Sohan, Eric Spada, Eric Davis, Mark Handley, Idan Burstein, Tony Hurson, Jithin Jose, Vivek Kashyap, Rong Pan, Sayantan Sur, Sreevatsa Anantharamu, Aviv Barnea, Adrian Caulfield, Elazar Cohen, Elliot Edmunds, Yamin Friedman, Mahdieh Ghazi, Murali Guramali, Torsten Hoefler, Vipin Jain, Abdul Kabbani, Noam Katz, Yanfang Le, Charlie Mbariky, Guglielmo Morandin, Masoud Moshref, Shane O'Neil, Michael Papamichael, Jonas Pfefferle, Siva Santosh Pyla, Costin Raiciu, David Riddoch, Karen Schramm, Yuval Shpigelman, Shahaf Shuler, Shy Shyman, Raghava Sivaramu, Amin Tootoonchian, Yang Wang2026-06-16下载MRC is an open, production-grade transport designed for large-scale AI/ML training over best-effort Ethernet. It extends RoCEv2 with explicit, composable primitives for per-packet multipath and sender...
OmniPlan: An Adaptive Framework for Timely and Near-Optimal Network Planning OptimizationLonglong Zhu, Jiashuo Yu, Zedi Chen, Yuhan Wu, Zhifan Jiang, Yuchen Xian, Yimeng Liu, Jiajie Su, Shaopeng Zhou, Xingyuan Li, Hongyan Liu, Xuan Liu, Dong Zhang, Chunming Wu, Xiang Chen2026-06-16下载Network planning optimization is a fundamental problem across diverse domains, including transportation systems, communication networks, and power grids.
Energy-Efficient FSO Reconfiguration under User Mobility in Hybrid Fiber-IAB BackhaulPiotr Lechowicz, Charitha Madapatha, Carlos Natalino, Tommy Svensson, Paolo Monti2026-06-16下载User mobility creates stochastic, time-varying backhaul demand that static capacity provisioning cannot match. We propose a closed-loop, load-aware hysteresis controller for hybrid fiber-IAB-FSO backh...
User-Mobility-Aware Optimization of Fiber Placement in Hybrid Fiber-IAB NetworksPiotr Lechowicz, Charitha Madapatha, Carlos Natalino, Tommy Svensson, Paolo Monti2026-06-16下载Metaheuristic optimization of hybrid fiber-IAB networks demonstrates that integrating user dynamics into topology design enables more adaptive and cost-efficient backhaul architectures, contributing t...
A T-API-Compliant ReAct Agentic Loop for Optical Networks: Generic vs. Domain-Specific Tool AbstractionsSeyed Morteza Ahmadian, Paolo Monti, Carlos Natalino2026-06-16下载Optical networks need intent-driven, closed-loop agentic management, a key enabler for higher autonomy levels. We present the first T-API-compliant reasoning and act (ReAct) loop.
Security-Induced Braess Paradoxes in Service Function Chain OrchestrationDaniel Commey, Bin Mai2026-06-16下载NFV/SDN orchestration lets operators instantiate and steer traffic through virtual firewalls, IDS/IPS replicas, WAF clusters, zero-trust gateways, backup inspection paths, and migration targets on dem...
UAV-CAS: A Calibrated Digital-Twin Dataset for Intrusion Detection in UAV Swarm NetworksSripath Mishra, Bharat Bhargava, Zizheng Liu, Shafkat Islam2026-06-16下载Intrusion detection systems (IDS) trained on wired-network benchmarks degrade sharply in real-world unmanned aerial vehicle (UAV) swarms, where mobility, fluctuating link quality, and decentralized ro...
FlowCLIP: Contrastive Pretraining Using Domain Names for Encrypted Traffic ClassificationEun Hun Choi2026-06-16下载Network traffic classification enables website fingerprinting, intrusion detection, and Quality of Service management. However, developing methods that capture stable and generalizable traffic pattern...
DPDS: A DPDK-Based Packet Delayer and SpacerEtienne Zink, Fabian Ihle, Michael Menth2026-06-16下载In this paper we tackle the problem of adding varying delay to packets for link emulation. Naive approaches either add more delay than desired or cause packet reordering, both of which are undesirable...
Integration of 5G and Industrial Digital Models: A Case Study with AGVsJ. Cañete-Martín, J. Gómez-Jerez, M. C. Lucas-Estañ, J. Gozálvez2026-06-16下载5G is a fundamental technology for the digitalization of smart manufacturing. Smart manufacturing relies on the use of digital models to optimize industrial processes before implementation on the manu...
5G Network Architecture and Configuration Choices to Support Teleoperated Driving at ScaleM. C. Lucas-Estañ, B. Coll-Perales, M. I. Khan, J. Gozálvez, S. S. Avedisov, O. Altintas, M. Sepulcre2026-06-16下载Teleoperated driving (ToD) enables the remote driving or control of vehicles. For this purpose, vehicles must transmit video feeds to the ToD control center so that the remote operator is fully aware ...
Predictive Configured Grant Scheduling for Deterministic Wireless CommunicationsSyed Morsleen Riaz, M. Carmen Lucas-Estañ, Baldomero Coll-Perales, Javier Gozalvez2026-06-16下载Future wireless networks must enhance their capacity to sustain deterministic service levels and support emerging time-sensitive services in key verticals.
Multi-Orientation Edge-Minimum Repair for Non-Redundant Fault-Tolerant Broadcasting in Dense Gaussian NetworksBader Albader2026-06-16下载Dense Gaussian networks are degree-four algebraic interconnection networks with compact diameter and simple modular routing. This paper studies non-redundant one-to-all broadcast repair in the dense G...
Local Fault Repair of Perfect Resource Placements in Dense Gaussian NetworksBader Albader2026-06-16下载Perfect resource placement in dense Gaussian networks partitions the network into Lee balls centered at resource nodes. The fault-free placement problem is already classified; this paper studies the c...
RATIO: Redundancy-Controlled Stochastic Routing for Reliable Vehicular Multi-Hop NetworkingLei Lei, Xudong Wang2026-06-16下载Reliable, low-latency multi-hop data delivery in vehicular networks is increasingly demanded, yet remains challenging due to frequent route failures caused by high mobility and intermittent blockage.
ResAware: Cross-Environment Website Fingerprinting via Resource-Privileged DistillationChongru Fan, Wei Wang, Wentao Huang, Zhenquan Ding, Jinqiao Shi, Lei Cui, Zhiyu Hao, Xiaochun Yun2026-06-16下载While Website Fingerprinting (WF) attacks achieve high accuracy in controlled laboratory settings, they often degrade substantially in real-world environments due to spatio-temporal drift, browser het...

cs.OS - Operating Systems ​

标题作者发布日期PDF摘要
CloakLM: Obfuscating GPU Memory Layout to Mitigate Model Ex-filtration for ServingKunal Jain, Seokjin Go, Divya Mahajan2026-06-16下载Large foundation models deployed on third-party and shared accelerator infrastructure face a practical risk of model exfiltration that existing defenses do not fully address.
Cordon: Semantic Transactions for Tool-Using LLM AgentsZheng Chen, Hanqing Liu, Duling Xu, Dong Dong, Jialin Li, Bangzheng Pu, Jidong Zhai2026-06-16下载Tool-using LLM agents are shifting the unit of computation from explicit human-issued commands to model-driven tasks with stateful consequences.

cs.PF - Performance ​

标题作者发布日期PDF摘要
Group Commit Self-Clocks: Why Tuning Is Unnecessary Above a Device-Set Load ThresholdMadhulatha Mandarapu, Sandeep Kunkunuru2026-06-16下载Group commit amortizes the fixed cost of a durable log flush across many committing transactions; the release rule - a timer, a batch size, or an adaptive policy - is a classic tuning knob.
Optimal Calibration of Quantum Network LinksVinay Kumar, Claudio Cicconetti, Marco Conti, Andrea Passarella2026-06-16下载The reliable distribution of entanglement is essential for the effective operation of quantum networks. Due to fundamental differences between quantum and classical communication systems, it is necess...

基于 VitePress 构建 · 使用本地搜索查找论文