Skip to content

2026-04-26 ​

cs.AR - Architecture ​

标题作者发布日期PDF摘要
Architectural Isolation as a Timing Safety Primitive for Edge AI Medical Devices: Controlled Experimental Evidence on a Shared-Silicon PlatformAkul Mallayya Swami2026-04-26下载A system can satisfy accuracy-based validation, maintain output stability (Safety-Threshold Exceedance Rate, STER, equal to zero), and still violate timing constraints under deployment load.
FlowPlace: Flow Matching for Chip PlacementPeng Xie, Ke Xue, Yunqi Shi, Ruo-Tong Chen, Chengrui Gao, Siyuan Xu, Chenjian Ding, Mingxuan Yuan, Chao Qian2026-04-26下载Chip placement plays an important role in physical design. While generative models like diffusion models offer promising learning-based solutions, current methods have the following limitations: they ...
Hardware-Efficient Softmax and Layer Normalization with Guaranteed Normalization for Edge DevicesDawon Choi, Hana Kim, Ji-Hoon Kim2026-04-26下载In Transformer models, non-GEMM (non-General Matrix Multiplication) operations -- especially Softmax and Layer Normalization (LayerNorm) -- often dominate hardware cost due to their nonlinear nature.
TimingLLM: A Two-Stage Retrieval-Augmented Framework for Pre-Synthesis Timing Prediction from VerilogArmin Abdollahi, Negin Ashrafi, Mehdi Kamal, Massoud Pedram2026-04-26下载Early, tool-free prediction of post-synthesis timing remains a key obstacle to rapid RTL iteration. We introduce TimingLLM, a two-stage retrieval-augmented LLM pipeline that estimates worst negative s...
Hardware-Efficient FPGA Implementation of Sigmoid Function Using Mixed-Radix Hyperbolic Rotation CORDICChintan Panchal, Ankur Changela, Mohendra Roy2026-04-26下载Efficient hardware implementation of nonlinear activation functions is a crucial task in deploying artificial neural networks on resource-constrained and edge devices such as Field-Programmable Gate A...

cs.DC - Distributed, Parallel, and Cluster Computing ​

标题作者发布日期PDF摘要
Towards System-Oriented Formal Verification of Local-First Access ControlFlorian Jacob, Johanna Stuber, Hannes Hartenstein2026-04-26下载Conflict-free replicated data types (CRDTs) and the local-first concept are increasingly employed not only in small-scale collaboration systems among few users who trust each other, but also in large-...
ClusterFusion++: Expanding Cluster-Level Fusion to Full Transformer-Block DecodingChiHeng Jin, Hongche Yu, Xihui Chen2026-04-26下载Large language model (LLM) decoding is latency-sensitive and often bottlenecked by fragmented operator execution and repeated off-chip materialization of intermediate tensors.

cs.NI - Networking and Internet Architecture ​

标题作者发布日期PDF摘要
Optimizing Information Freshness for Wireless Local Area Networks with Multiple APsAnanth Ram Rajagopalan, Jiahui Ni, Vishrant Tripathi2026-04-26下载Dense indoor WLANs increasingly rely on multiple access points (APs) operating over partially overlapping spectrum to support latency-sensitive applications.
NODE: Network Wide Top-K Flows in the Data PlaneEitan Stein, Lior Zeno, Shir Landau Feibish2026-04-26下载Monitoring network traffic is crucial for most network tasks, such as, identifying and blocking attacks, pinpointing failures and engineering and rerouting heavy traffic to maintain high throughput.
Adaptive Swin Transformer Partitioning over AI-RAN NetworksTam Thanh Nguyen, Yong Hao Pua, Tuan Van Ngo, Mao V. Ngo, Jihong Park, Binbin Chen, Tony Q. S. Quek2026-04-26下载This paper demonstrates the feasibility of transformer-based split inference for real-time video object detection over dynamic 5G AI-RAN networks.
PILOT: One Physics-Integrated Generation Framework to Unify 2D and 3D Radio Map ConstructionWeiming Huang, Hao Sun, Junting Chen2026-04-26下载Unified 2D and 3D radio map construction supports network planning, wireless digital twins, and unmanned aerial vehicle (UAV) applications. In urban environments, blockage, reflection, and diffraction...
Multi-Plane HyperX: A Low-Latency and Cost-Effective Network for Large-Scale AI and HPC SystemsZiyu Wang, Fei Lei, Dezun Dong2026-04-26下载Multi-plane architectures have become increasingly prevalent in the Fat-Tree networks of AI data centers. By leveraging multiple ports on a single network interface card (NIC) or multiple NICs within ...

cs.PF - Performance ​

标题作者发布日期PDF摘要
Optimas: An Intelligent Analytics-Informed Generative AI Framework for Performance OptimizationMohammad Zaeed, Tanzima Z. Islam, Vladimir Indic2026-04-26下载Large language models (LLMs) show promise for automated code optimization. However, without performance context, they struggle to produce correct and effective code transformations.

基于 VitePress 构建 · 使用本地搜索查找论文