Skip to content

2026-08-22 ​

cs.AR - Architecture ​

标题作者发布日期PDF摘要
What actually runs: a measurement study of language model placement and decode speed on the Apple Neural EngineShahir M A2026-08-22下载We ask what gets a language model onto the Apple Neural Engine (ANE) and what makes it fast there, and we answer with three measurements. We sweep a 64-shape matrix of LLM primitives that varies how a...
Constraint-Driven Modeling Enabling Dual Model Checking and Simulation for Discrete Event SystemsSoroosh Gholami, Hessam S. Sarjoughian2026-08-22下载Verification and validation (V&V) are crucial methods for evaluating the requirements and specifications of dynamical models that fulfill their intended purposes.
SweepLSD: A One-Pass, O(width)-Memory Line Segment Detector with an Integer-Only Streaming Core and a Real-Time FPGA RealizationYoshiyasu Shimizu2026-08-22下载We present SweepLSD, a line segment detector that reads the image exactly once and emits each segment within a few rows of its last pixel passing the scan line.
NoTB: Oracle-Free Triage of LLM-Generated RTL via Cross-Model Formal ConsensusElisavet Lydia Alvanaki, Je Yang, Biruk Seyoum, Luca P. Carloni2026-08-22下载Large language models (LLMs) are increasingly used to generate register-transfer-level (RTL) designs from natural-language specifications. However, assessing functional correctness at early stages rem...
TherMapNet Attention-Guided Runtime Full-Chip Thermal Map Prediction from Performance MetricsQin Gu, Chaofang Ma, Mingyu Yang, Yipu Zhang, Jiliang Zhang, Wei Zhang, Lin Jiang2026-08-22下载Runtime thermal management of high-performance chips depends on fast and accurate full-chip thermal maps. Conventional simulators typically estimate power traces from performance metrics first, which ...

cs.DC - Distributed, Parallel, and Cluster Computing ​

标题作者发布日期PDF摘要
Systematization of Knowledge: Formal Verification of Consensus ProtocolsNikita Bondarev, Kirill Ziborov, Yury Yanovich2026-08-22下载Formal verification is increasingly critical for blockchain consensus protocols, where subtle bugs can cause irreversible financial loss and network failure.
PRISM: Predictive Runtime In-place Scaling and Model Selection for Edge MicroservicesUwe Gropengießer, Thomas Reuter, Dominik Schön, Osama Abboud, Xun Xiao, Max Mühlhäuser2026-08-22下载Latency-sensitive edge AI services must balance strict deadlines, output quality, and limited compute and energy budgets. However, static CPU provisioning wastes resources because inference cost varie...
How Far Can You Do Nothing On a Quantum Computer?Nitay Mayo, Tal Mor, Aryeh Lev Zabokritskiy2026-08-22下载We present a route-resolved comparative assessment of Rigetti's Cepheus-1-108Q and IBM Heron-r2 processors using the established 'do-nothing' state-transfer protocol.
A Concurrent Queue System for Multi-GPU Platforms: Application to Bellman-Ford SSSPBeyza Cavusoglu2026-08-22下载This paper presents the design and implementation of a multi-GPU concurrent queue system using NVIDIA's NVSHMEM. The Bellman-Ford algorithm is used as a case study to evaluate the performance of the p...
FlashReg: GPU-Accelerated 3-Clique Point Cloud Registration for Real-Time Correspondence-to-Pose EstimationZiyang Yu, Xiang Li, Qiong Chang, Jun Miyazaki2026-08-22下载Graph-based point cloud registration achieves high robustness by identifying geometrically consistent correspondence sets, but constructing second-order compatibility graphs and enumerating candidate ...
Rapid Earthquake-to-Tsunami Waveform Generation via Large-Scale Multi-GPU FFT Convolution Applied to the Cascadia Subduction ZoneBowen Shi, Sreeram Venkat, Stefan Henneking, Omar Ghattas2026-08-22下载Data-driven methods for earthquake and tsunami early warning rely on large ensembles of rupture scenarios and their resulting waveforms, but generating such datasets with repeated high-fidelity seismi...
PowerSlider: Exploiting Phase Asymmetry for LLM Serving under Demand ResponseYueying Li, Jiayang Chen, Yuanfan Chen, Leo Han, Haoran Qiu, Esha Choukse, Rodrigo Fonseca, Udit Gupta2026-08-22下载AI inference clusters are increasingly constrained by instantaneous power, not just energy: grid operators condition new capacity on demand response, imposing time-varying power caps.

cs.NI - Networking and Internet Architecture ​

标题作者发布日期PDF摘要
TessIndex: Capability Verified Identity System for the Agent EconomyMehul Goenka, Tejas Pathak, Siddharth Asthana2026-08-22下载Software systems have traditionally been organized around applications where human users act as principal decision-makers. Recent developments in agentic capabilities alter this paradigm: software age...
Pruned Traffic Trees: Native Semantic Compression with a Protocol-Structured Model Family for Encrypted Traffic ClassificationYuantu Luo, Jun Tao, Xiangyu Xu, Linxiao Yu, Kangying Li2026-08-22下载Deep learning has achieved strong performance in encrypted traffic classification (ETC), yet its computational cost limits deployment on resource-constrained network devices such as routers and middle...
Reliability-Aware Scheduling for Digital Twin MaintenanceMilica Jankov, Carlo Fischione2026-08-22下载In Industrial Internet of Things systems, learning-enabled Digital Twins (DTs) support remote monitoring by using data reported by distributed devices to maintain digital representations of physical p...
Building A CSFQ-Inspired Transport for Switched CXL Memory PoolingZerui Guo, Emily Shriver, Ming Liu2026-08-22下载Emerging switched CXL memory pooling systems, albeit promising, suffer from significant performance interference due to the shared but performance-uncontrolled data path among concurrent memory stream...

cs.PF - Performance ​

标题作者发布日期PDF摘要
What actually runs: a measurement study of language model placement and decode speed on the Apple Neural EngineShahir M A2026-08-22下载We ask what gets a language model onto the Apple Neural Engine (ANE) and what makes it fast there, and we answer with three measurements. We sweep a 64-shape matrix of LLM primitives that varies how a...
When Structure is Silent: Opportunities for Algorithmic Dispatch in Linear AlgebraEmmanuel Lujan, Alan Edelman2026-08-22下载Algorithmic dispatch is essential for performance in linear-algebra-intensive systems. A persistent challenge lies in the treatment of structured matrices.

基于 VitePress 构建 · 使用本地搜索查找论文