2026-08-22
cs.AR - Architecture
| 标题 | 作者 | 发布日期 | 摘要 | |
|---|---|---|---|---|
| What actually runs: a measurement study of language model placement and decode speed on the Apple Neural Engine | Shahir M A | 2026-08-22 | 下载 | We ask what gets a language model onto the Apple Neural Engine (ANE) and what makes it fast there, and we answer with three measurements. We sweep a 64-shape matrix of LLM primitives that varies how a... |
| Constraint-Driven Modeling Enabling Dual Model Checking and Simulation for Discrete Event Systems | Soroosh Gholami, Hessam S. Sarjoughian | 2026-08-22 | 下载 | Verification and validation (V&V) are crucial methods for evaluating the requirements and specifications of dynamical models that fulfill their intended purposes. |
| SweepLSD: A One-Pass, O(width)-Memory Line Segment Detector with an Integer-Only Streaming Core and a Real-Time FPGA Realization | Yoshiyasu Shimizu | 2026-08-22 | 下载 | We present SweepLSD, a line segment detector that reads the image exactly once and emits each segment within a few rows of its last pixel passing the scan line. |
| NoTB: Oracle-Free Triage of LLM-Generated RTL via Cross-Model Formal Consensus | Elisavet Lydia Alvanaki, Je Yang, Biruk Seyoum, Luca P. Carloni | 2026-08-22 | 下载 | Large language models (LLMs) are increasingly used to generate register-transfer-level (RTL) designs from natural-language specifications. However, assessing functional correctness at early stages rem... |
| TherMapNet Attention-Guided Runtime Full-Chip Thermal Map Prediction from Performance Metrics | Qin Gu, Chaofang Ma, Mingyu Yang, Yipu Zhang, Jiliang Zhang, Wei Zhang, Lin Jiang | 2026-08-22 | 下载 | Runtime thermal management of high-performance chips depends on fast and accurate full-chip thermal maps. Conventional simulators typically estimate power traces from performance metrics first, which ... |
cs.DC - Distributed, Parallel, and Cluster Computing
| 标题 | 作者 | 发布日期 | 摘要 | |
|---|---|---|---|---|
| Systematization of Knowledge: Formal Verification of Consensus Protocols | Nikita Bondarev, Kirill Ziborov, Yury Yanovich | 2026-08-22 | 下载 | Formal verification is increasingly critical for blockchain consensus protocols, where subtle bugs can cause irreversible financial loss and network failure. |
| PRISM: Predictive Runtime In-place Scaling and Model Selection for Edge Microservices | Uwe Gropengießer, Thomas Reuter, Dominik Schön, Osama Abboud, Xun Xiao, Max Mühlhäuser | 2026-08-22 | 下载 | Latency-sensitive edge AI services must balance strict deadlines, output quality, and limited compute and energy budgets. However, static CPU provisioning wastes resources because inference cost varie... |
| How Far Can You Do Nothing On a Quantum Computer? | Nitay Mayo, Tal Mor, Aryeh Lev Zabokritskiy | 2026-08-22 | 下载 | We present a route-resolved comparative assessment of Rigetti's Cepheus-1-108Q and IBM Heron-r2 processors using the established 'do-nothing' state-transfer protocol. |
| A Concurrent Queue System for Multi-GPU Platforms: Application to Bellman-Ford SSSP | Beyza Cavusoglu | 2026-08-22 | 下载 | This paper presents the design and implementation of a multi-GPU concurrent queue system using NVIDIA's NVSHMEM. The Bellman-Ford algorithm is used as a case study to evaluate the performance of the p... |
| FlashReg: GPU-Accelerated 3-Clique Point Cloud Registration for Real-Time Correspondence-to-Pose Estimation | Ziyang Yu, Xiang Li, Qiong Chang, Jun Miyazaki | 2026-08-22 | 下载 | Graph-based point cloud registration achieves high robustness by identifying geometrically consistent correspondence sets, but constructing second-order compatibility graphs and enumerating candidate ... |
| Rapid Earthquake-to-Tsunami Waveform Generation via Large-Scale Multi-GPU FFT Convolution Applied to the Cascadia Subduction Zone | Bowen Shi, Sreeram Venkat, Stefan Henneking, Omar Ghattas | 2026-08-22 | 下载 | Data-driven methods for earthquake and tsunami early warning rely on large ensembles of rupture scenarios and their resulting waveforms, but generating such datasets with repeated high-fidelity seismi... |
| PowerSlider: Exploiting Phase Asymmetry for LLM Serving under Demand Response | Yueying Li, Jiayang Chen, Yuanfan Chen, Leo Han, Haoran Qiu, Esha Choukse, Rodrigo Fonseca, Udit Gupta | 2026-08-22 | 下载 | AI inference clusters are increasingly constrained by instantaneous power, not just energy: grid operators condition new capacity on demand response, imposing time-varying power caps. |
cs.NI - Networking and Internet Architecture
| 标题 | 作者 | 发布日期 | 摘要 | |
|---|---|---|---|---|
| TessIndex: Capability Verified Identity System for the Agent Economy | Mehul Goenka, Tejas Pathak, Siddharth Asthana | 2026-08-22 | 下载 | Software systems have traditionally been organized around applications where human users act as principal decision-makers. Recent developments in agentic capabilities alter this paradigm: software age... |
| Pruned Traffic Trees: Native Semantic Compression with a Protocol-Structured Model Family for Encrypted Traffic Classification | Yuantu Luo, Jun Tao, Xiangyu Xu, Linxiao Yu, Kangying Li | 2026-08-22 | 下载 | Deep learning has achieved strong performance in encrypted traffic classification (ETC), yet its computational cost limits deployment on resource-constrained network devices such as routers and middle... |
| Reliability-Aware Scheduling for Digital Twin Maintenance | Milica Jankov, Carlo Fischione | 2026-08-22 | 下载 | In Industrial Internet of Things systems, learning-enabled Digital Twins (DTs) support remote monitoring by using data reported by distributed devices to maintain digital representations of physical p... |
| Building A CSFQ-Inspired Transport for Switched CXL Memory Pooling | Zerui Guo, Emily Shriver, Ming Liu | 2026-08-22 | 下载 | Emerging switched CXL memory pooling systems, albeit promising, suffer from significant performance interference due to the shared but performance-uncontrolled data path among concurrent memory stream... |
cs.PF - Performance
| 标题 | 作者 | 发布日期 | 摘要 | |
|---|---|---|---|---|
| What actually runs: a measurement study of language model placement and decode speed on the Apple Neural Engine | Shahir M A | 2026-08-22 | 下载 | We ask what gets a language model onto the Apple Neural Engine (ANE) and what makes it fast there, and we answer with three measurements. We sweep a 64-shape matrix of LLM primitives that varies how a... |
| When Structure is Silent: Opportunities for Algorithmic Dispatch in Linear Algebra | Emmanuel Lujan, Alan Edelman | 2026-08-22 | 下载 | Algorithmic dispatch is essential for performance in linear-algebra-intensive systems. A persistent challenge lies in the treatment of structured matrices. |