2026-06-07
cs.AR - Architecture
| 标题 | 作者 | 发布日期 | 摘要 | |
|---|---|---|---|---|
| Accuracy-Configurable Floating-Point Multiplier Design for SRAM-Based Compute-in-Memory | Yiqi Zhou, Junhao Lu, Jiale Yu, Zhuo Xu, Yang He, Yue Yuan, Shan Shen, Daying Sun | 2026-06-07 | 下载 | Digital Compute-in-Memory (DCiM) reduces data movement and has become a promising solution for energy-efficient edge AI. However, most existing DCiM frameworks still primarily target integer or fixed-... |
| Programming Domain-Specific FPGA Hardblocks from HLS: An RTL Blackbox Approach | Ruthwik Reddy Sunketa, Jeevesh Choudhury, Aman Arora | 2026-06-07 | 下载 | Domain-specific Field Programmable Gate Array (FPGA) architectures increasingly integrate specialized hardblocks, such as Tensor Slices, to accelerate artificial intelligence and machine learning work... |
cs.DC - Distributed, Parallel, and Cluster Computing
| 标题 | 作者 | 发布日期 | 摘要 | |
|---|---|---|---|---|
| A Low-Latency Semantic State Estimator using Latent Predictive Learning for Dynamic Network Monitoring and Orchestration | Hari Madhukumar, Haiyuan Li, Xiaolan Liu, Andy Corston-Petrie, Dimitra Simeonidou | 2026-06-07 | 下载 | Closed-loop network monitoring and orchestration increasingly require semantic interpretations of live telemetry beyond raw counter collection. |
| Parallel SMT Solving via Dynamic Partitioning, Core-Guided Pruning, and Online Backbone Detection | Ilana Shapiro, Sorin Lerner, Nikolaj Bjørner | 2026-06-07 | 下载 | Exploiting parallelism in modern CPU architectures remains a longstanding challenge in optimizing SMT solvers. We introduce a novel parallel framework that dynamically builds a binary partition tree o... |
| Aperon Technical Report: Hierarchical No-Pointer Tangent-Local Search for High-Dimensional Approximate Nearest Neighbors | Yong Fu | 2026-06-07 | 下载 | We present HNTL (Hierarchical No-pointer Tangent-Local), the core vector indexing and candidate generation framework of the Aperon vector memory system. Proximity graphs (e.g. |
| APEX4: Efficient Pure W4A4 LLM Inference via Intra-SM Compute Rebalancing | Hong Guo, Nianhui Guo, Weixing Wang, Jona Otholt, Christoph Meinel, Haojin Yang | 2026-06-07 | 下载 | W4A4 quantization promises full utilization of INT4 Tensor Cores, yet group dequantization overhead on CUDA Cores has driven existing systems to mixed-precision fallbacks. |
| SpectrumKV: Per-Token Mixed-Precision KV Cache Transfer for Prefill-Decode Disaggregated LLM Serving | Yang Pengju | 2026-06-07 | 下载 | Prefill-decode (PD) disaggregation decouples prompt processing from token generation, but it also turns the key-value (KV) cache into a network payload. |
| Auditable Graph-Guided Root Cause Analysis for Kubernetes Incidents | Anastasiia Kuvshinova, Seungmin Jin | 2026-06-07 | 下载 | Kubernetes incidents are diagnosed reliably only when a root-cause system's reported gains come from incident evidence rather than scenario-specific shortcuts. |
| Unifying von-Neumann HPC and Neuromorphic Acceleration via the EBRAINS Research Infrastructure: A Framework for High-Performance Workflows | Krishna Kant Singh, Charl Linssen, Eric Müller, Eleni Mathioulaki, Wouter Klijn, Lena Oden | 2026-06-07 | 下载 | Modern scientific workflows increasingly span diverse computing architectures, yet executing a single computational model across disparate systems often forces researchers to maintain fragmented, site... |
| FlashCP: Load-Balanced Communication-Efficient Context Parallelism for LLM Training | Zheng Wang, Eric Liu, Linan Jiang, Zhongkai Yu, Zaifeng Pan, Yue Guan, Yuke Wang, Yufei Ding | 2026-06-07 | 下载 | Context parallelism (CP) is essential for training large-scale, long-context language models, as it partitions sequences to reduce memory overhead. |
cs.NI - Networking and Internet Architecture
| 标题 | 作者 | 发布日期 | 摘要 | |
|---|---|---|---|---|
| SCOPE: A Syndrome-Driven Control Plane for QEC-Enabled Quantum Networks | Xiaojie Fan, Zian Wang, Ashutosh Tiwari, Himanshu Gupta | 2026-06-07 | 下载 | As quantum networks evolve from experimental testbeds to fault-tolerant systems, the primary performance metric shifts from physical link fidelity to end-to-end logical error rate. |
| Systems-Level Planning and Coordination of Truck-Drone Collaborative Delivery Networks | Didem Cicek, Burak Kantarci | 2026-06-07 | 下载 | Urban last-mile parcel delivery increasingly relies on heterogeneous fleets whose performance depends on timely coordination, reliable communication, and scalable control. |
| Block coordinate descent for joint delay-energy optimization in multi-hop D2D networks | Kai-Xiang Hu, Jacek Gondzio, Caixia Kou | 2026-06-07 | 下载 | In multi-hop device-to-device (D2D) networks, the optimization of network-level metrics is particularly difficult due to the tight coupling between network-layer routing and physical-layer resource al... |
cs.PF - Performance
| 标题 | 作者 | 发布日期 | 摘要 | |
|---|---|---|---|---|
| An Empirical Comparison of General Context-Free Parsers | Huan Vo, Danushka Liyanage, Hong Jin Kang, Sasha Rubin, Rahul Gopinath | 2026-06-07 | 下载 | Parsing underpins a vast range of software engineering tasks, from compilers and static analyzers to language servers and fuzz testing tools. Yet most parsers deployed in practice are deterministic (L... |