2026-04-18
cs.AR - Architecture
| 标题 | 作者 | 发布日期 | 摘要 | |
|---|---|---|---|---|
| Configuration Over Selection: Hyperparameter Sensitivity Exceeds Model Differences in Open-Source LLMs for RTL Generation | Minghao Shao, Zeng Wang, Weimin Fu, Xiaolong Guo, Johann Knechtel, Ozgur Sinanoglu, Ramesh Karri, Muhammad Shafique | 2026-04-18 | 下载 | Benchmarking of open-source LLMs for hardware design focuses on which LLMs to use, while treating inference-time decoding configuration as a secondary concern. |
| From Natural Language to Silicon: The Representation Bottleneck in LLM Hardware Design | Weimin Fu, Zeng Wang, Minghao Shao, Johann Knechtel, Ozgur Sinanoglu, Ramesh Karri, Muhammad Shafique, Xiaolong Guo | 2026-04-18 | 下载 | Edge applications increasingly demand custom hardware, yet Field-Programmable Gate Array (FPGA) design requires expertise that domain engineers lack. |
| When Spike Sparsity Does Not Translate to Deployed Cost: VS-WNO on Jetson Orin Nano | Jason Yoo, Shailesh Garg, Souvik Chakraborty, Syed Bahauddin Alam | 2026-04-18 | 下载 | Spiking neural operators are appealing for neuromorphic edge computing because event-driven substrates can, in principle, translate sparse activity into lower latency and energy. |
| Different Perspectives of Memory System Simulation | Pouya Esmaili-Dokht, Arash Yadegari, Victor Xirau, Julian Pavon, Adrian Cristal, Eduard Ayguade, Petar Radojkovic | 2026-04-18 | 下载 | Memory simulators are used to estimate application performance on advanced memory systems, yet they may exhibit significant discrepancies compared to real hardware. |
| E2AFS: Energy-Efficient Approximate Floating Point Square Rooter for Error Tolerant Computing | Prateek Goyal, Jatin Kumar Reddy Mothe, Swara Rajesh Shelke, Sujit Kumar Sahoo | 2026-04-18 | 下载 | Floating-point square-root computation is a power- and delay-critical operation in edge-AI, signal-processing, and embedded systems. Conventional implementations typically rely on multipliers or itera... |
| HieraSparse: Hierarchical Semi-Structured Sparse KV Attention | Haoxuan Wang, Chen Wang | 2026-04-18 | 下载 | The deployment of long-context Large Language Models (LLMs) poses significant challenges due to the intense computational cost of self-attention and the substantial memory overhead of the Key-Value Ca... |
cs.DC - Distributed, Parallel, and Cluster Computing
| 标题 | 作者 | 发布日期 | 摘要 | |
|---|---|---|---|---|
| HiveMind: OS-Inspired Scheduling for Concurrent LLM Agent Workloads | Justice Owusu Agyemang, Jerry John Kponyo, Obed Kwasi Somuah, Elliot Amponsah, Godfred Manu Addo Boakye, Kwame Opuni-Boachie Obour Agyekum | 2026-04-18 | 下载 | When multiple LLM coding agents share a rate-limited API endpoint, they exhibit resource contention patterns analogous to unscheduled OS processes competing for CPU, memory, and I/O. |
| TensorHub: Rethinking AI Model Hub with Tensor-Centric Compression | Tingfeng Lan, Zirui Wang, Yunjia Zheng, Zhaoyuan Su, Juncheng Yang, Yue Cheng | 2026-04-18 | 下载 | Modern AI models are growing rapidly in size and redundancy, leading to significant storage and distribution challenges in model hubs. We present TensorHub, a tensor-centric system for reducing storag... |
| Sarus Suite: Cloud-native Containers for HPC | Alberto Madonna, Matteo Chesi, Gwangmu Lee, Michele Brambilla, Fawzi Roberto Mohamed, Felipe A. Cruz | 2026-04-18 | 下载 | High-performance computing (HPC) systems must support fast-moving software stacks, especially in AI/ML, while preserving scheduler control, scalable startup, and production performance. |
| Predictive Sectorization and Bayesian Optimized Consensus for Admission Control in Autonomous Airspace Operations | Aditya Dhodapkar, Avery Smidt, Aaron Verkleeren, Stacy Patterson, Carlos A. Varela | 2026-04-18 | 下载 | Conventional air traffic control divides airspace into specific regions, creating a scaling bottleneck as traffic grows. Choosing how to partition airspace is not straightforward because grid size aff... |
| The Cognitive Penalty: Ablating System 1 and System 2 Reasoning in Edge-Native SLMs for Decentralized Consensus | Syed Muhammad Aqdas Rizvi | 2026-04-18 | 下载 | Decentralized Autonomous Organizations (DAOs) are inclined explore Small Language Models (SLMs) as edge-native constitutional firewalls to vet proposals and mitigate semantic social engineering. |
| From Swap Axioms to Weighted Geometric Means: A Characterization of AMMs | Björn Assmann, Ulan Degenbaev | 2026-04-18 | 下载 | Many automated market makers can be understood through the geometry of their trading orbits, the sets of states reachable from one another through swaps. |
| HieraSparse: Hierarchical Semi-Structured Sparse KV Attention | Haoxuan Wang, Chen Wang | 2026-04-18 | 下载 | The deployment of long-context Large Language Models (LLMs) poses significant challenges due to the intense computational cost of self-attention and the substantial memory overhead of the Key-Value Ca... |
cs.NI - Networking and Internet Architecture
| 标题 | 作者 | 发布日期 | 摘要 | |
|---|---|---|---|---|
| Symphony: Taming Step Misalignments in the Network for Ring-based Collective Operations | Yuze Jin, Xin Zhe Khooi, Ruyi Yao, Mun Choon Chan | 2026-04-18 | 下载 | Ring-based collective operations are widely used in distributed AI training due to their efficient bandwidth utilization. While ring communication excels at pipelining, its performance is heavily depe... |
cs.OS - Operating Systems
| 标题 | 作者 | 发布日期 | 摘要 | |
|---|---|---|---|---|
| Governed MCP: Kernel-Level Tool Governance for AI Agents via Logit-Based Safety Primitives | Daeyeon Son | 2026-04-18 | 下载 | AI agents increasingly call external tools (file system, network, APIs) through the Model Context Protocol (MCP). These tool calls are the agent's syscalls -- privileged operations with side effects o... |