Skip to content

2026-04-22 ​

cs.AR - Architecture ​

标题作者发布日期PDF摘要
Enabling Mixed criticality applications for the Versal AI-EnginesVincent Sprave, Martin Wilhelm, Daniele Passaretti, Alberto Garcia-Ortiz, Thilo Pionteck2026-04-22下载Adaptive Systems-on-Chips (SoCs) are increasingly being used in mixed criticality systems (MCSs), such as in autonomous driving, aviation and medical systems.
Efficient Batch Search Algorithm for B+ Tree Index Structures with Level-Wise Traversal on FPGAsMax Tzschoppe, Martin Wilhelm, Sven Groppe, Thilo Pionteck2026-04-22下载This paper introduces a search algorithm for index structures based on a B+ tree, specifically optimized for execution on a field-programmable gate array (FPGA).
Evaluating Computing Platforms for Sustainability: A Comparative Analysis of FPGAs against ASICs, GPUs, and CPUsChetan Choppali Sudarshan, Aman Arora, Vidya A Chhabria2026-04-22下载Climate change concerns emphasize the need for sustainable computing. Modeling the carbon footprint (CFP), including operational and embodied CFP from semiconductor use, manufacture and design, is ess...
PVAC: A RowHammer Mitigation Architecture Exploiting Per-victim-row CountingJumin Kim, Seungmin Baek, Hwayong Nam, Minbok Wi, Nam Sung Kim, Jung Ho Ahn2026-04-22下载As DRAM scaling exacerbates RowHammer, DDR5 introduces per-row activation counting (PRAC) to track aggressor activity. However, PRAC indiscriminately increments counters on every activation -- includi...
A Novel Low-Power Cache Architecture Based on 6-Transistor SRAM CellsNaser Khatti Dizabadi, Ceyda Elcin Kaya2026-04-22下载This paper presents a low-power cache architecture based on the series interconnection of conventional 6-transistor static random-access memory (6T SRAM) cells.
AnalogMaster: Large Language Model-based Automated Analog IC Design Framework from Image to LayoutXian Rong Qin, Yong Zhang, Ying Hu, Tao Su, Bo-Wen Jia, Ning Xu2026-04-22下载Design automation has the potential to substantially improve the efficiency of analog integrated circuit (IC) design. However, existing algorithms and tools typically focus on individual stages, such ...
EnergAIzer: Fast and Accurate GPU Power Estimation Framework for AI WorkloadsKyungmi Lee, Zhiye Song, Eun Kyung Lee, Xin Zhang, Tamar Eilam, Anantha P. Chandrakasan2026-04-22下载As AI workloads drive increases in datacenter power consumption, accurate GPU power estimation is critical for proactive power management. However, existing power models face a scalability bottleneck ...

cs.DC - Distributed, Parallel, and Cluster Computing ​

标题作者发布日期PDF摘要
AGNT2: Autonomous Agent Economies on Interaction-Optimized Layer 2 InfrastructureAnbang Ruan, Xing Zhang2026-04-22下载Current blockchain Layer 2 solutions, including Optimism, Arbitrum, zkSync, and their derivatives, optimize for human-initiated financial transactions.
A Cloud-Native Architecture for Human-in-Control LLM-Assisted OpenSearch in Investigative SettingsBenjamin Puhani, Kai Brehmer, Malte Prieß2026-04-22下载Complex criminal investigations are often hindered by large volumes of unstructured evidence and by the semantic gap between natural language investigative intent and technical search logic.
Enabling Mixed criticality applications for the Versal AI-EnginesVincent Sprave, Martin Wilhelm, Daniele Passaretti, Alberto Garcia-Ortiz, Thilo Pionteck2026-04-22下载Adaptive Systems-on-Chips (SoCs) are increasingly being used in mixed criticality systems (MCSs), such as in autonomous driving, aviation and medical systems.
Efficient Batch Search Algorithm for B+ Tree Index Structures with Level-Wise Traversal on FPGAsMax Tzschoppe, Martin Wilhelm, Sven Groppe, Thilo Pionteck2026-04-22下载This paper introduces a search algorithm for index structures based on a B+ tree, specifically optimized for execution on a field-programmable gate array (FPGA).
TorchGWAS : GPU-accelerated GWAS for thousands of quantitative phenotypesXingzhong Zhao, Ziqian Xie, Islam, Sheikh Muhammad Saiful, Tian Xia, Chen, Cheng, Degui Zhi2026-04-22下载Motivation: Modern bioinformatics workflows, particularly in imaging and representation learning, can generate thousands to tens of thousands of quantitative phenotypes from a single cohort.
Distributed Generative Inference of LLM at Internet Scales with Multi-Dimensional Communication OptimizationJiu Chen, Shuangyan Yang, Xu Xiong, Hexiao Duan, Xinran Zhang, Jie Ren, Dong Li2026-04-22下载Decentralized LLM inference distributes computation among heterogeneous nodes across the internet, offering a performant and cost-efficient solution, alternative to traditional centralized inference.
FedSIR: Spectral Client Identification and Relabeling for Federated Learning with Noisy LabelsSina Gholami, Abdulmoneam Ali, Tania Haghighi, Ahmed Arafa, Minhaj Nur Alam2026-04-22下载Federated learning (FL) enables collaborative model training without sharing raw data; however, the presence of noisy labels across distributed clients can severely degrade the learning performance.
Stream-CQSA: Avoiding Out-of-Memory in Attention Computation via Flexible Workload SchedulingYiming Bian, Joshua M. Akey2026-04-22下载The scalability of long-context large language models is fundamentally limited by the quadratic memory cost of exact self-attention, which often leads to out-of-memory (OOM) failures on modern hardwar...
Distributed Quantum-Enhanced Optimization: A Topographical Preconditioning Approach for High-Dimensional SearchDominik Soós, Marc Paterno, John Stenger, Nikos Chrisochoides2026-04-22下载Optimization problems become fundamentally challenging as the number of variables increases. Because the volume of the search space grows exponentially, classical algorithms frequently fail to locate ...
Distributed Quantum Optimization for Large-Scale Higher-Order Problems with Dense InteractionsSeongmin Kim, Vincent R. Pascuzzi, Travis S. Humble, Thomas Beck, Sanghyo Hwang, Tengfei Luo, Eungkyu Lee, In-Saeng Suh2026-04-22下载Many real-world problems are naturally formulated as higher-order optimization (HUBO) tasks involving dense, multi-variable interactions, which are challenging to solve with classical methods.
FASER: Fine-Grained Phase Management for Speculative Decoding in Dynamic LLM ServingWenyan Chen, Chengzhi Lu, Yanying Lin, Dmitrii Ustiugov2026-04-22下载Speculative decoding (SD) is a widely used approach for accelerating decode-heavy LLM inference workloads. While online inference workloads are highly dynamic, existing SD systems are rigid and take a...
Extending Contract Verification for Parallel Programming Models to FortranYussur Mustafa Oraji, Christian Bischof2026-04-22下载High-performance computing often relies on parallel programming models such as MPI for distributed-memory systems. While powerful, these models are prone to subtle programming errors, leading to devel...
e112: A Context-Aware Mobile Emergency Communication Platform Leveraging Smartphone Sensing and Cloud ServicesKaterina Ioannidou, Marios D. Dikaiakos, Athena Stassopoulou2026-04-22下载This paper presents e112, a context-aware mobile emergency response application designed to strengthen communication between citizens and authorities during disasters.
A Delta-Aware Orchestration Framework for Scalable Multi-Agent Edge ComputingSamaresh Kumar Singh, Joyjit Roy2026-04-22下载The Synergistic Collapse occurs when scaling beyond 100 agents causes superlinear performance degradation that individual optimizations cannot prevent.
Quantum-HPC Software Stacks and the openQSE Reference Architecture: A SurveyAmir Shehata, Brian Austin, Tom Beck, Lukas Burgholzer, Alex Chernoguzov, Spencer Churchill, Andrea Delgado, Yasuko Eckert, Jeffery Heckey, Kevin Kissell, Katherine Klymko, Josh Moles, Thomas Naughton, Lee James O'Riordan, Christian Ortiz Pauyac, Guen Prawiroatmodjo, Ermal Rrapaj, Jiri Schindler, Laura Schulz, Sebastian Stern, Tyler Takeshita, Miwako Tsuji, Aleksander Wennersteen, Travis Humble, Martin Schulz2026-04-22下载Quantum resources are increasingly integrated into high-performance computing (HPC) and cloud environments, but quantum high-performance computing (QHPC) software stacks remain isolated, often proprie...
Characterizing and Fixing Silent Data Loss in Spark-on-AWS-Lambda with Open Table FormatsSrujan Kumar Gandla2026-04-22下载AWS Lambda terminates containers with an uncatchable SIGKILL signal when a function exceeds its configured timeout. When a Spark-on-AWS-Lambda (SoAL) job is killed between Phase 1 (data upload) and Ph...

cs.NI - Networking and Internet Architecture ​

标题作者发布日期PDF摘要
StarLoc: Pinpointing Transmitting LEO Satellites from a Single Passive ArrayIshani Janveja, Jida Zhang, Emerson Sie, Deepak Vasisht2026-04-22下载This paper focuses on 3D localization of transmitting satellites in low Earth orbits (LEO). 3D localization of transmitters in low orbits is an important emerging problem for many applications such as...
Behavioral Consistency and Transparency Analysis on Large Language Model API GatewaysGuanjie Lin, Yinxin Wan, Shichao Pei, Ting Xu, Kuai Xu, Guoliang Xue2026-04-22下载Third-party Large Language Model (LLM) API gateways are rapidly emerging as unified access points to models offered by multiple vendors. However, the internal routing, caching, and billing policies of...
Sema: Semantic Transport for Real-Time Multimodal AgentsJiaying Meng, Bojie Li2026-04-22下载Real-time multimodal agents transport raw audio and screenshots using networking stacks designed for human receivers, which optimize for perceptual fidelity and smooth playout.
Assessing the Challenges of Collective Perception via V2I Communications in High-Speed Scenarios with Open Road TestingJon Ander Iñiguez de Gordoa, Iker Alkorta, Itziar Urbieta, Gorka Velez, Andoni Mujika2026-04-22下载This paper presents a comprehensive end-to-end evaluation of an infrastructure-assisted collective perception (ICP) system deployed on a highway using ITS-G5 technology.
Forecasting Individual NetFlows using a Predictive Masked Graph AutoencoderGeorgios Anyfantis, Pere Barlet-Ros2026-04-22下载In this paper, we propose a proof-of-concept Graph Neural Network model that can successfully predict network flow-level traffic (NetFlow) by accurately modelling the graph structure and the connectio...
Interconnecting Regional QKD Networks: Hybrid Key Delivery Across Quantum DomainsDavid Barral, Aitor Brazaola-Vicario, Diego Cifrián, Natalia Costas, Gonzalo Blázquez, Ana Fernández-Vilas, Iago F. Llovo, Pedro Otero-García, Pablo P. Rejo, Alejandra Ruiz, Juan Villasuso, Manuel Fernández-Veiga2026-04-22下载QKD technology is being increasingly adopted inside the network core for protecting information transport against any form of computational attacks.
Column Generation for the Optimization of Switching in Repeaterless Quantum NetworksÁlvaro Troyano Olivas, Andrés Agustí Casado, Hans H. Brunner, Chi-Hang Fred Fung, Momtchil Peev, Laura Ortiz, Vicente Martin2026-04-22下载Efficient resource allocation and optical switching promise high key rates, network adaptability, and cost reduction in repeaterless quantum communication networks.

cs.PF - Performance ​

标题作者发布日期PDF摘要
A Delta-Aware Orchestration Framework for Scalable Multi-Agent Edge ComputingSamaresh Kumar Singh, Joyjit Roy2026-04-22下载The Synergistic Collapse occurs when scaling beyond 100 agents causes superlinear performance degradation that individual optimizations cannot prevent.

基于 VitePress 构建 · 使用本地搜索查找论文