2026-04-23
cs.AR - Architecture
| 标题 | 作者 | 发布日期 | 摘要 | |
|---|---|---|---|---|
| SPAC: Automating FPGA-based Network Switches with Protocol Adaptive Customization | Guoyu Li, Yang Cao, Lucas H L Ng, Alexander Charlton, Qianzhou Wang, Will Punter, Philippos Papaphilippou, Ce Guo, Hongxiang Fan, Wayne Luk, Saman Amarasinghe, Ajay Brahmakshatriya | 2026-04-23 | 下载 | With network requirements diverging across emerging applications, latency-critical services demand minimal logic delay, while hyperscale training and collectives require sustained line-rate throughput... |
| On the Role of Preprocessing and Memristor Dynamics in Reservoir Computing for Image Classification | Rishona Daniels, Duna Wattad, Ronny Ronen, David Saad, Shahar Kvatinsky | 2026-04-23 | 下载 | Reservoir computing (RC) is an emerging recurrent neural network architecture that has attracted growing attention for its low training cost and modest hardware requirements. |
| Leveraging SIMD for Accelerating Large-number Arithmetic | Subhrajit Das, Abhishek Bichhawat, Yuvraj Patel | 2026-04-23 | 下载 | Large-number arithmetic, widely used in scientific computing and cryptography, has seen limited adoption of single instruction, multiple data (SIMD) parallelism on modern CPUs due to the inherent depe... |
| Suppressing the Erasure Error of Fusion Operation in Photonic Quantum Computing | Xiangyu Ren, Yuexun Huang, Zhemin Zhang, Yuchen Zhu, Tsung-Yi Ho, Antonio Barbalace, Zhiding Liang | 2026-04-23 | 下载 | Photonic quantum computing provides a promising route toward quantum computation by naturally supporting the measurement-based quantum computation (MBQC) model. |
| Focus Session: Hardware and Software Techniques for Accelerating Multimodal Foundation Models | Muhammad Shafique, Abdul Basit, Muhammad Abdullah Hanif, Alberto Marchisio, Rachmad Vidya Wicaksana Putra, Minghao Shao | 2026-04-23 | 下载 | This work presents a multi-layered methodology for efficiently accelerating multimodal foundation models (MFMs). It combines hardware and software co-design of transformer blocks with an optimization ... |
cs.DC - Distributed, Parallel, and Cluster Computing
| 标题 | 作者 | 发布日期 | 摘要 | |
|---|---|---|---|---|
| FlashSpread: IO-Aware GPU Simulation of Non-Markovian Epidemic Dynamics via Kernel Fusion | Heman Shakeri, Behnaz Moradi-Jamei, Aram Vajdi, Ehsan Ardjmand | 2026-04-23 | 下载 | Non-Markovian (renewal) epidemic simulation on multi-million-node contact networks is essential for realistic forecasting under general age-dependent holding-time distributions (log-normal, Weibull, E... |
| Shard the Gradient, Scale the Model: Serverless Federated Aggregation via Gradient Partitioning | Amine Barrak | 2026-04-23 | 下载 | Federated learning (FL) aggregation on serverless platforms faces a hard scalability ceiling: existing architectures (lambda-FL, LIFL) partition clients across aggregators, but every aggregator must h... |
| Promoting Simple Agents: Ensemble Methods for Event-Log Prediction | Benedikt Bollig, Matthias Függer, Thomas Nowak, Paul Zeinaty | 2026-04-23 | 下载 | We compare lightweight automata-based models (n-grams) with neural architectures (LSTM, Transformer) for next-activity prediction in streaming event logs. |
| Leveraging SIMD for Accelerating Large-number Arithmetic | Subhrajit Das, Abhishek Bichhawat, Yuvraj Patel | 2026-04-23 | 下载 | Large-number arithmetic, widely used in scientific computing and cryptography, has seen limited adoption of single instruction, multiple data (SIMD) parallelism on modern CPUs due to the inherent depe... |
| Systematizing Blockchain Research Themes and Design Patterns: Insights from the University Blockchain Research Initiative (UBRI) | Chien-Chih Chen, Yitian Wang, Emma Nasseri, Yebo Feng, Lauren Weymouth | 2026-04-23 | 下载 | The rapid expansion of blockchain and digital asset ecosystems has intensified the challenge of translating academic research into deployable systems and regulatory frameworks. |
| Risk-Aware and Stable Edge Server Selection Under Network Latency SLOs | Mohan Liyanage, Arnova Abdullah, Eldiyar Zhantileuov, Rolf Schuster | 2026-04-23 | 下载 | We present a lightweight and interpretable decision framework for dynamic edge server selection in latency-critical applications that explicitly accounts for tail risk and switching stability. |
| Research on the efficiency of data loading and storage in Data Lakehouse architectures for the formation of analytical data systems | Ivan Borodii, Halyna Osukhivska | 2026-04-23 | 下载 | The paper presents a study of the efficiency of loading and storing data in the three most common Data Lakehouse systems, including Apache Hudi, Apache Iceberg, and Delta Lake, using Apache Spark as a... |
| A Task Decomposition and Planning Framework for Efficient LLM Inference in AI-Enabled WiFi-Offload Networks | Mingqi Han, Xinghua Sun | 2026-04-23 | 下载 | AI WiFi offload is emerging as a promising approach for providing large language model (LLM) services to resource-constrained wireless devices. |
| GraphLeap: Decoupling Graph Construction and Convolution for Vision GNN Acceleration on FPGA | Anvitha Ramachandran, Dhruv Parikh, Viktor Prasanna | 2026-04-23 | 下载 | Vision Graph Neural Networks (ViGs) represent an image as a graph of patch tokens, enabling adaptive, feature-driven neighborhoods. Unlike CNNs with fixed grid biases or Vision Transformers with globa... |
| Optimizing High-Throughput Distributed Data Pipelines for Reproducible Deep Learning at Scale | Kashish Mittal, Di Yu, Roozbeh Ketabi, Arushi Arora, Brendon Lapp, Peng Zhang | 2026-04-23 | 下载 | Training massive-scale deep learning models on datasets spanning tens of terabytes presents critical challenges in hardware utilization and training reproducibility. |
cs.NI - Networking and Internet Architecture
| 标题 | 作者 | 发布日期 | 摘要 | |
|---|---|---|---|---|
| Learning Coverage- and Power-Optimal Transmitter Placement from Building Maps: A Comparative Study of Direct and Indirect Neural Approaches | Çağkan Yapar | 2026-04-23 | 下载 | Optimal wireless transmitter placement is a central task in radio-network planning, yet exhaustive search becomes prohibitively expensive at scale. |
| SPAC: Automating FPGA-based Network Switches with Protocol Adaptive Customization | Guoyu Li, Yang Cao, Lucas H L Ng, Alexander Charlton, Qianzhou Wang, Will Punter, Philippos Papaphilippou, Ce Guo, Hongxiang Fan, Wayne Luk, Saman Amarasinghe, Ajay Brahmakshatriya | 2026-04-23 | 下载 | With network requirements diverging across emerging applications, latency-critical services demand minimal logic delay, while hyperscale training and collectives require sustained line-rate throughput... |
| Iterative Receiver Processing at Relays in PNC-Enabled Multi-Hop Underwater Acoustic Networks | Gewei Zhang, Deqing Wang, Lizhao You, Xiangming Cai, Liqun Fu | 2026-04-23 | 下载 | Physical-layer network coding (PNC) can increase end-to-end throughput in bi-directional multi-hop underwater acoustic (UWA) networks. However, multipath delay spread and Doppler-induced inter-carrier... |
| Risk-Aware and Stable Edge Server Selection Under Network Latency SLOs | Mohan Liyanage, Arnova Abdullah, Eldiyar Zhantileuov, Rolf Schuster | 2026-04-23 | 下载 | We present a lightweight and interpretable decision framework for dynamic edge server selection in latency-critical applications that explicitly accounts for tail risk and switching stability. |
| A Task Decomposition and Planning Framework for Efficient LLM Inference in AI-Enabled WiFi-Offload Networks | Mingqi Han, Xinghua Sun | 2026-04-23 | 下载 | AI WiFi offload is emerging as a promising approach for providing large language model (LLM) services to resource-constrained wireless devices. |
| An Efficient Wireless iBCI Headstage with Adaptive ADC Sample Rate | Hongyao Liu, Junyi Wang, Jinglong Chen, Liuqun Zhai | 2026-04-23 | 下载 | Implantable Brain-Computer Interfaces (iBCIs) are increasingly pivotal in clinical and daily applications. However, wireless iBCIs face severe constraints in power consumption and data throughput. |
| SparKV: Overhead-Aware KV Cache Loading for Efficient On-Device LLM Inference | Hongyao Liu, Liuqun Zhai, Junyi Wang, Zhengru Fang | 2026-04-23 | 下载 | Efficient inference for on-device Large Language Models (LLMs) remains challenging due to limited hardware resources and the high cost of the prefill stage, which processes the full input context to c... |
cs.PF - Performance
| 标题 | 作者 | 发布日期 | 摘要 | |
|---|---|---|---|---|
| Large-Scale Data Parallelization of Product Quantization and Inverted Indexing Using Dask | Ashley N. Abraham, Andrew Strelzoff, Haley R. Dozier, Althea C. Henslee, Mark A. Chappell | 2026-04-23 | 下载 | Large-scale Nearest Neighbor (NN) search, though widely utilized in the similarity search field, remains challenged by the computational limitations inherent in processing large scale data. |
| An Efficient Wireless iBCI Headstage with Adaptive ADC Sample Rate | Hongyao Liu, Junyi Wang, Jinglong Chen, Liuqun Zhai | 2026-04-23 | 下载 | Implantable Brain-Computer Interfaces (iBCIs) are increasingly pivotal in clinical and daily applications. However, wireless iBCIs face severe constraints in power consumption and data throughput. |
| SparKV: Overhead-Aware KV Cache Loading for Efficient On-Device LLM Inference | Hongyao Liu, Liuqun Zhai, Junyi Wang, Zhengru Fang | 2026-04-23 | 下载 | Efficient inference for on-device Large Language Models (LLMs) remains challenging due to limited hardware resources and the high cost of the prefill stage, which processes the full input context to c... |