Skip to content

2026-04-23 ​

cs.AR - Architecture ​

标题作者发布日期PDF摘要
SPAC: Automating FPGA-based Network Switches with Protocol Adaptive CustomizationGuoyu Li, Yang Cao, Lucas H L Ng, Alexander Charlton, Qianzhou Wang, Will Punter, Philippos Papaphilippou, Ce Guo, Hongxiang Fan, Wayne Luk, Saman Amarasinghe, Ajay Brahmakshatriya2026-04-23下载With network requirements diverging across emerging applications, latency-critical services demand minimal logic delay, while hyperscale training and collectives require sustained line-rate throughput...
On the Role of Preprocessing and Memristor Dynamics in Reservoir Computing for Image ClassificationRishona Daniels, Duna Wattad, Ronny Ronen, David Saad, Shahar Kvatinsky2026-04-23下载Reservoir computing (RC) is an emerging recurrent neural network architecture that has attracted growing attention for its low training cost and modest hardware requirements.
Leveraging SIMD for Accelerating Large-number ArithmeticSubhrajit Das, Abhishek Bichhawat, Yuvraj Patel2026-04-23下载Large-number arithmetic, widely used in scientific computing and cryptography, has seen limited adoption of single instruction, multiple data (SIMD) parallelism on modern CPUs due to the inherent depe...
Suppressing the Erasure Error of Fusion Operation in Photonic Quantum ComputingXiangyu Ren, Yuexun Huang, Zhemin Zhang, Yuchen Zhu, Tsung-Yi Ho, Antonio Barbalace, Zhiding Liang2026-04-23下载Photonic quantum computing provides a promising route toward quantum computation by naturally supporting the measurement-based quantum computation (MBQC) model.
Focus Session: Hardware and Software Techniques for Accelerating Multimodal Foundation ModelsMuhammad Shafique, Abdul Basit, Muhammad Abdullah Hanif, Alberto Marchisio, Rachmad Vidya Wicaksana Putra, Minghao Shao2026-04-23下载This work presents a multi-layered methodology for efficiently accelerating multimodal foundation models (MFMs). It combines hardware and software co-design of transformer blocks with an optimization ...

cs.DC - Distributed, Parallel, and Cluster Computing ​

标题作者发布日期PDF摘要
FlashSpread: IO-Aware GPU Simulation of Non-Markovian Epidemic Dynamics via Kernel FusionHeman Shakeri, Behnaz Moradi-Jamei, Aram Vajdi, Ehsan Ardjmand2026-04-23下载Non-Markovian (renewal) epidemic simulation on multi-million-node contact networks is essential for realistic forecasting under general age-dependent holding-time distributions (log-normal, Weibull, E...
Shard the Gradient, Scale the Model: Serverless Federated Aggregation via Gradient PartitioningAmine Barrak2026-04-23下载Federated learning (FL) aggregation on serverless platforms faces a hard scalability ceiling: existing architectures (lambda-FL, LIFL) partition clients across aggregators, but every aggregator must h...
Promoting Simple Agents: Ensemble Methods for Event-Log PredictionBenedikt Bollig, Matthias Függer, Thomas Nowak, Paul Zeinaty2026-04-23下载We compare lightweight automata-based models (n-grams) with neural architectures (LSTM, Transformer) for next-activity prediction in streaming event logs.
Leveraging SIMD for Accelerating Large-number ArithmeticSubhrajit Das, Abhishek Bichhawat, Yuvraj Patel2026-04-23下载Large-number arithmetic, widely used in scientific computing and cryptography, has seen limited adoption of single instruction, multiple data (SIMD) parallelism on modern CPUs due to the inherent depe...
Systematizing Blockchain Research Themes and Design Patterns: Insights from the University Blockchain Research Initiative (UBRI)Chien-Chih Chen, Yitian Wang, Emma Nasseri, Yebo Feng, Lauren Weymouth2026-04-23下载The rapid expansion of blockchain and digital asset ecosystems has intensified the challenge of translating academic research into deployable systems and regulatory frameworks.
Risk-Aware and Stable Edge Server Selection Under Network Latency SLOsMohan Liyanage, Arnova Abdullah, Eldiyar Zhantileuov, Rolf Schuster2026-04-23下载We present a lightweight and interpretable decision framework for dynamic edge server selection in latency-critical applications that explicitly accounts for tail risk and switching stability.
Research on the efficiency of data loading and storage in Data Lakehouse architectures for the formation of analytical data systemsIvan Borodii, Halyna Osukhivska2026-04-23下载The paper presents a study of the efficiency of loading and storing data in the three most common Data Lakehouse systems, including Apache Hudi, Apache Iceberg, and Delta Lake, using Apache Spark as a...
A Task Decomposition and Planning Framework for Efficient LLM Inference in AI-Enabled WiFi-Offload NetworksMingqi Han, Xinghua Sun2026-04-23下载AI WiFi offload is emerging as a promising approach for providing large language model (LLM) services to resource-constrained wireless devices.
GraphLeap: Decoupling Graph Construction and Convolution for Vision GNN Acceleration on FPGAAnvitha Ramachandran, Dhruv Parikh, Viktor Prasanna2026-04-23下载Vision Graph Neural Networks (ViGs) represent an image as a graph of patch tokens, enabling adaptive, feature-driven neighborhoods. Unlike CNNs with fixed grid biases or Vision Transformers with globa...
Optimizing High-Throughput Distributed Data Pipelines for Reproducible Deep Learning at ScaleKashish Mittal, Di Yu, Roozbeh Ketabi, Arushi Arora, Brendon Lapp, Peng Zhang2026-04-23下载Training massive-scale deep learning models on datasets spanning tens of terabytes presents critical challenges in hardware utilization and training reproducibility.

cs.NI - Networking and Internet Architecture ​

标题作者发布日期PDF摘要
Learning Coverage- and Power-Optimal Transmitter Placement from Building Maps: A Comparative Study of Direct and Indirect Neural ApproachesÇağkan Yapar2026-04-23下载Optimal wireless transmitter placement is a central task in radio-network planning, yet exhaustive search becomes prohibitively expensive at scale.
SPAC: Automating FPGA-based Network Switches with Protocol Adaptive CustomizationGuoyu Li, Yang Cao, Lucas H L Ng, Alexander Charlton, Qianzhou Wang, Will Punter, Philippos Papaphilippou, Ce Guo, Hongxiang Fan, Wayne Luk, Saman Amarasinghe, Ajay Brahmakshatriya2026-04-23下载With network requirements diverging across emerging applications, latency-critical services demand minimal logic delay, while hyperscale training and collectives require sustained line-rate throughput...
Iterative Receiver Processing at Relays in PNC-Enabled Multi-Hop Underwater Acoustic NetworksGewei Zhang, Deqing Wang, Lizhao You, Xiangming Cai, Liqun Fu2026-04-23下载Physical-layer network coding (PNC) can increase end-to-end throughput in bi-directional multi-hop underwater acoustic (UWA) networks. However, multipath delay spread and Doppler-induced inter-carrier...
Risk-Aware and Stable Edge Server Selection Under Network Latency SLOsMohan Liyanage, Arnova Abdullah, Eldiyar Zhantileuov, Rolf Schuster2026-04-23下载We present a lightweight and interpretable decision framework for dynamic edge server selection in latency-critical applications that explicitly accounts for tail risk and switching stability.
A Task Decomposition and Planning Framework for Efficient LLM Inference in AI-Enabled WiFi-Offload NetworksMingqi Han, Xinghua Sun2026-04-23下载AI WiFi offload is emerging as a promising approach for providing large language model (LLM) services to resource-constrained wireless devices.
An Efficient Wireless iBCI Headstage with Adaptive ADC Sample RateHongyao Liu, Junyi Wang, Jinglong Chen, Liuqun Zhai2026-04-23下载Implantable Brain-Computer Interfaces (iBCIs) are increasingly pivotal in clinical and daily applications. However, wireless iBCIs face severe constraints in power consumption and data throughput.
SparKV: Overhead-Aware KV Cache Loading for Efficient On-Device LLM InferenceHongyao Liu, Liuqun Zhai, Junyi Wang, Zhengru Fang2026-04-23下载Efficient inference for on-device Large Language Models (LLMs) remains challenging due to limited hardware resources and the high cost of the prefill stage, which processes the full input context to c...

cs.PF - Performance ​

标题作者发布日期PDF摘要
Large-Scale Data Parallelization of Product Quantization and Inverted Indexing Using DaskAshley N. Abraham, Andrew Strelzoff, Haley R. Dozier, Althea C. Henslee, Mark A. Chappell2026-04-23下载Large-scale Nearest Neighbor (NN) search, though widely utilized in the similarity search field, remains challenged by the computational limitations inherent in processing large scale data.
An Efficient Wireless iBCI Headstage with Adaptive ADC Sample RateHongyao Liu, Junyi Wang, Jinglong Chen, Liuqun Zhai2026-04-23下载Implantable Brain-Computer Interfaces (iBCIs) are increasingly pivotal in clinical and daily applications. However, wireless iBCIs face severe constraints in power consumption and data throughput.
SparKV: Overhead-Aware KV Cache Loading for Efficient On-Device LLM InferenceHongyao Liu, Liuqun Zhai, Junyi Wang, Zhengru Fang2026-04-23下载Efficient inference for on-device Large Language Models (LLMs) remains challenging due to limited hardware resources and the high cost of the prefill stage, which processes the full input context to c...

基于 VitePress 构建 · 使用本地搜索查找论文