2026-08-08
cs.AR - Architecture
| 标题 | 作者 | 发布日期 | 摘要 | |
|---|---|---|---|---|
| Ego-OSCAR: Egocentric Open source Stereo CAptuRe System | Gunjan Paul, Senthil Palanisamy, Satpal Singh Rathore, Pratyush Kumar Patnaik, Shubhanshu Khatana, Abhishek Anand | 2026-08-08 | 下载 | We present Ego-OSCAR, an open-hardware, low-cost, head-mounted stereo-inertial capture device for egocentric data collection in the wild. EgoOSCAR pairs a hardware-synchronized global-shutter stereo c... |
| Hybrid ASIC-FPAA Fabric for Performance Security Trade-off | Ziyi Chen, Vaibhav Venugopal Rao, Kyle Juretus, Ioannis Savidis | 2026-08-08 | 下载 | In this paper, a hybrid architecture that combines an application specific integrated circuit (ASIC) and a field-programmable analog array (ASIC-FPAA) is proposed to address the performance-security t... |
cs.DC - Distributed, Parallel, and Cluster Computing
| 标题 | 作者 | 发布日期 | 摘要 | |
|---|---|---|---|---|
| OpRAG: A Resource-Deterministic Runtime for GPU-Backed Multi-Stage RAG Workflows | Arup Kumar Sarker, Mills Staylor, Aymen Alsaadi, Gregor von Laszewski, Shantenu Jha, Geoffrey Fox | 2026-08-08 | 下载 | Agentic retrieval-augmented generation (RAG) systems combine preprocessing, embedding, retrieval, memory access, context construction, generation, and vector-index updates. |
| Truly Work-efficient Parallel Deterministic (Δ+1)-coloring and Maximal Independent Set | Chase Hutton, Adam Melrod | 2026-08-08 | 下载 | We give deterministic parallel algorithms that compute a (Δ+1)-coloring and a maximal independent set for a simple graph with vertices and edges in work and ... |
| What Irregularity Costs: CUDA C++, Rust, and Triton on a Hash-Blocked GPU Workload | Petr Korolev | 2026-08-08 | 下载 | GPU language comparisons are almost always run on tiled dense linear algebra, where every toolchain is good and the differences are small. We implement the same hash-blocked TSDF fusion kernel in CUDA... |
| SAGE: SLO-Aware Adaptive Retrieval for Production RAG Systems | Muhammad Faizan Raza, Shuo, Yang, Satish Mahadevan Srinivasan | 2026-08-08 | 下载 | Retrieval-Augmented Generation (RAG) systems in production operate under strict service level objectives (SLOs) on tail latency and infrastructure cost. |
| Evaluating and Improving Weak Scalability Analysis of Visualization Algorithms | Marvin Petersen, Jonas Lukasczyk, Christoph Garth | 2026-08-08 | 下载 | Research on visualizing large-scale datasets traditionally relies on empirical evaluation of scalability, determining the effectiveness of specific computation methods, algorithmic strategies, or impl... |
| Hierarchical Multi-Task Federated Learning in VANETs | M. Saeid HaghighiFard, Sinem Coleri | 2026-08-08 | 下载 | Vehicular Ad hoc Networks (VANETs) increasingly rely on federated learning (FL) to enable collaborative intelligence without sharing raw sensory data. |
| OasisKV: Scaling In-Decode KV Cache Beyond HBM with Lookahead Sparse Prefetching | Can Xiao, Sukmin Cho, Junbong We, Zhixiong Niu, Jianyi Cheng, Yiren Zhao, Youngjin Kwon, Yongqiang Xiong, Rui Ma, Junyi Liu | 2026-08-08 | 下载 | Large language model (LLM) inference serving is increasingly constrained by memory rather than compute. As long-context and long-form reasoning workloads become more prevalent, the key-value (KV) cach... |
| Effect of Abstractions and Prompting Strategies on LLM-Guided High-Performance Optimizations | Jiří Klepl, Maty'aš Brabec, Martin Kruliš | 2026-08-08 | 下载 | Code performance optimization is a vital aspect of modern software development, as it enables faster response times and reduced resource usage. |
| eIRWR: Enhanced Iterative Random Walk with Restart for Scalable Root Cause Analysis in Microservices | Saiful Khan, Afrah Farea | 2026-08-08 | 下载 | Root cause analysis (RCA) in microservice architectures needs to pinpoint the originating faulty service responsible for the cascading symptoms seen across hundreds or thousands of interdependent serv... |
| ZeroLock: Concurrent Memory-Efficient LLM Training via Modular Update Decoupling | Wentao Dai, Xuanran Li, Yuxiang Zhang, Ming Tang, Chao Huang | 2026-08-08 | 下载 | Large language model (LLM) fine-tuning at the edge adapts the model to scenario-specific data while preserving privacy. Although existing studies proposed pipeline parallelism to address the limited m... |
| ElastiCo: Elastic Configuration and Interference-Aware Orchestration for GPU Clusters | Jinghao Wang, Yihang Zhou, Xiaoyang Sun, Chunming Hu, Tianyu Wo, Xu Wang, Albert Y. Zomaya, Renyu Yang | 2026-08-08 | 下载 | Modern GPU clusters must simultaneously serve deep learning training and offline large language model inference workloads, yet existing schedulers treat these as isolated resource consumers with rigid... |
| Directed Neuro-Symbolic Stochastic Execution for Verification of Distributed Parallel AI Programs | Gautham Koorma, Vikas Sharma, George Edwards, Mahdi Eslamimehr | 2026-08-08 | 下载 | Distributed parallel Artificial Intelligence (AI) programs expose reliability gaps that conventional testing cannot close: parallel executions are non-deterministic, and AI workloads bring high-dimens... |
| ScaleSense: Cost-Intelligent Scaling Framework via Learned Resource Estimation in Alibaba AnalyticDB | Yifan Wu, Yuhan Li, Zhenhua Wang, Ke Chen, Lidan Shou, Zonghao Chen, Liang Lin, Huan Li, Gang Chen | 2026-08-08 | 下载 | Cloud-native serverless data warehouses achieve fine-grained elasticity by decoupling storage from compute, yet determining the optimal resource allocation for highly heterogeneous ad-hoc queries rema... |
cs.NI - Networking and Internet Architecture
| 标题 | 作者 | 发布日期 | 摘要 | |
|---|---|---|---|---|
| WirelessOpsAgent: A Benchmark and Agent Design for Action Assurance in Wireless Networks | Zijian Lu, Yiping Zuo, Hao Xu, Weicong Chen, Xin He, Jiajia Guo, Shi Jin | 2026-08-08 | 下载 | Large language model (LLM) agents are emerging as planners for autonomous wireless network operations. Yet a task answer that is correct at proposal time can still be unsafe at execution time if suppo... |
| Orchestrated Vulnerability Management for Heterogeneous Networks: Adaptive Two-Stage Vulnerability Assessment, Context-Aware Risk Prioritization, and Automated Mitigation | Ricardo Lopes, Jose Moura, Rui Neto Marinheiro | 2026-08-08 | 下载 | Heterogeneous networks pose significant security challenges due to device diversity, fragile operating conditions, and heterogeneous firmware and service configurations. |
| Hierarchical Multi-Task Federated Learning in VANETs | M. Saeid HaghighiFard, Sinem Coleri | 2026-08-08 | 下载 | Vehicular Ad hoc Networks (VANETs) increasingly rely on federated learning (FL) to enable collaborative intelligence without sharing raw sensory data. |
| Decentralized Multi-Agent Urban Traffic Management via Spatio-Temporal Mobility Profile Planning | Lorenzo Mario Amorosa, Lorenzo Farina, Vittorio Todisco, Alessandro Bazzi | 2026-08-08 | 下载 | As modern cities face increasingly severe traffic congestion, connected and autonomous vehicles (CAVs) have emerged as a crucial enabling technology for next-generation intelligent traffic management. |
cs.OS - Operating Systems
| 标题 | 作者 | 发布日期 | 摘要 | |
|---|---|---|---|---|
| Velosiraptor: Code Synthesis for Memory Translation | Reto Achermann, Em Chu, Ryan Mehri, Ilias Karimalis, Margo Seltzer | 2026-08-08 | 下载 | Security is among the top concerns of operating system (OS) developers. A secure runtime environment relies on the OS to correctly configure the memory hardware on which it runs. |
cs.PF - Performance
| 标题 | 作者 | 发布日期 | 摘要 | |
|---|---|---|---|---|
| What Irregularity Costs: CUDA C++, Rust, and Triton on a Hash-Blocked GPU Workload | Petr Korolev | 2026-08-08 | 下载 | GPU language comparisons are almost always run on tiled dense linear algebra, where every toolchain is good and the differences are small. We implement the same hash-blocked TSDF fusion kernel in CUDA... |
| When Does Trace-Driven Evaluation Mislead MoE Expert Caching? Replay Semantics, Workload Contamination, and Operating Regimes | Yu Zhang | 2026-08-08 | 下载 | Mixture-of-Experts (MoE) models have outgrown accelerator memory, and offloading expert weights to host memory is now standard. This makes expert cache management an attractive lever: a policy that ra... |