Skip to content

2026-08-08 ​

cs.AR - Architecture ​

标题作者发布日期PDF摘要
Ego-OSCAR: Egocentric Open source Stereo CAptuRe SystemGunjan Paul, Senthil Palanisamy, Satpal Singh Rathore, Pratyush Kumar Patnaik, Shubhanshu Khatana, Abhishek Anand2026-08-08下载We present Ego-OSCAR, an open-hardware, low-cost, head-mounted stereo-inertial capture device for egocentric data collection in the wild. EgoOSCAR pairs a hardware-synchronized global-shutter stereo c...
Hybrid ASIC-FPAA Fabric for Performance Security Trade-offZiyi Chen, Vaibhav Venugopal Rao, Kyle Juretus, Ioannis Savidis2026-08-08下载In this paper, a hybrid architecture that combines an application specific integrated circuit (ASIC) and a field-programmable analog array (ASIC-FPAA) is proposed to address the performance-security t...

cs.DC - Distributed, Parallel, and Cluster Computing ​

标题作者发布日期PDF摘要
OpRAG: A Resource-Deterministic Runtime for GPU-Backed Multi-Stage RAG WorkflowsArup Kumar Sarker, Mills Staylor, Aymen Alsaadi, Gregor von Laszewski, Shantenu Jha, Geoffrey Fox2026-08-08下载Agentic retrieval-augmented generation (RAG) systems combine preprocessing, embedding, retrieval, memory access, context construction, generation, and vector-index updates.
Truly Work-efficient Parallel Deterministic (Δ+1)-coloring and Maximal Independent SetChase Hutton, Adam Melrod2026-08-08下载We give deterministic parallel algorithms that compute a (Δ+1)-coloring and a maximal independent set for a simple graph with nn vertices and mm edges in O(n+m)O(n+m) work and O(polylogn)O(\mathrm{poly}\log n)...
What Irregularity Costs: CUDA C++, Rust, and Triton on a Hash-Blocked GPU WorkloadPetr Korolev2026-08-08下载GPU language comparisons are almost always run on tiled dense linear algebra, where every toolchain is good and the differences are small. We implement the same hash-blocked TSDF fusion kernel in CUDA...
SAGE: SLO-Aware Adaptive Retrieval for Production RAG SystemsMuhammad Faizan Raza, Shuo, Yang, Satish Mahadevan Srinivasan2026-08-08下载Retrieval-Augmented Generation (RAG) systems in production operate under strict service level objectives (SLOs) on tail latency and infrastructure cost.
Evaluating and Improving Weak Scalability Analysis of Visualization AlgorithmsMarvin Petersen, Jonas Lukasczyk, Christoph Garth2026-08-08下载Research on visualizing large-scale datasets traditionally relies on empirical evaluation of scalability, determining the effectiveness of specific computation methods, algorithmic strategies, or impl...
Hierarchical Multi-Task Federated Learning in VANETsM. Saeid HaghighiFard, Sinem Coleri2026-08-08下载Vehicular Ad hoc Networks (VANETs) increasingly rely on federated learning (FL) to enable collaborative intelligence without sharing raw sensory data.
OasisKV: Scaling In-Decode KV Cache Beyond HBM with Lookahead Sparse PrefetchingCan Xiao, Sukmin Cho, Junbong We, Zhixiong Niu, Jianyi Cheng, Yiren Zhao, Youngjin Kwon, Yongqiang Xiong, Rui Ma, Junyi Liu2026-08-08下载Large language model (LLM) inference serving is increasingly constrained by memory rather than compute. As long-context and long-form reasoning workloads become more prevalent, the key-value (KV) cach...
Effect of Abstractions and Prompting Strategies on LLM-Guided High-Performance OptimizationsJiří Klepl, Maty'aš Brabec, Martin Kruliš2026-08-08下载Code performance optimization is a vital aspect of modern software development, as it enables faster response times and reduced resource usage.
eIRWR: Enhanced Iterative Random Walk with Restart for Scalable Root Cause Analysis in MicroservicesSaiful Khan, Afrah Farea2026-08-08下载Root cause analysis (RCA) in microservice architectures needs to pinpoint the originating faulty service responsible for the cascading symptoms seen across hundreds or thousands of interdependent serv...
ZeroLock: Concurrent Memory-Efficient LLM Training via Modular Update DecouplingWentao Dai, Xuanran Li, Yuxiang Zhang, Ming Tang, Chao Huang2026-08-08下载Large language model (LLM) fine-tuning at the edge adapts the model to scenario-specific data while preserving privacy. Although existing studies proposed pipeline parallelism to address the limited m...
ElastiCo: Elastic Configuration and Interference-Aware Orchestration for GPU ClustersJinghao Wang, Yihang Zhou, Xiaoyang Sun, Chunming Hu, Tianyu Wo, Xu Wang, Albert Y. Zomaya, Renyu Yang2026-08-08下载Modern GPU clusters must simultaneously serve deep learning training and offline large language model inference workloads, yet existing schedulers treat these as isolated resource consumers with rigid...
Directed Neuro-Symbolic Stochastic Execution for Verification of Distributed Parallel AI ProgramsGautham Koorma, Vikas Sharma, George Edwards, Mahdi Eslamimehr2026-08-08下载Distributed parallel Artificial Intelligence (AI) programs expose reliability gaps that conventional testing cannot close: parallel executions are non-deterministic, and AI workloads bring high-dimens...
ScaleSense: Cost-Intelligent Scaling Framework via Learned Resource Estimation in Alibaba AnalyticDBYifan Wu, Yuhan Li, Zhenhua Wang, Ke Chen, Lidan Shou, Zonghao Chen, Liang Lin, Huan Li, Gang Chen2026-08-08下载Cloud-native serverless data warehouses achieve fine-grained elasticity by decoupling storage from compute, yet determining the optimal resource allocation for highly heterogeneous ad-hoc queries rema...

cs.NI - Networking and Internet Architecture ​

标题作者发布日期PDF摘要
WirelessOpsAgent: A Benchmark and Agent Design for Action Assurance in Wireless NetworksZijian Lu, Yiping Zuo, Hao Xu, Weicong Chen, Xin He, Jiajia Guo, Shi Jin2026-08-08下载Large language model (LLM) agents are emerging as planners for autonomous wireless network operations. Yet a task answer that is correct at proposal time can still be unsafe at execution time if suppo...
Orchestrated Vulnerability Management for Heterogeneous Networks: Adaptive Two-Stage Vulnerability Assessment, Context-Aware Risk Prioritization, and Automated MitigationRicardo Lopes, Jose Moura, Rui Neto Marinheiro2026-08-08下载Heterogeneous networks pose significant security challenges due to device diversity, fragile operating conditions, and heterogeneous firmware and service configurations.
Hierarchical Multi-Task Federated Learning in VANETsM. Saeid HaghighiFard, Sinem Coleri2026-08-08下载Vehicular Ad hoc Networks (VANETs) increasingly rely on federated learning (FL) to enable collaborative intelligence without sharing raw sensory data.
Decentralized Multi-Agent Urban Traffic Management via Spatio-Temporal Mobility Profile PlanningLorenzo Mario Amorosa, Lorenzo Farina, Vittorio Todisco, Alessandro Bazzi2026-08-08下载As modern cities face increasingly severe traffic congestion, connected and autonomous vehicles (CAVs) have emerged as a crucial enabling technology for next-generation intelligent traffic management.

cs.OS - Operating Systems ​

标题作者发布日期PDF摘要
Velosiraptor: Code Synthesis for Memory TranslationReto Achermann, Em Chu, Ryan Mehri, Ilias Karimalis, Margo Seltzer2026-08-08下载Security is among the top concerns of operating system (OS) developers. A secure runtime environment relies on the OS to correctly configure the memory hardware on which it runs.

cs.PF - Performance ​

标题作者发布日期PDF摘要
What Irregularity Costs: CUDA C++, Rust, and Triton on a Hash-Blocked GPU WorkloadPetr Korolev2026-08-08下载GPU language comparisons are almost always run on tiled dense linear algebra, where every toolchain is good and the differences are small. We implement the same hash-blocked TSDF fusion kernel in CUDA...
When Does Trace-Driven Evaluation Mislead MoE Expert Caching? Replay Semantics, Workload Contamination, and Operating RegimesYu Zhang2026-08-08下载Mixture-of-Experts (MoE) models have outgrown accelerator memory, and offloading expert weights to host memory is now standard. This makes expert cache management an attractive lever: a policy that ra...

基于 VitePress 构建 · 使用本地搜索查找论文