2026-06-01
cs.DC - Distributed, Parallel, and Cluster Computing
| 标题 | 作者 | 发布日期 | 摘要 | |
|---|---|---|---|---|
| GreenGNN: Energy-Aware Windowed Communication Optimization for Distributed GNN Training | Arefin Niam, Tevfik Kosar, M. S. Q. Zulkar Nine | 2026-06-01 | 下载 | Large-scale graph neural network (GNN) training often requires distributed clusters because graph structure and feature tensors no longer fit in a single node's memory. |
| AURA: Action-Gated Memory for Robot Policies at Constant VRAM | Josef Chen | 2026-06-01 | 下载 | The KV-cache is the right memory for datacenters but the wrong memory for robots. Datacenter inference batches many short requests and resets them, amortizing an attention cache across a crowd. |
| IntraShuffler: A Privacy Preserving Framework for Heterogeneous DP Federated Learning | Farhin Farhad Riya, Olivera Kotevska, Jinyuan Stella Sun | 2026-06-01 | 下载 | Heterogeneous Differential Privacy (HDP) in Federated Learning (FL) allows clients to select individual privacy budgets () according to institutional policies and data sensitivity. |
| Not All Errors Are Equal: A Systematic Study of Error Propagation in Large Language Model Inference | Yafan Huang, Sheng Di, Guanpeng Li | 2026-06-01 | 下载 | Large language models (LLMs) are increasingly integrated into high-performance computing (HPC) workflows, accelerating scientific discovery through diverse perspectives such as code generation and dom... |
| Strategies for Molecular Dynamics using Hybrid Systems: LAMMPS Use Case | Paulo Henrique Leme Ramalho, Dennis Alves Pedersen, Fábio Andrijauskas | 2026-06-01 | 下载 | The complexity of biomolecular simulations has substantially increased the demand for High-Performance Computing (HPC) infrastructures, particularly in molecular dynamics and coarse-grained modeling. |
| EES-CND: Collaborative Neural Decision-Making for Drift-Aware Fault-Tolerant Edge-Cloud Service Placement | Mohammadsadeq Garshasbi Herabad, Javid Taheri, Bestoun S. Ahmed, Calin Curescu | 2026-06-01 | 下载 | The edge-cloud paradigm improves service delivery by orchestrating resources across edge nodes and cloud data centres. These environments consist of heterogeneous, interconnected computing nodes that ... |
| TAPAAL SMC: Statistical Model Checking of Stochastic Timed-Arc Petri Nets | Tanguy Dubois, Kim G. Larsen, Jiri Srba | 2026-06-01 | 下载 | Timed-Arc Petri net (TAPN) is a timed extension of the classical Petri net model where tokens have their age and input arcs are associated with time intervals restricting the ages of tokens available ... |
| Scaling LLM Inference Beyond Amdahl`s Limits via Eliminating Non-Scalable Overheads | Alan Zhao, Cyril Y. He, Wei Xu | 2026-06-01 | 下载 | Deployers of online LLM services usually seek to maximize cluster-wide performance given a fixed number of GPUs. Tensor parallelism (TP) is necessary to fit modern models but scales sub-linearly as th... |
| Boosting Multimodal Federated Learning via Chained Modality Optimization | Zixin Zhang, Fan Qi, Shuai Li, Xiaoshan Yang, Changsheng Xu | 2026-06-01 | 下载 | Multimodal Federated Learning (MMFL) enables privacy-preserving collaborative learning across decentralized clients with heterogeneous data and modality availability. |
| Parallelizing Large-Scale Tensor Network Contraction on Multiple GPUs | Feng Pan, Hanfeng Gu, Paul Springer, Xipeng Li | 2026-06-01 | 下载 | Exact tensor network contraction underpins quantum circuit simulation, quantum error correction, combinatorial optimization, and many-body dynamics. |
| Observation, Not Prediction: Conversation-Level Disaggregated Scheduling for Agentic Serving | Jianru Ding, Ryien Hosseini, Pouya Mahdi Gholami, Mingyuan Xiang, Henry Hoffmann | 2026-06-01 | 下载 | LLM-based agents resolve a user task through many turns of dependent inference and tool calls, producing a workload whose total cost is unknown when the task arrives. |
| Night-Window Batching versus Carbon-Aware Scheduling for Clinical AI GPU Workloads | Nishi Doshi, Shrey Shah | 2026-06-01 | 下载 | Hospitals run more machine learning on GPUs while the carbon footprint of grid electricity rises and falls through the day. Using a computer simulation, we compare scheduling rules on mixed GPU h... |
| Post-Deterministic Distributed Systems: A New Foundation for Trustworthy Autonomous Infrastructure | Jun He, Deying Yu | 2026-06-01 | 下载 | For decades, distributed systems have typically assumed that correct participants execute protocol-specified behavior with stable, externally defined, and deterministic semantics. |
| Scalable Concurrent Queues for GPU | Pratheek Prakash Shetty, Thomas R. W. Scogland, Wu-chun Feng | 2026-06-01 | 下载 | Concurrent queues can significantly impact supercomputing performance by being critical bottlenecks for task distribution, load balancing, and resource utilization. |
| Don't Let a Few Network Failures Slow the Entire AllReduce | Peiqing Chen, Jiedong Jiang, Nengneng Yu, Yuefeng Wang, Sixian Xiong, Wei Wang, Zaoxing Liu | 2026-06-01 | 下载 | Network failures are among the most frequent hardware faults in large-scale GPU clusters and a leading cause of training-job interruptions. Modern collective communication libraries such as NCCL mitig... |
| A Sheaf Framework for Strategic Multi-Agent Systems: From Consensus to Nash Equilibria | Manuel Hernández, Eduardo Sánchez-Soto | 2026-06-01 | 下载 | The coordination of heterogeneous autonomous agents in dynamic, adversarial environments requires simultaneous satisfaction of geometric constraints, logical consistency, temporal reasoning, and strat... |
| TwinQuant: Learnable Subspace Decomposition for 4-Bit LLM Quantization | Haodong Wang, Junjie Liu, Zicong Hong, Qianli Liu, Jian Lin, Song Guo, Xu Chen | 2026-06-01 | 下载 | 4-bit quantization reduces the memory footprint and latency of large language model inference, but its aggressive precision reduction can severely degrade accuracy. |
| Self-Conditioned Positional HNSW for Overlap-Aware Retrieval in Chunked-Document RAG Systems: Method and Industrial Evidence-Quality Audit | Nataraj Agaram Sundar, Tejas Morabia | 2026-06-01 | 下载 | Chunked-document retrieval is a common component of retrieval-augmented generation (RAG) systems. Documents are split into overlapping chunks, embedded, and indexed with approximate nearest-neighbor s... |
| Compliance-Scored Best-of-N Guardrail Orchestration for Multimodal Document Generation in Payments Dispute Defense | Nataraj Agaram Sundar, Tejas Morabia | 2026-06-01 | 下载 | High-stakes enterprise document generation, including financial dispute narratives, compliance notices, and audit summaries, demands schema correctness, policy compliance, and low-latency operation at... |