Distributed and Cluster Computing
Papers filed under cs.DC on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.
Search paper metadata (including unsummarized papers)
301 to 360 of 1,072
Toward Edge General Intelligence with Multiple-Large Language Model (Multi-LLM): Architecture, Trust, and Orchestration
Haoxiang Luo, Yinqiu Liu, Ruichen Zhang +7
cs.NIcs.DCarXiv:2507.00672v12025Federated Learning for Cyber Physical Systems: A Comprehensive Survey
Minh K. Quan, Pubudu N. Pathirana, Mayuri Wijayasundara +5
cs.LGcs.AIcs.CRarXiv:2505.04873v12025AI Flow: Perspectives, Scenarios, and Approaches
Hongjun An, Wenhan Hu, Sida Huang +11
cs.AIcs.CLcs.CVarXiv:2506.12479v32025Federated Learning for Healthcare Domain - Pipeline, Applications and Challenges
Madhura Joshi, Ankit Pal, Malaikannan Sankarasubbu
cs.LGcs.AIcs.CRarXiv:2211.07893v22022Distributed Deep Learning Using Synchronous Stochastic Gradient Descent
Dipankar Das, Sasikanth Avancha, Dheevatsa Mudigere +5
cs.DCcs.LGarXiv:1602.06709v12016A Time-driven Data Placement Strategy for a Scientific Workflow Combining Edge Computing and Cloud Computing
Bing Lin, Fangning Zhu, Jianshan Zhang +4
cs.DCarXiv:1901.07216v22019Deep Learning on FPGAs: Past, Present, and Future
Griffin Lacey, Graham W. Taylor, Shawki Areibi
cs.DCcs.LGstat.MLarXiv:1602.04283v12016Towards Efficient Generative Large Language Model Serving: A Survey from Algorithms to Systems
Xupeng Miao, Gabriele Oliaro, Zhihao Zhang +4
cs.LGcs.AIcs.DCarXiv:2312.15234v22023RT-HiSS: Ray Tracing Accelerated High Dimensional Vector Similarity Searches
Revanth Reddy Munugala, Michael Gowanlock
cs.DCcs.DBarXiv:2609.01975v12026LongCat-Flash Technical Report
Meituan LongCat Team, Bayan, Bei Li +179
cs.CLcs.AIcs.DCarXiv:2509.01322v22025Distributionally Robust Optimization for Aerial Multi-access Edge Computing via Cooperation of UAVs and HAPs
Ziye Jia, Can Cui, Chao Dong +4
cs.DCcs.ITeess.SParXiv:2506.01972v12025Performance Evaluation of RED-ONION: A High-Speed Disk-to-Disk Transfer System
Keichi Takahashi, Hiroaki Kataoka, Takeo Hosomi +5
cs.DCcs.NIarXiv:2608.29053v12026LoongServe: Efficiently Serving Long-Context Large Language Models with Elastic Sequence Parallelism
Bingyang Wu, Shengyu Liu, Yinmin Zhong +3
cs.DCcs.LGarXiv:2404.09526v22024Asynchronous Distributed Optimization using a Randomized Alternating Direction Method of Multipliers
Franck Iutzeler, Pascal Bianchi, Philippe Ciblat +1
cs.DCmath.OCarXiv:1303.2837v12013MedCache: Efficient and Temporally Valid Memory for Longitudinal Clinical Agents
Hei Ting, Chan, Chenwei Wu +5
cs.LGcs.DCcs.MAarXiv:2608.29528v12026SZ3: A Modular Framework for Composing Prediction-Based Error-Bounded Lossy Compressors
Xin Liang, Kai Zhao, Sheng Di +9
cs.DCarXiv:2111.02925v22021MNN: A Universal and Efficient Inference Engine
Xiaotang Jiang, Huan Wang, Yiliu Chen +9
cs.CVcs.DCcs.LGarXiv:2002.12418v12020Automatic Optimization for MapReduce Programs
Eaman Jahani, Michael J. Cafarella, Christopher Ré
cs.DBcs.DCarXiv:1104.3217v12011Astra: A Multi-Agent System for GPU Kernel Performance Optimization
Anjiang Wei, Tianran Sun, Yogesh Seenichamy +5
cs.DCcs.AIcs.CLarXiv:2509.07506v22025BlendCAC: A BLockchain-ENabled Decentralized Capability-based Access Control for IoTs
Ronghua Xu, Yu Chen, Erik Blasch +1
cs.NIcs.CRcs.DCarXiv:1804.09267v12018TAPAS: Thermal- and Power-Aware Scheduling for LLM Inference in Cloud Platforms
Jovan Stojkovic, Chaojie Zhang, Íñigo Goiri +5
cs.DCcs.AIarXiv:2501.02600v12025Load Balancing for MapReduce-based Entity Resolution
Lars Kolb, Andreas Thor, Erhard Rahm
cs.DCarXiv:1108.1631v12011B$^3$-PWL: GPU-Batched Branch-and-Bound for Piecewise-Linear Optimization with SOS2 Constraints
Yilin Guan, Shuqing Luo, Pingzhi Li +2
math.OCcs.DCarXiv:2608.28988v12026Blink: Fast and Generic Collectives for Distributed ML
Guanhua Wang, Shivaram Venkataraman, Amar Phanishayee +3
cs.DCcs.LGarXiv:1910.04940v12019KVFlow: Efficient Prefix Caching for Accelerating LLM-Based Multi-Agent Workflows
Zaifeng Pan, Ajjkumar Patel, Zhengding Hu +6
cs.DCcs.MAarXiv:2507.07400v12025RECIPE : Converting Concurrent DRAM Indexes to Persistent-Memory Indexes
Se Kwon Lee, Jayashree Mohan, Sanidhya Kashyap +2
cs.DCcs.DBcs.DSarXiv:1909.13670v42019An Exponential Separation Between Randomized and Deterministic Complexity in the LOCAL Model
Yi-Jun Chang, Tsvi Kopelowitz, Seth Pettie
cs.CCcs.DCcs.DSarXiv:1602.08166v22016DILAND: An Algorithm for Distributed Sensor Localization with Noisy Distance Measurements
Usman A. Khan, Soummya Kar, Jose M. F. Moura
cs.DCcs.ITarXiv:0910.2743v12009A Performance Analysis of You Only Look Once Models for Deployment on Constrained Computational Edge Devices in Drone Applications
Lucas Rey, Ana M. Bernardos, Andrzej D. Dobrzycki +4
cs.DCcs.AIcs.CVarXiv:2502.15737v12025AInfer-PD: Communication-Safe In-Place Prefill-Decode Multiplexing for Distributed MoE Rollouts
Guowei Wang, Chaokun Yang, Zhenxuan Pan +5
cs.DCarXiv:2609.00993v12026Serving deep learning models in a serverless platform
Vatche Ishakian, Vinod Muthusamy, Aleksander Slominski
cs.DCarXiv:1710.08460v22017Distributed Online Convex Optimization with Time-Varying Coupled Inequality Constraints
Xinlei Yi, Xiuxian Li, Lihua Xie +1
math.OCcs.DCcs.LGarXiv:1903.04277v22019A Unified Theory of Decentralized SGD with Changing Topology and Local Updates
Anastasia Koloskova, Nicolas Loizou, Sadra Boreiri +2
cs.LGcs.DCmath.OCarXiv:2003.10422v32020Federated Learning of a Mixture of Global and Local Models
Filip Hanzely, Peter Richtárik
cs.LGcs.DCmath.OCarXiv:2002.05516v32020PAPI: Exploiting Dynamic Parallelism in Large Language Model Decoding with a Processing-In-Memory-Enabled Computing System
Yintao He, Haiyu Mao, Christina Giannoula +6
cs.ARcs.AIcs.DCarXiv:2502.15470v22025A Smallest-Need-First Job Scheduling Framework with Adaptive Optimization of Idle Node Counts for Energy-Efficient HPC Systems
Reza Pulungan, Raka Satya Prasasta, Santana Yuda Pradata +3
cs.DCarXiv:2608.29656v12026Speeding Up Distributed Machine Learning Using Codes
Kangwook Lee, Maximilian Lam, Ramtin Pedarsani +2
cs.DCcs.ITcs.LGarXiv:1512.02673v32015DynamoLLM: Designing LLM Inference Clusters for Performance and Energy Efficiency
Jovan Stojkovic, Chaojie Zhang, Íñigo Goiri +2
cs.AIcs.ARcs.DCarXiv:2408.00741v12024Summaries:한국어AIOpsLab: A Holistic Framework to Evaluate AI Agents for Enabling Autonomous Clouds
Yinfang Chen, Manish Shetty, Gagan Somashekar +6
cs.AIcs.DCcs.MAarXiv:2501.06706v12025On-Device Machine Learning: An Algorithms and Learning Theory Perspective
Sauptik Dhar, Junyao Guo, Jiayi Liu +3
cs.LGcs.DCstat.MLarXiv:1911.00623v22019Comet: Fine-grained Computation-communication Overlapping for Mixture-of-Experts
Shulai Zhang, Ningxin Zheng, Haibin Lin +9
cs.DCcs.AIcs.LGarXiv:2502.19811v32025KVCache Cache in the Wild: Characterizing and Optimizing KVCache Cache at a Large Cloud Provider
Jiahao Wang, Jinbo Han, Xingda Wei +6
cs.DCcs.AIarXiv:2506.02634v52025ITBench: Evaluating AI Agents across Diverse Real-World IT Automation Tasks
Saurabh Jha, Rohan Arora, Yuji Watanabe +40
cs.AIcs.DCcs.MAarXiv:2502.05352v12025ServerlessLLM: Low-Latency Serverless Inference for Large Language Models
Yao Fu, Leyang Xue, Yeqi Huang +4
cs.LGcs.DCarXiv:2401.14351v22024Ithemal: Accurate, Portable and Fast Basic Block Throughput Estimation using Deep Neural Networks
Charith Mendis, Alex Renda, Saman Amarasinghe +1
cs.DCcs.LGstat.MLarXiv:1808.07412v22018Blockchain-Based Decentralized Energy Management Platform for Residential Distributed Energy Resources in A Virtual Power Plant
Qing Yang, Hao Wang, Taotao Wang +3
eess.SYcs.CRcs.DCarXiv:2105.00174v22021Computing Resource Allocation in Three-Tier IoT Fog Networks: a Joint Optimization Approach Combining Stackelberg Game and Matching
Huaqing Zhang, Yong Xiao, Shengrong Bu +3
cs.GTcs.DCarXiv:1701.03922v12017Parallel Algorithms for Geometric Graph Problems
Alexandr Andoni, Aleksandar Nikolov, Krzysztof Onak +1
cs.DScs.DCarXiv:1401.0042v22013Scaling Nakamoto Consensus to Thousands of Transactions per Second
Chenxing Li, Peilun Li, Dong Zhou +3
cs.DCarXiv:1805.03870v42018Deploying a Top-100 Supercomputer for Large Parallel Workloads: the Niagara Supercomputer
Marcelo Ponce, Ramses van Zon, Scott Northrup +14
cs.DCarXiv:1907.13600v12019Autellix: An Efficient Serving Engine for LLM Agents as General Programs
Michael Luo, Xiaoxiang Shi, Colin Cai +8
cs.LGcs.AIcs.DCarXiv:2502.13965v12025Faster Convergence of Multidimensional Approximate Agreement via Smallest Enclosing Balls
Darya Melnyk
cs.DCarXiv:2609.01490v12026Demystifying NCCL: An In-depth Analysis of GPU Communication Protocols and Algorithms
Zhiyi Hu, Siyuan Shen, Tommaso Bonato +6
cs.DCarXiv:2507.04786v32025CUDA-L1: Improving CUDA Optimization via Contrastive Reinforcement Learning
Xiaoya Li, Albert Wang, Guoyin Wang +2
cs.AIcs.DCcs.LGarXiv:2507.14111v122025Comparison of Algebraic Block Multi-Coloring and Leiden Methods for Parallel Preconditioning in the ICCG Method
Tomohiro Suzuki
cs.DCmath.NAarXiv:2609.00561v12026A Survey of Distributed Data Aggregation Algorithms
Paulo Jesus, Carlos Baquero, Paulo Sérgio Almeida
cs.DCcs.DScs.IRarXiv:1110.0725v12011GSPMD: General and Scalable Parallelization for ML Computation Graphs
Yuanzhong Xu, HyoukJoong Lee, Dehao Chen +13
cs.DCcs.LGarXiv:2105.04663v22021Lambada: Interactive Data Analytics on Cold Data using Serverless Cloud Infrastructure
Ingo Müller, Renato Marroquín, Gustavo Alonso
cs.DBcs.DCarXiv:1912.00937v12019Linear Convergence in Federated Learning: Tackling Client Heterogeneity and Sparse Gradients
Aritra Mitra, Rayana Jaafar, George J. Pappas +1
cs.LGcs.DCeess.SYarXiv:2102.07053v22021Prediction-Robust Service Deployment with Capacity-Aware Edge Admission
Hailiang Zhao, Ziqi Wang, Yifei Zhang +4
cs.DCarXiv:2609.00877v12026