Distributed and Cluster Computing

Papers filed under cs.DC on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.

Search paper metadata (including unsummarized papers)

301 to 360 of 1,072

  1. Toward Edge General Intelligence with Multiple-Large Language Model (Multi-LLM): Architecture, Trust, and Orchestration

    Haoxiang Luo, Yinqiu Liu, Ruichen Zhang +7

    cs.NIcs.DCarXiv:2507.00672v12025
  2. Federated Learning for Cyber Physical Systems: A Comprehensive Survey

    Minh K. Quan, Pubudu N. Pathirana, Mayuri Wijayasundara +5

    cs.LGcs.AIcs.CRarXiv:2505.04873v12025
  3. AI Flow: Perspectives, Scenarios, and Approaches

    Hongjun An, Wenhan Hu, Sida Huang +11

    cs.AIcs.CLcs.CVarXiv:2506.12479v32025
  4. Federated Learning for Healthcare Domain - Pipeline, Applications and Challenges

    Madhura Joshi, Ankit Pal, Malaikannan Sankarasubbu

    cs.LGcs.AIcs.CRarXiv:2211.07893v22022
  5. Distributed Deep Learning Using Synchronous Stochastic Gradient Descent

    Dipankar Das, Sasikanth Avancha, Dheevatsa Mudigere +5

    cs.DCcs.LGarXiv:1602.06709v12016
  6. A Time-driven Data Placement Strategy for a Scientific Workflow Combining Edge Computing and Cloud Computing

    Bing Lin, Fangning Zhu, Jianshan Zhang +4

    cs.DCarXiv:1901.07216v22019
  7. Deep Learning on FPGAs: Past, Present, and Future

    Griffin Lacey, Graham W. Taylor, Shawki Areibi

    cs.DCcs.LGstat.MLarXiv:1602.04283v12016
  8. Towards Efficient Generative Large Language Model Serving: A Survey from Algorithms to Systems

    Xupeng Miao, Gabriele Oliaro, Zhihao Zhang +4

    cs.LGcs.AIcs.DCarXiv:2312.15234v22023
  9. RT-HiSS: Ray Tracing Accelerated High Dimensional Vector Similarity Searches

    Revanth Reddy Munugala, Michael Gowanlock

    cs.DCcs.DBarXiv:2609.01975v12026
  10. LongCat-Flash Technical Report

    Meituan LongCat Team, Bayan, Bei Li +179

    cs.CLcs.AIcs.DCarXiv:2509.01322v22025
  11. Distributionally Robust Optimization for Aerial Multi-access Edge Computing via Cooperation of UAVs and HAPs

    Ziye Jia, Can Cui, Chao Dong +4

    cs.DCcs.ITeess.SParXiv:2506.01972v12025
  12. Performance Evaluation of RED-ONION: A High-Speed Disk-to-Disk Transfer System

    Keichi Takahashi, Hiroaki Kataoka, Takeo Hosomi +5

    cs.DCcs.NIarXiv:2608.29053v12026
  13. LoongServe: Efficiently Serving Long-Context Large Language Models with Elastic Sequence Parallelism

    Bingyang Wu, Shengyu Liu, Yinmin Zhong +3

    cs.DCcs.LGarXiv:2404.09526v22024
  14. Asynchronous Distributed Optimization using a Randomized Alternating Direction Method of Multipliers

    Franck Iutzeler, Pascal Bianchi, Philippe Ciblat +1

    cs.DCmath.OCarXiv:1303.2837v12013
  15. MedCache: Efficient and Temporally Valid Memory for Longitudinal Clinical Agents

    Hei Ting, Chan, Chenwei Wu +5

    cs.LGcs.DCcs.MAarXiv:2608.29528v12026
  16. SZ3: A Modular Framework for Composing Prediction-Based Error-Bounded Lossy Compressors

    Xin Liang, Kai Zhao, Sheng Di +9

    cs.DCarXiv:2111.02925v22021
  17. MNN: A Universal and Efficient Inference Engine

    Xiaotang Jiang, Huan Wang, Yiliu Chen +9

    cs.CVcs.DCcs.LGarXiv:2002.12418v12020
  18. Automatic Optimization for MapReduce Programs

    Eaman Jahani, Michael J. Cafarella, Christopher Ré

    cs.DBcs.DCarXiv:1104.3217v12011
  19. Astra: A Multi-Agent System for GPU Kernel Performance Optimization

    Anjiang Wei, Tianran Sun, Yogesh Seenichamy +5

    cs.DCcs.AIcs.CLarXiv:2509.07506v22025
  20. BlendCAC: A BLockchain-ENabled Decentralized Capability-based Access Control for IoTs

    Ronghua Xu, Yu Chen, Erik Blasch +1

    cs.NIcs.CRcs.DCarXiv:1804.09267v12018
  21. TAPAS: Thermal- and Power-Aware Scheduling for LLM Inference in Cloud Platforms

    Jovan Stojkovic, Chaojie Zhang, Íñigo Goiri +5

    cs.DCcs.AIarXiv:2501.02600v12025
  22. Load Balancing for MapReduce-based Entity Resolution

    Lars Kolb, Andreas Thor, Erhard Rahm

    cs.DCarXiv:1108.1631v12011
  23. B$^3$-PWL: GPU-Batched Branch-and-Bound for Piecewise-Linear Optimization with SOS2 Constraints

    Yilin Guan, Shuqing Luo, Pingzhi Li +2

    math.OCcs.DCarXiv:2608.28988v12026
  24. Blink: Fast and Generic Collectives for Distributed ML

    Guanhua Wang, Shivaram Venkataraman, Amar Phanishayee +3

    cs.DCcs.LGarXiv:1910.04940v12019
  25. KVFlow: Efficient Prefix Caching for Accelerating LLM-Based Multi-Agent Workflows

    Zaifeng Pan, Ajjkumar Patel, Zhengding Hu +6

    cs.DCcs.MAarXiv:2507.07400v12025
  26. RECIPE : Converting Concurrent DRAM Indexes to Persistent-Memory Indexes

    Se Kwon Lee, Jayashree Mohan, Sanidhya Kashyap +2

    cs.DCcs.DBcs.DSarXiv:1909.13670v42019
  27. An Exponential Separation Between Randomized and Deterministic Complexity in the LOCAL Model

    Yi-Jun Chang, Tsvi Kopelowitz, Seth Pettie

    cs.CCcs.DCcs.DSarXiv:1602.08166v22016
  28. DILAND: An Algorithm for Distributed Sensor Localization with Noisy Distance Measurements

    Usman A. Khan, Soummya Kar, Jose M. F. Moura

    cs.DCcs.ITarXiv:0910.2743v12009
  29. A Performance Analysis of You Only Look Once Models for Deployment on Constrained Computational Edge Devices in Drone Applications

    Lucas Rey, Ana M. Bernardos, Andrzej D. Dobrzycki +4

    cs.DCcs.AIcs.CVarXiv:2502.15737v12025
  30. AInfer-PD: Communication-Safe In-Place Prefill-Decode Multiplexing for Distributed MoE Rollouts

    Guowei Wang, Chaokun Yang, Zhenxuan Pan +5

    cs.DCarXiv:2609.00993v12026
  31. Serving deep learning models in a serverless platform

    Vatche Ishakian, Vinod Muthusamy, Aleksander Slominski

    cs.DCarXiv:1710.08460v22017
  32. Distributed Online Convex Optimization with Time-Varying Coupled Inequality Constraints

    Xinlei Yi, Xiuxian Li, Lihua Xie +1

    math.OCcs.DCcs.LGarXiv:1903.04277v22019
  33. A Unified Theory of Decentralized SGD with Changing Topology and Local Updates

    Anastasia Koloskova, Nicolas Loizou, Sadra Boreiri +2

    cs.LGcs.DCmath.OCarXiv:2003.10422v32020
  34. Federated Learning of a Mixture of Global and Local Models

    Filip Hanzely, Peter Richtárik

    cs.LGcs.DCmath.OCarXiv:2002.05516v32020
  35. PAPI: Exploiting Dynamic Parallelism in Large Language Model Decoding with a Processing-In-Memory-Enabled Computing System

    Yintao He, Haiyu Mao, Christina Giannoula +6

    cs.ARcs.AIcs.DCarXiv:2502.15470v22025
  36. A Smallest-Need-First Job Scheduling Framework with Adaptive Optimization of Idle Node Counts for Energy-Efficient HPC Systems

    Reza Pulungan, Raka Satya Prasasta, Santana Yuda Pradata +3

    cs.DCarXiv:2608.29656v12026
  37. Speeding Up Distributed Machine Learning Using Codes

    Kangwook Lee, Maximilian Lam, Ramtin Pedarsani +2

    cs.DCcs.ITcs.LGarXiv:1512.02673v32015
  38. DynamoLLM: Designing LLM Inference Clusters for Performance and Energy Efficiency

    Jovan Stojkovic, Chaojie Zhang, Íñigo Goiri +2

    cs.AIcs.ARcs.DCarXiv:2408.00741v12024
    Summaries:한국어
  39. AIOpsLab: A Holistic Framework to Evaluate AI Agents for Enabling Autonomous Clouds

    Yinfang Chen, Manish Shetty, Gagan Somashekar +6

    cs.AIcs.DCcs.MAarXiv:2501.06706v12025
  40. On-Device Machine Learning: An Algorithms and Learning Theory Perspective

    Sauptik Dhar, Junyao Guo, Jiayi Liu +3

    cs.LGcs.DCstat.MLarXiv:1911.00623v22019
  41. Comet: Fine-grained Computation-communication Overlapping for Mixture-of-Experts

    Shulai Zhang, Ningxin Zheng, Haibin Lin +9

    cs.DCcs.AIcs.LGarXiv:2502.19811v32025
  42. KVCache Cache in the Wild: Characterizing and Optimizing KVCache Cache at a Large Cloud Provider

    Jiahao Wang, Jinbo Han, Xingda Wei +6

    cs.DCcs.AIarXiv:2506.02634v52025
  43. ITBench: Evaluating AI Agents across Diverse Real-World IT Automation Tasks

    Saurabh Jha, Rohan Arora, Yuji Watanabe +40

    cs.AIcs.DCcs.MAarXiv:2502.05352v12025
  44. ServerlessLLM: Low-Latency Serverless Inference for Large Language Models

    Yao Fu, Leyang Xue, Yeqi Huang +4

    cs.LGcs.DCarXiv:2401.14351v22024
  45. Ithemal: Accurate, Portable and Fast Basic Block Throughput Estimation using Deep Neural Networks

    Charith Mendis, Alex Renda, Saman Amarasinghe +1

    cs.DCcs.LGstat.MLarXiv:1808.07412v22018
  46. Blockchain-Based Decentralized Energy Management Platform for Residential Distributed Energy Resources in A Virtual Power Plant

    Qing Yang, Hao Wang, Taotao Wang +3

    eess.SYcs.CRcs.DCarXiv:2105.00174v22021
  47. Computing Resource Allocation in Three-Tier IoT Fog Networks: a Joint Optimization Approach Combining Stackelberg Game and Matching

    Huaqing Zhang, Yong Xiao, Shengrong Bu +3

    cs.GTcs.DCarXiv:1701.03922v12017
  48. Parallel Algorithms for Geometric Graph Problems

    Alexandr Andoni, Aleksandar Nikolov, Krzysztof Onak +1

    cs.DScs.DCarXiv:1401.0042v22013
  49. Scaling Nakamoto Consensus to Thousands of Transactions per Second

    Chenxing Li, Peilun Li, Dong Zhou +3

    cs.DCarXiv:1805.03870v42018
  50. Deploying a Top-100 Supercomputer for Large Parallel Workloads: the Niagara Supercomputer

    Marcelo Ponce, Ramses van Zon, Scott Northrup +14

    cs.DCarXiv:1907.13600v12019
  51. Autellix: An Efficient Serving Engine for LLM Agents as General Programs

    Michael Luo, Xiaoxiang Shi, Colin Cai +8

    cs.LGcs.AIcs.DCarXiv:2502.13965v12025
  52. Faster Convergence of Multidimensional Approximate Agreement via Smallest Enclosing Balls

    Darya Melnyk

    cs.DCarXiv:2609.01490v12026
  53. Demystifying NCCL: An In-depth Analysis of GPU Communication Protocols and Algorithms

    Zhiyi Hu, Siyuan Shen, Tommaso Bonato +6

    cs.DCarXiv:2507.04786v32025
  54. CUDA-L1: Improving CUDA Optimization via Contrastive Reinforcement Learning

    Xiaoya Li, Albert Wang, Guoyin Wang +2

    cs.AIcs.DCcs.LGarXiv:2507.14111v122025
  55. Comparison of Algebraic Block Multi-Coloring and Leiden Methods for Parallel Preconditioning in the ICCG Method

    Tomohiro Suzuki

    cs.DCmath.NAarXiv:2609.00561v12026
  56. A Survey of Distributed Data Aggregation Algorithms

    Paulo Jesus, Carlos Baquero, Paulo Sérgio Almeida

    cs.DCcs.DScs.IRarXiv:1110.0725v12011
  57. GSPMD: General and Scalable Parallelization for ML Computation Graphs

    Yuanzhong Xu, HyoukJoong Lee, Dehao Chen +13

    cs.DCcs.LGarXiv:2105.04663v22021
  58. Lambada: Interactive Data Analytics on Cold Data using Serverless Cloud Infrastructure

    Ingo Müller, Renato Marroquín, Gustavo Alonso

    cs.DBcs.DCarXiv:1912.00937v12019
  59. Linear Convergence in Federated Learning: Tackling Client Heterogeneity and Sparse Gradients

    Aritra Mitra, Rayana Jaafar, George J. Pappas +1

    cs.LGcs.DCeess.SYarXiv:2102.07053v22021
  60. Prediction-Robust Service Deployment with Capacity-Aware Edge Admission

    Hailiang Zhao, Ziqi Wang, Yifei Zhang +4

    cs.DCarXiv:2609.00877v12026