Distributed and Cluster Computing
Papers filed under cs.DC on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.
Search paper metadata (including unsummarized papers)
241 to 300 of 1,071
Para-Pipe: Exploiting Hierarchical Operator Parallelism of ML Computational Graphs on SoCs
Yujie Zhang, Huiying Lan, Ehsan Aghapour +5
cs.DCcs.LGcs.PFarXiv:2609.04168v12026Improving Progressive Compression with Adaptive Interpolation and Coefficient Decomposition
Wenbo Li, Xuan Wu, Qian Gong +6
cs.DCarXiv:2609.04573v12026Heterogeneous LoRA for Federated Fine-tuning of On-Device Foundation Models
Yae Jee Cho, Luyang Liu, Zheng Xu +2
cs.LGcs.DCarXiv:2401.06432v22024Collaborative On-Sensor Array Cameras
Jipeng Sun, Kaixuan Wei, Thomas Eboli +6
physics.opticscs.CVcs.DCarXiv:2506.04061v12025CloudGenius: Decision Support for Web Server Cloud Migration
Michael Menzel, Rajiv Ranjan
cs.DCcs.SEarXiv:1203.3997v12012Thinking Like a Vertex: a Survey of Vertex-Centric Frameworks for Distributed Graph Processing
Robert Ryan McCune, Tim Weninger, Gregory Madey
cs.DCarXiv:1507.04405v12015CONCUR: High-Throughput Agentic Batch Inference of LLM via Congestion-Based Concurrency Control
Qiaoling Chen, Zhisheng Ye, Tian Tang +7
cs.DCarXiv:2601.22705v12026Graph-Structured Deep Learning Framework for Multi-task Contention Identification with High-dimensional Metrics
Xiao Yang, Yinan Ni, Yuqi Tang +3
cs.DCcs.LGarXiv:2601.20389v12026Modern Computing: Vision and Challenges
Sukhpal Singh Gill, Huaming Wu, Panos Patros +22
cs.DCarXiv:2401.02469v12024Same Request, Different Answer: Quantization Amplifies Cache-Induced Divergence in LLM Serving
Aditi Patodiya
cs.SEcs.DCcs.LGarXiv:2609.04748v12026Solution-space heterogeneity shapes federated learning dynamics across partial differential equations
Ping Luo, Jiahuan Wang, Ziqing Wen +2
cs.LGcs.DCarXiv:2609.05012v12026Resilience Beyond Stationary Client Unavailability: Unlocking Efficient and Unbiased Federated Learning
Ming Xiang, Stratis Ioannidis, Edmund Yeh +2
cs.LGcs.DCmath.OCarXiv:2609.04763v12026Optimizing CNN Model Inference on CPUs
Yizhi Liu, Yao Wang, Ruofei Yu +3
cs.DCarXiv:1809.02697v32018Edge Impulse: An MLOps Platform for Tiny Machine Learning
Shawn Hymel, Colby Banbury, Daniel Situnayake +13
cs.DCcs.LGcs.SEarXiv:2212.03332v32022Scalable Training of Mixture-of-Experts Models with Megatron Core
Zijie Yan, Hongxiao Bai, Xin Yao +42
cs.DCcs.CLcs.LGarXiv:2603.07685v22026SparkNet: Training Deep Networks in Spark
Philipp Moritz, Robert Nishihara, Ion Stoica +1
stat.MLcs.DCcs.LGarXiv:1511.06051v42015ECHO: Elastic Speculative Decoding with Sparse Gating for High-Concurrency Scenarios
Xinyi Hu, Yuhao Shen, Baolin Zhang +6
cs.DCcs.AIcs.LGarXiv:2604.09603v22026HetPipe: Enabling Large DNN Training on (Whimpy) Heterogeneous GPU Clusters through Integration of Pipelined Model Parallelism and Data Parallelism
Jay H. Park, Gyeongchan Yun, Chang M. Yi +5
cs.DCarXiv:2005.14038v12020A Unifying Framework for Parallel and Distributed Processing in R using Futures
Henrik Bengtsson
cs.DCstat.COarXiv:2008.00553v42020FedTGP: Trainable Global Prototypes with Adaptive-Margin-Enhanced Contrastive Learning for Data and Model Heterogeneity in Federated Learning
Jianqing Zhang, Yang Liu, Yang Hua +1
cs.LGcs.CRcs.DCarXiv:2401.03230v12024Optimal Deterministic Routing and Sorting on the Congested Clique
Christoph Lenzen
cs.DCarXiv:1207.1852v42012Application Management in Fog Computing Environments: A Taxonomy, Review and Future Directions
Redowan Mahmud, Kotagiri Ramamohanarao, Rajkumar Buyya
cs.DCeess.SParXiv:2005.10460v12020Split Learning for collaborative deep learning in healthcare
Maarten G. Poirot, Praneeth Vepakomma, Ken Chang +3
cs.LGcs.DCstat.MLarXiv:1912.12115v12019Model Accuracy and Runtime Tradeoff in Distributed Deep Learning:A Systematic Study
Suyog Gupta, Wei Zhang, Fei Wang
stat.MLcs.DCcs.LGarXiv:1509.04210v32015LLM-42: Enabling Determinism in LLM Inference with Verified Speculation
Raja Gond, Aditya K Kamath, Ramachandran Ramjee +1
cs.LGcs.AIcs.DCarXiv:2601.17768v22026Decision Support Tools for Cloud Migration in the Enterprise
Ali Khajeh-Hosseini, Ian Sommerville, Jurgen Bogaerts +1
cs.DCarXiv:1105.0149v12011Joint Temporal-Structural Representation Learning for Distributed Fault Discrimination in Microservice Architectures
Yihan Xue, Yuxiao Wang, Ao Zhu +2
cs.DCarXiv:2605.01776v12026GNNAdvisor: An Adaptive and Efficient Runtime System for GNN Acceleration on GPUs
Yuke Wang, Boyuan Feng, Gushu Li +4
cs.DCarXiv:2006.06608v32020Communication-Computation Efficient Gradient Coding
Min Ye, Emmanuel Abbe
stat.MLcs.DCcs.ITarXiv:1802.03475v12018Reference Architecture of a Quantum-Centric Supercomputer
Seetharami Seelam, Jerry M. Chow, Antonio Córcoles +10
cs.ETcs.ARcs.DCarXiv:2603.10970v22026Prefill-as-a-Service: KVCache of Next-Generation Models Could Go Cross-Datacenter
Ruoyu Qin, Weiran He, Yaoyu Wang +5
cs.DCarXiv:2604.15039v22026First Analysis of Local GD on Heterogeneous Data
Ahmed Khaled, Konstantin Mishchenko, Peter Richtárik
cs.LGcs.DCmath.NAarXiv:1909.04715v22019LLMServingSim 2.0: A Unified Simulator for Heterogeneous and Disaggregated LLM Serving Infrastructure
Jaehong Cho, Hyunmin Choi, Guseul Heo +1
cs.DCcs.AIarXiv:2602.23036v22026LightLDA: Big Topic Models on Modest Compute Clusters
Jinhui Yuan, Fei Gao, Qirong Ho +6
stat.MLcs.DCcs.IRarXiv:1412.1576v12014KernelFoundry: Hardware-aware evolutionary GPU kernel optimization
Nina Wiedemann, Quentin Leboutet, Michael Paulitsch +2
cs.DCcs.LGarXiv:2603.12440v22026Compute and Storage Clouds Using Wide Area High Performance Networks
Robert L. Grossman, Yunhong Gu, Michael Sabala +1
cs.DCarXiv:0808.1802v12008AIConfigurator: Lightning-Fast Configuration Optimization for Multi-Framework LLM Serving
Tianhao Xu, Yiming Liu, Xianglong Lu +18
cs.LGcs.AIcs.DCarXiv:2601.06288v12026Characterization of Large Language Model Development in the Datacenter
Qinghao Hu, Zhisheng Ye, Zerui Wang +9
cs.DCcs.LGarXiv:2403.07648v22024Ankh: Optimized Protein Language Model Unlocks General-Purpose Modelling
Ahmed Elnaggar, Hazem Essam, Wafaa Salah-Eldin +4
cs.LGcs.CLcs.DCarXiv:2301.06568v12023Parallelizing Tool Execution and LLM Generation for Low-Latency Agent Serving
Yifan Sui, Han Zhao, Rui Ma +6
cs.DCcs.AIarXiv:2603.18897v32026RLlib: Abstractions for Distributed Reinforcement Learning
Eric Liang, Richard Liaw, Philipp Moritz +6
cs.AIcs.DCcs.LGarXiv:1712.09381v42017Effectively Prefetching Remote Memory with Leap
Hasan Al Maruf, Mosharaf Chowdhury
cs.DCcs.OSarXiv:1911.09829v12019Jolteon and Ditto: Network-Adaptive Efficient Consensus with Asynchronous Fallback
Rati Gelashvili, Lefteris Kokoris-Kogias, Alberto Sonnino +2
cs.DCcs.CRarXiv:2106.10362v42021Where Do the Joules Go? Diagnosing Inference Energy Consumption
Jae-Won Chung, Ruofan Wu, Jeff J. Ma +1
cs.LGcs.DCarXiv:2601.22076v22026PowerSlider: Exploiting Phase Asymmetry for LLM Serving under Demand Response
Yueying Li, Jiayang Chen, Yuanfan Chen +5
cs.DCcs.AIarXiv:2608.21719v12026Mobility-Aware Cooperative Caching in Vehicular Edge Computing Based on Asynchronous Federated and Deep Reinforcement Learning
Qiong Wu, Yu Zhao, Qiang Fan +3
cs.DCcs.LGarXiv:2208.01219v12022Algorithmic Simplification for Million-Vertex Diffusion History Reconstruction
Gökhan Göktürk
cs.SIcs.DCarXiv:2608.28955v12026EC-SAGINs: Edge Computing-enhanced Space-Air-Ground Integrated Networks for Internet of Vehicles
Shuai Yu, Xiaowen Gong, Qian Shi +2
cs.NIcs.AIcs.DCarXiv:2101.06056v12021Energy-Aware Load Balancing in Content Delivery Networks
Vimal Mathew, Ramesh K. Sitaraman, Prashant Shenoy
cs.NIcs.DCarXiv:1109.5641v12011Mind the Memory Gap: Unveiling GPU Bottlenecks in Large-Batch LLM Inference
Pol G. Recasens, Ferran Agullo, Yue Zhu +5
cs.DCcs.LGarXiv:2503.08311v22025Atomix: Timely, Transactional Tool Use for Reliable Agentic Workflows
Bardia Mohammadi, Nearchos Potamitis, Lars Klein +2
cs.LGcs.AIcs.DCarXiv:2602.14849v22026MedPerf: Open Benchmarking Platform for Medical Artificial Intelligence using Federated Evaluation
Alexandros Karargyris, Renato Umeton, Micah J. Sheller +39
cs.LGcs.DCcs.PFarXiv:2110.01406v32021Datacenter Traffic Control: Understanding Techniques and Trade-offs
Mohammad Noormohammadpour, Cauligi S. Raghavendra
cs.NIcs.DCcs.PFarXiv:1712.03530v12017A Unified Coding Framework for Distributed Computing with Straggling Servers
Songze Li, Mohammad Ali Maddah-Ali, A. Salman Avestimehr
cs.ITcs.DCarXiv:1609.01690v12016TurboTransformers: An Efficient GPU Serving System For Transformer Models
Jiarui Fang, Yang Yu, Chengduo Zhao +1
cs.DCcs.AIcs.LGarXiv:2010.05680v42020vLLM-Omni: Fully Disaggregated Serving for Any-to-Any Multimodal Models
Peiqi Yin, Jiangyun Zhu, Han Gao +13
cs.DCarXiv:2602.02204v12026Theoretically Efficient Parallel Graph Algorithms Can Be Fast and Scalable
Laxman Dhulipala, Guy E. Blelloch, Julian Shun
cs.DScs.DCarXiv:1805.05208v42018Dorylus: Affordable, Scalable, and Accurate GNN Training with Distributed CPU Servers and Serverless Threads
John Thorpe, Yifan Qiao, Jonathan Eyolfson +8
cs.DCcs.LGarXiv:2105.11118v22021Measurement of Generative AI Workload Power Profiles for Whole-Facility Data Center Infrastructure Planning
Roberto Vercellino, Jared Willard, Gustavo Campos +4
eess.SYcs.DCcs.LGarXiv:2604.07345v12026Toward Edge General Intelligence with Multiple-Large Language Model (Multi-LLM): Architecture, Trust, and Orchestration
Haoxiang Luo, Yinqiu Liu, Ruichen Zhang +7
cs.NIcs.DCarXiv:2507.00672v12025