Machine Learning

Papers filed under cs.LG on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.

Search paper metadata (including unsummarized papers)

5,161 to 5,220 of 20,175

  1. Overfitting Mechanism and Avoidance in Deep Neural Networks

    Shaeke Salman, Xiuwen Liu

    cs.LGcs.NEstat.MLarXiv:1901.06566v12019
  2. Wireless Communications for Collaborative Federated Learning

    Mingzhe Chen, H. Vincent Poor, Walid Saad +1

    cs.ITcs.LGarXiv:2006.02499v22020
  3. A Unified Framework for Sparse Relaxed Regularized Regression: SR3

    Peng Zheng, Travis Askham, Steven L. Brunton +2

    stat.MLcs.LGmath.OCarXiv:1807.05411v42018
  4. Detection of Coronavirus (COVID-19) Associated Pneumonia based on Generative Adversarial Networks and a Fine-Tuned Deep Transfer Learning Model using Chest X-ray Dataset

    Nour Eldeen M. Khalifa, Mohamed Hamed N. Taha, Aboul Ella Hassanien +1

    eess.IVcs.CVcs.LGarXiv:2004.01184v12020
  5. Deep Reinforcement Learning for Solving the Heterogeneous Capacitated Vehicle Routing Problem

    Jingwen Li, Yining Ma, Ruize Gao +4

    cs.LGmath.OCarXiv:2110.02629v22021
  6. Parameter-Efficient Fine-Tuning for Foundation Models

    Dan Zhang, Tao Feng, Lilong Xue +3

    cs.CLcs.AIcs.LGarXiv:2501.13787v12025
  7. NVIDIA FLARE: Federated Learning from Simulation to Real-World

    Holger R. Roth, Yan Cheng, Yuhong Wen +20

    cs.LGcs.AIcs.CVarXiv:2210.13291v32022
  8. TIPS: Turn-Level Information-Potential Reward Shaping for Search-Augmented LLMs

    Yutao Xie, Nathaniel Thomas, Nicklas Hansen +3

    cs.CLcs.AIcs.LGarXiv:2603.22293v12026
  9. Incompressible Knowledge Probes: Estimating Black-Box LLM Parameter Counts via Factual Capacity

    Bojie Li

    cs.LGcs.AIarXiv:2604.24827v22026
  10. Machine learning approach for early detection of autism by combining questionnaire and home video screening

    Halim Abbas, Ford Garberson, Eric Glover +1

    cs.CYcs.LGarXiv:1703.06076v12017
  11. Optimizing Large Language Model Training Using FP4 Quantization

    Ruizhe Wang, Yeyun Gong, Xiao Liu +5

    cs.LGcs.CLarXiv:2501.17116v22025
  12. Astra: A Multi-Agent System for GPU Kernel Performance Optimization

    Anjiang Wei, Tianran Sun, Yogesh Seenichamy +5

    cs.DCcs.AIcs.CLarXiv:2509.07506v22025
  13. General In-Hand Object Rotation with Vision and Touch

    Haozhi Qi, Brent Yi, Sudharshan Suresh +4

    cs.ROcs.AIcs.CVarXiv:2309.09979v22023
  14. Position: Graph Learning Will Lose Relevance Due To Poor Benchmarks

    Maya Bechler-Speicher, Ben Finkelshtein, Fabrizio Frasca +9

    cs.LGcs.AIcs.NEarXiv:2502.14546v12025
  15. Variational Federated Multi-Task Learning

    Luca Corinzia, Ami Beuret, Joachim M. Buhmann

    cs.LGstat.MLarXiv:1906.06268v22019
  16. Towards Efficient Large Language Model Serving: A Survey on System-Aware KV Cache Optimization

    Jiantong Jiang, Peiyu Yang, Rui Zhang +1

    cs.LGcs.AIcs.CLarXiv:2607.08057v12026
  17. A Generative Deep Learning Approach to Stochastic Downscaling of Precipitation Forecasts

    Lucy Harris, Andrew T. T. McRae, Matthew Chantry +2

    physics.ao-phcs.AIcs.CVarXiv:2204.02028v22022
  18. A Survey of Research in Large Language Models for Electronic Design Automation

    Jingyu Pan, Guanglei Zhou, Chen-Chia Chang +3

    cs.LGarXiv:2501.09655v12025
  19. Graph Neural Networks in Modern AI-aided Drug Discovery

    Odin Zhang, Haitao Lin, Xujun Zhang +9

    q-bio.BMcs.LGarXiv:2506.06915v12025
  20. Oculi: A Conversational Agentic Platform for Automated Credit Risk Analysis

    Vennise Ho, Kristian Diana, Sandy Mourad +3

    cs.AIcs.LGarXiv:2608.28944v12026
  21. Adaptive and Safe Bayesian Optimization in High Dimensions via One-Dimensional Subspaces

    Johannes Kirschner, Mojmír Mutný, Nicole Hiller +2

    cs.LGstat.MLarXiv:1902.03229v22019
  22. Linear Mode Connectivity in Multitask and Continual Learning

    Seyed Iman Mirzadeh, Mehrdad Farajtabar, Dilan Gorur +2

    cs.LGcs.AIcs.CVarXiv:2010.04495v12020
  23. Robust Ensemble Clustering Using Probability Trajectories

    Dong Huang, Jian-Huang Lai, Chang-Dong Wang

    stat.MLcs.LGarXiv:1606.01160v12016
  24. V2TATC: A Joint Voice-Trajectory Embedding Framework and Dataset for Air Traffic Controller Situational Awareness

    Louis Brusset, Mathurin Petit, Jordan Kam +1

    cs.LGeess.ASarXiv:2608.28981v12026
  25. OCGQuant: Outlier-Companion Grouping for NVFP4 Quantization

    Yishan Yao, Binjun Li, Hanling Yi +5

    cs.CLcs.AIcs.LGarXiv:2609.00066v12026
  26. Poly-YOLO: higher speed, more precise detection and instance segmentation for YOLOv3

    Petr Hurtik, Vojtech Molek, Jan Hula +3

    cs.CVcs.LGeess.IVarXiv:2005.13243v22020
  27. Riemannian Continuous Normalizing Flows

    Emile Mathieu, Maximilian Nickel

    stat.MLcs.LGarXiv:2006.10605v22020
  28. A Federated Learning Approach to Anomaly Detection in Smart Buildings

    Raed Abdel Sater, A. Ben Hamza

    cs.LGarXiv:2010.10293v32020
  29. LLaVE: Large Language and Vision Embedding Models with Hardness-Weighted Contrastive Learning

    Zhibin Lan, Liqiang Niu, Fandong Meng +2

    cs.CVcs.AIcs.CLarXiv:2503.04812v22025
  30. PET-MAD, a lightweight universal interatomic potential for advanced materials modeling

    Arslan Mazitov, Filippo Bigi, Matthias Kellner +6

    cond-mat.mtrl-scics.LGphysics.chem-pharXiv:2503.14118v22025
  31. Test-time regression: a unifying framework for designing sequence models with associative memory

    Ke Alexander Wang, Jiaxin Shi, Emily B. Fox

    cs.LGcs.AIcs.NEarXiv:2501.12352v32025
  32. Self-Supervised Learning of State Estimation for Manipulating Deformable Linear Objects

    Mengyuan Yan, Yilin Zhu, Ning Jin +1

    cs.ROcs.CVcs.LGarXiv:1911.06283v32019
  33. Generative Teaching Networks: Accelerating Neural Architecture Search by Learning to Generate Synthetic Training Data

    Felipe Petroski Such, Aditya Rawal, Joel Lehman +2

    cs.LGstat.MLarXiv:1912.07768v12019
  34. Confidence Is All You Need: Few-Shot RL Fine-Tuning of Language Models

    Pengyi Li, Matvey Skripkin, Alexander Zubrey +2

    cs.CLcs.LGarXiv:2506.06395v32025
  35. Multi-Agent Reinforcement Learning via Double Averaging Primal-Dual Optimization

    Hoi-To Wai, Zhuoran Yang, Zhaoran Wang +1

    cs.LGmath.OCstat.MLarXiv:1806.00877v42018
  36. Reducing SO(3) Convolutions to SO(2) for Efficient Equivariant GNNs

    Saro Passaro, C. Lawrence Zitnick

    cs.LGphysics.chem-phphysics.comp-pharXiv:2302.03655v22023
  37. Hardness-Aware Deep Metric Learning

    Wenzhao Zheng, Zhaodong Chen, Jiwen Lu +1

    cs.CVcs.LGarXiv:1903.05503v22019
  38. Single Model Deep Learning on Imbalanced Small Datasets for Skin Lesion Classification

    Peng Yao, Shuwei Shen, Mengjuan Xu +6

    cs.CVcs.LGarXiv:2102.01284v22021
  39. Large Language Models to Enhance Bayesian Optimization

    Tennison Liu, Nicolás Astorga, Nabeel Seedat +1

    cs.LGcs.AIarXiv:2402.03921v22024
  40. UserBench: An Interactive Gym Environment for User-Centric Agents

    Cheng Qian, Zuxin Liu, Akshara Prabhakar +9

    cs.AIcs.CLcs.LGarXiv:2507.22034v12025
  41. Categorical Flow Maps

    Daan Roos, Oscar Davis, Floor Eijkelboom +5

    cs.LGarXiv:2602.12233v12026
  42. PruneShift: A Framework for Evaluating Decision Reliability in Structured Pruning

    Hao Ye, Gaopeng Zhang

    cs.LGcs.NEarXiv:2608.29765v12026
  43. Natural Compression for Distributed Deep Learning

    Samuel Horvath, Chen-Yu Ho, Ludovit Horvath +3

    cs.LGmath.OCstat.MLarXiv:1905.10988v32019
  44. Unsupervised Learning by Competing Hidden Units

    Dmitry Krotov, John Hopfield

    cs.LGcs.CVcs.NEarXiv:1806.10181v22018
  45. Purified OPSD: On-Policy Self-Distillation Without Losing How to Think

    Zhanming Shen, Jintao Tong, Shaotian Yan +9

    cs.AIcs.LGarXiv:2607.02234v12026
  46. Tsunami: A Learned Multi-dimensional Index for Correlated Data and Skewed Workloads

    Jialin Ding, Vikram Nathan, Mohammad Alizadeh +1

    cs.DBcs.LGarXiv:2006.13282v12020
  47. Are Reasoning Models More Prone to Hallucination?

    Zijun Yao, Yantao Liu, Yanxu Chen +5

    cs.CLcs.LGarXiv:2505.23646v12025
  48. Robust Prompt Optimization for Defending Language Models Against Jailbreaking Attacks

    Andy Zhou, Bo Li, Haohan Wang

    cs.LGcs.AIcs.CLarXiv:2401.17263v52024
  49. Generative Pre-Training for Speech with Autoregressive Predictive Coding

    Yu-An Chung, James Glass

    eess.AScs.CLcs.LGarXiv:1910.12607v22019
  50. Is Chain-of-Thought Reasoning of LLMs a Mirage? A Data Distribution Lens

    Chengshuai Zhao, Zhen Tan, Pingchuan Ma +5

    cs.AIcs.CLcs.LGarXiv:2508.01191v62025
  51. LGGNet: Learning from Local-Global-Graph Representations for Brain-Computer Interface

    Yi Ding, Neethu Robinson, Chengxuan Tong +2

    cs.NEcs.LGeess.SParXiv:2105.02786v32021
  52. DeepEMD: Differentiable Earth Mover's Distance for Few-Shot Learning

    Chi Zhang, Yujun Cai, Guosheng Lin +1

    cs.CVcs.LGeess.IVarXiv:2003.06777v52020
  53. Steering Large Language Model Activations in Sparse Spaces

    Reza Bayat, Ali Rahimi-Kalahroudi, Mohammad Pezeshki +2

    cs.LGcs.AIarXiv:2503.00177v12025
  54. BEACON: Behavioral and Semantic Enrichment of AlphaEarth Embeddings through Tri-Modal Contrastive Learning

    Hao Tian, Heng Cai, Yifan Yang

    cs.LGarXiv:2608.29553v12026
  55. Contrastive Code Representation Learning

    Paras Jain, Ajay Jain, Tianjun Zhang +3

    cs.LGcs.AIcs.PLarXiv:2007.04973v42020
  56. The HSIC Bottleneck: Deep Learning without Back-Propagation

    Wan-Duo Kurt Ma, J. P. Lewis, W. Bastiaan Kleijn

    cs.LGstat.MLarXiv:1908.01580v32019
  57. SPA-RL: Reinforcing LLM Agents via Stepwise Progress Attribution

    Hanlin Wang, Chak Tou Leong, Jiashuo Wang +2

    cs.CLcs.LGarXiv:2505.20732v12025
  58. Robustness of Graph Neural Networks at Scale

    Simon Geisler, Tobias Schmidt, Hakan Şirin +3

    cs.LGstat.MLarXiv:2110.14038v42021
  59. Steer LLM Latents for Hallucination Detection

    Seongheon Park, Xuefeng Du, Min-Hsuan Yeh +2

    cs.LGcs.AIcs.CLarXiv:2503.01917v22025
  60. A Note on Shumailov et al. (2024): `AI Models Collapse When Trained on Recursively Generated Data'

    Ali Borji

    cs.LGcs.AIarXiv:2410.12954v22024