Machine Learning

Papers filed under cs.LG on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.

Search paper metadata (including unsummarized papers)

541 to 600 of 20,225

  1. A Near-Linear Time Algorithm for the Chamfer Distance

    Ainesh Bakshi, Piotr Indyk, Rajesh Jayaram +2

    cs.DScs.CGcs.GRarXiv:2307.03043v12023
  2. Lipschitz Bandits with Stochastic Delayed Feedback

    Zhongxuan Liu, Yue Kang, Thomas C. M. Lee

    cs.LGstat.MLarXiv:2510.00309v22025
  3. Measuring training variability from stochastic optimization using robust nonparametric testing

    Sinjini Banerjee, Tim Marrinan, Reilly Cannon +2

    stat.MLcs.LGarXiv:2406.08307v22024
  4. Normalization Propagation: A Parametric Technique for Removing Internal Covariate Shift in Deep Networks

    Devansh Arpit, Yingbo Zhou, Bhargava U. Kota +1

    stat.MLcs.LGarXiv:1603.01431v62016
  5. DiagrammerGPT: Generating Open-Domain, Open-Platform Diagrams via LLM Planning

    Abhay Zala, Han Lin, Jaemin Cho +1

    cs.CVcs.AIcs.CLarXiv:2310.12128v22023
  6. Weight decay induces low-rank attention layers

    Seijin Kobayashi, Yassir Akram, Johannes Von Oswald

    cs.LGarXiv:2410.23819v12024
  7. Think in English, Answer in Korean: Efficient Adaptation of Multilingual Tool-Using Agents

    Utsav Garg, Sungjin Hong, Jason Jung +6

    cs.AIcs.LGarXiv:2606.31648v12026
  8. Online convex optimization in the bandit setting: gradient descent without a gradient

    Abraham D. Flaxman, Adam Tauman Kalai, H. Brendan McMahan

    cs.LGcs.CCarXiv:cs/0408007v12004
  9. Tight Differential Privacy for Discrete-Valued Mechanisms and for the Subsampled Gaussian Mechanism Using FFT

    Antti Koskela, Joonas Jälkö, Lukas Prediger +1

    stat.MLcs.CRcs.LGarXiv:2006.07134v32020
  10. Addressing Some Limitations of Transformers with Feedback Memory

    Angela Fan, Thibaut Lavril, Edouard Grave +2

    cs.LGcs.CLstat.MLarXiv:2002.09402v32020
  11. Fair Adversarial Gradient Tree Boosting

    Vincent Grari, Boris Ruf, Sylvain Lamprier +1

    cs.LGcs.AIcs.CYarXiv:1911.05369v22019
  12. Estimating individual treatment effect: generalization bounds and algorithms

    Uri Shalit, Fredrik D. Johansson, David Sontag

    stat.MLcs.AIcs.LGarXiv:1606.03976v52016
  13. Reinforcement Learning with Verifiable yet Noisy Rewards under Imperfect Verifiers

    Xin-Qiang Cai, Wei Wang, Feng Liu +3

    cs.LGcs.AIarXiv:2510.00915v42025
  14. Geometric Operator Learning with Optimal Transport

    Xinyi Li, Zongyi Li, Nikola Kovachki +1

    cs.LGarXiv:2507.20065v12025
  15. OceanLight: Efficient Global Ocean Forecasting via Geometry-Adaptive Unstructured Mesh Representation

    Wei Wu, Xiang Wang, Hongze Leng +3

    cs.LGcs.AIarXiv:2608.16070v12026
  16. Paper2Agent: Reimagining Research Papers As Interactive and Reliable AI Agents

    Jiacheng Miao, Joe R. Davis, Yaohui Zhang +2

    cs.AIcs.CLcs.LGarXiv:2509.06917v22025
  17. Multimodal Whole Slide Foundation Model for Pathology

    Tong Ding, Sophia J. Wagner, Andrew H. Song +20

    eess.IVcs.AIcs.CVarXiv:2411.19666v12024
  18. DeBERTa: Decoding-enhanced BERT with Disentangled Attention

    Pengcheng He, Xiaodong Liu, Jianfeng Gao +1

    cs.CLcs.LGarXiv:2006.03654v62020
  19. On the Value of Out-of-Distribution Testing: An Example of Goodhart's Law

    Damien Teney, Kushal Kafle, Robik Shrestha +3

    cs.CVcs.LGarXiv:2005.09241v12020
  20. Smoothed Dilated Convolutions for Improved Dense Prediction

    Zhengyang Wang, Shuiwang Ji

    cs.CVcs.LGarXiv:1808.08931v22018
  21. Sleeper Agents: Training Deceptive LLMs that Persist Through Safety Training

    Evan Hubinger, Carson Denison, Jesse Mu +36

    cs.CRcs.AIcs.CLarXiv:2401.05566v32024
  22. Addressing the Item Cold-start Problem by Attribute-driven Active Learning

    Yu Zhu, Jinhao Lin, Shibi He +4

    cs.IRcs.LGstat.MLarXiv:1805.09023v12018
  23. MotifNet: a motif-based Graph Convolutional Network for directed graphs

    Federico Monti, Karl Otness, Michael M. Bronstein

    cs.LGarXiv:1802.01572v12018
  24. Wav2Letter: an End-to-End ConvNet-based Speech Recognition System

    Ronan Collobert, Christian Puhrsch, Gabriel Synnaeve

    cs.LGcs.AIcs.CLarXiv:1609.03193v22016
  25. Sparse MoEs meet Efficient Ensembles

    James Urquhart Allingham, Florian Wenzel, Zelda E Mariet +10

    cs.LGcs.CVstat.MLarXiv:2110.03360v22021
  26. Graph of Thoughts: Solving Elaborate Problems with Large Language Models

    Maciej Besta, Nils Blach, Ales Kubicek +8

    cs.CLcs.AIcs.LGarXiv:2308.09687v42023
  27. The Impact of Positional Encoding on Length Generalization in Transformers

    Amirhossein Kazemnejad, Inkit Padhi, Karthikeyan Natesan Ramamurthy +2

    cs.CLcs.AIcs.LGarXiv:2305.19466v22023
  28. ProlificDreamer: High-Fidelity and Diverse Text-to-3D Generation with Variational Score Distillation

    Zhengyi Wang, Cheng Lu, Yikai Wang +4

    cs.LGcs.CVarXiv:2305.16213v22023
  29. Learning Fine-Grained Bimanual Manipulation with Low-Cost Hardware

    Tony Z. Zhao, Vikash Kumar, Sergey Levine +1

    cs.ROcs.LGarXiv:2304.13705v12023
  30. StoRM: A Diffusion-based Stochastic Regeneration Model for Speech Enhancement and Dereverberation

    Jean-Marie Lemercier, Julius Richter, Simon Welker +1

    eess.AScs.LGcs.SDarXiv:2212.11851v22022
  31. Beyond neural scaling laws: beating power law scaling via data pruning

    Ben Sorscher, Robert Geirhos, Shashank Shekhar +2

    cs.LGcs.AIcs.CVarXiv:2206.14486v62022
  32. FedBABU: Towards Enhanced Representation for Federated Image Classification

    Jaehoon Oh, Sangmook Kim, Se-Young Yun

    cs.LGarXiv:2106.06042v32021
  33. Back into Plato's Cave: Examining Cross-modal Representational Convergence at Scale

    A. Sophia Koepke, Daniil Zverev, Shiry Ginosar +1

    cs.CVcs.AIcs.LGarXiv:2604.18572v22026
  34. Loop, Think, & Generalize: Implicit Reasoning in Recurrent-Depth Transformers

    Harsh Kohli, Srinivasan Parthasarathy, Huan Sun +1

    cs.CLcs.AIcs.LGarXiv:2604.07822v22026
  35. PLDR-LLMs Reason At Self-Organized Criticality

    Burc Gokden

    cs.AIcs.CLcs.LGarXiv:2603.23539v12026
  36. A Neurosymbolic Approach for Constructing Planning Domain Models from Clinical Narratives

    Ranveer Singh, Saurabh Mathur, Michael Skinner +3

    cs.LGarXiv:2608.21186v12026
  37. NerVE: Nonlinear Eigenspectrum Dynamics in LLM Feed-Forward Networks

    Nandan Kumar Jha, Brandon Reagen

    cs.LGarXiv:2603.06922v22026
  38. Reinforcement Learning for Code Optimization

    Pierre Chambon, Kunhao Zheng, Juliette Decugis +2

    cs.LGcs.AIarXiv:2607.25970v12026
  39. A Survey of Convolutional Neural Networks: Analysis, Applications, and Prospects

    Zewen Li, Wenjie Yang, Shouheng Peng +1

    cs.CVcs.LGeess.IVarXiv:2004.02806v12020
  40. SpatiaLab: Can Vision-Language Models Perform Spatial Reasoning in the Wild?

    Azmine Toushik Wasi, Wahid Faisal, Abdur Rahman +12

    cs.CVcs.CEcs.CLarXiv:2602.03916v32026
  41. SINDy-PI: A Robust Algorithm for Parallel Implicit Sparse Identification of Nonlinear Dynamics

    Kadierdan Kaheman, J. Nathan Kutz, Steven L. Brunton

    cs.LGphysics.comp-phstat.MLarXiv:2004.02322v22020
  42. Meta Label Correction for Noisy Label Learning

    Guoqing Zheng, Ahmed Hassan Awadallah, Susan Dumais

    cs.LGstat.MLarXiv:1911.03809v22019
  43. The Born Supremacy: Quantum Advantage and Training of an Ising Born Machine

    Brian Coyle, Daniel Mills, Vincent Danos +1

    quant-phcs.LGarXiv:1904.02214v42019
  44. Federated Optimization in Heterogeneous Networks

    Tian Li, Anit Kumar Sahu, Manzil Zaheer +3

    cs.LGstat.MLarXiv:1812.06127v52018
  45. Automating the Design of Embodied Agent Architectures

    Jian Zhou, Sihao Lin, Jin Li +3

    cs.ROcs.AIcs.LGarXiv:2606.30111v22026
  46. TROPT: An Open Framework for Unifying and Advancing Discrete Text Optimization

    Matan Ben-Tov, Mahmood Sharif

    cs.LGcs.CRarXiv:2606.23496v12026
  47. Rethinking Psychometric Evaluation of LLMs: When and Why Self-Reports Predict Behavior

    Rafal Kocielnik, Pengrui Han, Peiyang Song +5

    cs.AIcs.CLcs.CYarXiv:2606.12730v12026
  48. World Models

    David Ha, Jürgen Schmidhuber

    cs.LGstat.MLarXiv:1803.10122v42018
  49. DynaFLIP: Rethinking Robotics Perception via Tri-Modal-Dynamics Guided Representation

    Jusuk Lee, Seungjae Lee, Jonghun Shin +6

    cs.ROcs.LGarXiv:2605.30350v12026
  50. Probabilistic Fair Clustering

    Seyed A. Esmaeili, Brian Brubach, Leonidas Tsepenekas +1

    cs.LGcs.AIcs.DSarXiv:2006.10916v32020
  51. CoSPlay: Cooperative Self-Play at Test-Time with Self-Generated Code and Unit Test

    Zhangyi Hu, Chenhui Liu, Tian Huang +6

    cs.LGcs.AIcs.CLarXiv:2605.23491v22026
  52. Learn to Reason Efficiently with Adaptive Length-based Reward Shaping

    Wei Liu, Ruochen Zhou, Yiyun Deng +5

    cs.CLcs.AIcs.LGarXiv:2505.15612v12025
  53. Deep Activity Recognition Models with Triaxial Accelerometers

    Mohammad Abu Alsheikh, Ahmed Selim, Dusit Niyato +3

    cs.LGcs.HCcs.NEarXiv:1511.04664v22015
  54. Emergent Abilities in Large Language Models: A Survey

    Leonardo Berti, Flavio Giorgi, Gjergji Kasneci

    cs.LGcs.AIcs.CLarXiv:2503.05788v32025
  55. Utility Engineering: Analyzing and Controlling Emergent Value Systems in AIs

    Mantas Mazeika, Xuwang Yin, Rishub Tamirisa +8

    cs.LGcs.AIcs.CLarXiv:2502.08640v22025
  56. Finite Versus Infinite Neural Networks: an Empirical Study

    Jaehoon Lee, Samuel S. Schoenholz, Jeffrey Pennington +4

    cs.LGstat.MLarXiv:2007.15801v22020
  57. Propagating Confidences through CNNs for Sparse Data Regression

    Abdelrahman Eldesokey, Michael Felsberg, Fahad Shahbaz Khan

    cs.CVcs.LGarXiv:1805.11913v32018
  58. Trajectory Balance with Asynchrony: Decoupling Exploration and Learning for Fast, Scalable LLM Post-Training

    Brian Bartoldson, Siddarth Venkatraman, James Diffenderfer +7

    cs.LGarXiv:2503.18929v22025
  59. Graph Oracle Models, Lower Bounds, and Gaps for Parallel Stochastic Optimization

    Blake Woodworth, Jialei Wang, Adam Smith +2

    math.OCcs.LGstat.MLarXiv:1805.10222v32018
  60. Efficient Differentiable Simulation of Articulated Bodies

    Yi-Ling Qiao, Junbang Liang, Vladlen Koltun +1

    cs.LGcs.GRcs.ROarXiv:2109.07719v12021