Machine Learning

Papers filed under cs.LG on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.

Search paper metadata (including unsummarized papers)

6,181 to 6,240 of 20,193

  1. Conformal Inference for Online Prediction with Arbitrary Distribution Shifts

    Isaac Gibbs, Emmanuel Candès

    stat.MEcs.LGarXiv:2208.08401v32022
  2. EQ-VAE: Equivariance Regularized Latent Space for Improved Generative Image Modeling

    Theodoros Kouzelis, Ioannis Kakogeorgiou, Spyros Gidaris +1

    cs.LGarXiv:2502.09509v32025
  3. Agent Lightning: Train ANY AI Agents with Reinforcement Learning

    Xufang Luo, Yuge Zhang, Zhiyuan He +5

    cs.AIcs.LGarXiv:2508.03680v12025
  4. Sparse coding for multitask and transfer learning

    Andreas Maurer, Massimiliano Pontil, Bernardino Romera-Paredes

    cs.LGstat.MLarXiv:1209.0738v32012
  5. The Variational Gaussian Process

    Dustin Tran, Rajesh Ranganath, David M. Blei

    stat.MLcs.LGcs.NEarXiv:1511.06499v42015
  6. Deep Lattice Networks and Partial Monotonic Functions

    Seungil You, David Ding, Kevin Canini +2

    stat.MLcs.LGarXiv:1709.06680v12017
  7. Composite Task-Completion Dialogue Policy Learning via Hierarchical Deep Reinforcement Learning

    Baolin Peng, Xiujun Li, Lihong Li +4

    cs.CLcs.AIcs.LGarXiv:1704.03084v32017
  8. CausaLM: Causal Model Explanation Through Counterfactual Language Models

    Amir Feder, Nadav Oved, Uri Shalit +1

    cs.CLcs.AIcs.LGarXiv:2005.13407v52020
  9. ShareGPT-4o-Image: Aligning Multimodal Models with GPT-4o-Level Image Generation

    Junying Chen, Zhenyang Cai, Pengcheng Chen +5

    cs.CVcs.AIcs.LGarXiv:2506.18095v12025
  10. Understanding Gradient Orthogonalization for Deep Learning via Non-Euclidean Trust-Region Optimization

    Dmitry Kovalev

    cs.LGmath.OCstat.MLarXiv:2503.12645v22025
  11. Continual Learning for Large Language Models: A Survey

    Tongtong Wu, Linhao Luo, Yuan-Fang Li +3

    cs.CLcs.LGarXiv:2402.01364v22024
  12. Self-Tuning Networks: Bilevel Optimization of Hyperparameters using Structured Best-Response Functions

    Matthew MacKay, Paul Vicol, Jon Lorraine +2

    cs.LGstat.MLarXiv:1903.03088v12019
  13. Big-Math: A Large-Scale, High-Quality Math Dataset for Reinforcement Learning in Language Models

    Alon Albalak, Duy Phung, Nathan Lile +8

    cs.LGcs.AIcs.CLarXiv:2502.17387v12025
  14. Neural Symbollic Regression Using Deep Learning and Sparse Modelling

    Ravi Kumar U, Sumitra S

    cs.LGcs.NEcs.SCarXiv:2609.01102v12026
  15. Masked Autoencoders Are Effective Tokenizers for Diffusion Models

    Hao Chen, Yujin Han, Fangyi Chen +7

    cs.CVcs.AIcs.LGarXiv:2502.03444v22025
  16. MDocAgent: A Multi-Modal Multi-Agent Framework for Document Understanding

    Siwei Han, Peng Xia, Ruiyi Zhang +4

    cs.LGarXiv:2503.13964v12025
  17. Chain-of-Visual-Thought: Teaching VLMs to See and Think Better with Continuous Visual Tokens

    Yiming Qin, Bomin Wei, Jiaxin Ge +4

    cs.CVcs.AIcs.LGarXiv:2511.19418v32025
  18. Goal-Conditioned Reinforcement Learning with Imagined Subgoals

    Elliot Chane-Sane, Cordelia Schmid, Ivan Laptev

    cs.LGcs.ROarXiv:2107.00541v12021
  19. Ad Headline Generation using Self-Critical Masked Language Model

    Yashal Shakti Kanungo, Sumit Negi, Aruna Rajan

    cs.CLcs.AIcs.LGarXiv:2607.06818v12026
  20. GenONet: A Generative operator Network for High-Resolution Precipitation Nowcasting

    Mohammad Kian Golkar, Luciano Alves de Oliveira, Mohammad Khanjani

    cs.LGphysics.ao-pharXiv:2609.00544v12026
  21. G-Memory: Tracing Hierarchical Memory for Multi-Agent Systems

    Guibin Zhang, Muxin Fu, Guancheng Wan +3

    cs.MAcs.CLcs.LGarXiv:2506.07398v22025
  22. Cryo-CARE: Content-Aware Image Restoration for Cryo-Transmission Electron Microscopy Data

    Tim-Oliver Buchholz, Mareike Jordan, Gaia Pigino +1

    cs.CVcs.LGarXiv:1810.05420v22018
  23. Apple Intelligence Foundation Language Models: Tech Report 2025

    Ethan Li, Anders Boesen Lindbo Larsen, Chen Zhang +395

    cs.LGcs.AIarXiv:2507.13575v32025
  24. Domino: Discovering Systematic Errors with Cross-Modal Embeddings

    Sabri Eyuboglu, Maya Varma, Khaled Saab +5

    cs.LGcs.AIarXiv:2203.14960v32022
  25. Solving 3D Inverse Problems using Pre-trained 2D Diffusion Models

    Hyungjin Chung, Dohoon Ryu, Michael T. McCann +2

    cs.CVcs.AIcs.LGarXiv:2211.10655v12022
  26. MATH-Perturb: Benchmarking LLMs' Math Reasoning Abilities against Hard Perturbations

    Kaixuan Huang, Jiacheng Guo, Zihao Li +15

    cs.LGcs.AIcs.CLarXiv:2502.06453v22025
  27. Personalized Transformer for Explainable Recommendation

    Lei Li, Yongfeng Zhang, Li Chen

    cs.IRcs.AIcs.CLarXiv:2105.11601v22021
  28. ResearchClawBench: A Benchmark for End-to-End Autonomous Scientific Research

    Wanghan Xu, Shuo Li, Tianlin Ye +48

    cs.LGcs.AIcs.CLarXiv:2606.07591v52026
  29. Image Hijacks: Adversarial Images can Control Generative Models at Runtime

    Luke Bailey, Euan Ong, Stuart Russell +1

    cs.LGcs.CLcs.CRarXiv:2309.00236v42023
  30. Efficient Reinforcement Finetuning via Adaptive Curriculum Learning

    Taiwei Shi, Yiyang Wu, Linxin Song +2

    cs.LGcs.CLarXiv:2504.05520v42025
  31. NorMuon: Making Muon more efficient and scalable

    Zichong Li, Liming Liu, Chen Liang +2

    cs.LGcs.CLarXiv:2510.05491v12025
  32. Facet-0: A Robotic Foundation Model for Contact-Rich Precise Manipulation

    Haoyuan Deng, Haichao Liu, Wenkai Guo +6

    cs.ROcs.LGarXiv:2609.01596v12026
  33. Towards a universal meta-optics solver via large language models

    Huanshu Zhang, Lei Kang, Yuyan Chen +3

    physics.opticscs.LGarXiv:2608.26417v12026
  34. Llama-Nemotron: Efficient Reasoning Models

    Akhiad Bercovich, Itay Levy, Izik Golan +133

    cs.CLcs.AIcs.LGarXiv:2505.00949v52025
  35. Soft Active Electromyography Interface for Machine Learning-Enabled Silent Speech Recognition

    Yuta Kurotaki, Shusuke Yamakoshi, Reitaro Yoshida +6

    cs.LGcs.HCcs.SDarXiv:2608.27048v12026
  36. Principles and Guidelines for Evaluating Social Robot Navigation Algorithms

    Anthony Francis, Claudia Pérez-D'Arpino, Chengshu Li +28

    cs.ROcs.AIcs.HCarXiv:2306.16740v42023
  37. NeuroLKH: Combining Deep Learning Model with Lin-Kernighan-Helsgaun Heuristic for Solving the Traveling Salesman Problem

    Liang Xin, Wen Song, Zhiguang Cao +1

    cs.AIcs.LGarXiv:2110.07983v12021
  38. DiffSTG: Probabilistic Spatio-Temporal Graph Forecasting with Denoising Diffusion Models

    Haomin Wen, Youfang Lin, Yutong Xia +4

    cs.LGarXiv:2301.13629v42023
  39. Helping or Herding? Reward Model Ensembles Mitigate but do not Eliminate Reward Hacking

    Jacob Eisenstein, Chirag Nagpal, Alekh Agarwal +9

    cs.LGarXiv:2312.09244v32023
  40. Diffuse and Disperse: Image Generation with Representation Regularization

    Runqian Wang, Kaiming He

    cs.CVcs.AIcs.LGarXiv:2506.09027v22025
  41. Reconciling Reality through Simulation: A Real-to-Sim-to-Real Approach for Robust Manipulation

    Marcel Torne, Anthony Simeonov, Zechu Li +4

    cs.ROcs.AIcs.LGarXiv:2403.03949v32024
  42. A Random Matrix Approach to Neural Networks

    Cosme Louart, Zhenyu Liao, Romain Couillet

    math.PRcs.LGarXiv:1702.05419v22017
  43. From neural PCA to deep unsupervised learning

    Harri Valpola

    stat.MLcs.LGcs.NEarXiv:1411.7783v22014
  44. Capability-Gated Language Models: Security Composes, Utility Does Not

    Patrikas Vanagas, Augustas Mačijauskas, Laurynas Lopata

    cs.CRcs.AIcs.LGarXiv:2609.00445v12026
  45. Rethinking Rubric Generation for Improving LLM Judge and Reward Modeling for Open-ended Tasks

    William F. Shen, Xinchi Qiu, Chenxi Whitehouse +6

    cs.LGcs.AIarXiv:2602.05125v12026
  46. DARec: Deep Domain Adaptation for Cross-Domain Recommendation via Transferring Rating Patterns

    Feng Yuan, Lina Yao, Boualem Benatallah

    cs.LGcs.IRstat.MLarXiv:1905.10760v12019
  47. Convergence of the Deep BSDE Method for Coupled FBSDEs

    Jiequn Han, Jihao Long

    math.PRcs.LGmath.NAarXiv:1811.01165v42018
  48. Pass@K Policy Optimization: Solving Harder Reinforcement Learning Problems

    Christian Walder, Deep Karkhanis

    cs.LGcs.AIcs.CLarXiv:2505.15201v52025
  49. LightThinker: Thinking Step-by-Step Compression

    Jintian Zhang, Yuqi Zhu, Mengshu Sun +6

    cs.CLcs.AIcs.IRarXiv:2502.15589v22025
  50. How to build a consistency model: Learning flow maps via self-distillation

    Nicholas M. Boffi, Michael S. Albergo, Eric Vanden-Eijnden

    cs.LGcs.CVarXiv:2505.18825v22025
  51. Disentanglement via Mechanism Sparsity Regularization: A New Principle for Nonlinear ICA

    Sébastien Lachapelle, Pau Rodríguez López, Yash Sharma +4

    stat.MLcs.LGarXiv:2107.10098v32021
  52. Global Sparse Momentum SGD for Pruning Very Deep Neural Networks

    Xiaohan Ding, Guiguang Ding, Xiangxin Zhou +3

    cs.LGcs.CVstat.MLarXiv:1909.12778v32019
  53. A Comprehensive Survey on Long Context Language Modeling

    Jiaheng Liu, Dawei Zhu, Zhiqi Bai +34

    cs.CLcs.LGarXiv:2503.17407v22025
  54. Sharp Mixed Spectral Barron Regularity of Coulombic Many-Electron Wave Functions

    Pingbing Ming, Hao Yu

    math.APcs.LGmath.NAarXiv:2609.00872v12026
  55. Learning to Prove Theorems via Interacting with Proof Assistants

    Kaiyu Yang, Jia Deng

    cs.LOcs.AIcs.LGarXiv:1905.09381v12019
  56. Elite-Weighted Supervised Fine-tuning for Goal-Directed Molecular Optimization

    Shiyun Wa, Yifei Wang, Anna G. Green +2

    cs.LGarXiv:2609.00189v12026
  57. On The Reasons Behind Decisions

    Adnan Darwiche, Auguste Hirth

    cs.AIcs.LGarXiv:2002.09284v12020
  58. OpenMM 8: Molecular Dynamics Simulation with Machine Learning Potentials

    Peter Eastman, Raimondas Galvelis, Raúl P. Peláez +22

    physics.chem-phcs.LGarXiv:2310.03121v22023
  59. Neural 3D Morphable Models: Spiral Convolutional Networks for 3D Shape Representation Learning and Generation

    Giorgos Bouritsas, Sergiy Bokhnyak, Stylianos Ploumpis +2

    cs.CVcs.AIcs.GRarXiv:1905.02876v32019
  60. Compressing DMA Engine: Leveraging Activation Sparsity for Training Deep Neural Networks

    Minsoo Rhu, Mike O'Connor, Niladrish Chatterjee +2

    cs.LGcs.ARarXiv:1705.01626v12017