Machine Learning

Papers filed under cs.LG on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.

Search paper metadata (including unsummarized papers)

661 to 720 of 20,454

  1. Transition Matching Distillation for Fast Video Generation

    Weili Nie, Julius Berner, Nanye Ma +3

    cs.CVcs.AIcs.LGarXiv:2601.09881v22026
  2. MirrorBench: A Benchmark to Evaluate Conversational User-Proxy Agents for Human-Likeness

    Ashutosh Hathidara, Julien Yu, Vaishali Senthil +2

    cs.AIcs.LGarXiv:2601.08118v32026
  3. Llama-Mobile: Efficient 2.7-Bit Quantization of VLMs

    Luka Ribar, Jeevan Bhoot, Douglas Orr

    cs.CVcs.LGarXiv:2608.21134v12026
  4. Comparison of Bayesian predictive methods for model selection

    Juho Piironen, Aki Vehtari

    stat.MEcs.LGarXiv:1503.08650v42015
  5. CDRL: Certification-Driven Reinforcement Learning for Neutrino Flavor Model Discovery

    Piyush Jha, Jake Rudolph, Victoria Knapp-Pérez +3

    cs.AIcs.LGcs.LOarXiv:2608.20686v12026
  6. The Principles of Diffusion Models

    Chieh-Hsin Lai, Yang Song, Dongjun Kim +2

    cs.LGcs.AIcs.GRarXiv:2510.21890v32025
  7. End-to-end Continuous Speech Recognition using Attention-based Recurrent NN: First Results

    Jan Chorowski, Dzmitry Bahdanau, Kyunghyun Cho +1

    cs.NEcs.LGstat.MLarXiv:1412.1602v12014
  8. What is Missing from AI Post-Training AI: An Empirical Analysis

    Joy Jia Yin Lim, Xin Huang, Hao Peng +5

    cs.AIcs.CLcs.LGarXiv:2608.19072v12026
  9. Flama: a Python framework for development and deployment of production-ready APIs, machine learning, and LLM services

    José A. Perdiguero López, Miguel A. Durán-Olivencia

    cs.SEcs.AIcs.LGarXiv:2608.18733v12026
  10. Partition the Support, Reconstruct the Residual: Training-Free Sparse Attention for Video Generation and World Models

    Pardis Taghavi, Reza Langari, Gaurav Pandey

    cs.CVcs.AIcs.LGarXiv:2608.18484v12026
  11. Gradient Descent on Neural Networks Typically Occurs at the Edge of Stability

    Jeremy M. Cohen, Simran Kaur, Yuanzhi Li +2

    cs.LGstat.MLarXiv:2103.00065v32021
  12. MapAnything: Universal Feed-Forward Metric 3D Reconstruction

    Nikhil Keetha, Norman Müller, Johannes Schönberger +14

    cs.CVcs.AIcs.LGarXiv:2509.13414v32025
  13. A Survey of Reinforcement Learning for Large Reasoning Models

    Kaiyan Zhang, Yuxin Zuo, Bingxiang He +36

    cs.CLcs.AIcs.LGarXiv:2509.08827v32025
  14. Deep Think with Confidence

    Yichao Fu, Xuewei Wang, Yuandong Tian +1

    cs.LGarXiv:2508.15260v12025
  15. Degradation-Aligned Self-Supervised Learning for State of Health Estimation of Lithium-Ion Batteries under Label Sparsity

    Jiaqi Yao, Julia Kowal

    eess.SPcs.AIcs.LGarXiv:2608.16612v12026
  16. Discrete Diffusion in Large Language and Multimodal Models: A Survey

    Runpeng Yu, Qi Li, Xinchao Wang

    cs.LGcs.AIarXiv:2506.13759v52025
  17. AlphaEvolve: A coding agent for scientific and algorithmic discovery

    Alexander Novikov, Ngân Vũ, Marvin Eisenberger +15

    cs.AIcs.LGcs.NEarXiv:2506.13131v12025
  18. Semi-supervised clustering methods

    Eric Bair

    stat.MEcs.LGstat.MLarXiv:1307.0252v12013
  19. Mapping the Space of Chemical Reactions Using Attention-Based Neural Networks

    Philippe Schwaller, Daniel Probst, Alain C. Vaucher +4

    physics.chem-phcs.CLcs.LGarXiv:2012.06051v12020
  20. Confidence Intervals and Hypothesis Testing for High-Dimensional Regression

    Adel Javanmard, Andrea Montanari

    stat.MEcs.ITcs.LGarXiv:1306.3171v22013
  21. J1: Incentivizing Thinking in LLM-as-a-Judge via Reinforcement Learning

    Chenxi Whitehouse, Tianlu Wang, Ping Yu +4

    cs.CLcs.AIcs.LGarXiv:2505.10320v32025
  22. Absolute Zero: Reinforced Self-play Reasoning with Zero Data

    Andrew Zhao, Yiran Wu, Yang Yue +8

    cs.LGcs.AIcs.CLarXiv:2505.03335v32025
  23. Measuring Structured Predictability in Neural Training Dynamics: A Cross-Regime Study

    Fanqi Wang, Weisheng Tang, Hairong Qi

    cs.LGarXiv:2608.15483v12026
  24. Interpreting Graph Neural Networks for NLP With Differentiable Edge Masking

    Michael Sejr Schlichtkrull, Nicola De Cao, Ivan Titov

    cs.CLcs.LGstat.MLarXiv:2010.00577v32020
  25. A Survey on Multi-view Learning

    Chang Xu, Dacheng Tao, Chao Xu

    cs.LGarXiv:1304.5634v12013
  26. Language models suffer from a curse of ambiguity

    Nicolas Zucchet, Hyun Dong Lee, Scott Linderman

    cs.CLcs.LGcs.NEarXiv:2608.15448v12026
  27. Contrastive Self-supervised Learning for Graph Classification

    Jiaqi Zeng, Pengtao Xie

    cs.LGstat.MLarXiv:2009.05923v12020
  28. Diagnosing and Mitigating Perception-Decision Misalignment in Omni-LLMs via Modality Subspace Activation

    Hongbo Jiang, Jie Li, Yunhang Shen +2

    cs.LGcs.CVarXiv:2608.14655v12026
  29. Calibrated Trust, Not Sharper Prediction: An Empirical Test of Uncertainty Fusion

    Surya Saka

    cs.LGcs.AIcs.CLarXiv:2608.14617v12026
  30. NSGANetV2: Evolutionary Multi-Objective Surrogate-Assisted Neural Architecture Search

    Zhichao Lu, Kalyanmoy Deb, Erik Goodman +2

    cs.CVcs.LGcs.NEarXiv:2007.10396v12020
  31. Distribution Aligning Refinery of Pseudo-label for Imbalanced Semi-supervised Learning

    Jaehyung Kim, Youngbum Hur, Sejun Park +3

    cs.LGstat.MLarXiv:2007.08844v22020
  32. CardioState-JEPA: Delay-Aware Cross-Modal Learning of a Shared Cardiac Representation

    Hamza Shafiq, Hung Manh Pham, Bin Zhu +3

    cs.LGeess.IVstat.MLarXiv:2608.12944v12026
  33. Hands-on Bayesian Neural Networks -- a Tutorial for Deep Learning Users

    Laurent Valentin Jospin, Wray Buntine, Farid Boussaid +2

    cs.LGstat.MLarXiv:2007.06823v32020
  34. Multiscale Simulations of Complex Systems by Learning their Effective Dynamics

    Pantelis R. Vlachas, Georgios Arampatzis, Caroline Uhler +1

    physics.comp-phcs.LGnlin.CDarXiv:2006.13431v32020
  35. A Bayesian Approach to Robust Inverse Reinforcement Learning

    Ran Wei, Siliang Zeng, Chenliang Li +3

    cs.LGarXiv:2309.08571v22023
  36. TokenSqueeze: Performance-Preserving Compression for Reasoning LLMs

    Yuxiang Zhang, Zhengxu Yu, Weihang Pan +5

    cs.LGcs.AIarXiv:2511.13223v12025
  37. TxAgent: An AI Agent for Therapeutic Reasoning Across a Universe of Tools

    Shanghua Gao, Richard Zhu, Zhenglun Kong +5

    cs.AIcs.LGarXiv:2503.10970v12025
  38. Byzantine-Robust Learning on Heterogeneous Datasets via Bucketing

    Sai Praneeth Karimireddy, Lie He, Martin Jaggi

    cs.LGstat.MLarXiv:2006.09365v62020
  39. Training Generative Adversarial Networks with Limited Data

    Tero Karras, Miika Aittala, Janne Hellsten +3

    cs.CVcs.LGcs.NEarXiv:2006.06676v22020
  40. Audio Flamingo 2: An Audio-Language Model with Long-Audio Understanding and Expert Reasoning Abilities

    Sreyan Ghosh, Zhifeng Kong, Sonal Kumar +6

    cs.SDcs.CLcs.LGarXiv:2503.03983v12025
  41. Deep Learning is Not So Mysterious or Different

    Andrew Gordon Wilson

    cs.LGstat.MLarXiv:2503.02113v22025
  42. Retrieval-Augmented Generation for Knowledge-Intensive NLP Tasks

    Patrick Lewis, Ethan Perez, Aleksandra Piktus +9

    cs.CLcs.LGarXiv:2005.11401v42020
    Summaries:한국어
  43. Dataset Distillation with Neural Characteristic Function: A Minmax Perspective

    Shaobo Wang, Yicun Yang, Zhiyuan Liu +4

    cs.CVcs.AIcs.LGarXiv:2502.20653v12025
    Summaries:한국어
  44. Universal Sparse Autoencoders: Interpretable Cross-Model Concept Alignment

    Harrish Thasarathan, Julian Forsyth, Thomas Fel +2

    cs.CVcs.LGarXiv:2502.03714v22025
  45. Maximum Density Divergence for Domain Adaptation

    Li Jingjing, Chen Erpeng, Ding Zhengming +3

    cs.CVcs.LGcs.MMarXiv:2004.12615v12020
  46. Reasoning Denoiser: Denoising Reasoning Traces for Hallucination Detection in Large Reasoning Models

    Junlin Fang, Do Nguyen-Thanh, Xiaogang Xu +2

    cs.AIcs.LGarXiv:2607.22098v12026
  47. Don't Judge an Object by Its Context: Learning to Overcome Contextual Bias

    Krishna Kumar Singh, Dhruv Mahajan, Kristen Grauman +3

    cs.CVcs.LGarXiv:2001.03152v22020
  48. Improving Medical Large Vision-Language Models with Abnormal-Aware Feedback

    Yucheng Zhou, Lingran Song, Jianbing Shen

    cs.CLcs.AIcs.CVarXiv:2501.01377v22025
  49. Frontier Models are Capable of In-context Scheming

    Alexander Meinke, Bronson Schoen, Jérémy Scheurer +3

    cs.AIcs.LGarXiv:2412.04984v22024
  50. Is One Layer Enough? Training A Single Transformer Layer Can Match Full-Parameter RL Training

    Zijian Zhang, Rizhen Hu, Athanasios Glentis +4

    cs.LGcs.CLarXiv:2607.01232v22026
  51. Omni-Scale CNNs: a simple and effective kernel size configuration for time series classification

    Wensi Tang, Guodong Long, Lu Liu +3

    cs.LGstat.MLarXiv:2002.10061v32020
  52. Beyond IID: How General Are Tabular Foundation Models, Really?

    Lennart Purucker, Andrej Tschalzev, Nick Erickson +7

    cs.LGcs.AIarXiv:2606.30410v12026
    Summaries:한국어
  53. Explaining Explanations: Axiomatic Feature Interactions for Deep Networks

    Joseph D. Janizek, Pascal Sturmfels, Su-In Lee

    cs.LGstat.MLarXiv:2002.04138v32020
  54. PhysGen: Rigid-Body Physics-Grounded Image-to-Video Generation

    Shaowei Liu, Zhongzheng Ren, Saurabh Gupta +1

    cs.CVcs.AIcs.LGarXiv:2409.18964v12024
  55. LogicVista: Multimodal LLM Logical Reasoning Benchmark in Visual Contexts

    Yijia Xiao, Edward Sun, Tianyu Liu +1

    cs.AIcs.CLcs.CVarXiv:2407.04973v12024
  56. Step-DPO: Step-wise Preference Optimization for Long-chain Reasoning of LLMs

    Xin Lai, Zhuotao Tian, Yukang Chen +3

    cs.LGcs.AIcs.CLarXiv:2406.18629v12024
  57. SWE-agent: Agent-Computer Interfaces Enable Automated Software Engineering

    John Yang, Carlos E. Jimenez, Alexander Wettig +4

    cs.SEcs.AIcs.CLarXiv:2405.15793v32024
  58. Variational Mixture-of-Experts Autoencoders for Multi-Modal Deep Generative Models

    Yuge Shi, N. Siddharth, Brooks Paige +1

    stat.MLcs.LGarXiv:1911.03393v12019
  59. The Heidelberg spiking datasets for the systematic evaluation of spiking neural networks

    Benjamin Cramer, Yannik Stradmann, Johannes Schemmel +1

    cs.NEcs.LGq-bio.NCarXiv:1910.07407v32019
  60. CAVEWOMAN: How Large Language Models Behave Under Linguistic Input and Output Compression

    Morayo Danielle Adeyemi, Ryan A. Rossi, Franck Dernoncourt

    cs.CLcs.AIcs.LGarXiv:2606.24083v12026