Machine Learning

Papers filed under cs.LG on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.

Search paper metadata (including unsummarized papers)

6,961 to 7,020 of 20,199

  1. Optimal Errors and Phase Transitions in High-Dimensional Generalized Linear Models

    Jean Barbier, Florent Krzakala, Nicolas Macris +2

    cs.ITcond-mat.dis-nncs.AIarXiv:1708.03395v32017
  2. Fairness Beyond Disparate Treatment & Disparate Impact: Learning Classification without Disparate Mistreatment

    Muhammad Bilal Zafar, Isabel Valera, Manuel Gomez Rodriguez +1

    stat.MLcs.LGarXiv:1610.08452v22016
  3. Quantized Neural Networks: Training Neural Networks with Low Precision Weights and Activations

    Itay Hubara, Matthieu Courbariaux, Daniel Soudry +2

    cs.NEcs.LGarXiv:1609.07061v12016
  4. Stealing Machine Learning Models via Prediction APIs

    Florian Tramèr, Fan Zhang, Ari Juels +2

    cs.CRcs.LGstat.MLarXiv:1609.02943v22016
  5. Data Programming: Creating Large Training Sets, Quickly

    Alexander Ratner, Christopher De Sa, Sen Wu +2

    stat.MLcs.AIcs.LGarXiv:1605.07723v32016
  6. Thought Anchors: Which LLM Reasoning Steps Matter?

    Paul C. Bogdan, Uzay Macar, Neel Nanda +1

    cs.LGcs.AIcs.CLarXiv:2506.19143v42025
  7. GDPval: Evaluating AI Model Performance on Real-World Economically Valuable Tasks

    Tejal Patwardhan, Rachel Dias, Elizabeth Proehl +16

    cs.LGcs.AIcs.CYarXiv:2510.04374v12025
  8. Fast Convergence of Regularized Learning in Games

    Vasilis Syrgkanis, Alekh Agarwal, Haipeng Luo +1

    cs.GTcs.AIcs.LGarXiv:1507.00407v52015
  9. The Surprising Effectiveness of Negative Reinforcement in LLM Reasoning

    Xinyu Zhu, Mengzhou Xia, Zhepei Wei +3

    cs.CLcs.LGarXiv:2506.01347v22025
  10. Convex Optimization: Algorithms and Complexity

    Sébastien Bubeck

    math.OCcs.CCcs.LGarXiv:1405.4980v22014
  11. Tuned Models of Peer Assessment in MOOCs

    Chris Piech, Jonathan Huang, Zhenghao Chen +3

    cs.LGcs.AIcs.HCarXiv:1307.2579v12013
  12. On the Theoretical Limitations of Embedding-Based Retrieval

    Orion Weller, Michael Boratko, Iftekhar Naim +1

    cs.IRcs.CLcs.LGarXiv:2508.21038v22025
    Summaries:한국어
  13. Reactive Diffusion Policy: Slow-Fast Visual-Tactile Policy Learning for Contact-Rich Manipulation

    Han Xue, Jieji Ren, Wendi Chen +5

    cs.ROcs.AIcs.LGarXiv:2503.02881v32025
  14. Thompson Sampling for Contextual Bandits with Linear Payoffs

    Shipra Agrawal, Navin Goyal

    cs.LGcs.DSstat.MLarXiv:1209.3352v42012
  15. Adversarial-Learned Loss for Domain Adaptation

    Minghao Chen, Shuai Zhao, Haifeng Liu +1

    cs.CVcs.LGarXiv:2001.01046v12020
  16. Thompson Sampling: An Asymptotically Optimal Finite Time Analysis

    Emilie Kaufmann, Nathaniel Korda, Rémi Munos

    stat.MLcs.LGarXiv:1205.4217v22012
  17. Confidence-Aware Learning for Deep Neural Networks

    Jooyoung Moon, Jihyo Kim, Younghak Shin +1

    cs.LGstat.MLarXiv:2007.01458v32020
  18. A Text Classification Framework for Simple and Effective Early Depression Detection Over Social Media Streams

    Sergio G. Burdisso, Marcelo Errecalde, Manuel Montes-y-Gómez

    cs.CYcs.CLcs.IRarXiv:1905.08772v22019
  19. Light-R1: Curriculum SFT, DPO and RL for Long COT from Scratch and Beyond

    Liang Wen, Yunke Cai, Fenrui Xiao +11

    cs.CLcs.LGarXiv:2503.10460v42025
  20. IAIA-BL: A Case-based Interpretable Deep Learning Model for Classification of Mass Lesions in Digital Mammography

    Alina Jade Barnett, Fides Regina Schwartz, Chaofan Tao +4

    cs.LGcs.AIcs.CVarXiv:2103.12308v12021
  21. Language Modeling with Deep Transformers

    Kazuki Irie, Albert Zeyer, Ralf Schlüter +1

    cs.CLcs.LGarXiv:1905.04226v22019
  22. RM-R1: Reward Modeling as Reasoning

    Xiusi Chen, Gaotang Li, Ziqi Wang +9

    cs.CLcs.AIcs.LGarXiv:2505.02387v42025
  23. On Variance Reduction in Stochastic Gradient Descent and its Asynchronous Variants

    Sashank J. Reddi, Ahmed Hefny, Suvrit Sra +2

    cs.LGstat.MLarXiv:1506.06840v22015
  24. Spectral Convergence of Random Feature Method in Multiple Dimensions

    Pingbing Ming, Hao Yu

    math.NAcs.AIcs.LGarXiv:2609.03401v12026
  25. Bayes-Optimal BER and AUC: Estimation and Evaluation of Estimators

    Ryota Ushio, Takashi Ishida, Masashi Sugiyama

    cs.LGstat.MLarXiv:2609.02304v12026
  26. LLaDA-V: Large Language Diffusion Models with Visual Instruction Tuning

    Zebin You, Shen Nie, Xiaolu Zhang +5

    cs.LGcs.CLcs.CVarXiv:2505.16933v22025
  27. Preference Fine-Tuning of LLMs Should Leverage Suboptimal, On-Policy Data

    Fahim Tajwar, Anikait Singh, Archit Sharma +6

    cs.LGarXiv:2404.14367v32024
  28. Adversarial camera stickers: A physical camera-based attack on deep learning systems

    Juncheng Li, Frank R. Schmidt, J. Zico Kolter

    cs.CVcs.CRcs.LGarXiv:1904.00759v42019
  29. RL's Razor: Why Online Reinforcement Learning Forgets Less

    Idan Shenfeld, Jyothish Pari, Pulkit Agrawal

    cs.LGarXiv:2509.04259v12025
  30. Pseudo-Simulation for Autonomous Driving

    Wei Cao, Marcel Hallgarten, Tianyu Li +11

    cs.ROcs.AIcs.CVarXiv:2506.04218v32025
  31. Diffusion-Based Refinement for Kilometer-Scale Probabilistic Precipitation Nowcasting

    Dohyun Park, Changhoon Song, Tengyuan Chang +2

    cs.LGphysics.ao-pharXiv:2608.30205v12026
  32. Multiclass Linear Perceptrons with Multiplicative Margins

    Dmitri Rachkovskij, Evgeny Osipov, Olexander Volkov +2

    cs.LGcs.NEarXiv:2608.30028v12026
  33. Puro-2B: Poor Lab's Qwen2-1.5B Trained on RTX 5090 within $5090

    Kairong Luo, Jiarui Cui, Yaorui Yin +8

    cs.CLcs.LGarXiv:2608.27370v12026
  34. Simple Actors and Deep Critics for Scalable Reinforcement Learning

    Guhyeon Kang, Jaehwi Lee, Minhae Kwon

    cs.LGarXiv:2608.26659v12026
  35. Predicting Quantifiability from Primary Screens to Prioritize Dose-Response Profiling

    Sean Lim

    cs.LGq-bio.BMarXiv:2608.26538v12026
  36. CG4AI: A Column Generation Framework for Training AI Models Under Constraints

    Youcef Magnouche, Abderrahmane Driouch, Sébastien Martin +1

    cs.LGcs.AIcs.DMarXiv:2608.26375v12026
  37. Why ML-based cough models do not generalize: a systematic cross-dataset evaluation for tuberculosis screening

    Wensi Zhang, Tomas Teijeiro, Jérôme Thevenot +1

    eess.AScs.AIcs.LGarXiv:2608.25846v12026
  38. MetaSieve: Faster Relational Deep Learning through SQL-Based Metapath Selection

    Fahim Shahriar Khan, Ashraf Aboulnaga

    cs.DBcs.LGarXiv:2608.25903v12026
  39. Frequency-aware forecasting for short-term typhoon gust prediction

    Xuefei Wang, Tingyi Liu, Heng Zhang +1

    cs.LGarXiv:2608.25604v12026
  40. Functional linear regression from sparse to dense designs: a pooling-ridge method and minimax optimality

    Shunxing Yan, Fang Yao

    stat.MEcs.LGmath.STarXiv:2608.25468v12026
  41. Sundial: A Family of Highly Capable Time Series Foundation Models

    Yong Liu, Guo Qin, Zhiyuan Shi +5

    cs.LGarXiv:2502.00816v42025
  42. Multi-Modal Anomaly Detection: A Survey

    Xudong Mou, Zexin Wu, Chuan Luo +4

    cs.LGcs.AIarXiv:2608.24937v12026
  43. Compression Trinity: Exploring Sparsity, Quantization, and Low-Rank Approximations for LLM Compression

    Mohammad Mozaffari

    cs.AIcs.DCcs.LGarXiv:2608.24070v12026
  44. Provable Quantum--Classical Separation for Continuous Gibbs Sampling

    Enrico Olivucci, Mariia Sobchuk, Sehmimul Hoque +4

    quant-phcs.DScs.ETarXiv:2608.24527v12026
  45. A Tensorized Transformer for Language Modeling

    Xindian Ma, Peng Zhang, Shuai Zhang +4

    cs.CLcs.LGarXiv:1906.09777v32019
  46. A Theory of Speciation in Generative Diffusion Models on Compact Riemannian Manifolds

    Alessio Marta, Paola Causin

    cs.LGarXiv:2608.23798v12026
  47. Text Processing Like Humans Do: Visually Attacking and Shielding NLP Systems

    Steffen Eger, Gözde Gül Şahin, Andreas Rücklé +6

    cs.CLcs.CRcs.CVarXiv:1903.11508v22019
  48. The Axiomatic Trader: Latent Regularity, Information Budgets, and the Canonical Form of a Quantitative Investment System

    Jiayu Li

    cs.LGq-fin.PMarXiv:2608.23416v12026
    Summaries:한국어
  49. Constitutional Classifiers: Defending against Universal Jailbreaks across Thousands of Hours of Red Teaming

    Mrinank Sharma, Meg Tong, Jesse Mu +40

    cs.CLcs.AIcs.CRarXiv:2501.18837v12025
  50. Efficient Learning of Generalized Linear and Single Index Models with Isotonic Regression

    Sham Kakade, Adam Tauman Kalai, Varun Kanade +1

    cs.AIcs.LGstat.MLarXiv:1104.2018v12011
  51. Interpretable Deep Learning under Fire

    Xinyang Zhang, Ningfei Wang, Hua Shen +3

    cs.CRcs.LGarXiv:1812.00891v32018
  52. DiffusionSat: A Generative Foundation Model for Satellite Imagery

    Samar Khanna, Patrick Liu, Linqi Zhou +5

    cs.CVcs.AIcs.LGarXiv:2312.03606v22023
  53. Show, Attend and Distill:Knowledge Distillation via Attention-based Feature Matching

    Mingi Ji, Byeongho Heo, Sungrae Park

    cs.LGarXiv:2102.02973v12021
  54. Seed Diffusion: A Large-Scale Diffusion Language Model with High-Speed Inference

    Yuxuan Song, Zheng Zhang, Cheng Luo +19

    cs.CLcs.LGarXiv:2508.02193v12025
  55. Primal--Dual Alternating Neural Learning for Timely Classification with Performance Guarantees

    Jiaming Qiu, Yingye Zheng, Ying-Qi Zhao

    stat.MLcs.LGstat.MEarXiv:2608.23480v12026
  56. Mamba-3: Improved Sequence Modeling using State Space Principles

    Aakash Lahoti, Kevin Y. Li, Berlin Chen +5

    cs.LGarXiv:2603.15569v12026
  57. Poisson Subspace Clustering: Focusing on the Essentials in Count Data

    Collin Leiber, Kai Puolamäki, Heikki Mannila

    cs.LGarXiv:2608.23287v12026
  58. Symbolic Neural ODEs: Learning interpretable models from time-series data

    Nibodh Boddupalli, Jeff Moehlis

    cs.LGeess.SYmath.DSarXiv:2608.22112v12026
  59. FreKoo++: Learning Continuous Spectral Dynamics for Temporal Domain Generalization

    En Yu, Xiaoyu Yang, Wei Duan +2

    cs.LGcs.AIarXiv:2608.22224v12026
  60. Two-level domain-decomposition AdaGrad method for scalable training of graph neural networks

    Laurynas Varnas, Julien Herrmann, Alexander Heinlein +2

    math.NAcs.LGarXiv:2608.22575v12026