Machine Learning
Papers filed under cs.LG on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.
Search paper metadata (including unsummarized papers)
6,961 to 7,020 of 20,199
Optimal Errors and Phase Transitions in High-Dimensional Generalized Linear Models
Jean Barbier, Florent Krzakala, Nicolas Macris +2
cs.ITcond-mat.dis-nncs.AIarXiv:1708.03395v32017Fairness Beyond Disparate Treatment & Disparate Impact: Learning Classification without Disparate Mistreatment
Muhammad Bilal Zafar, Isabel Valera, Manuel Gomez Rodriguez +1
stat.MLcs.LGarXiv:1610.08452v22016Quantized Neural Networks: Training Neural Networks with Low Precision Weights and Activations
Itay Hubara, Matthieu Courbariaux, Daniel Soudry +2
cs.NEcs.LGarXiv:1609.07061v12016Stealing Machine Learning Models via Prediction APIs
Florian Tramèr, Fan Zhang, Ari Juels +2
cs.CRcs.LGstat.MLarXiv:1609.02943v22016Data Programming: Creating Large Training Sets, Quickly
Alexander Ratner, Christopher De Sa, Sen Wu +2
stat.MLcs.AIcs.LGarXiv:1605.07723v32016Thought Anchors: Which LLM Reasoning Steps Matter?
Paul C. Bogdan, Uzay Macar, Neel Nanda +1
cs.LGcs.AIcs.CLarXiv:2506.19143v42025GDPval: Evaluating AI Model Performance on Real-World Economically Valuable Tasks
Tejal Patwardhan, Rachel Dias, Elizabeth Proehl +16
cs.LGcs.AIcs.CYarXiv:2510.04374v12025Fast Convergence of Regularized Learning in Games
Vasilis Syrgkanis, Alekh Agarwal, Haipeng Luo +1
cs.GTcs.AIcs.LGarXiv:1507.00407v52015The Surprising Effectiveness of Negative Reinforcement in LLM Reasoning
Xinyu Zhu, Mengzhou Xia, Zhepei Wei +3
cs.CLcs.LGarXiv:2506.01347v22025Convex Optimization: Algorithms and Complexity
Sébastien Bubeck
math.OCcs.CCcs.LGarXiv:1405.4980v22014Tuned Models of Peer Assessment in MOOCs
Chris Piech, Jonathan Huang, Zhenghao Chen +3
cs.LGcs.AIcs.HCarXiv:1307.2579v12013On the Theoretical Limitations of Embedding-Based Retrieval
Orion Weller, Michael Boratko, Iftekhar Naim +1
cs.IRcs.CLcs.LGarXiv:2508.21038v22025Summaries:한국어Reactive Diffusion Policy: Slow-Fast Visual-Tactile Policy Learning for Contact-Rich Manipulation
Han Xue, Jieji Ren, Wendi Chen +5
cs.ROcs.AIcs.LGarXiv:2503.02881v32025Thompson Sampling for Contextual Bandits with Linear Payoffs
Shipra Agrawal, Navin Goyal
cs.LGcs.DSstat.MLarXiv:1209.3352v42012Adversarial-Learned Loss for Domain Adaptation
Minghao Chen, Shuai Zhao, Haifeng Liu +1
cs.CVcs.LGarXiv:2001.01046v12020Thompson Sampling: An Asymptotically Optimal Finite Time Analysis
Emilie Kaufmann, Nathaniel Korda, Rémi Munos
stat.MLcs.LGarXiv:1205.4217v22012Confidence-Aware Learning for Deep Neural Networks
Jooyoung Moon, Jihyo Kim, Younghak Shin +1
cs.LGstat.MLarXiv:2007.01458v32020A Text Classification Framework for Simple and Effective Early Depression Detection Over Social Media Streams
Sergio G. Burdisso, Marcelo Errecalde, Manuel Montes-y-Gómez
cs.CYcs.CLcs.IRarXiv:1905.08772v22019Light-R1: Curriculum SFT, DPO and RL for Long COT from Scratch and Beyond
Liang Wen, Yunke Cai, Fenrui Xiao +11
cs.CLcs.LGarXiv:2503.10460v42025IAIA-BL: A Case-based Interpretable Deep Learning Model for Classification of Mass Lesions in Digital Mammography
Alina Jade Barnett, Fides Regina Schwartz, Chaofan Tao +4
cs.LGcs.AIcs.CVarXiv:2103.12308v12021Language Modeling with Deep Transformers
Kazuki Irie, Albert Zeyer, Ralf Schlüter +1
cs.CLcs.LGarXiv:1905.04226v22019RM-R1: Reward Modeling as Reasoning
Xiusi Chen, Gaotang Li, Ziqi Wang +9
cs.CLcs.AIcs.LGarXiv:2505.02387v42025On Variance Reduction in Stochastic Gradient Descent and its Asynchronous Variants
Sashank J. Reddi, Ahmed Hefny, Suvrit Sra +2
cs.LGstat.MLarXiv:1506.06840v22015Spectral Convergence of Random Feature Method in Multiple Dimensions
Pingbing Ming, Hao Yu
math.NAcs.AIcs.LGarXiv:2609.03401v12026Bayes-Optimal BER and AUC: Estimation and Evaluation of Estimators
Ryota Ushio, Takashi Ishida, Masashi Sugiyama
cs.LGstat.MLarXiv:2609.02304v12026LLaDA-V: Large Language Diffusion Models with Visual Instruction Tuning
Zebin You, Shen Nie, Xiaolu Zhang +5
cs.LGcs.CLcs.CVarXiv:2505.16933v22025Preference Fine-Tuning of LLMs Should Leverage Suboptimal, On-Policy Data
Fahim Tajwar, Anikait Singh, Archit Sharma +6
cs.LGarXiv:2404.14367v32024Adversarial camera stickers: A physical camera-based attack on deep learning systems
Juncheng Li, Frank R. Schmidt, J. Zico Kolter
cs.CVcs.CRcs.LGarXiv:1904.00759v42019RL's Razor: Why Online Reinforcement Learning Forgets Less
Idan Shenfeld, Jyothish Pari, Pulkit Agrawal
cs.LGarXiv:2509.04259v12025Pseudo-Simulation for Autonomous Driving
Wei Cao, Marcel Hallgarten, Tianyu Li +11
cs.ROcs.AIcs.CVarXiv:2506.04218v32025Diffusion-Based Refinement for Kilometer-Scale Probabilistic Precipitation Nowcasting
Dohyun Park, Changhoon Song, Tengyuan Chang +2
cs.LGphysics.ao-pharXiv:2608.30205v12026Multiclass Linear Perceptrons with Multiplicative Margins
Dmitri Rachkovskij, Evgeny Osipov, Olexander Volkov +2
cs.LGcs.NEarXiv:2608.30028v12026Puro-2B: Poor Lab's Qwen2-1.5B Trained on RTX 5090 within $5090
Kairong Luo, Jiarui Cui, Yaorui Yin +8
cs.CLcs.LGarXiv:2608.27370v12026Simple Actors and Deep Critics for Scalable Reinforcement Learning
Guhyeon Kang, Jaehwi Lee, Minhae Kwon
cs.LGarXiv:2608.26659v12026Predicting Quantifiability from Primary Screens to Prioritize Dose-Response Profiling
Sean Lim
cs.LGq-bio.BMarXiv:2608.26538v12026CG4AI: A Column Generation Framework for Training AI Models Under Constraints
Youcef Magnouche, Abderrahmane Driouch, Sébastien Martin +1
cs.LGcs.AIcs.DMarXiv:2608.26375v12026Why ML-based cough models do not generalize: a systematic cross-dataset evaluation for tuberculosis screening
Wensi Zhang, Tomas Teijeiro, Jérôme Thevenot +1
eess.AScs.AIcs.LGarXiv:2608.25846v12026MetaSieve: Faster Relational Deep Learning through SQL-Based Metapath Selection
Fahim Shahriar Khan, Ashraf Aboulnaga
cs.DBcs.LGarXiv:2608.25903v12026Frequency-aware forecasting for short-term typhoon gust prediction
Xuefei Wang, Tingyi Liu, Heng Zhang +1
cs.LGarXiv:2608.25604v12026Functional linear regression from sparse to dense designs: a pooling-ridge method and minimax optimality
Shunxing Yan, Fang Yao
stat.MEcs.LGmath.STarXiv:2608.25468v12026Sundial: A Family of Highly Capable Time Series Foundation Models
Yong Liu, Guo Qin, Zhiyuan Shi +5
cs.LGarXiv:2502.00816v42025Multi-Modal Anomaly Detection: A Survey
Xudong Mou, Zexin Wu, Chuan Luo +4
cs.LGcs.AIarXiv:2608.24937v12026Compression Trinity: Exploring Sparsity, Quantization, and Low-Rank Approximations for LLM Compression
Mohammad Mozaffari
cs.AIcs.DCcs.LGarXiv:2608.24070v12026Provable Quantum--Classical Separation for Continuous Gibbs Sampling
Enrico Olivucci, Mariia Sobchuk, Sehmimul Hoque +4
quant-phcs.DScs.ETarXiv:2608.24527v12026A Tensorized Transformer for Language Modeling
Xindian Ma, Peng Zhang, Shuai Zhang +4
cs.CLcs.LGarXiv:1906.09777v32019A Theory of Speciation in Generative Diffusion Models on Compact Riemannian Manifolds
Alessio Marta, Paola Causin
cs.LGarXiv:2608.23798v12026Text Processing Like Humans Do: Visually Attacking and Shielding NLP Systems
Steffen Eger, Gözde Gül Şahin, Andreas Rücklé +6
cs.CLcs.CRcs.CVarXiv:1903.11508v22019The Axiomatic Trader: Latent Regularity, Information Budgets, and the Canonical Form of a Quantitative Investment System
Jiayu Li
cs.LGq-fin.PMarXiv:2608.23416v12026Summaries:한국어Constitutional Classifiers: Defending against Universal Jailbreaks across Thousands of Hours of Red Teaming
Mrinank Sharma, Meg Tong, Jesse Mu +40
cs.CLcs.AIcs.CRarXiv:2501.18837v12025Efficient Learning of Generalized Linear and Single Index Models with Isotonic Regression
Sham Kakade, Adam Tauman Kalai, Varun Kanade +1
cs.AIcs.LGstat.MLarXiv:1104.2018v12011Interpretable Deep Learning under Fire
Xinyang Zhang, Ningfei Wang, Hua Shen +3
cs.CRcs.LGarXiv:1812.00891v32018DiffusionSat: A Generative Foundation Model for Satellite Imagery
Samar Khanna, Patrick Liu, Linqi Zhou +5
cs.CVcs.AIcs.LGarXiv:2312.03606v22023Show, Attend and Distill:Knowledge Distillation via Attention-based Feature Matching
Mingi Ji, Byeongho Heo, Sungrae Park
cs.LGarXiv:2102.02973v12021Seed Diffusion: A Large-Scale Diffusion Language Model with High-Speed Inference
Yuxuan Song, Zheng Zhang, Cheng Luo +19
cs.CLcs.LGarXiv:2508.02193v12025Primal--Dual Alternating Neural Learning for Timely Classification with Performance Guarantees
Jiaming Qiu, Yingye Zheng, Ying-Qi Zhao
stat.MLcs.LGstat.MEarXiv:2608.23480v12026Mamba-3: Improved Sequence Modeling using State Space Principles
Aakash Lahoti, Kevin Y. Li, Berlin Chen +5
cs.LGarXiv:2603.15569v12026Poisson Subspace Clustering: Focusing on the Essentials in Count Data
Collin Leiber, Kai Puolamäki, Heikki Mannila
cs.LGarXiv:2608.23287v12026Symbolic Neural ODEs: Learning interpretable models from time-series data
Nibodh Boddupalli, Jeff Moehlis
cs.LGeess.SYmath.DSarXiv:2608.22112v12026FreKoo++: Learning Continuous Spectral Dynamics for Temporal Domain Generalization
En Yu, Xiaoyu Yang, Wei Duan +2
cs.LGcs.AIarXiv:2608.22224v12026Two-level domain-decomposition AdaGrad method for scalable training of graph neural networks
Laurynas Varnas, Julien Herrmann, Alexander Heinlein +2
math.NAcs.LGarXiv:2608.22575v12026