Machine Learning
Papers filed under cs.LG on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.
Search paper metadata (including unsummarized papers)
361 to 420 of 20,193
A Unified Perspective on Conformal Prediction and Wasserstein Distributionally Robust Optimization for Uncertainty Quantification
Kehan Long, Yiqi Zhao, Pol Mestres +3
math.OCcs.LGeess.SYarXiv:2608.29789v12026Machine learning for neural decoding
Joshua I. Glaser, Ari S. Benjamin, Raeed H. Chowdhury +3
q-bio.NCcs.LGstat.MLarXiv:1708.00909v42017BERTology of Molecular Property Prediction
Mohammad Mostafanejad, Paul Saxe, T. Daniel Crawford
cs.LGcs.CLarXiv:2603.13627v12026Beyond Non-IID: Learner--Client Distribution Mismatch in Federated Learning
Yiming Xie, Lili Su, Ningfang Mi
cs.LGarXiv:2608.27715v12026Equal Ranking Quality, Different Decisions: Training Order-Consistent LLM Scorers
Markus Frohmann, Mahdiyar Alavi, Elizabeth Lingg +1
cs.CLcs.IRcs.LGarXiv:2608.26762v12026Daydreaming: Stealing Hidden Agent Skills through Black-Box Task Interaction
Yu-Lin Tsai, Yu-An Lu, Ci-Yang Tsai +3
cs.CRcs.AIcs.LGarXiv:2608.26733v12026Evaluating Memory Structure in LLM Agents
Alina Shutova, Alexandra Olenina, Ivan Vinogradov +1
cs.LGcs.CLarXiv:2602.11243v22026A meta-algorithm for ab initio reconstruction of complex mixtures in cryo-EM
Alkin Kaz, Arda Kaz, Ellen D. Zhong
q-bio.BMcs.LGarXiv:2608.25388v12026GENIUS: Generative Fluid Intelligence Evaluation Suite
Ruichuan An, Sihan Yang, Ziyu Guo +8
cs.LGcs.AIcs.CVarXiv:2602.11144v12026FuzzingBrain-Bench V1: Evaluating Open-Ended Bug Discovery by LLMs
Ze Sheng, Aleksandar Kezic, Zhicheng Chen +1
cs.AIcs.CRcs.LGarXiv:2608.25158v12026Scalable Self-Supervised Learning for Multiphase AC-OPF in Distribution Systems with Topology Reconfiguration
Hoang T. Nguyen, Shaohui Liu, Reetam Sen Biswas +4
eess.SYcs.LGmath.OCarXiv:2608.25095v12026Reasoning Cache: Continual Improvement Over Long Horizons via Short-Horizon RL
Ian Wu, Yuxiao Qu, Amrith Setlur +1
cs.LGarXiv:2602.03773v22026From Numerical Simulators of PDEs to Neural Emulators and Back
Felix Koehler
cs.LGarXiv:2608.24547v12026An Empirical Study of World Model Quantization
Zhongqian Fu, Tianyi Zhao, Kai Han +3
cs.LGcs.CVarXiv:2602.02110v12026Unsupervised Monocular Depth Estimation with Left-Right Consistency
Clément Godard, Oisin Mac Aodha, Gabriel J. Brostow
cs.CVcs.LGstat.MLarXiv:1609.03677v32016It depends: Incorporating correlations for joint aleatoric and epistemic uncertainties of high-dimensional output spaces
Leonhard F. Feiner, Manuel Nickel, Martin Menten +6
cs.LGcs.CVarXiv:2608.24518v12026Improved Techniques for Training GANs
Tim Salimans, Ian Goodfellow, Wojciech Zaremba +3
cs.LGcs.CVcs.NEarXiv:1606.03498v12016LUCAID: Agentic Multimodal AI for Lung Cancer Precision Pathology
Marie-Lisa Eich, Kai Standvoss, Timo Milbich +30
cs.CVcs.AIcs.LGarXiv:2608.23803v12026Photorealistic Novel View Synthesis of Human Faces using Next-Scale Transformers
Federico Stella, Fei Jiang, Zhongshi Jiang +4
cs.CVcs.LGarXiv:2608.23410v12026Less Noise, More Voice: Reinforcement Learning for Reasoning via Instruction Purification
Yiju Guo, Tianyi Hu, Zexu Sun +1
cs.LGcs.AIcs.CLarXiv:2601.21244v32026Deep Knowledge Tracing
Chris Piech, Jonathan Spencer, Jonathan Huang +4
cs.AIcs.CYcs.LGarXiv:1506.05908v12015Fundamental Limitations of Favorable Privacy-Utility Guarantees for DP-SGD
Murat Bilgehan Ertan, Marten van Dijk
cs.LGcs.CRarXiv:2601.10237v32026Barycentric Fused Gromov-Wasserstein Balancing for Causal Inference under Multiple Treatments
Yuki Murakami, Takumi Hattori, Kohsuke Kubota
stat.MEcs.AIcs.LGarXiv:2608.22024v12026Transition Matching Distillation for Fast Video Generation
Weili Nie, Julius Berner, Nanye Ma +3
cs.CVcs.AIcs.LGarXiv:2601.09881v22026MirrorBench: A Benchmark to Evaluate Conversational User-Proxy Agents for Human-Likeness
Ashutosh Hathidara, Julien Yu, Vaishali Senthil +2
cs.AIcs.LGarXiv:2601.08118v32026Llama-Mobile: Efficient 2.7-Bit Quantization of VLMs
Luka Ribar, Jeevan Bhoot, Douglas Orr
cs.CVcs.LGarXiv:2608.21134v12026Comparison of Bayesian predictive methods for model selection
Juho Piironen, Aki Vehtari
stat.MEcs.LGarXiv:1503.08650v42015CDRL: Certification-Driven Reinforcement Learning for Neutrino Flavor Model Discovery
Piyush Jha, Jake Rudolph, Victoria Knapp-Pérez +3
cs.AIcs.LGcs.LOarXiv:2608.20686v12026The Principles of Diffusion Models
Chieh-Hsin Lai, Yang Song, Dongjun Kim +2
cs.LGcs.AIcs.GRarXiv:2510.21890v32025End-to-end Continuous Speech Recognition using Attention-based Recurrent NN: First Results
Jan Chorowski, Dzmitry Bahdanau, Kyunghyun Cho +1
cs.NEcs.LGstat.MLarXiv:1412.1602v12014What is Missing from AI Post-Training AI: An Empirical Analysis
Joy Jia Yin Lim, Xin Huang, Hao Peng +5
cs.AIcs.CLcs.LGarXiv:2608.19072v12026Flama: a Python framework for development and deployment of production-ready APIs, machine learning, and LLM services
José A. Perdiguero López, Miguel A. Durán-Olivencia
cs.SEcs.AIcs.LGarXiv:2608.18733v12026Partition the Support, Reconstruct the Residual: Training-Free Sparse Attention for Video Generation and World Models
Pardis Taghavi, Reza Langari, Gaurav Pandey
cs.CVcs.AIcs.LGarXiv:2608.18484v12026Gradient Descent on Neural Networks Typically Occurs at the Edge of Stability
Jeremy M. Cohen, Simran Kaur, Yuanzhi Li +2
cs.LGstat.MLarXiv:2103.00065v32021MapAnything: Universal Feed-Forward Metric 3D Reconstruction
Nikhil Keetha, Norman Müller, Johannes Schönberger +14
cs.CVcs.AIcs.LGarXiv:2509.13414v32025A Survey of Reinforcement Learning for Large Reasoning Models
Kaiyan Zhang, Yuxin Zuo, Bingxiang He +36
cs.CLcs.AIcs.LGarXiv:2509.08827v32025Deep Think with Confidence
Yichao Fu, Xuewei Wang, Yuandong Tian +1
cs.LGarXiv:2508.15260v12025Degradation-Aligned Self-Supervised Learning for State of Health Estimation of Lithium-Ion Batteries under Label Sparsity
Jiaqi Yao, Julia Kowal
eess.SPcs.AIcs.LGarXiv:2608.16612v12026Discrete Diffusion in Large Language and Multimodal Models: A Survey
Runpeng Yu, Qi Li, Xinchao Wang
cs.LGcs.AIarXiv:2506.13759v52025AlphaEvolve: A coding agent for scientific and algorithmic discovery
Alexander Novikov, Ngân Vũ, Marvin Eisenberger +15
cs.AIcs.LGcs.NEarXiv:2506.13131v12025Semi-supervised clustering methods
Eric Bair
stat.MEcs.LGstat.MLarXiv:1307.0252v12013Mapping the Space of Chemical Reactions Using Attention-Based Neural Networks
Philippe Schwaller, Daniel Probst, Alain C. Vaucher +4
physics.chem-phcs.CLcs.LGarXiv:2012.06051v12020Confidence Intervals and Hypothesis Testing for High-Dimensional Regression
Adel Javanmard, Andrea Montanari
stat.MEcs.ITcs.LGarXiv:1306.3171v22013J1: Incentivizing Thinking in LLM-as-a-Judge via Reinforcement Learning
Chenxi Whitehouse, Tianlu Wang, Ping Yu +4
cs.CLcs.AIcs.LGarXiv:2505.10320v32025Absolute Zero: Reinforced Self-play Reasoning with Zero Data
Andrew Zhao, Yiran Wu, Yang Yue +8
cs.LGcs.AIcs.CLarXiv:2505.03335v32025Measuring Structured Predictability in Neural Training Dynamics: A Cross-Regime Study
Fanqi Wang, Weisheng Tang, Hairong Qi
cs.LGarXiv:2608.15483v12026Interpreting Graph Neural Networks for NLP With Differentiable Edge Masking
Michael Sejr Schlichtkrull, Nicola De Cao, Ivan Titov
cs.CLcs.LGstat.MLarXiv:2010.00577v32020A Survey on Multi-view Learning
Chang Xu, Dacheng Tao, Chao Xu
cs.LGarXiv:1304.5634v12013Language models suffer from a curse of ambiguity
Nicolas Zucchet, Hyun Dong Lee, Scott Linderman
cs.CLcs.LGcs.NEarXiv:2608.15448v12026Contrastive Self-supervised Learning for Graph Classification
Jiaqi Zeng, Pengtao Xie
cs.LGstat.MLarXiv:2009.05923v12020Diagnosing and Mitigating Perception-Decision Misalignment in Omni-LLMs via Modality Subspace Activation
Hongbo Jiang, Jie Li, Yunhang Shen +2
cs.LGcs.CVarXiv:2608.14655v12026Calibrated Trust, Not Sharper Prediction: An Empirical Test of Uncertainty Fusion
Surya Saka
cs.LGcs.AIcs.CLarXiv:2608.14617v12026NSGANetV2: Evolutionary Multi-Objective Surrogate-Assisted Neural Architecture Search
Zhichao Lu, Kalyanmoy Deb, Erik Goodman +2
cs.CVcs.LGcs.NEarXiv:2007.10396v12020Distribution Aligning Refinery of Pseudo-label for Imbalanced Semi-supervised Learning
Jaehyung Kim, Youngbum Hur, Sejun Park +3
cs.LGstat.MLarXiv:2007.08844v22020CardioState-JEPA: Delay-Aware Cross-Modal Learning of a Shared Cardiac Representation
Hamza Shafiq, Hung Manh Pham, Bin Zhu +3
cs.LGeess.IVstat.MLarXiv:2608.12944v12026Hands-on Bayesian Neural Networks -- a Tutorial for Deep Learning Users
Laurent Valentin Jospin, Wray Buntine, Farid Boussaid +2
cs.LGstat.MLarXiv:2007.06823v32020Multiscale Simulations of Complex Systems by Learning their Effective Dynamics
Pantelis R. Vlachas, Georgios Arampatzis, Caroline Uhler +1
physics.comp-phcs.LGnlin.CDarXiv:2006.13431v32020A Bayesian Approach to Robust Inverse Reinforcement Learning
Ran Wei, Siliang Zeng, Chenliang Li +3
cs.LGarXiv:2309.08571v22023TokenSqueeze: Performance-Preserving Compression for Reasoning LLMs
Yuxiang Zhang, Zhengxu Yu, Weihang Pan +5
cs.LGcs.AIarXiv:2511.13223v12025TxAgent: An AI Agent for Therapeutic Reasoning Across a Universe of Tools
Shanghua Gao, Richard Zhu, Zhenglun Kong +5
cs.AIcs.LGarXiv:2503.10970v12025