Machine Learning

Papers filed under cs.LG on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.

Search paper metadata (including unsummarized papers)

361 to 420 of 20,193

  1. A Unified Perspective on Conformal Prediction and Wasserstein Distributionally Robust Optimization for Uncertainty Quantification

    Kehan Long, Yiqi Zhao, Pol Mestres +3

    math.OCcs.LGeess.SYarXiv:2608.29789v12026
  2. Machine learning for neural decoding

    Joshua I. Glaser, Ari S. Benjamin, Raeed H. Chowdhury +3

    q-bio.NCcs.LGstat.MLarXiv:1708.00909v42017
  3. BERTology of Molecular Property Prediction

    Mohammad Mostafanejad, Paul Saxe, T. Daniel Crawford

    cs.LGcs.CLarXiv:2603.13627v12026
  4. Beyond Non-IID: Learner--Client Distribution Mismatch in Federated Learning

    Yiming Xie, Lili Su, Ningfang Mi

    cs.LGarXiv:2608.27715v12026
  5. Equal Ranking Quality, Different Decisions: Training Order-Consistent LLM Scorers

    Markus Frohmann, Mahdiyar Alavi, Elizabeth Lingg +1

    cs.CLcs.IRcs.LGarXiv:2608.26762v12026
  6. Daydreaming: Stealing Hidden Agent Skills through Black-Box Task Interaction

    Yu-Lin Tsai, Yu-An Lu, Ci-Yang Tsai +3

    cs.CRcs.AIcs.LGarXiv:2608.26733v12026
  7. Evaluating Memory Structure in LLM Agents

    Alina Shutova, Alexandra Olenina, Ivan Vinogradov +1

    cs.LGcs.CLarXiv:2602.11243v22026
  8. A meta-algorithm for ab initio reconstruction of complex mixtures in cryo-EM

    Alkin Kaz, Arda Kaz, Ellen D. Zhong

    q-bio.BMcs.LGarXiv:2608.25388v12026
  9. GENIUS: Generative Fluid Intelligence Evaluation Suite

    Ruichuan An, Sihan Yang, Ziyu Guo +8

    cs.LGcs.AIcs.CVarXiv:2602.11144v12026
  10. FuzzingBrain-Bench V1: Evaluating Open-Ended Bug Discovery by LLMs

    Ze Sheng, Aleksandar Kezic, Zhicheng Chen +1

    cs.AIcs.CRcs.LGarXiv:2608.25158v12026
  11. Scalable Self-Supervised Learning for Multiphase AC-OPF in Distribution Systems with Topology Reconfiguration

    Hoang T. Nguyen, Shaohui Liu, Reetam Sen Biswas +4

    eess.SYcs.LGmath.OCarXiv:2608.25095v12026
  12. Reasoning Cache: Continual Improvement Over Long Horizons via Short-Horizon RL

    Ian Wu, Yuxiao Qu, Amrith Setlur +1

    cs.LGarXiv:2602.03773v22026
  13. From Numerical Simulators of PDEs to Neural Emulators and Back

    Felix Koehler

    cs.LGarXiv:2608.24547v12026
  14. An Empirical Study of World Model Quantization

    Zhongqian Fu, Tianyi Zhao, Kai Han +3

    cs.LGcs.CVarXiv:2602.02110v12026
  15. Unsupervised Monocular Depth Estimation with Left-Right Consistency

    Clément Godard, Oisin Mac Aodha, Gabriel J. Brostow

    cs.CVcs.LGstat.MLarXiv:1609.03677v32016
  16. It depends: Incorporating correlations for joint aleatoric and epistemic uncertainties of high-dimensional output spaces

    Leonhard F. Feiner, Manuel Nickel, Martin Menten +6

    cs.LGcs.CVarXiv:2608.24518v12026
  17. Improved Techniques for Training GANs

    Tim Salimans, Ian Goodfellow, Wojciech Zaremba +3

    cs.LGcs.CVcs.NEarXiv:1606.03498v12016
  18. LUCAID: Agentic Multimodal AI for Lung Cancer Precision Pathology

    Marie-Lisa Eich, Kai Standvoss, Timo Milbich +30

    cs.CVcs.AIcs.LGarXiv:2608.23803v12026
  19. Photorealistic Novel View Synthesis of Human Faces using Next-Scale Transformers

    Federico Stella, Fei Jiang, Zhongshi Jiang +4

    cs.CVcs.LGarXiv:2608.23410v12026
  20. Less Noise, More Voice: Reinforcement Learning for Reasoning via Instruction Purification

    Yiju Guo, Tianyi Hu, Zexu Sun +1

    cs.LGcs.AIcs.CLarXiv:2601.21244v32026
  21. Deep Knowledge Tracing

    Chris Piech, Jonathan Spencer, Jonathan Huang +4

    cs.AIcs.CYcs.LGarXiv:1506.05908v12015
  22. Fundamental Limitations of Favorable Privacy-Utility Guarantees for DP-SGD

    Murat Bilgehan Ertan, Marten van Dijk

    cs.LGcs.CRarXiv:2601.10237v32026
  23. Barycentric Fused Gromov-Wasserstein Balancing for Causal Inference under Multiple Treatments

    Yuki Murakami, Takumi Hattori, Kohsuke Kubota

    stat.MEcs.AIcs.LGarXiv:2608.22024v12026
  24. Transition Matching Distillation for Fast Video Generation

    Weili Nie, Julius Berner, Nanye Ma +3

    cs.CVcs.AIcs.LGarXiv:2601.09881v22026
  25. MirrorBench: A Benchmark to Evaluate Conversational User-Proxy Agents for Human-Likeness

    Ashutosh Hathidara, Julien Yu, Vaishali Senthil +2

    cs.AIcs.LGarXiv:2601.08118v32026
  26. Llama-Mobile: Efficient 2.7-Bit Quantization of VLMs

    Luka Ribar, Jeevan Bhoot, Douglas Orr

    cs.CVcs.LGarXiv:2608.21134v12026
  27. Comparison of Bayesian predictive methods for model selection

    Juho Piironen, Aki Vehtari

    stat.MEcs.LGarXiv:1503.08650v42015
  28. CDRL: Certification-Driven Reinforcement Learning for Neutrino Flavor Model Discovery

    Piyush Jha, Jake Rudolph, Victoria Knapp-Pérez +3

    cs.AIcs.LGcs.LOarXiv:2608.20686v12026
  29. The Principles of Diffusion Models

    Chieh-Hsin Lai, Yang Song, Dongjun Kim +2

    cs.LGcs.AIcs.GRarXiv:2510.21890v32025
  30. End-to-end Continuous Speech Recognition using Attention-based Recurrent NN: First Results

    Jan Chorowski, Dzmitry Bahdanau, Kyunghyun Cho +1

    cs.NEcs.LGstat.MLarXiv:1412.1602v12014
  31. What is Missing from AI Post-Training AI: An Empirical Analysis

    Joy Jia Yin Lim, Xin Huang, Hao Peng +5

    cs.AIcs.CLcs.LGarXiv:2608.19072v12026
  32. Flama: a Python framework for development and deployment of production-ready APIs, machine learning, and LLM services

    José A. Perdiguero López, Miguel A. Durán-Olivencia

    cs.SEcs.AIcs.LGarXiv:2608.18733v12026
  33. Partition the Support, Reconstruct the Residual: Training-Free Sparse Attention for Video Generation and World Models

    Pardis Taghavi, Reza Langari, Gaurav Pandey

    cs.CVcs.AIcs.LGarXiv:2608.18484v12026
  34. Gradient Descent on Neural Networks Typically Occurs at the Edge of Stability

    Jeremy M. Cohen, Simran Kaur, Yuanzhi Li +2

    cs.LGstat.MLarXiv:2103.00065v32021
  35. MapAnything: Universal Feed-Forward Metric 3D Reconstruction

    Nikhil Keetha, Norman Müller, Johannes Schönberger +14

    cs.CVcs.AIcs.LGarXiv:2509.13414v32025
  36. A Survey of Reinforcement Learning for Large Reasoning Models

    Kaiyan Zhang, Yuxin Zuo, Bingxiang He +36

    cs.CLcs.AIcs.LGarXiv:2509.08827v32025
  37. Deep Think with Confidence

    Yichao Fu, Xuewei Wang, Yuandong Tian +1

    cs.LGarXiv:2508.15260v12025
  38. Degradation-Aligned Self-Supervised Learning for State of Health Estimation of Lithium-Ion Batteries under Label Sparsity

    Jiaqi Yao, Julia Kowal

    eess.SPcs.AIcs.LGarXiv:2608.16612v12026
  39. Discrete Diffusion in Large Language and Multimodal Models: A Survey

    Runpeng Yu, Qi Li, Xinchao Wang

    cs.LGcs.AIarXiv:2506.13759v52025
  40. AlphaEvolve: A coding agent for scientific and algorithmic discovery

    Alexander Novikov, Ngân Vũ, Marvin Eisenberger +15

    cs.AIcs.LGcs.NEarXiv:2506.13131v12025
  41. Semi-supervised clustering methods

    Eric Bair

    stat.MEcs.LGstat.MLarXiv:1307.0252v12013
  42. Mapping the Space of Chemical Reactions Using Attention-Based Neural Networks

    Philippe Schwaller, Daniel Probst, Alain C. Vaucher +4

    physics.chem-phcs.CLcs.LGarXiv:2012.06051v12020
  43. Confidence Intervals and Hypothesis Testing for High-Dimensional Regression

    Adel Javanmard, Andrea Montanari

    stat.MEcs.ITcs.LGarXiv:1306.3171v22013
  44. J1: Incentivizing Thinking in LLM-as-a-Judge via Reinforcement Learning

    Chenxi Whitehouse, Tianlu Wang, Ping Yu +4

    cs.CLcs.AIcs.LGarXiv:2505.10320v32025
  45. Absolute Zero: Reinforced Self-play Reasoning with Zero Data

    Andrew Zhao, Yiran Wu, Yang Yue +8

    cs.LGcs.AIcs.CLarXiv:2505.03335v32025
  46. Measuring Structured Predictability in Neural Training Dynamics: A Cross-Regime Study

    Fanqi Wang, Weisheng Tang, Hairong Qi

    cs.LGarXiv:2608.15483v12026
  47. Interpreting Graph Neural Networks for NLP With Differentiable Edge Masking

    Michael Sejr Schlichtkrull, Nicola De Cao, Ivan Titov

    cs.CLcs.LGstat.MLarXiv:2010.00577v32020
  48. A Survey on Multi-view Learning

    Chang Xu, Dacheng Tao, Chao Xu

    cs.LGarXiv:1304.5634v12013
  49. Language models suffer from a curse of ambiguity

    Nicolas Zucchet, Hyun Dong Lee, Scott Linderman

    cs.CLcs.LGcs.NEarXiv:2608.15448v12026
  50. Contrastive Self-supervised Learning for Graph Classification

    Jiaqi Zeng, Pengtao Xie

    cs.LGstat.MLarXiv:2009.05923v12020
  51. Diagnosing and Mitigating Perception-Decision Misalignment in Omni-LLMs via Modality Subspace Activation

    Hongbo Jiang, Jie Li, Yunhang Shen +2

    cs.LGcs.CVarXiv:2608.14655v12026
  52. Calibrated Trust, Not Sharper Prediction: An Empirical Test of Uncertainty Fusion

    Surya Saka

    cs.LGcs.AIcs.CLarXiv:2608.14617v12026
  53. NSGANetV2: Evolutionary Multi-Objective Surrogate-Assisted Neural Architecture Search

    Zhichao Lu, Kalyanmoy Deb, Erik Goodman +2

    cs.CVcs.LGcs.NEarXiv:2007.10396v12020
  54. Distribution Aligning Refinery of Pseudo-label for Imbalanced Semi-supervised Learning

    Jaehyung Kim, Youngbum Hur, Sejun Park +3

    cs.LGstat.MLarXiv:2007.08844v22020
  55. CardioState-JEPA: Delay-Aware Cross-Modal Learning of a Shared Cardiac Representation

    Hamza Shafiq, Hung Manh Pham, Bin Zhu +3

    cs.LGeess.IVstat.MLarXiv:2608.12944v12026
  56. Hands-on Bayesian Neural Networks -- a Tutorial for Deep Learning Users

    Laurent Valentin Jospin, Wray Buntine, Farid Boussaid +2

    cs.LGstat.MLarXiv:2007.06823v32020
  57. Multiscale Simulations of Complex Systems by Learning their Effective Dynamics

    Pantelis R. Vlachas, Georgios Arampatzis, Caroline Uhler +1

    physics.comp-phcs.LGnlin.CDarXiv:2006.13431v32020
  58. A Bayesian Approach to Robust Inverse Reinforcement Learning

    Ran Wei, Siliang Zeng, Chenliang Li +3

    cs.LGarXiv:2309.08571v22023
  59. TokenSqueeze: Performance-Preserving Compression for Reasoning LLMs

    Yuxiang Zhang, Zhengxu Yu, Weihang Pan +5

    cs.LGcs.AIarXiv:2511.13223v12025
  60. TxAgent: An AI Agent for Therapeutic Reasoning Across a Universe of Tools

    Shanghua Gao, Richard Zhu, Zhenglun Kong +5

    cs.AIcs.LGarXiv:2503.10970v12025