Machine Learning (stat)

Papers filed under stat.ML on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.

Search paper metadata (including unsummarized papers)

1,681 to 1,740 of 6,774

  1. What Makes a Reward Model a Good Teacher? An Optimization Perspective

    Noam Razin, Zixuan Wang, Hubert Strauss +3

    cs.LGcs.AIcs.CLarXiv:2503.15477v42025
  2. Measuring consistency via ensemble margin and local prediction variability: Auditing decision systems in the presence of predictive multiplicity

    Sinjini Banerjee, Tim Marrinan, Anand D. Sarwate

    stat.MLcs.AIcs.LGarXiv:2609.01397v12026
  3. Spatial Broadcast Decoder: A Simple Architecture for Learning Disentangled Representations in VAEs

    Nicholas Watters, Loic Matthey, Christopher P. Burgess +1

    cs.LGcs.CVstat.MLarXiv:1901.07017v22019
  4. A Wholistic View of Continual Learning with Deep Neural Networks: Forgotten Lessons and the Bridge to Active and Open World Learning

    Martin Mundt, Yongwon Hong, Iuliia Pliushch +1

    cs.LGstat.MLarXiv:2009.01797v32020
  5. Embedded Conditional Independence Tests for Large Language Model Generated Text with an Application to German Parliament Speeches

    Marco Simnacher, Georg Keilbar, Benjamin König +2

    stat.MLcs.AIcs.LGarXiv:2609.00946v12026
  6. Learning the Travelling Salesperson Problem Requires Rethinking Generalization

    Chaitanya K. Joshi, Quentin Cappart, Louis-Martin Rousseau +1

    cs.LGstat.MLarXiv:2006.07054v62020
  7. How to Make Causal Inferences Using Texts

    Naoki Egami, Christian J. Fong, Justin Grimmer +2

    stat.MLcs.CLstat.MEarXiv:1802.02163v12018
  8. Inference-Time Alignment in Diffusion Models with Reward-Guided Generation: Tutorial and Review

    Masatoshi Uehara, Yulai Zhao, Chenyu Wang +4

    cs.AIcs.LGq-bio.QMarXiv:2501.09685v22025
  9. Reconstructing Training Data from Trained Neural Networks

    Niv Haim, Gal Vardi, Gilad Yehudai +2

    cs.LGcs.CRcs.CVarXiv:2206.07758v32022
  10. AdaComp : Adaptive Residual Gradient Compression for Data-Parallel Distributed Training

    Chia-Yu Chen, Jungwook Choi, Daniel Brand +3

    cs.LGstat.MLarXiv:1712.02679v12017
  11. Sample Complexity Bounds for Stochastic Shortest Path with a Generative Model

    Jean Tarbouriech, Matteo Pirotta, Michal Valko +1

    cs.LGstat.MLarXiv:2604.16111v12026
  12. A Study of Reinforcement Learning for Neural Machine Translation

    Lijun Wu, Fei Tian, Tao Qin +2

    cs.LGcs.AIstat.MLarXiv:1808.08866v12018
  13. The challenge of realistic music generation: modelling raw audio at scale

    Sander Dieleman, Aäron van den Oord, Karen Simonyan

    cs.SDcs.LGeess.ASarXiv:1806.10474v12018
  14. Online EM Algorithm for Hidden Markov Models

    Olivier Cappé

    stat.COstat.MLarXiv:0908.2359v22009
  15. Iterative Amortized Inference

    Joseph Marino, Yisong Yue, Stephan Mandt

    cs.LGstat.MLarXiv:1807.09356v12018
  16. Effective Diversity in Population Based Reinforcement Learning

    Jack Parker-Holder, Aldo Pacchiano, Krzysztof Choromanski +1

    cs.LGstat.MLarXiv:2002.00632v32020
  17. A feature agnostic approach for glaucoma detection in OCT volumes

    Stefan Maetschke, Bhavna Antony, Hiroshi Ishikawa +3

    cs.CVcs.LGstat.MLarXiv:1807.04855v42018
  18. Parametrized quantum policies for reinforcement learning

    Sofiene Jerbi, Casper Gyurik, Simon C. Marshall +2

    quant-phcs.AIcs.LGarXiv:2103.05577v22021
  19. Data Banzhaf: A Robust Data Valuation Framework for Machine Learning

    Jiachen T. Wang, Ruoxi Jia

    cs.LGcs.GTstat.MLarXiv:2205.15466v72022
  20. SS-ESOAP: Self-Scaled Adaptive Preconditioning for Physics-Informed Learning

    Guangyuan Wang, Mads Toftrup, Sebastian Loeschcke +2

    cs.LGcs.AImath.OCarXiv:2608.29448v12026
  21. Hybrid Block Successive Approximation for One-Sided Non-Convex Min-Max Problems: Algorithms and Applications

    Songtao Lu, Ioannis Tsaknakis, Mingyi Hong +1

    math.OCstat.MLarXiv:1902.08294v22019
  22. Anchored Correlation Explanation: Topic Modeling with Minimal Domain Knowledge

    Ryan J. Gallagher, Kyle Reing, David Kale +1

    cs.CLcs.IRcs.ITarXiv:1611.10277v42016
  23. Reducing Overestimation Bias in Multi-Agent Domains Using Double Centralized Critics

    Johannes Ackermann, Volker Gabler, Takayuki Osa +1

    cs.LGcs.AIcs.MAarXiv:1910.01465v22019
  24. Leveraging the Feature Distribution in Transfer-based Few-Shot Learning

    Yuqing Hu, Vincent Gripon, Stéphane Pateux

    cs.LGstat.MLarXiv:2006.03806v32020
  25. Convergence of Gradient Descent on Separable Data

    Mor Shpigel Nacson, Jason D. Lee, Suriya Gunasekar +3

    stat.MLcs.LGarXiv:1803.01905v32018
  26. Pruning the Pilots: Deep Learning-Based Pilot Design and Channel Estimation for MIMO-OFDM Systems

    Mahdi Boloursaz Mashhadi, Deniz Gunduz

    cs.ITeess.SPstat.MLarXiv:2006.11796v32020
  27. Benchmarking Reinforcement Learning Algorithms on Real-World Robots

    A. Rupam Mahmood, Dmytro Korenkevych, Gautham Vasan +2

    cs.LGcs.AIcs.ROarXiv:1809.07731v12018
  28. Neural Embeddings of Graphs in Hyperbolic Space

    Benjamin Paul Chamberlain, James Clough, Marc Peter Deisenroth

    stat.MLcs.LGarXiv:1705.10359v12017
  29. Proximity Forest: An effective and scalable distance-based classifier for time series

    Benjamin Lucas, Ahmed Shifaz, Charlotte Pelletier +5

    cs.LGstat.MLarXiv:1808.10594v22018
  30. PhaseLink: A Deep Learning Approach to Seismic Phase Association

    Zachary E. Ross, Yisong Yue, Men-Andrin Meier +2

    cs.LGphysics.geo-phstat.MLarXiv:1809.02880v22018
  31. Application of machine learning for hematological diagnosis

    Gregor Gunčar, Matjaž Kukar, Mateja Notar +4

    stat.MLarXiv:1708.00253v12017
  32. Learning Neural Causal Models from Unknown Interventions

    Nan Rosemary Ke, Olexa Bilaniuk, Anirudh Goyal +6

    stat.MLcs.AIcs.LGarXiv:1910.01075v22019
  33. GenDICE: Generalized Offline Estimation of Stationary Values

    Ruiyi Zhang, Bo Dai, Lihong Li +1

    stat.MLcs.LGarXiv:2002.09072v12020
  34. Preventing Posterior Collapse with delta-VAEs

    Ali Razavi, Aäron van den Oord, Ben Poole +1

    cs.LGstat.MLarXiv:1901.03416v12019
  35. Imitation Learning as $f$-Divergence Minimization

    Liyiming Ke, Sanjiban Choudhury, Matt Barnes +3

    cs.LGcs.ITcs.ROarXiv:1905.12888v22019
  36. From Entropy to Epiplexity: Rethinking Information for Computationally Bounded Intelligence

    Marc Finzi, Shikai Qiu, Yiding Jiang +3

    cs.LGstat.MLarXiv:2601.03220v22026
  37. Learning Optimal and Fair Decision Trees for Non-Discriminative Decision-Making

    Sina Aghaei, Mohammad Javad Azizi, Phebe Vayanos

    cs.LGstat.MLarXiv:1903.10598v12019
  38. Adapting Auxiliary Losses Using Gradient Similarity

    Yunshu Du, Wojciech M. Czarnecki, Siddhant M. Jayakumar +3

    stat.MLcs.LGarXiv:1812.02224v22018
  39. Scaling Video Analytics on Constrained Edge Nodes

    Christopher Canel, Thomas Kim, Giulio Zhou +5

    cs.CVcs.LGcs.PFarXiv:1905.13536v12019
  40. Matched Queries for Curvature and Density at Branching Junctions

    Ziqi Zhao, Qingjian Ni

    stat.MLcs.LGarXiv:2609.01319v12026
  41. Variational inference for large-scale models of discrete choice

    Michael Braun, Jon McAuliffe

    stat.MEstat.COstat.MLarXiv:0712.2526v32007
  42. Tunable Efficient Unitary Neural Networks (EUNN) and their application to RNNs

    Li Jing, Yichen Shen, Tena Dubček +5

    cs.LGcs.NEstat.MLarXiv:1612.05231v32016
  43. Natural Neural Networks

    Guillaume Desjardins, Karen Simonyan, Razvan Pascanu +1

    stat.MLcs.LGcs.NEarXiv:1507.00210v12015
  44. Estimation from Pairwise Comparisons: Sharp Minimax Bounds with Topology Dependence

    Nihar B. Shah, Sivaraman Balakrishnan, Joseph Bradley +3

    cs.LGcs.ITstat.MLarXiv:1505.01462v12015
  45. Fully Parameterized Quantile Function for Distributional Reinforcement Learning

    Derek Yang, Li Zhao, Zichuan Lin +3

    cs.LGcs.AIstat.MLarXiv:1911.02140v32019
  46. A survey of algorithmic recourse: definitions, formulations, solutions, and prospects

    Amir-Hossein Karimi, Gilles Barthe, Bernhard Schölkopf +1

    cs.LGcs.AIstat.MLarXiv:2010.04050v22020
  47. Gluon: Making Muon & Scion Great Again! (Bridging Theory and Practice of LMO-based Optimizers for LLMs)

    Artem Riabinin, Egor Shulgin, Kaja Gruntkowska +1

    cs.LGmath.OCstat.MLarXiv:2505.13416v12025
  48. Voice Separation with an Unknown Number of Multiple Speakers

    Eliya Nachmani, Yossi Adi, Lior Wolf

    eess.AScs.LGcs.SDarXiv:2003.01531v42020
  49. Missing Data Imputation using Optimal Transport

    Boris Muzellec, Julie Josse, Claire Boyer +1

    stat.MLcs.LGarXiv:2002.03860v32020
  50. Solving Schrödinger Bridges via Maximum Likelihood

    Francisco Vargas, Pierre Thodoroff, Neil D. Lawrence +1

    stat.MLcs.LGarXiv:2106.02081v92021
  51. Learning with Good Feature Representations in Bandits and in RL with a Generative Model

    Tor Lattimore, Csaba Szepesvari, Gellert Weisz

    stat.MLcs.LGarXiv:1911.07676v22019
  52. From Pixels to Torques: Policy Learning with Deep Dynamical Models

    Niklas Wahlström, Thomas B. Schön, Marc Peter Deisenroth

    stat.MLcs.LGcs.ROarXiv:1502.02251v32015
  53. A PSO and Pattern Search based Memetic Algorithm for SVMs Parameters Optimization

    Yukun Bao, Zhongyi Hu, Tao Xiong

    cs.LGcs.AIcs.NEarXiv:1401.1926v12014
  54. Differentiable Ranks and Sorting using Optimal Transport

    Marco Cuturi, Olivier Teboul, Jean-Philippe Vert

    cs.LGstat.MLarXiv:1905.11885v22019
  55. Deep Neural Networks with Random Gaussian Weights: A Universal Classification Strategy?

    Raja Giryes, Guillermo Sapiro, Alex M. Bronstein

    cs.NEcs.LGstat.MLarXiv:1504.08291v52015
  56. HyperMC: Multi-Fidelity Hyperparameter Tuning for Stochastic Gradient MCMC

    Ming Tan, Xiyun Jiao

    stat.MLcs.LGstat.COarXiv:2609.02138v12026
  57. Influence Function based Data Poisoning Attacks to Top-N Recommender Systems

    Minghong Fang, Neil Zhenqiang Gong, Jia Liu

    cs.CRcs.IRcs.LGarXiv:2002.08025v32020
  58. Algorithms and Theory for Multiple-Source Adaptation

    Judy Hoffman, Mehryar Mohri, Ningshan Zhang

    cs.LGstat.MLarXiv:1805.08727v12018
  59. Approximate Nearest Neighbor Search in High Dimensions

    Alexandr Andoni, Piotr Indyk, Ilya Razenshteyn

    cs.DScs.CGcs.DBarXiv:1806.09823v12018
  60. Reconstructing subclonal composition and evolution from whole genome sequencing of tumors

    Amit G. Deshwar, Shankar Vembu, Christina K. Yung +3

    q-bio.PEcs.LGstat.MLarXiv:1406.7250v32014