Machine Learning (stat)
Papers filed under stat.ML on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.
Search paper metadata (including unsummarized papers)
1,681 to 1,740 of 6,774
What Makes a Reward Model a Good Teacher? An Optimization Perspective
Noam Razin, Zixuan Wang, Hubert Strauss +3
cs.LGcs.AIcs.CLarXiv:2503.15477v42025Measuring consistency via ensemble margin and local prediction variability: Auditing decision systems in the presence of predictive multiplicity
Sinjini Banerjee, Tim Marrinan, Anand D. Sarwate
stat.MLcs.AIcs.LGarXiv:2609.01397v12026Spatial Broadcast Decoder: A Simple Architecture for Learning Disentangled Representations in VAEs
Nicholas Watters, Loic Matthey, Christopher P. Burgess +1
cs.LGcs.CVstat.MLarXiv:1901.07017v22019A Wholistic View of Continual Learning with Deep Neural Networks: Forgotten Lessons and the Bridge to Active and Open World Learning
Martin Mundt, Yongwon Hong, Iuliia Pliushch +1
cs.LGstat.MLarXiv:2009.01797v32020Embedded Conditional Independence Tests for Large Language Model Generated Text with an Application to German Parliament Speeches
Marco Simnacher, Georg Keilbar, Benjamin König +2
stat.MLcs.AIcs.LGarXiv:2609.00946v12026Learning the Travelling Salesperson Problem Requires Rethinking Generalization
Chaitanya K. Joshi, Quentin Cappart, Louis-Martin Rousseau +1
cs.LGstat.MLarXiv:2006.07054v62020How to Make Causal Inferences Using Texts
Naoki Egami, Christian J. Fong, Justin Grimmer +2
stat.MLcs.CLstat.MEarXiv:1802.02163v12018Inference-Time Alignment in Diffusion Models with Reward-Guided Generation: Tutorial and Review
Masatoshi Uehara, Yulai Zhao, Chenyu Wang +4
cs.AIcs.LGq-bio.QMarXiv:2501.09685v22025Reconstructing Training Data from Trained Neural Networks
Niv Haim, Gal Vardi, Gilad Yehudai +2
cs.LGcs.CRcs.CVarXiv:2206.07758v32022AdaComp : Adaptive Residual Gradient Compression for Data-Parallel Distributed Training
Chia-Yu Chen, Jungwook Choi, Daniel Brand +3
cs.LGstat.MLarXiv:1712.02679v12017Sample Complexity Bounds for Stochastic Shortest Path with a Generative Model
Jean Tarbouriech, Matteo Pirotta, Michal Valko +1
cs.LGstat.MLarXiv:2604.16111v12026A Study of Reinforcement Learning for Neural Machine Translation
Lijun Wu, Fei Tian, Tao Qin +2
cs.LGcs.AIstat.MLarXiv:1808.08866v12018The challenge of realistic music generation: modelling raw audio at scale
Sander Dieleman, Aäron van den Oord, Karen Simonyan
cs.SDcs.LGeess.ASarXiv:1806.10474v12018Online EM Algorithm for Hidden Markov Models
Olivier Cappé
stat.COstat.MLarXiv:0908.2359v22009Iterative Amortized Inference
Joseph Marino, Yisong Yue, Stephan Mandt
cs.LGstat.MLarXiv:1807.09356v12018Effective Diversity in Population Based Reinforcement Learning
Jack Parker-Holder, Aldo Pacchiano, Krzysztof Choromanski +1
cs.LGstat.MLarXiv:2002.00632v32020A feature agnostic approach for glaucoma detection in OCT volumes
Stefan Maetschke, Bhavna Antony, Hiroshi Ishikawa +3
cs.CVcs.LGstat.MLarXiv:1807.04855v42018Parametrized quantum policies for reinforcement learning
Sofiene Jerbi, Casper Gyurik, Simon C. Marshall +2
quant-phcs.AIcs.LGarXiv:2103.05577v22021Data Banzhaf: A Robust Data Valuation Framework for Machine Learning
Jiachen T. Wang, Ruoxi Jia
cs.LGcs.GTstat.MLarXiv:2205.15466v72022SS-ESOAP: Self-Scaled Adaptive Preconditioning for Physics-Informed Learning
Guangyuan Wang, Mads Toftrup, Sebastian Loeschcke +2
cs.LGcs.AImath.OCarXiv:2608.29448v12026Hybrid Block Successive Approximation for One-Sided Non-Convex Min-Max Problems: Algorithms and Applications
Songtao Lu, Ioannis Tsaknakis, Mingyi Hong +1
math.OCstat.MLarXiv:1902.08294v22019Anchored Correlation Explanation: Topic Modeling with Minimal Domain Knowledge
Ryan J. Gallagher, Kyle Reing, David Kale +1
cs.CLcs.IRcs.ITarXiv:1611.10277v42016Reducing Overestimation Bias in Multi-Agent Domains Using Double Centralized Critics
Johannes Ackermann, Volker Gabler, Takayuki Osa +1
cs.LGcs.AIcs.MAarXiv:1910.01465v22019Leveraging the Feature Distribution in Transfer-based Few-Shot Learning
Yuqing Hu, Vincent Gripon, Stéphane Pateux
cs.LGstat.MLarXiv:2006.03806v32020Convergence of Gradient Descent on Separable Data
Mor Shpigel Nacson, Jason D. Lee, Suriya Gunasekar +3
stat.MLcs.LGarXiv:1803.01905v32018Pruning the Pilots: Deep Learning-Based Pilot Design and Channel Estimation for MIMO-OFDM Systems
Mahdi Boloursaz Mashhadi, Deniz Gunduz
cs.ITeess.SPstat.MLarXiv:2006.11796v32020Benchmarking Reinforcement Learning Algorithms on Real-World Robots
A. Rupam Mahmood, Dmytro Korenkevych, Gautham Vasan +2
cs.LGcs.AIcs.ROarXiv:1809.07731v12018Neural Embeddings of Graphs in Hyperbolic Space
Benjamin Paul Chamberlain, James Clough, Marc Peter Deisenroth
stat.MLcs.LGarXiv:1705.10359v12017Proximity Forest: An effective and scalable distance-based classifier for time series
Benjamin Lucas, Ahmed Shifaz, Charlotte Pelletier +5
cs.LGstat.MLarXiv:1808.10594v22018PhaseLink: A Deep Learning Approach to Seismic Phase Association
Zachary E. Ross, Yisong Yue, Men-Andrin Meier +2
cs.LGphysics.geo-phstat.MLarXiv:1809.02880v22018Application of machine learning for hematological diagnosis
Gregor Gunčar, Matjaž Kukar, Mateja Notar +4
stat.MLarXiv:1708.00253v12017Learning Neural Causal Models from Unknown Interventions
Nan Rosemary Ke, Olexa Bilaniuk, Anirudh Goyal +6
stat.MLcs.AIcs.LGarXiv:1910.01075v22019GenDICE: Generalized Offline Estimation of Stationary Values
Ruiyi Zhang, Bo Dai, Lihong Li +1
stat.MLcs.LGarXiv:2002.09072v12020Preventing Posterior Collapse with delta-VAEs
Ali Razavi, Aäron van den Oord, Ben Poole +1
cs.LGstat.MLarXiv:1901.03416v12019Imitation Learning as $f$-Divergence Minimization
Liyiming Ke, Sanjiban Choudhury, Matt Barnes +3
cs.LGcs.ITcs.ROarXiv:1905.12888v22019From Entropy to Epiplexity: Rethinking Information for Computationally Bounded Intelligence
Marc Finzi, Shikai Qiu, Yiding Jiang +3
cs.LGstat.MLarXiv:2601.03220v22026Learning Optimal and Fair Decision Trees for Non-Discriminative Decision-Making
Sina Aghaei, Mohammad Javad Azizi, Phebe Vayanos
cs.LGstat.MLarXiv:1903.10598v12019Adapting Auxiliary Losses Using Gradient Similarity
Yunshu Du, Wojciech M. Czarnecki, Siddhant M. Jayakumar +3
stat.MLcs.LGarXiv:1812.02224v22018Scaling Video Analytics on Constrained Edge Nodes
Christopher Canel, Thomas Kim, Giulio Zhou +5
cs.CVcs.LGcs.PFarXiv:1905.13536v12019Matched Queries for Curvature and Density at Branching Junctions
Ziqi Zhao, Qingjian Ni
stat.MLcs.LGarXiv:2609.01319v12026Variational inference for large-scale models of discrete choice
Michael Braun, Jon McAuliffe
stat.MEstat.COstat.MLarXiv:0712.2526v32007Tunable Efficient Unitary Neural Networks (EUNN) and their application to RNNs
Li Jing, Yichen Shen, Tena Dubček +5
cs.LGcs.NEstat.MLarXiv:1612.05231v32016Natural Neural Networks
Guillaume Desjardins, Karen Simonyan, Razvan Pascanu +1
stat.MLcs.LGcs.NEarXiv:1507.00210v12015Estimation from Pairwise Comparisons: Sharp Minimax Bounds with Topology Dependence
Nihar B. Shah, Sivaraman Balakrishnan, Joseph Bradley +3
cs.LGcs.ITstat.MLarXiv:1505.01462v12015Fully Parameterized Quantile Function for Distributional Reinforcement Learning
Derek Yang, Li Zhao, Zichuan Lin +3
cs.LGcs.AIstat.MLarXiv:1911.02140v32019A survey of algorithmic recourse: definitions, formulations, solutions, and prospects
Amir-Hossein Karimi, Gilles Barthe, Bernhard Schölkopf +1
cs.LGcs.AIstat.MLarXiv:2010.04050v22020Gluon: Making Muon & Scion Great Again! (Bridging Theory and Practice of LMO-based Optimizers for LLMs)
Artem Riabinin, Egor Shulgin, Kaja Gruntkowska +1
cs.LGmath.OCstat.MLarXiv:2505.13416v12025Voice Separation with an Unknown Number of Multiple Speakers
Eliya Nachmani, Yossi Adi, Lior Wolf
eess.AScs.LGcs.SDarXiv:2003.01531v42020Missing Data Imputation using Optimal Transport
Boris Muzellec, Julie Josse, Claire Boyer +1
stat.MLcs.LGarXiv:2002.03860v32020Solving Schrödinger Bridges via Maximum Likelihood
Francisco Vargas, Pierre Thodoroff, Neil D. Lawrence +1
stat.MLcs.LGarXiv:2106.02081v92021Learning with Good Feature Representations in Bandits and in RL with a Generative Model
Tor Lattimore, Csaba Szepesvari, Gellert Weisz
stat.MLcs.LGarXiv:1911.07676v22019From Pixels to Torques: Policy Learning with Deep Dynamical Models
Niklas Wahlström, Thomas B. Schön, Marc Peter Deisenroth
stat.MLcs.LGcs.ROarXiv:1502.02251v32015A PSO and Pattern Search based Memetic Algorithm for SVMs Parameters Optimization
Yukun Bao, Zhongyi Hu, Tao Xiong
cs.LGcs.AIcs.NEarXiv:1401.1926v12014Differentiable Ranks and Sorting using Optimal Transport
Marco Cuturi, Olivier Teboul, Jean-Philippe Vert
cs.LGstat.MLarXiv:1905.11885v22019Deep Neural Networks with Random Gaussian Weights: A Universal Classification Strategy?
Raja Giryes, Guillermo Sapiro, Alex M. Bronstein
cs.NEcs.LGstat.MLarXiv:1504.08291v52015HyperMC: Multi-Fidelity Hyperparameter Tuning for Stochastic Gradient MCMC
Ming Tan, Xiyun Jiao
stat.MLcs.LGstat.COarXiv:2609.02138v12026Influence Function based Data Poisoning Attacks to Top-N Recommender Systems
Minghong Fang, Neil Zhenqiang Gong, Jia Liu
cs.CRcs.IRcs.LGarXiv:2002.08025v32020Algorithms and Theory for Multiple-Source Adaptation
Judy Hoffman, Mehryar Mohri, Ningshan Zhang
cs.LGstat.MLarXiv:1805.08727v12018Approximate Nearest Neighbor Search in High Dimensions
Alexandr Andoni, Piotr Indyk, Ilya Razenshteyn
cs.DScs.CGcs.DBarXiv:1806.09823v12018Reconstructing subclonal composition and evolution from whole genome sequencing of tumors
Amit G. Deshwar, Shankar Vembu, Christina K. Yung +3
q-bio.PEcs.LGstat.MLarXiv:1406.7250v32014