Machine Learning (stat)
Papers filed under stat.ML on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.
Search paper metadata (including unsummarized papers)
6,721 to 6,780 of 6,782
Dropout as a Bayesian Approximation: Representing Model Uncertainty in Deep Learning
Yarin Gal, Zoubin Ghahramani
stat.MLcs.LGarXiv:1506.02142v62015Diffusion Models Beat GANs on Image Synthesis
Prafulla Dhariwal, Alex Nichol
cs.LGcs.AIcs.CVarXiv:2105.05233v42021BART: Denoising Sequence-to-Sequence Pre-training for Natural Language Generation, Translation, and Comprehension
Mike Lewis, Yinhan Liu, Naman Goyal +5
cs.CLcs.LGstat.MLarXiv:1910.13461v12019Summaries:한국어mixup: Beyond Empirical Risk Minimization
Hongyi Zhang, Moustapha Cisse, Yann N. Dauphin +1
cs.LGstat.MLarXiv:1710.09412v22017Photo-Realistic Single Image Super-Resolution Using a Generative Adversarial Network
Christian Ledig, Lucas Theis, Ferenc Huszar +8
cs.CVstat.MLarXiv:1609.04802v52016node2vec: Scalable Feature Learning for Networks
Aditya Grover, Jure Leskovec
cs.SIcs.LGstat.MLarXiv:1607.00653v12016Feature Priming in Online Linear Regression: Sparse-Regret Lower Bounds and a Tight Univariate Rate
Huibo Xu, Shi Fu, Qixin Zhang +1
stat.MLcs.LGstat.AParXiv:2608.17573v12026Layer Normalization
Jimmy Lei Ba, Jamie Ryan Kiros, Geoffrey E. Hinton
stat.MLcs.LGarXiv:1607.06450v12016A Style-Based Generator Architecture for Generative Adversarial Networks
Tero Karras, Samuli Laine, Timo Aila
cs.NEcs.LGstat.MLarXiv:1812.04948v32018Towards Deep Learning Models Resistant to Adversarial Attacks
Aleksander Madry, Aleksandar Makelov, Ludwig Schmidt +2
stat.MLcs.LGcs.NEarXiv:1706.06083v42017Continuous control with deep reinforcement learning
Timothy P. Lillicrap, Jonathan J. Hunt, Alexander Pritzel +5
cs.LGstat.MLarXiv:1509.02971v62015Tight Bounds for Data-driven Multiple Hyper-parameter Tuning with Structured Loss Function
Anh Tuan Nguyen, Viet Anh Nguyen
cs.LGstat.MLarXiv:2608.17343v12026EfficientNet: Rethinking Model Scaling for Convolutional Neural Networks
Mingxing Tan, Quoc V. Le
cs.LGcs.CVstat.MLarXiv:1905.11946v52019Inductive Representation Learning on Large Graphs
William L. Hamilton, Rex Ying, Jure Leskovec
cs.SIcs.LGstat.MLarXiv:1706.02216v42017Explaining and Harnessing Adversarial Examples
Ian J. Goodfellow, Jonathon Shlens, Christian Szegedy
stat.MLcs.LGarXiv:1412.6572v32014Learning Phrase Representations using RNN Encoder-Decoder for Statistical Machine Translation
Kyunghyun Cho, Bart van Merrienboer, Caglar Gulcehre +4
cs.CLcs.LGcs.NEarXiv:1406.1078v32014Information fusion and machine learning for sensitivity analysis using physics knowledge and experimental data
Berkcan Kapusuzoglu, Sankaran Mahadevan
cs.CEcs.LGstat.MEarXiv:2608.17248v12026Graph Attention Networks
Petar Veličković, Guillem Cucurull, Arantxa Casanova +3
stat.MLcs.AIcs.LGarXiv:1710.10903v32017A Unified Approach to Interpreting Model Predictions
Scott Lundberg, Su-In Lee
cs.AIcs.LGstat.MLarXiv:1705.07874v22017Distributed Representations of Words and Phrases and their Compositionality
Tomas Mikolov, Ilya Sutskever, Kai Chen +2
cs.CLcs.LGstat.MLarXiv:1310.4546v12013Benchmarking Quantum Machine Learning for Power-System Attack Detection: Evaluation Choices Decide the Outcome Before the Models Do
Md Rezwanul Islam
cs.LGcs.CRstat.MLarXiv:2608.15617v12026Semi-Supervised Classification with Graph Convolutional Networks
Thomas N. Kipf, Max Welling
cs.LGstat.MLarXiv:1609.02907v42016Thinking in a Low-Resource Language: What SFT Builds, What RL Fixes, What Accuracy Cannot See
Ayoub Kirouane, Christos Petrocheilos
cs.CLcs.LGcs.ROarXiv:2608.17744v12026The Data Manifold under the Microscope
Marios Koulakis, Constantin Seibold
cs.LGstat.MLarXiv:2606.15760v12026Rethinking Reverse KL as Adaptive Entropy Distillation
Shizhen Li, Zhiyu Shen, Yuyin Lu +4
cs.LGstat.MLarXiv:2608.14685v12026Convolution Smoothed Quantile Regression for XGBoost
Mandy Yao, Meredith Franklin
stat.MLcs.LGarXiv:2608.15290v12026Cross-Entropy Risk Estimation for Language Models: Inconsistency Must Be Dense, and the Holdout Method Is No Exception
Hanti Lin
cs.LGstat.MLarXiv:2608.15798v12026Queryable LoRA: Instruction-Regularized Routing Over Shared Low-Rank Update Atoms
Omatharv Bharat Vaidya, Connor T. Jerzak, Nhat Ho +1
cs.LGcs.CLstat.MLarXiv:2605.08423v12026The Distributional View of Knowledge Distillation
Gordei Verbii, Juho Lee
stat.MLcs.LGarXiv:2608.15215v12026Summaries:한국어Generative Learning of Separatrices
Ellis R. Crabtree, Dimitris G. Giovanis, Anastasia Georgiou +2
cs.LGmath.DSstat.MLarXiv:2608.14743v12026Shape Operator PCA: Curvature-Aware Projections for Geometric Machine Learning
Alexandre L. M. Levada
cs.LGcs.AIcs.CVarXiv:2608.15313v12026Adaptive surrogate modeling for high-dimensional spatio-temporal output
Berkcan Kapusuzoglu, Shunsaku Matsumoto, Yoshitomo Miyagi +2
cs.CEcs.AIcs.LGarXiv:2608.17250v12026A Deep Learning Model for Spatially Clustered Data via Differentiable Cluster Assignment
Kexuan Li, Weidong Ma
stat.MLcs.LGarXiv:2608.14968v12026ARISE: An adaptive residual-informed stability ensemble for feature selection in small-sample biomedical omics
Zardad Khan, Amjad Ali, Naz Gul +2
stat.MLcs.LGarXiv:2608.14866v12026How Many Samples Are Needed to Determine Causal Direction? Sharp Minimax Bounds for Bivariate LiNGAM
Jikai Jin
math.STcs.LGecon.EMarXiv:2608.15840v12026Inferential Evaluation of Surrogate-Derived Models under Covariate Shift
Longtian Shi, Molei Liu, Doudou Zhou
stat.MLcs.LGstat.AParXiv:2608.15783v12026Self-Supervised Auxiliary Task Discovery for Stable Reinforcement Learning in Stock Trading
Arishi Orra, Himanshu Choudhary, Manoj Thakur
cs.LGq-fin.CPstat.MLarXiv:2608.15841v12026Learning Stock Trading Policies via Barycenter-Based Adversarial Inverse Reinforcement Learning
Arishi Orra, Himanshu Choudhary, Manoj Thakur
cs.LGstat.MLarXiv:2608.15770v12026FirstDiff: One-Step Diffusion-Based Anomaly Detection for Multivariate Time Series via Initial Noise Prediction
Ali Boudaghi, Alireza Nemati, Hadi Zare
cs.LGcs.AIstat.MLarXiv:2608.15727v12026PERO: Efficient Robust Post-Training Foundation Models for Encrypted Traffic Classification
Wumei Du, Jiarong Wen, Kaiyu Zhang +5
cs.LGstat.MLarXiv:2608.15504v12026Conditional Evaluation of Language Models with Cheap Auxiliary Signals
Zhi Zhang, Lingfeng Lyu, Yue Kang +1
cs.LGstat.MLarXiv:2608.16210v12026Pion: A Spectrum-Preserving Optimizer via Orthogonal Equivalence Transformation
Kexuan Shi, Hanxuan Li, Zeju Qiu +3
cs.LGstat.MLarXiv:2605.12492v12026Hide&Seek: Learning to Explain in an End-to-End Differentiable Network
Tal Ellinson, Hadi Mohasel Afshar, Sally Cripps
stat.MLcs.LGarXiv:2608.16689v12026A Unified Geometric Framework for Developmental Analysis of Spatial Transcriptomic Data
Mary Chriselda Antony Oliver, Kaitlyn Hohmeier, Tuyen Tran +3
stat.MLcs.LGmath.MGarXiv:2608.15306v12026Coded Hankel Polynomial Chaos: Spectral Identification of Dominant Polynomial-Chaos Modes
Zhiliang Deng, Xiaomei Yang
stat.MLcs.LGarXiv:2608.16126v12026Scale-Consistent Posterior Dynamics for Diffusion Inverse Problems
Zhaoqiang Liu, Tongyao Pang, Ruibing Wang +1
stat.MLcs.AIcs.LGarXiv:2608.15144v12026Beyond Effective Sample Size: Effective Number of Proposals for Adaptive Importance Sampling
Ali Mousavi, Victor Elvira
stat.MLcs.LGarXiv:2608.15154v12026PathFinder: Joint Decompositions of Linked Multimodal Datasets
Ying-Qiu Zheng, Alex Fung, Stephen M Smith +2
cs.LGeess.IVq-bio.QMarXiv:2608.14951v12026GRPO, Dr. GRPO, and DAPO Are Three Operations on One Number: The Group-Standard-Deviation Identity
Yong Yi Bay, Kathleen A. Yearick
cs.LGcs.AIcs.CLarXiv:2607.00152v12026SiamJEPA: On the Role of Siamese Student Encoders in JEPA
Makoto Yamada
cs.CVstat.MLarXiv:2607.04044v22026TREK: Distill to Explore, Reinforce to Refine
Yuanda Xu, Zhengze Zhou, Kayhan Behdin +10
cs.LGcs.AIstat.MLarXiv:2607.05339v12026When More Sampling Hurts: The Modal Ceiling and Correlation Ceiling of Test-Time Scaling
Yong Yi Bay, Kathleen A. Yearick
cs.LGcs.AIcs.CLarXiv:2606.28661v12026High-dimensional nonparametric changepoint detection via low-rank degree-two density projection
Guoqing Zhang, Zhaixin Chen
cs.LGstat.MLarXiv:2608.13922v12026Multi-Turn On-Policy Distillation with Prefix Replay
Baohao Liao, Hanze Dong, Christof Monz +3
cs.LGcs.AIcs.CLarXiv:2607.04763v32026Round-Trip Consistency: Bidirectional Diffusion Models Can Predict Their Own Rollout Errors
Alexander Scheinker
stat.MLcs.LGphysics.comp-pharXiv:2608.00675v22026Learning Unsteady Aneurysm Hemodynamics with Physics-Informed DeepONets
Oscar L. Cruz-Gonzalez, Valérie Deplano, Badih Ghattas
stat.MLcs.LGphysics.flu-dynarXiv:2608.13629v12026On the Brittleness of Maximum Likelihood Estimation for Gaussian Process Hyperparameter Optimization
Tyler R. Johnson, Kian Ben-Jacob, Christopher P. Muller +1
stat.MLcs.LGstat.MEarXiv:2608.13793v12026L-FNO: Lorentzian Fourier Neural Operator for Stochastic Event Dynamics
Songhee Kang, Jihoon Kang
cs.LGstat.MLarXiv:2608.13562v12026When Does More Correct Data Hurt? Insertion-Stability and the Limits of Dimension-Based Theory
Joseph Sankoorikal Johny
cs.LGstat.MLarXiv:2608.14020v12026Forecast Collapse in Time-Series Foundation Models
Shu Wan, Miles Ma, Hank Zhu +4
cs.LGcs.AIcs.CEarXiv:2608.14106v12026