Machine Learning
Papers filed under cs.LG on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.
Search paper metadata (including unsummarized papers)
2,341 to 2,400 of 20,024
Toward Interpretable Deep Reinforcement Learning with Linear Model U-Trees
Guiliang Liu, Oliver Schulte, Wang Zhu +1
cs.LGstat.MLarXiv:1807.05887v12018DiffuseVAE: Efficient, Controllable and High-Fidelity Generation from Low-Dimensional Latents
Kushagra Pandey, Avideep Mukherjee, Piyush Rai +1
cs.LGcs.CVarXiv:2201.00308v32022Self-Consistent Trajectory Autoencoder: Hierarchical Reinforcement Learning with Trajectory Embeddings
John D. Co-Reyes, YuXuan Liu, Abhishek Gupta +3
cs.LGcs.AIstat.MLarXiv:1806.02813v12018EEG-based Cross-Subject Driver Drowsiness Recognition with an Interpretable Convolutional Neural Network
Jian Cui, Zirui Lan, Olga Sourina +1
eess.SPcs.LGcs.NEarXiv:2107.09507v42021Distribution-Consistent Inference for Dynamic Sparse Mixture-of-Experts
Dohyeon Kim, Bedionita Soro, Sung Ju Hwang
cs.LGcs.AIcs.CLarXiv:2609.09241v12026Gradient Descent Only Converges to Minimizers: Non-Isolated Critical Points and Invariant Regions
Ioannis Panageas, Georgios Piliouras
math.DScs.LGarXiv:1605.00405v22016Scaling Post-Training Ternarisation to Qwen3-8B Capability Retention, Reproduction, Lossless Packing, and Packed Execution
Anirudh Malik, M Sparsh Mehra, Poojith Devan
cs.LGcs.AIarXiv:2609.09240v12026Adversarial Examples that Fool Detectors
Jiajun Lu, Hussein Sibai, Evan Fabry
cs.CVcs.AIcs.GRarXiv:1712.02494v12017Model-based Pricing for Machine Learning in a Data Marketplace
Lingjiao Chen, Paraschos Koutris, Arun Kumar
cs.DBcs.GTcs.LGarXiv:1805.11450v12018How Important Is a Neuron?
Kedar Dhamdhere, Mukund Sundararajan, Qiqi Yan
cs.LGstat.MLarXiv:1805.12233v12018EnsembleDAgger: A Bayesian Approach to Safe Imitation Learning
Kunal Menda, Katherine Driggs-Campbell, Mykel J. Kochenderfer
cs.LGcs.AIarXiv:1807.08364v32018Infinite attention: NNGP and NTK for deep attention networks
Jiri Hron, Yasaman Bahri, Jascha Sohl-Dickstein +1
stat.MLcs.LGarXiv:2006.10540v12020Input Sparsity Time Low-Rank Approximation via Ridge Leverage Score Sampling
Michael B. Cohen, Cameron Musco, Christopher Musco
cs.DScs.LGarXiv:1511.07263v22015BOND: Benchmarking Unsupervised Outlier Node Detection on Static Attributed Graphs
Kay Liu, Yingtong Dou, Yue Zhao +12
cs.LGcs.SIarXiv:2206.10071v22022TACCL: Guiding Collective Algorithm Synthesis using Communication Sketches
Aashaka Shah, Vijay Chidambaram, Meghan Cowan +6
cs.DCcs.LGarXiv:2111.04867v42021Complexity Theoretic Limitations on Learning Halfspaces
Amit Daniely
cs.CCcs.LGarXiv:1505.05800v22015THE COLOSSEUM: A Benchmark for Evaluating Generalization for Robotic Manipulation
Wilbert Pumacay, Ishika Singh, Jiafei Duan +3
cs.ROcs.AIcs.LGarXiv:2402.08191v22024An Effective Approach to Unsupervised Machine Translation
Mikel Artetxe, Gorka Labaka, Eneko Agirre
cs.CLcs.AIcs.LGarXiv:1902.01313v22019Directed Graph Convolutional Network
Zekun Tong, Yuxuan Liang, Changsheng Sun +2
cs.LGstat.MLarXiv:2004.13970v12020Contextualizing Hate Speech Classifiers with Post-hoc Explanation
Brendan Kennedy, Xisen Jin, Aida Mostafazadeh Davani +2
cs.CLcs.IRcs.LGarXiv:2005.02439v32020LM-Polygraph: Uncertainty Estimation for Language Models
Ekaterina Fadeeva, Roman Vashurin, Akim Tsvigun +9
cs.CLcs.LGarXiv:2311.07383v12023Active and passive learning of linear separators under log-concave distributions
Maria Florina Balcan, Philip M. Long
cs.LGmath.STstat.MLarXiv:1211.1082v32012TRACE: Training Reasoning Agents for Causal Exploration with Synthesized Rewards
Rui Sun, Zhan Shi, Bing He
cs.AIcs.LGarXiv:2609.10315v12026On the Exploitability of Instruction Tuning
Manli Shu, Jiongxiao Wang, Chen Zhu +3
cs.CRcs.CLcs.LGarXiv:2306.17194v22023Effectiveness of self-supervised pre-training for speech recognition
Alexei Baevski, Michael Auli, Abdelrahman Mohamed
cs.CLcs.LGarXiv:1911.03912v32019Live Face De-Identification in Video
Oran Gafni, Lior Wolf, Yaniv Taigman
cs.LGcs.CVcs.GRarXiv:1911.08348v12019DouZero: Mastering DouDizhu with Self-Play Deep Reinforcement Learning
Daochen Zha, Jingru Xie, Wenye Ma +4
cs.AIcs.LGarXiv:2106.06135v12021Learning Search Space Partition for Black-box Optimization using Monte Carlo Tree Search
Linnan Wang, Rodrigo Fonseca, Yuandong Tian
cs.LGcs.AIcs.ROarXiv:2007.00708v22020Kernel-Managed Shared Memory for System-Wide Personalization
Ryan Lum, Yongfeng Zhang
cs.AIcs.LGarXiv:2609.10144v12026Agent-Based ML-LLM Fusion with Self-Optimizing Prompts for Plateau Weather Alerts
Shuai Yan, Yang Xu, Shan He
cs.AIcs.LGarXiv:2609.10135v12026On Contrastive Learning for Likelihood-free Inference
Conor Durkan, Iain Murray, George Papamakarios
stat.MLcs.LGarXiv:2002.03712v22020Out-Of-Distribution Generalization on Graphs: A Survey
Haoyang Li, Xin Wang, Ziwei Zhang +1
cs.LGarXiv:2202.07987v22022Convergence and sample complexity of gradient methods for the model-free linear quadratic regulator problem
Hesameddin Mohammadi, Armin Zare, Mahdi Soltanolkotabi +1
math.OCcs.AIcs.LGarXiv:1912.11899v32019Belief-State Engine: Augmenting LLMs for Principled Planning Under Partial Observability
Arnab Chattopadhayay, Debdipta Halder
cs.AIcs.LGcs.ROarXiv:2609.10036v12026A Tensor Approach to Learning Mixed Membership Community Models
Anima Anandkumar, Rong Ge, Daniel Hsu +1
cs.LGcs.SIstat.MLarXiv:1302.2684v42013Omnigrok: Grokking Beyond Algorithmic Data
Ziming Liu, Eric J. Michaud, Max Tegmark
cs.LGcs.AIphysics.data-anarXiv:2210.01117v22022Nonparametric variational inference
Samuel Gershman, Matt Hoffman, David Blei
cs.LGstat.MLarXiv:1206.4665v12012Quantifying Logical Consistency in Transformers via Query-Key Alignment
Eduard Tulchinskii, Anastasia Voznyuk, Laida Kushnareva +4
cs.CLcs.AIcs.ITarXiv:2502.17017v12025Quasi-hyperbolic momentum and Adam for deep learning
Jerry Ma, Denis Yarats
cs.LGstat.MLarXiv:1810.06801v42018Deep ConvLSTM with self-attention for human activity decoding using wearables
Satya P. Singh, Aimé Lay-Ekuakille, Deepak Gangwar +2
cs.HCcs.LGeess.SParXiv:2005.00698v22020Certifiably Robust RAG against Retrieval Corruption
Chong Xiang, Tong Wu, Zexuan Zhong +3
cs.LGcs.CLcs.CRarXiv:2405.15556v22024Fair Mixup: Fairness via Interpolation
Ching-Yao Chuang, Youssef Mroueh
cs.LGcs.CYstat.MLarXiv:2103.06503v12021A Systematic Literature Review on the Use of Deep Learning in Software Engineering Research
Cody Watson, Nathan Cooper, David Nader Palacio +2
cs.SEcs.AIcs.LGarXiv:2009.06520v22020Emotion Recognition from Multiple Modalities: Fundamentals and Methodologies
Sicheng Zhao, Guoli Jia, Jufeng Yang +2
eess.SPcs.AIcs.LGarXiv:2108.10152v12021Topology Optimization via Machine Learning and Deep Learning: A Review
Seungyeon Shin, Dongju Shin, Namwoo Kang
cs.LGarXiv:2210.10782v22022Bridging Theory and Data: Correcting Nuclear Mass Models with Interpretable Machine Learning
Yanhua Lu, Tianshuai Shang, Pengxiang Du +2
nucl-thcs.LGnucl-exarXiv:2603.15203v12026Correlation-aware Adversarial Domain Adaptation and Generalization
Mohammad Mahfujur Rahman, Clinton Fookes, Mahsa Baktashmotlagh +1
cs.CVcs.LGarXiv:1911.12983v12019POWERPLAY: Training an Increasingly General Problem Solver by Continually Searching for the Simplest Still Unsolvable Problem
Jürgen Schmidhuber
cs.AIcs.LGarXiv:1112.5309v22011Deep Learning in Cardiology
Paschalis Bizopoulos, Dimitrios Koutsouris
cs.CVcs.AIcs.LGarXiv:1902.11122v52019Driving Behavior Analysis through CAN Bus Data in an Uncontrolled Environment
Umberto Fugiglando, Emanuele Massaro, Paolo Santi +5
cs.LGcs.CYphysics.data-anarXiv:1710.04133v12017A Survey on Large-scale Machine Learning
Meng Wang, Weijie Fu, Xiangnan He +2
cs.LGstat.MLarXiv:2008.03911v12020A cross Transformer for image denoising
Chunwei Tian, Menghua Zheng, Wangmeng Zuo +3
eess.IVcs.CVcs.LGarXiv:2310.10408v12023Gaussian Processes for Nonlinear Signal Processing
Fernando Pérez-Cruz, Steven Van Vaerenbergh, Juan José Murillo-Fuentes +2
cs.LGcs.ITstat.MLarXiv:1303.2823v22013From One Hand to Multiple Hands: Imitation Learning for Dexterous Manipulation from Single-Camera Teleoperation
Yuzhe Qin, Hao Su, Xiaolong Wang
cs.ROcs.CVcs.LGarXiv:2204.12490v22022Influence of Initialization on the Performance of Metaheuristic Optimizers
Qian Li, San-Yang Liu, Xin-She Yang
cs.NEcs.LGmath.OCarXiv:2003.03789v12020MOFormer: Self-Supervised Transformer model for Metal-Organic Framework Property Prediction
Zhonglin Cao, Rishikesh Magar, Yuyang Wang +1
cs.LGphysics.chem-pharXiv:2210.14188v12022A Lightweight Concept Drift Detection and Adaptation Framework for IoT Data Streams
Li Yang, Abdallah Shami
cs.LGcs.AIarXiv:2104.10529v12021NN-EUCLID: deep-learning hyperelasticity without stress data
Prakash Thakolkaran, Akshay Joshi, Yiwen Zheng +3
cs.LGcs.CEarXiv:2205.06664v22022Lipschitz constant estimation of Neural Networks via sparse polynomial optimization
Fabian Latorre, Paul Rolland, Volkan Cevher
cs.LGstat.MLarXiv:2004.08688v12020SwitchNet: a neural network model for forward and inverse scattering problems
Yuehaw Khoo, Lexing Ying
math.NAcs.LGarXiv:1810.09675v12018