Machine Learning
Papers filed under cs.LG on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.
Search paper metadata (including unsummarized papers)
18,121 to 18,180 of 20,198
Gradient Descent Finds Global Minima of Deep Neural Networks
Simon S. Du, Jason D. Lee, Haochuan Li +2
cs.LGcs.AIcs.CVarXiv:1811.03804v42018Defense-GAN: Protecting Classifiers Against Adversarial Attacks Using Generative Models
Pouya Samangouei, Maya Kabkab, Rama Chellappa
cs.CVcs.LGstat.MLarXiv:1805.06605v22018Discovering Reinforcement Learning Interfaces with Large Language Models
Akshat Singh Jaswal, Ashish Baghel, Paras Chopra
cs.LGcs.AIarXiv:2605.03408v12026The StarCraft Multi-Agent Challenge
Mikayel Samvelyan, Tabish Rashid, Christian Schroeder de Witt +7
cs.LGcs.MAstat.MLarXiv:1902.04043v52019Deep Captioning with Multimodal Recurrent Neural Networks (m-RNN)
Junhua Mao, Wei Xu, Yi Yang +3
cs.CVcs.CLcs.LGarXiv:1412.6632v52014Adversarial Attacks on Neural Networks for Graph Data
Daniel Zügner, Amir Akbarnejad, Stephan Günnemann
stat.MLcs.CRcs.LGarXiv:1805.07984v42018APEX: Large-scale Multi-task Aesthetic-Informed Popularity Prediction for AI-Generated Music
Jaavid Aktar Husain, Dorien Herremans
cs.SDcs.AIcs.LGarXiv:2605.03395v22026Characterizing possible failure modes in physics-informed neural networks
Aditi S. Krishnapriyan, Amir Gholami, Shandian Zhe +2
cs.LGcs.AImath.NAarXiv:2109.01050v22021FourCastNet: A Global Data-driven High-resolution Weather Model using Adaptive Fourier Neural Operators
Jaideep Pathak, Shashank Subramanian, Peter Harrington +10
physics.ao-phcs.LGarXiv:2202.11214v12022Recurrent Neural Networks for Time Series Forecasting: Current Status and Future Directions
Hansika Hewamalage, Christoph Bergmeir, Kasun Bandara
cs.LGcs.NEstat.MLarXiv:1909.00590v52019Local SGD Converges Fast and Communicates Little
Sebastian U. Stich
math.OCcs.DCcs.LGarXiv:1805.09767v32018When to Trust Your Model: Model-Based Policy Optimization
Michael Janner, Justin Fu, Marvin Zhang +1
cs.LGcs.AIstat.MLarXiv:1906.08253v32019What is the State of Neural Network Pruning?
Davis Blalock, Jose Javier Gonzalez Ortiz, Jonathan Frankle +1
cs.LGstat.MLarXiv:2003.03033v12020Blended Diffusion for Text-driven Editing of Natural Images
Omri Avrahami, Dani Lischinski, Ohad Fried
cs.CVcs.GRcs.LGarXiv:2111.14818v22021Semi-supervised Sequence Learning
Andrew M. Dai, Quoc V. Le
cs.LGcs.CLarXiv:1511.01432v12015Poison Frogs! Targeted Clean-Label Poisoning Attacks on Neural Networks
Ali Shafahi, W. Ronny Huang, Mahyar Najibi +4
cs.LGcs.CRcs.CVarXiv:1804.00792v22018A Reductions Approach to Fair Classification
Alekh Agarwal, Alina Beygelzimer, Miroslav Dudík +2
cs.LGarXiv:1803.02453v32018Hierarchical Deep Reinforcement Learning: Integrating Temporal Abstraction and Intrinsic Motivation
Tejas D. Kulkarni, Karthik R. Narasimhan, Ardavan Saeedi +1
cs.LGcs.AIcs.CVarXiv:1604.06057v22016Stein Variational Gradient Descent: A General Purpose Bayesian Inference Algorithm
Qiang Liu, Dilin Wang
stat.MLcs.LGarXiv:1608.04471v32016A Sensitivity Analysis of (and Practitioners' Guide to) Convolutional Neural Networks for Sentence Classification
Ye Zhang, Byron Wallace
cs.CLcs.LGcs.NEarXiv:1510.03820v42015Square Attack: a query-efficient black-box adversarial attack via random search
Maksym Andriushchenko, Francesco Croce, Nicolas Flammarion +1
cs.LGcs.CRcs.CVarXiv:1912.00049v32019Quantizing deep convolutional networks for efficient inference: A whitepaper
Raghuraman Krishnamoorthi
cs.LGcs.CVstat.MLarXiv:1806.08342v12018struc2vec: Learning Node Representations from Structural Identity
Leonardo F. R. Ribeiro, Pedro H. P. Savarese, Daniel R. Figueiredo
cs.SIcs.LGstat.MLarXiv:1704.03165v32017MaskGAN: Towards Diverse and Interactive Facial Image Manipulation
Cheng-Han Lee, Ziwei Liu, Lingyun Wu +1
cs.CVcs.GRcs.LGarXiv:1907.11922v22019Matrix Completion from a Few Entries
Raghunandan H. Keshavan, Andrea Montanari, Sewoong Oh
cs.LGstat.MLarXiv:0901.3150v42009WebShop: Towards Scalable Real-World Web Interaction with Grounded Language Agents
Shunyu Yao, Howard Chen, John Yang +1
cs.CLcs.AIcs.LGarXiv:2207.01206v42022SEGAN: Speech Enhancement Generative Adversarial Network
Santiago Pascual, Antonio Bonafonte, Joan Serrà
cs.LGcs.NEcs.SDarXiv:1703.09452v32017Learning Traffic as Images: A Deep Convolutional Neural Network for Large-Scale Transportation Network Speed Prediction
Xiaolei Ma, Zhuang Dai, Zhengbing He +3
cs.LGstat.MLarXiv:1701.04245v42017Overcoming catastrophic forgetting with hard attention to the task
Joan Serrà, Dídac Surís, Marius Miron +1
cs.LGcs.AIcs.NEarXiv:1801.01423v32018OptNet: Differentiable Optimization as a Layer in Neural Networks
Brandon Amos, J. Zico Kolter
cs.LGcs.AImath.OCarXiv:1703.00443v52017Mastering Atari with Discrete World Models
Danijar Hafner, Timothy Lillicrap, Mohammad Norouzi +1
cs.LGcs.AIstat.MLarXiv:2010.02193v42020Metrics for Multi-Class Classification: an Overview
Margherita Grandini, Enrico Bagli, Giorgio Visani
stat.MLcs.LGarXiv:2008.05756v12020LSTM Fully Convolutional Networks for Time Series Classification
Fazle Karim, Somshubra Majumdar, Houshang Darabi +1
cs.LGstat.MLarXiv:1709.05206v12017When No Benchmark Exists: Validating Comparative LLM Safety Scoring Without Ground-Truth Labels
Sushant Gautam, Finn Schwall, Annika Willoch Olstad +6
cs.LGcs.AIcs.CLarXiv:2605.06652v12026Towards Personalized Federated Learning
Alysa Ziying Tan, Han Yu, Lizhen Cui +1
cs.LGcs.AIcs.DCarXiv:2103.00710v32021Are We Making Progress in Multimodal Domain Generalization? A Comprehensive Benchmark Study
Hao Dong, Hongzhao Li, Shupan Li +3
cs.CVcs.AIcs.LGarXiv:2605.06643v12026MagNet: a Two-Pronged Defense against Adversarial Examples
Dongyu Meng, Hao Chen
cs.CRcs.LGarXiv:1705.09064v22017Fine-tuning Aligned Language Models Compromises Safety, Even When Users Do Not Intend To!
Xiangyu Qi, Yi Zeng, Tinghao Xie +4
cs.CLcs.AIcs.CRarXiv:2310.03693v12023CURL: Contrastive Unsupervised Representations for Reinforcement Learning
Aravind Srinivas, Michael Laskin, Pieter Abbeel
cs.LGcs.CVstat.MLarXiv:2004.04136v42020Deep Gaussian Processes
Andreas C. Damianou, Neil D. Lawrence
stat.MLcs.LGmath.PRarXiv:1211.0358v22012A Closed-Form Upper Bound for Admissible Learning-Rate Steps in Belief-Space Dynamics
Zixi Li, Youzhen Li
cs.LGarXiv:2605.06741v12026AI Feynman: a Physics-Inspired Method for Symbolic Regression
Silviu-Marian Udrescu, Max Tegmark
physics.comp-phcs.AIcs.LGarXiv:1905.11481v22019TIDE: Every Layer Knows the Token Beneath the Context
Ajay Jaiswal, Lauren Hannah, Han-Byul Kim +3
cs.CLcs.AIcs.LGarXiv:2605.06216v12026Nonsense Helps: Prompt Space Perturbation Broadens Reasoning Exploration
Langlin Huang, Chengsong Huang, Jinyuan Li +3
cs.AIcs.CLcs.LGarXiv:2605.05566v12026Empirical Evidence for Simply Connected Decision Regions in Image Classifiers
Arjhun Swaminathan, Mete Akgün
cs.CVcs.LGarXiv:2605.06380v12026PACEvolve++: Improving Test-time Learning for Evolutionary Search Agents
Minghao Yan, Bo Peng, Benjamin Coleman +11
cs.LGarXiv:2605.07039v12026Mastering Diverse Domains through World Models
Danijar Hafner, Jurgis Pasukonis, Jimmy Ba +1
cs.AIcs.LGstat.MLarXiv:2301.04104v22023Graph U-Nets
Hongyang Gao, Shuiwang Ji
cs.LGstat.MLarXiv:1905.05178v12019LiVeAction: a Lightweight, Versatile, and Asymmetric Neural Codec Design for Real-time Operation
Dan Jacobellis, Neeraja J. Yadwadkar
eess.IVcs.LGcs.MMarXiv:2605.06628v12026Reinforcement Learning with Unsupervised Auxiliary Tasks
Max Jaderberg, Volodymyr Mnih, Wojciech Marian Czarnecki +4
cs.LGcs.NEarXiv:1611.05397v12016Binarized Neural Networks: Training Deep Neural Networks with Weights and Activations Constrained to +1 or -1
Matthieu Courbariaux, Itay Hubara, Daniel Soudry +2
cs.LGarXiv:1602.02830v32016Stabilizing Off-Policy Q-Learning via Bootstrapping Error Reduction
Aviral Kumar, Justin Fu, George Tucker +1
cs.LGstat.MLarXiv:1906.00949v22019Planning with Diffusion for Flexible Behavior Synthesis
Michael Janner, Yilun Du, Joshua B. Tenenbaum +1
cs.LGcs.AIarXiv:2205.09991v22022Learning Sparse Neural Networks through $L_0$ Regularization
Christos Louizos, Max Welling, Diederik P. Kingma
stat.MLcs.LGarXiv:1712.01312v22017Adversarially Learned Inference
Vincent Dumoulin, Ishmael Belghazi, Ben Poole +4
stat.MLcs.LGarXiv:1606.00704v32016Twins: Revisiting the Design of Spatial Attention in Vision Transformers
Xiangxiang Chu, Zhi Tian, Yuqing Wang +5
cs.CVcs.AIcs.LGarXiv:2104.13840v42021Quantifying Attention Flow in Transformers
Samira Abnar, Willem Zuidema
cs.LGcs.AIcs.CLarXiv:2005.00928v22020Self-Attention Graph Pooling
Junhyun Lee, Inyeop Lee, Jaewoo Kang
cs.LGstat.MLarXiv:1904.08082v42019FastText.zip: Compressing text classification models
Armand Joulin, Edouard Grave, Piotr Bojanowski +3
cs.CLcs.LGarXiv:1612.03651v12016API design for machine learning software: experiences from the scikit-learn project
Lars Buitinck, Gilles Louppe, Mathieu Blondel +12
cs.LGcs.MSarXiv:1309.0238v12013