Machine Learning

Papers filed under cs.LG on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.

Search paper metadata (including unsummarized papers)

18,121 to 18,180 of 20,198

  1. Gradient Descent Finds Global Minima of Deep Neural Networks

    Simon S. Du, Jason D. Lee, Haochuan Li +2

    cs.LGcs.AIcs.CVarXiv:1811.03804v42018
  2. Defense-GAN: Protecting Classifiers Against Adversarial Attacks Using Generative Models

    Pouya Samangouei, Maya Kabkab, Rama Chellappa

    cs.CVcs.LGstat.MLarXiv:1805.06605v22018
  3. Discovering Reinforcement Learning Interfaces with Large Language Models

    Akshat Singh Jaswal, Ashish Baghel, Paras Chopra

    cs.LGcs.AIarXiv:2605.03408v12026
  4. The StarCraft Multi-Agent Challenge

    Mikayel Samvelyan, Tabish Rashid, Christian Schroeder de Witt +7

    cs.LGcs.MAstat.MLarXiv:1902.04043v52019
  5. Deep Captioning with Multimodal Recurrent Neural Networks (m-RNN)

    Junhua Mao, Wei Xu, Yi Yang +3

    cs.CVcs.CLcs.LGarXiv:1412.6632v52014
  6. Adversarial Attacks on Neural Networks for Graph Data

    Daniel Zügner, Amir Akbarnejad, Stephan Günnemann

    stat.MLcs.CRcs.LGarXiv:1805.07984v42018
  7. APEX: Large-scale Multi-task Aesthetic-Informed Popularity Prediction for AI-Generated Music

    Jaavid Aktar Husain, Dorien Herremans

    cs.SDcs.AIcs.LGarXiv:2605.03395v22026
  8. Characterizing possible failure modes in physics-informed neural networks

    Aditi S. Krishnapriyan, Amir Gholami, Shandian Zhe +2

    cs.LGcs.AImath.NAarXiv:2109.01050v22021
  9. FourCastNet: A Global Data-driven High-resolution Weather Model using Adaptive Fourier Neural Operators

    Jaideep Pathak, Shashank Subramanian, Peter Harrington +10

    physics.ao-phcs.LGarXiv:2202.11214v12022
  10. Recurrent Neural Networks for Time Series Forecasting: Current Status and Future Directions

    Hansika Hewamalage, Christoph Bergmeir, Kasun Bandara

    cs.LGcs.NEstat.MLarXiv:1909.00590v52019
  11. Local SGD Converges Fast and Communicates Little

    Sebastian U. Stich

    math.OCcs.DCcs.LGarXiv:1805.09767v32018
  12. When to Trust Your Model: Model-Based Policy Optimization

    Michael Janner, Justin Fu, Marvin Zhang +1

    cs.LGcs.AIstat.MLarXiv:1906.08253v32019
  13. What is the State of Neural Network Pruning?

    Davis Blalock, Jose Javier Gonzalez Ortiz, Jonathan Frankle +1

    cs.LGstat.MLarXiv:2003.03033v12020
  14. Blended Diffusion for Text-driven Editing of Natural Images

    Omri Avrahami, Dani Lischinski, Ohad Fried

    cs.CVcs.GRcs.LGarXiv:2111.14818v22021
  15. Semi-supervised Sequence Learning

    Andrew M. Dai, Quoc V. Le

    cs.LGcs.CLarXiv:1511.01432v12015
  16. Poison Frogs! Targeted Clean-Label Poisoning Attacks on Neural Networks

    Ali Shafahi, W. Ronny Huang, Mahyar Najibi +4

    cs.LGcs.CRcs.CVarXiv:1804.00792v22018
  17. A Reductions Approach to Fair Classification

    Alekh Agarwal, Alina Beygelzimer, Miroslav Dudík +2

    cs.LGarXiv:1803.02453v32018
  18. Hierarchical Deep Reinforcement Learning: Integrating Temporal Abstraction and Intrinsic Motivation

    Tejas D. Kulkarni, Karthik R. Narasimhan, Ardavan Saeedi +1

    cs.LGcs.AIcs.CVarXiv:1604.06057v22016
  19. Stein Variational Gradient Descent: A General Purpose Bayesian Inference Algorithm

    Qiang Liu, Dilin Wang

    stat.MLcs.LGarXiv:1608.04471v32016
  20. A Sensitivity Analysis of (and Practitioners' Guide to) Convolutional Neural Networks for Sentence Classification

    Ye Zhang, Byron Wallace

    cs.CLcs.LGcs.NEarXiv:1510.03820v42015
  21. Square Attack: a query-efficient black-box adversarial attack via random search

    Maksym Andriushchenko, Francesco Croce, Nicolas Flammarion +1

    cs.LGcs.CRcs.CVarXiv:1912.00049v32019
  22. Quantizing deep convolutional networks for efficient inference: A whitepaper

    Raghuraman Krishnamoorthi

    cs.LGcs.CVstat.MLarXiv:1806.08342v12018
  23. struc2vec: Learning Node Representations from Structural Identity

    Leonardo F. R. Ribeiro, Pedro H. P. Savarese, Daniel R. Figueiredo

    cs.SIcs.LGstat.MLarXiv:1704.03165v32017
  24. MaskGAN: Towards Diverse and Interactive Facial Image Manipulation

    Cheng-Han Lee, Ziwei Liu, Lingyun Wu +1

    cs.CVcs.GRcs.LGarXiv:1907.11922v22019
  25. Matrix Completion from a Few Entries

    Raghunandan H. Keshavan, Andrea Montanari, Sewoong Oh

    cs.LGstat.MLarXiv:0901.3150v42009
  26. WebShop: Towards Scalable Real-World Web Interaction with Grounded Language Agents

    Shunyu Yao, Howard Chen, John Yang +1

    cs.CLcs.AIcs.LGarXiv:2207.01206v42022
  27. SEGAN: Speech Enhancement Generative Adversarial Network

    Santiago Pascual, Antonio Bonafonte, Joan Serrà

    cs.LGcs.NEcs.SDarXiv:1703.09452v32017
  28. Learning Traffic as Images: A Deep Convolutional Neural Network for Large-Scale Transportation Network Speed Prediction

    Xiaolei Ma, Zhuang Dai, Zhengbing He +3

    cs.LGstat.MLarXiv:1701.04245v42017
  29. Overcoming catastrophic forgetting with hard attention to the task

    Joan Serrà, Dídac Surís, Marius Miron +1

    cs.LGcs.AIcs.NEarXiv:1801.01423v32018
  30. OptNet: Differentiable Optimization as a Layer in Neural Networks

    Brandon Amos, J. Zico Kolter

    cs.LGcs.AImath.OCarXiv:1703.00443v52017
  31. Mastering Atari with Discrete World Models

    Danijar Hafner, Timothy Lillicrap, Mohammad Norouzi +1

    cs.LGcs.AIstat.MLarXiv:2010.02193v42020
  32. Metrics for Multi-Class Classification: an Overview

    Margherita Grandini, Enrico Bagli, Giorgio Visani

    stat.MLcs.LGarXiv:2008.05756v12020
  33. LSTM Fully Convolutional Networks for Time Series Classification

    Fazle Karim, Somshubra Majumdar, Houshang Darabi +1

    cs.LGstat.MLarXiv:1709.05206v12017
  34. When No Benchmark Exists: Validating Comparative LLM Safety Scoring Without Ground-Truth Labels

    Sushant Gautam, Finn Schwall, Annika Willoch Olstad +6

    cs.LGcs.AIcs.CLarXiv:2605.06652v12026
  35. Towards Personalized Federated Learning

    Alysa Ziying Tan, Han Yu, Lizhen Cui +1

    cs.LGcs.AIcs.DCarXiv:2103.00710v32021
  36. Are We Making Progress in Multimodal Domain Generalization? A Comprehensive Benchmark Study

    Hao Dong, Hongzhao Li, Shupan Li +3

    cs.CVcs.AIcs.LGarXiv:2605.06643v12026
  37. MagNet: a Two-Pronged Defense against Adversarial Examples

    Dongyu Meng, Hao Chen

    cs.CRcs.LGarXiv:1705.09064v22017
  38. Fine-tuning Aligned Language Models Compromises Safety, Even When Users Do Not Intend To!

    Xiangyu Qi, Yi Zeng, Tinghao Xie +4

    cs.CLcs.AIcs.CRarXiv:2310.03693v12023
  39. CURL: Contrastive Unsupervised Representations for Reinforcement Learning

    Aravind Srinivas, Michael Laskin, Pieter Abbeel

    cs.LGcs.CVstat.MLarXiv:2004.04136v42020
  40. Deep Gaussian Processes

    Andreas C. Damianou, Neil D. Lawrence

    stat.MLcs.LGmath.PRarXiv:1211.0358v22012
  41. A Closed-Form Upper Bound for Admissible Learning-Rate Steps in Belief-Space Dynamics

    Zixi Li, Youzhen Li

    cs.LGarXiv:2605.06741v12026
  42. AI Feynman: a Physics-Inspired Method for Symbolic Regression

    Silviu-Marian Udrescu, Max Tegmark

    physics.comp-phcs.AIcs.LGarXiv:1905.11481v22019
  43. TIDE: Every Layer Knows the Token Beneath the Context

    Ajay Jaiswal, Lauren Hannah, Han-Byul Kim +3

    cs.CLcs.AIcs.LGarXiv:2605.06216v12026
  44. Nonsense Helps: Prompt Space Perturbation Broadens Reasoning Exploration

    Langlin Huang, Chengsong Huang, Jinyuan Li +3

    cs.AIcs.CLcs.LGarXiv:2605.05566v12026
  45. Empirical Evidence for Simply Connected Decision Regions in Image Classifiers

    Arjhun Swaminathan, Mete Akgün

    cs.CVcs.LGarXiv:2605.06380v12026
  46. PACEvolve++: Improving Test-time Learning for Evolutionary Search Agents

    Minghao Yan, Bo Peng, Benjamin Coleman +11

    cs.LGarXiv:2605.07039v12026
  47. Mastering Diverse Domains through World Models

    Danijar Hafner, Jurgis Pasukonis, Jimmy Ba +1

    cs.AIcs.LGstat.MLarXiv:2301.04104v22023
  48. Graph U-Nets

    Hongyang Gao, Shuiwang Ji

    cs.LGstat.MLarXiv:1905.05178v12019
  49. LiVeAction: a Lightweight, Versatile, and Asymmetric Neural Codec Design for Real-time Operation

    Dan Jacobellis, Neeraja J. Yadwadkar

    eess.IVcs.LGcs.MMarXiv:2605.06628v12026
  50. Reinforcement Learning with Unsupervised Auxiliary Tasks

    Max Jaderberg, Volodymyr Mnih, Wojciech Marian Czarnecki +4

    cs.LGcs.NEarXiv:1611.05397v12016
  51. Binarized Neural Networks: Training Deep Neural Networks with Weights and Activations Constrained to +1 or -1

    Matthieu Courbariaux, Itay Hubara, Daniel Soudry +2

    cs.LGarXiv:1602.02830v32016
  52. Stabilizing Off-Policy Q-Learning via Bootstrapping Error Reduction

    Aviral Kumar, Justin Fu, George Tucker +1

    cs.LGstat.MLarXiv:1906.00949v22019
  53. Planning with Diffusion for Flexible Behavior Synthesis

    Michael Janner, Yilun Du, Joshua B. Tenenbaum +1

    cs.LGcs.AIarXiv:2205.09991v22022
  54. Learning Sparse Neural Networks through $L_0$ Regularization

    Christos Louizos, Max Welling, Diederik P. Kingma

    stat.MLcs.LGarXiv:1712.01312v22017
  55. Adversarially Learned Inference

    Vincent Dumoulin, Ishmael Belghazi, Ben Poole +4

    stat.MLcs.LGarXiv:1606.00704v32016
  56. Twins: Revisiting the Design of Spatial Attention in Vision Transformers

    Xiangxiang Chu, Zhi Tian, Yuqing Wang +5

    cs.CVcs.AIcs.LGarXiv:2104.13840v42021
  57. Quantifying Attention Flow in Transformers

    Samira Abnar, Willem Zuidema

    cs.LGcs.AIcs.CLarXiv:2005.00928v22020
  58. Self-Attention Graph Pooling

    Junhyun Lee, Inyeop Lee, Jaewoo Kang

    cs.LGstat.MLarXiv:1904.08082v42019
  59. FastText.zip: Compressing text classification models

    Armand Joulin, Edouard Grave, Piotr Bojanowski +3

    cs.CLcs.LGarXiv:1612.03651v12016
  60. API design for machine learning software: experiences from the scikit-learn project

    Lars Buitinck, Gilles Louppe, Mathieu Blondel +12

    cs.LGcs.MSarXiv:1309.0238v12013