Machine Learning

Papers filed under cs.LG on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.

Search paper metadata (including unsummarized papers)

17,401 to 17,460 of 20,199

  1. Loss Surfaces, Mode Connectivity, and Fast Ensembling of DNNs

    Timur Garipov, Pavel Izmailov, Dmitrii Podoprikhin +2

    stat.MLcs.AIcs.LGarXiv:1802.10026v42018
  2. Transfer Learning from Speaker Verification to Multispeaker Text-To-Speech Synthesis

    Ye Jia, Yu Zhang, Ron J. Weiss +8

    cs.CLcs.LGcs.SDarXiv:1806.04558v42018
  3. Rewarding the Rare: Uniqueness-Aware RL for Creative Problem Solving in LLMs

    Zhiyuan Hu, Yucheng Wang, Yufei He +7

    cs.LGcs.CLarXiv:2601.08763v22026
  4. GraphAgents: Knowledge Graph-Guided Agentic AI for Cross-Domain Materials Design

    Isabella A. Stewart, Tarjei Paule Hage, Yu-Chuan Hsu +1

    cs.AIcond-mat.mes-hallcond-mat.mtrl-sciarXiv:2602.07491v12026
  5. SimVLM: Simple Visual Language Model Pretraining with Weak Supervision

    Zirui Wang, Jiahui Yu, Adams Wei Yu +3

    cs.CVcs.CLcs.LGarXiv:2108.10904v32021
  6. Network Trimming: A Data-Driven Neuron Pruning Approach towards Efficient Deep Architectures

    Hengyuan Hu, Rui Peng, Yu-Wing Tai +1

    cs.NEcs.CVcs.LGarXiv:1607.03250v12016
  7. Fast Algorithms for Convolutional Neural Networks

    Andrew Lavin, Scott Gray

    cs.NEcs.LGarXiv:1509.09308v22015
  8. Overfitting in adversarially robust deep learning

    Leslie Rice, Eric Wong, J. Zico Kolter

    cs.LGstat.MLarXiv:2002.11569v22020
  9. Model-Agnostic Interpretability of Machine Learning

    Marco Tulio Ribeiro, Sameer Singh, Carlos Guestrin

    stat.MLcs.LGarXiv:1606.05386v12016
  10. StarCraft II: A New Challenge for Reinforcement Learning

    Oriol Vinyals, Timo Ewalds, Sergey Bartunov +22

    cs.LGcs.AIarXiv:1708.04782v12017
  11. RMA: Rapid Motor Adaptation for Legged Robots

    Ashish Kumar, Zipeng Fu, Deepak Pathak +1

    cs.LGcs.AIcs.CVarXiv:2107.04034v12021
  12. Benign Overfitting in Linear Regression

    Peter L. Bartlett, Philip M. Long, Gábor Lugosi +1

    stat.MLcs.LGmath.STarXiv:1906.11300v32019
  13. The Vision Wormhole: Latent-Space Communication in Heterogeneous Multi-Agent Systems

    Xiaoze Liu, Ruowang Zhang, Weichen Yu +7

    cs.CLcs.CVcs.LGarXiv:2602.15382v22026
  14. Recommendation as Language Processing (RLP): A Unified Pretrain, Personalized Prompt & Predict Paradigm (P5)

    Shijie Geng, Shuchang Liu, Zuohui Fu +2

    cs.IRcs.AIcs.CLarXiv:2203.13366v72022
  15. Unity: A General Platform for Intelligent Agents

    Arthur Juliani, Vincent-Pierre Berges, Ervin Teng +8

    cs.LGcs.AIcs.NEarXiv:1809.02627v22018
  16. A Survey on the Explainability of Supervised Machine Learning

    Nadia Burkart, Marco F. Huber

    cs.LGcs.AIstat.MLarXiv:2011.07876v12020
  17. \$OneMillion-Bench: How Far are Language Agents from Human Experts?

    Qianyu Yang, Yang Liu, Jiaqi Li +20

    cs.LGcs.AIcs.CLarXiv:2603.07980v12026
  18. Understanding disentangling in $β$-VAE

    Christopher P. Burgess, Irina Higgins, Arka Pal +4

    stat.MLcs.AIcs.LGarXiv:1804.03599v12018
  19. Making Convolutional Networks Shift-Invariant Again

    Richard Zhang

    cs.CVcs.LGarXiv:1904.11486v22019
  20. Deep Learning Recommendation Model for Personalization and Recommendation Systems

    Maxim Naumov, Dheevatsa Mudigere, Hao-Jun Michael Shi +21

    cs.IRcs.LGarXiv:1906.00091v12019
  21. How does Disagreement Help Generalization against Label Corruption?

    Xingrui Yu, Bo Han, Jiangchao Yao +3

    cs.LGstat.MLarXiv:1901.04215v32019
  22. On the Cross-lingual Transferability of Monolingual Representations

    Mikel Artetxe, Sebastian Ruder, Dani Yogatama

    cs.CLcs.AIcs.LGarXiv:1910.11856v32019
  23. Deep Fragment Embeddings for Bidirectional Image Sentence Mapping

    Andrej Karpathy, Armand Joulin, Li Fei-Fei

    cs.CVcs.CLcs.LGarXiv:1406.5679v12014
  24. GuacaMol: Benchmarking Models for De Novo Molecular Design

    Nathan Brown, Marco Fiscato, Marwin H. S. Segler +1

    q-bio.QMcs.LGphysics.chem-pharXiv:1811.09621v22018
  25. QuantVLA: Scale-Calibrated Post-Training Quantization for Vision-Language-Action Models

    Jingxuan Zhang, Yunta Hsieh, Zhongwei Wan +5

    cs.LGarXiv:2602.20309v42026
  26. Split learning for health: Distributed deep learning without sharing raw patient data

    Praneeth Vepakomma, Otkrist Gupta, Tristan Swedish +1

    cs.LGstat.MLarXiv:1812.00564v12018
  27. In-context Learning and Induction Heads

    Catherine Olsson, Nelson Elhage, Neel Nanda +23

    cs.LGarXiv:2209.11895v12022
  28. Equivariant Diffusion for Molecule Generation in 3D

    Emiel Hoogeboom, Victor Garcia Satorras, Clément Vignac +1

    cs.LGq-bio.QMstat.MLarXiv:2203.17003v22022
  29. SpatialVLM: Endowing Vision-Language Models with Spatial Reasoning Capabilities

    Boyuan Chen, Zhuo Xu, Sean Kirmani +6

    cs.CVcs.CLcs.LGarXiv:2401.12168v12024
  30. Detecting Adversarial Samples from Artifacts

    Reuben Feinman, Ryan R. Curtin, Saurabh Shintre +1

    stat.MLcs.LGarXiv:1703.00410v32017
  31. Deep Reinforcement Learning at the Edge of the Statistical Precipice

    Rishabh Agarwal, Max Schwarzer, Pablo Samuel Castro +2

    cs.LGcs.AIstat.MEarXiv:2108.13264v42021
  32. CLIPort: What and Where Pathways for Robotic Manipulation

    Mohit Shridhar, Lucas Manuelli, Dieter Fox

    cs.ROcs.CLcs.CVarXiv:2109.12098v12021
  33. Understanding the Challenges in Iterative Generative Optimization with LLMs

    Allen Nie, Xavier Daull, Zhiyi Kuang +10

    cs.LGcs.AIarXiv:2603.23994v22026
  34. MADE: Masked Autoencoder for Distribution Estimation

    Mathieu Germain, Karol Gregor, Iain Murray +1

    cs.LGcs.NEstat.MLarXiv:1502.03509v22015
  35. GUI-Libra: Training Native GUI Agents to Reason and Act with Action-aware Supervision and Partially Verifiable RL

    Rui Yang, Qianhui Wu, Zhaoyang Wang +8

    cs.LGcs.AIcs.CLarXiv:2602.22190v22026
  36. Zero-Shot Learning by Convex Combination of Semantic Embeddings

    Mohammad Norouzi, Tomas Mikolov, Samy Bengio +5

    cs.LGarXiv:1312.5650v32013
  37. Actor-Attention-Critic for Multi-Agent Reinforcement Learning

    Shariq Iqbal, Fei Sha

    cs.LGcs.AIcs.MAarXiv:1810.02912v22018
  38. Learning Query-Aware Budget-Tier Routing for Runtime Agent Memory

    Haozhen Zhang, Haodong Yue, Tao Feng +8

    cs.CLcs.AIcs.LGarXiv:2602.06025v32026
  39. LK Losses: Direct Acceptance Rate Optimization for Speculative Decoding

    Alexander Samarin, Sergei Krutikov, Anton Shevtsov +3

    cs.LGcs.CLarXiv:2602.23881v22026
  40. "Zero-Shot" Super-Resolution using Deep Internal Learning

    Assaf Shocher, Nadav Cohen, Michal Irani

    cs.CVcs.LGcs.NEarXiv:1712.06087v12017
  41. GLIGEN: Open-Set Grounded Text-to-Image Generation

    Yuheng Li, Haotian Liu, Qingyang Wu +5

    cs.CVcs.AIcs.CLarXiv:2301.07093v22023
  42. Invariant Information Clustering for Unsupervised Image Classification and Segmentation

    Xu Ji, João F. Henriques, Andrea Vedaldi

    cs.CVcs.LGarXiv:1807.06653v42018
  43. Fast KVzip: Efficient and Accurate LLM Inference with Gated KV Eviction

    Jang-Hyun Kim, Dongyoon Han, Sangdoo Yun

    cs.LGcs.CLarXiv:2601.17668v22026
  44. Dual-path RNN: efficient long sequence modeling for time-domain single-channel speech separation

    Yi Luo, Zhuo Chen, Takuya Yoshioka

    eess.AScs.LGcs.SDarXiv:1910.06379v22019
  45. FedPAQ: A Communication-Efficient Federated Learning Method with Periodic Averaging and Quantization

    Amirhossein Reisizadeh, Aryan Mokhtari, Hamed Hassani +2

    cs.LGcs.DCmath.OCarXiv:1909.13014v42019
  46. Discovering Language Model Behaviors with Model-Written Evaluations

    Ethan Perez, Sam Ringer, Kamilė Lukošiūtė +60

    cs.CLcs.AIcs.LGarXiv:2212.09251v12022
  47. Detecting Backdoor Attacks on Deep Neural Networks by Activation Clustering

    Bryant Chen, Wilka Carvalho, Nathalie Baracaldo +5

    cs.LGcs.CRstat.MLarXiv:1811.03728v12018
  48. Network Embedding as Matrix Factorization: Unifying DeepWalk, LINE, PTE, and node2vec

    Jiezhong Qiu, Yuxiao Dong, Hao Ma +3

    cs.SIcs.LGstat.MLarXiv:1710.02971v42017
  49. Patient Knowledge Distillation for BERT Model Compression

    Siqi Sun, Yu Cheng, Zhe Gan +1

    cs.CLcs.LGarXiv:1908.09355v12019
  50. Examining Reasoning LLMs-as-Judges in Non-Verifiable LLM Post-Training

    Yixin Liu, Yue Yu, DiJia Su +7

    cs.AIcs.CLcs.LGarXiv:2603.12246v12026
  51. Good SFT Optimizes for SFT, Better SFT Prepares for Reinforcement Learning

    Dylan Zhang, Yufeng Xu, Haojin Wang +2

    cs.LGcs.AIcs.CLarXiv:2602.01058v22026
  52. A trans-disciplinary review of deep learning research for water resources scientists

    Chaopeng Shen

    stat.MLcs.LGarXiv:1712.02162v32017
  53. A Watermark for Large Language Models

    John Kirchenbauer, Jonas Geiping, Yuxin Wen +3

    cs.LGcs.CLcs.CRarXiv:2301.10226v42023
  54. Learning Continuous Image Representation with Local Implicit Image Function

    Yinbo Chen, Sifei Liu, Xiaolong Wang

    cs.CVcs.LGarXiv:2012.09161v22020
  55. Experiential Reinforcement Learning

    Taiwei Shi, Sihao Chen, Bowen Jiang +3

    cs.LGcs.AIarXiv:2602.13949v12026
  56. Latent Particle World Models: Self-supervised Object-centric Stochastic Dynamics Modeling

    Tal Daniel, Carl Qi, Dan Haramati +5

    cs.LGarXiv:2603.04553v12026
  57. What Matters in Learning from Offline Human Demonstrations for Robot Manipulation

    Ajay Mandlekar, Danfei Xu, Josiah Wong +7

    cs.ROcs.AIcs.LGarXiv:2108.03298v22021
  58. Budget-Constrained Agentic Large Language Models: Intention-Based Planning for Costly Tool Use

    Hanbing Liu, Chunhao Tian, Nan An +4

    cs.AIcs.LGarXiv:2602.11541v12026
  59. Non-stationary Transformers: Exploring the Stationarity in Time Series Forecasting

    Yong Liu, Haixu Wu, Jianmin Wang +1

    cs.LGeess.SParXiv:2205.14415v42022
  60. Making Reconstruction FID Predictive of Diffusion Generation FID

    Tongda Xu, Mingwei He, Shady Abu-Hussein +6

    cs.CVcs.LGarXiv:2603.05630v22026