Machine Learning
Papers filed under cs.LG on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.
Search paper metadata (including unsummarized papers)
17,401 to 17,460 of 20,199
Loss Surfaces, Mode Connectivity, and Fast Ensembling of DNNs
Timur Garipov, Pavel Izmailov, Dmitrii Podoprikhin +2
stat.MLcs.AIcs.LGarXiv:1802.10026v42018Transfer Learning from Speaker Verification to Multispeaker Text-To-Speech Synthesis
Ye Jia, Yu Zhang, Ron J. Weiss +8
cs.CLcs.LGcs.SDarXiv:1806.04558v42018Rewarding the Rare: Uniqueness-Aware RL for Creative Problem Solving in LLMs
Zhiyuan Hu, Yucheng Wang, Yufei He +7
cs.LGcs.CLarXiv:2601.08763v22026GraphAgents: Knowledge Graph-Guided Agentic AI for Cross-Domain Materials Design
Isabella A. Stewart, Tarjei Paule Hage, Yu-Chuan Hsu +1
cs.AIcond-mat.mes-hallcond-mat.mtrl-sciarXiv:2602.07491v12026SimVLM: Simple Visual Language Model Pretraining with Weak Supervision
Zirui Wang, Jiahui Yu, Adams Wei Yu +3
cs.CVcs.CLcs.LGarXiv:2108.10904v32021Network Trimming: A Data-Driven Neuron Pruning Approach towards Efficient Deep Architectures
Hengyuan Hu, Rui Peng, Yu-Wing Tai +1
cs.NEcs.CVcs.LGarXiv:1607.03250v12016Fast Algorithms for Convolutional Neural Networks
Andrew Lavin, Scott Gray
cs.NEcs.LGarXiv:1509.09308v22015Overfitting in adversarially robust deep learning
Leslie Rice, Eric Wong, J. Zico Kolter
cs.LGstat.MLarXiv:2002.11569v22020Model-Agnostic Interpretability of Machine Learning
Marco Tulio Ribeiro, Sameer Singh, Carlos Guestrin
stat.MLcs.LGarXiv:1606.05386v12016StarCraft II: A New Challenge for Reinforcement Learning
Oriol Vinyals, Timo Ewalds, Sergey Bartunov +22
cs.LGcs.AIarXiv:1708.04782v12017RMA: Rapid Motor Adaptation for Legged Robots
Ashish Kumar, Zipeng Fu, Deepak Pathak +1
cs.LGcs.AIcs.CVarXiv:2107.04034v12021Benign Overfitting in Linear Regression
Peter L. Bartlett, Philip M. Long, Gábor Lugosi +1
stat.MLcs.LGmath.STarXiv:1906.11300v32019The Vision Wormhole: Latent-Space Communication in Heterogeneous Multi-Agent Systems
Xiaoze Liu, Ruowang Zhang, Weichen Yu +7
cs.CLcs.CVcs.LGarXiv:2602.15382v22026Recommendation as Language Processing (RLP): A Unified Pretrain, Personalized Prompt & Predict Paradigm (P5)
Shijie Geng, Shuchang Liu, Zuohui Fu +2
cs.IRcs.AIcs.CLarXiv:2203.13366v72022Unity: A General Platform for Intelligent Agents
Arthur Juliani, Vincent-Pierre Berges, Ervin Teng +8
cs.LGcs.AIcs.NEarXiv:1809.02627v22018A Survey on the Explainability of Supervised Machine Learning
Nadia Burkart, Marco F. Huber
cs.LGcs.AIstat.MLarXiv:2011.07876v12020\$OneMillion-Bench: How Far are Language Agents from Human Experts?
Qianyu Yang, Yang Liu, Jiaqi Li +20
cs.LGcs.AIcs.CLarXiv:2603.07980v12026Understanding disentangling in $β$-VAE
Christopher P. Burgess, Irina Higgins, Arka Pal +4
stat.MLcs.AIcs.LGarXiv:1804.03599v12018Making Convolutional Networks Shift-Invariant Again
Richard Zhang
cs.CVcs.LGarXiv:1904.11486v22019Deep Learning Recommendation Model for Personalization and Recommendation Systems
Maxim Naumov, Dheevatsa Mudigere, Hao-Jun Michael Shi +21
cs.IRcs.LGarXiv:1906.00091v12019How does Disagreement Help Generalization against Label Corruption?
Xingrui Yu, Bo Han, Jiangchao Yao +3
cs.LGstat.MLarXiv:1901.04215v32019On the Cross-lingual Transferability of Monolingual Representations
Mikel Artetxe, Sebastian Ruder, Dani Yogatama
cs.CLcs.AIcs.LGarXiv:1910.11856v32019Deep Fragment Embeddings for Bidirectional Image Sentence Mapping
Andrej Karpathy, Armand Joulin, Li Fei-Fei
cs.CVcs.CLcs.LGarXiv:1406.5679v12014GuacaMol: Benchmarking Models for De Novo Molecular Design
Nathan Brown, Marco Fiscato, Marwin H. S. Segler +1
q-bio.QMcs.LGphysics.chem-pharXiv:1811.09621v22018QuantVLA: Scale-Calibrated Post-Training Quantization for Vision-Language-Action Models
Jingxuan Zhang, Yunta Hsieh, Zhongwei Wan +5
cs.LGarXiv:2602.20309v42026Split learning for health: Distributed deep learning without sharing raw patient data
Praneeth Vepakomma, Otkrist Gupta, Tristan Swedish +1
cs.LGstat.MLarXiv:1812.00564v12018In-context Learning and Induction Heads
Catherine Olsson, Nelson Elhage, Neel Nanda +23
cs.LGarXiv:2209.11895v12022Equivariant Diffusion for Molecule Generation in 3D
Emiel Hoogeboom, Victor Garcia Satorras, Clément Vignac +1
cs.LGq-bio.QMstat.MLarXiv:2203.17003v22022SpatialVLM: Endowing Vision-Language Models with Spatial Reasoning Capabilities
Boyuan Chen, Zhuo Xu, Sean Kirmani +6
cs.CVcs.CLcs.LGarXiv:2401.12168v12024Detecting Adversarial Samples from Artifacts
Reuben Feinman, Ryan R. Curtin, Saurabh Shintre +1
stat.MLcs.LGarXiv:1703.00410v32017Deep Reinforcement Learning at the Edge of the Statistical Precipice
Rishabh Agarwal, Max Schwarzer, Pablo Samuel Castro +2
cs.LGcs.AIstat.MEarXiv:2108.13264v42021CLIPort: What and Where Pathways for Robotic Manipulation
Mohit Shridhar, Lucas Manuelli, Dieter Fox
cs.ROcs.CLcs.CVarXiv:2109.12098v12021Understanding the Challenges in Iterative Generative Optimization with LLMs
Allen Nie, Xavier Daull, Zhiyi Kuang +10
cs.LGcs.AIarXiv:2603.23994v22026MADE: Masked Autoencoder for Distribution Estimation
Mathieu Germain, Karol Gregor, Iain Murray +1
cs.LGcs.NEstat.MLarXiv:1502.03509v22015GUI-Libra: Training Native GUI Agents to Reason and Act with Action-aware Supervision and Partially Verifiable RL
Rui Yang, Qianhui Wu, Zhaoyang Wang +8
cs.LGcs.AIcs.CLarXiv:2602.22190v22026Zero-Shot Learning by Convex Combination of Semantic Embeddings
Mohammad Norouzi, Tomas Mikolov, Samy Bengio +5
cs.LGarXiv:1312.5650v32013Actor-Attention-Critic for Multi-Agent Reinforcement Learning
Shariq Iqbal, Fei Sha
cs.LGcs.AIcs.MAarXiv:1810.02912v22018Learning Query-Aware Budget-Tier Routing for Runtime Agent Memory
Haozhen Zhang, Haodong Yue, Tao Feng +8
cs.CLcs.AIcs.LGarXiv:2602.06025v32026LK Losses: Direct Acceptance Rate Optimization for Speculative Decoding
Alexander Samarin, Sergei Krutikov, Anton Shevtsov +3
cs.LGcs.CLarXiv:2602.23881v22026"Zero-Shot" Super-Resolution using Deep Internal Learning
Assaf Shocher, Nadav Cohen, Michal Irani
cs.CVcs.LGcs.NEarXiv:1712.06087v12017GLIGEN: Open-Set Grounded Text-to-Image Generation
Yuheng Li, Haotian Liu, Qingyang Wu +5
cs.CVcs.AIcs.CLarXiv:2301.07093v22023Invariant Information Clustering for Unsupervised Image Classification and Segmentation
Xu Ji, João F. Henriques, Andrea Vedaldi
cs.CVcs.LGarXiv:1807.06653v42018Fast KVzip: Efficient and Accurate LLM Inference with Gated KV Eviction
Jang-Hyun Kim, Dongyoon Han, Sangdoo Yun
cs.LGcs.CLarXiv:2601.17668v22026Dual-path RNN: efficient long sequence modeling for time-domain single-channel speech separation
Yi Luo, Zhuo Chen, Takuya Yoshioka
eess.AScs.LGcs.SDarXiv:1910.06379v22019FedPAQ: A Communication-Efficient Federated Learning Method with Periodic Averaging and Quantization
Amirhossein Reisizadeh, Aryan Mokhtari, Hamed Hassani +2
cs.LGcs.DCmath.OCarXiv:1909.13014v42019Discovering Language Model Behaviors with Model-Written Evaluations
Ethan Perez, Sam Ringer, Kamilė Lukošiūtė +60
cs.CLcs.AIcs.LGarXiv:2212.09251v12022Detecting Backdoor Attacks on Deep Neural Networks by Activation Clustering
Bryant Chen, Wilka Carvalho, Nathalie Baracaldo +5
cs.LGcs.CRstat.MLarXiv:1811.03728v12018Network Embedding as Matrix Factorization: Unifying DeepWalk, LINE, PTE, and node2vec
Jiezhong Qiu, Yuxiao Dong, Hao Ma +3
cs.SIcs.LGstat.MLarXiv:1710.02971v42017Patient Knowledge Distillation for BERT Model Compression
Siqi Sun, Yu Cheng, Zhe Gan +1
cs.CLcs.LGarXiv:1908.09355v12019Examining Reasoning LLMs-as-Judges in Non-Verifiable LLM Post-Training
Yixin Liu, Yue Yu, DiJia Su +7
cs.AIcs.CLcs.LGarXiv:2603.12246v12026Good SFT Optimizes for SFT, Better SFT Prepares for Reinforcement Learning
Dylan Zhang, Yufeng Xu, Haojin Wang +2
cs.LGcs.AIcs.CLarXiv:2602.01058v22026A trans-disciplinary review of deep learning research for water resources scientists
Chaopeng Shen
stat.MLcs.LGarXiv:1712.02162v32017A Watermark for Large Language Models
John Kirchenbauer, Jonas Geiping, Yuxin Wen +3
cs.LGcs.CLcs.CRarXiv:2301.10226v42023Learning Continuous Image Representation with Local Implicit Image Function
Yinbo Chen, Sifei Liu, Xiaolong Wang
cs.CVcs.LGarXiv:2012.09161v22020Experiential Reinforcement Learning
Taiwei Shi, Sihao Chen, Bowen Jiang +3
cs.LGcs.AIarXiv:2602.13949v12026Latent Particle World Models: Self-supervised Object-centric Stochastic Dynamics Modeling
Tal Daniel, Carl Qi, Dan Haramati +5
cs.LGarXiv:2603.04553v12026What Matters in Learning from Offline Human Demonstrations for Robot Manipulation
Ajay Mandlekar, Danfei Xu, Josiah Wong +7
cs.ROcs.AIcs.LGarXiv:2108.03298v22021Budget-Constrained Agentic Large Language Models: Intention-Based Planning for Costly Tool Use
Hanbing Liu, Chunhao Tian, Nan An +4
cs.AIcs.LGarXiv:2602.11541v12026Non-stationary Transformers: Exploring the Stationarity in Time Series Forecasting
Yong Liu, Haixu Wu, Jianmin Wang +1
cs.LGeess.SParXiv:2205.14415v42022Making Reconstruction FID Predictive of Diffusion Generation FID
Tongda Xu, Mingwei He, Shady Abu-Hussein +6
cs.CVcs.LGarXiv:2603.05630v22026