Machine Learning
Papers filed under cs.LG on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.
Search paper metadata (including unsummarized papers)
17,461 to 17,520 of 20,193
Toy Models of Superposition
Nelson Elhage, Tristan Hume, Catherine Olsson +13
cs.LGarXiv:2209.10652v12022Simple Recipe Works: Vision-Language-Action Models are Natural Continual Learners with Reinforcement Learning
Jiaheng Hu, Jay Shim, Chen Tang +4
cs.LGcs.ROarXiv:2603.11653v32026Quantifying the Carbon Emissions of Machine Learning
Alexandre Lacoste, Alexandra Luccioni, Victor Schmidt +1
cs.CYcs.LGarXiv:1910.09700v22019Tensor Robust Principal Component Analysis with A New Tensor Nuclear Norm
Canyi Lu, Jiashi Feng, Yudong Chen +3
stat.MLcs.LGarXiv:1804.03728v22018Memex(RL): Scaling Long-Horizon LLM Agents via Indexed Experience Memory
Zhenting Wang, Huancheng Chen, Jiayun Wang +1
cs.CLcs.LGarXiv:2603.04257v12026Mano: Restriking Manifold Optimization for LLM Training
Yufei Gu, Zeke Xie
cs.LGcs.AIarXiv:2601.23000v12026The Diffusion Duality, Chapter II: $Ψ$-Samplers
Justin Deschenaux, Caglar Gulcehre, Subham Sekhar Sahoo
cs.LGarXiv:2602.21185v22026Sparsity in Deep Learning: Pruning and growth for efficient inference and training in neural networks
Torsten Hoefler, Dan Alistarh, Tal Ben-Nun +2
cs.LGcs.AIcs.ARarXiv:2102.00554v12021Vectorizing the Trie: Efficient Constrained Decoding for LLM-based Generative Retrieval on Accelerators
Zhengyang Su, Isay Katsman, Yueqi Wang +10
cs.IRcs.CLcs.LGarXiv:2602.22647v22026Multimodal Few-Shot Learning with Frozen Language Models
Maria Tsimpoukelli, Jacob Menick, Serkan Cabi +3
cs.CVcs.CLcs.LGarXiv:2106.13884v22021Heterogeneous Agent Collaborative Reinforcement Learning
Zhixia Zhang, Zixuan Huang, Gongxun Li +10
cs.LGarXiv:2603.02604v22026Mode Seeking meets Mean Seeking for Fast Long Video Generation
Shengqu Cai, Weili Nie, Chao Liu +8
cs.CVcs.LGarXiv:2602.24289v12026Model-Based Reinforcement Learning for Atari
Lukasz Kaiser, Mohammad Babaeizadeh, Piotr Milos +11
cs.LGstat.MLarXiv:1903.00374v52019Jet-RL: Enabling On-Policy FP8 Reinforcement Learning with Unified Training and Rollout Precision Flow
Haocheng Xi, Charlie Ruan, Peiyuan Liao +7
cs.LGcs.CLarXiv:2601.14243v22026AudioLM: a Language Modeling Approach to Audio Generation
Zalán Borsos, Raphaël Marinier, Damien Vincent +8
cs.SDcs.LGeess.ASarXiv:2209.03143v22022A high-bias, low-variance introduction to Machine Learning for physicists
Pankaj Mehta, Marin Bukov, Ching-Hao Wang +4
physics.comp-phcond-mat.stat-mechcs.LGarXiv:1803.08823v32018On Robustness and Chain-of-Thought Consistency of RL-Finetuned VLMs
Rosie Zhao, Anshul Shah, Xiaoyu Zhu +5
cs.LGarXiv:2602.12506v32026Does the Whole Exceed its Parts? The Effect of AI Explanations on Complementary Team Performance
Gagan Bansal, Tongshuang Wu, Joyce Zhou +5
cs.AIcs.CLcs.HCarXiv:2006.14779v32020Training Diffusion Models with Reinforcement Learning
Kevin Black, Michael Janner, Yilun Du +2
cs.LGcs.AIcs.CVarXiv:2305.13301v42023Ladder Variational Autoencoders
Casper Kaae Sønderby, Tapani Raiko, Lars Maaløe +2
stat.MLcs.LGarXiv:1602.02282v32016BBN: Bilateral-Branch Network with Cumulative Learning for Long-Tailed Visual Recognition
Boyan Zhou, Quan Cui, Xiu-Shen Wei +1
cs.CVcs.LGarXiv:1912.02413v42019Deep Predictive Coding Networks for Video Prediction and Unsupervised Learning
William Lotter, Gabriel Kreiman, David Cox
cs.LGcs.AIcs.CVarXiv:1605.08104v52016Visual7W: Grounded Question Answering in Images
Yuke Zhu, Oliver Groth, Michael Bernstein +1
cs.CVcs.LGcs.NEarXiv:1511.03416v42015Quartet II: Accurate LLM Pre-Training in NVFP4 by Improved Unbiased Gradient Estimation
Andrei Panferov, Erik Schultheis, Soroush Tabesh +1
cs.LGarXiv:2601.22813v22026Universal Transformers
Mostafa Dehghani, Stephan Gouws, Oriol Vinyals +2
cs.CLcs.LGstat.MLarXiv:1807.03819v32018Extracting Training Data from Diffusion Models
Nicholas Carlini, Jamie Hayes, Milad Nasr +6
cs.CRcs.CVcs.LGarXiv:2301.13188v12023Interpretable Machine Learning: Fundamental Principles and 10 Grand Challenges
Cynthia Rudin, Chaofan Chen, Zhi Chen +3
cs.LGstat.MLarXiv:2103.11251v22021Deep, Convolutional, and Recurrent Models for Human Activity Recognition using Wearables
Nils Y. Hammerla, Shane Halloran, Thomas Ploetz
cs.LGcs.AIcs.HCarXiv:1604.08880v12016Jukebox: A Generative Model for Music
Prafulla Dhariwal, Heewoo Jun, Christine Payne +3
eess.AScs.LGcs.SDarXiv:2005.00341v12020PromptRL: Prompt Matters in RL for Flow-Based Image Generation
Fu-Yun Wang, Han Zhang, Michael Gharbi +2
cs.CVcs.LGarXiv:2602.01382v12026Photographic Image Synthesis with Cascaded Refinement Networks
Qifeng Chen, Vladlen Koltun
cs.CVcs.AIcs.GRarXiv:1707.09405v12017A Survey of Data Augmentation Approaches for NLP
Steven Y. Feng, Varun Gangal, Jason Wei +4
cs.CLcs.AIcs.LGarXiv:2105.03075v52021Generation Enhances Understanding in Unified Multimodal Models via Multi-Representation Generation
Zihan Su, Hongyang Wei, Kangrui Cen +4
cs.CVcs.LGarXiv:2601.21406v32026KromHC: Manifold-Constrained Hyper-Connections with Kronecker-Product Residual Matrices
Wuyang Zhou, Yuxuan Gu, Giorgos Iacovides +1
cs.CLcs.LGarXiv:2601.21579v22026A Simple Baseline for Bayesian Uncertainty in Deep Learning
Wesley Maddox, Timur Garipov, Pavel Izmailov +2
cs.LGcs.AIcs.CVarXiv:1902.02476v22019Attention Sinks Are Provably Necessary in Softmax Transformers: Evidence from Trigger-Conditional Tasks
Yuval Ran-Milo
cs.LGarXiv:2603.11487v52026Certified Defenses against Adversarial Examples
Aditi Raghunathan, Jacob Steinhardt, Percy Liang
cs.LGarXiv:1801.09344v22018GraphVAE: Towards Generation of Small Graphs Using Variational Autoencoders
Martin Simonovsky, Nikos Komodakis
cs.LGcs.CVcs.NEarXiv:1802.03480v12018Questioning the AI: Informing Design Practices for Explainable AI User Experiences
Q. Vera Liao, Daniel Gruen, Sarah Miller
cs.HCcs.AIcs.LGarXiv:2001.02478v32020Safe Reinforcement Learning via Shielding
Mohammed Alshiekh, Roderick Bloem, Ruediger Ehlers +3
cs.LOcs.AIcs.LGarXiv:1708.08611v22017Stochastic Gradient Hamiltonian Monte Carlo
Tianqi Chen, Emily B. Fox, Carlos Guestrin
stat.MEcs.LGstat.MLarXiv:1402.4102v22014Attentional Factorization Machines: Learning the Weight of Feature Interactions via Attention Networks
Jun Xiao, Hao Ye, Xiangnan He +3
cs.LGarXiv:1708.04617v12017RLBench: The Robot Learning Benchmark & Learning Environment
Stephen James, Zicong Ma, David Rovick Arrojo +1
cs.ROcs.AIcs.CVarXiv:1909.12271v12019On Fairness and Calibration
Geoff Pleiss, Manish Raghavan, Felix Wu +2
cs.LGcs.CYstat.MLarXiv:1709.02012v22017Safety Verification of Deep Neural Networks
Xiaowei Huang, Marta Kwiatkowska, Sen Wang +1
cs.AIcs.LGstat.MLarXiv:1610.06940v32016RLAnything: Forge Environment, Policy, and Reward Model in Completely Dynamic RL System
Yinjie Wang, Tianbao Xie, Ke Shen +2
cs.LGcs.AIcs.CLarXiv:2602.02488v12026Autonomous Agents Coordinating Distributed Discovery Through Emergent Artifact Exchange
Fiona Y. Wang, Lee Marom, Subhadeep Pal +4
cs.AIcond-mat.dis-nncs.LGarXiv:2603.14312v12026Efficient Low-rank Multimodal Fusion with Modality-Specific Factors
Zhun Liu, Ying Shen, Varun Bharadhwaj Lakshminarasimhan +3
cs.AIcs.LGstat.MLarXiv:1806.00064v12018CogView: Mastering Text-to-Image Generation via Transformers
Ming Ding, Zhuoyi Yang, Wenyi Hong +8
cs.CVcs.LGarXiv:2105.13290v32021Meta-Learning with Implicit Gradients
Aravind Rajeswaran, Chelsea Finn, Sham Kakade +1
cs.LGcs.AImath.OCarXiv:1909.04630v12019GraphMAE: Self-Supervised Masked Graph Autoencoders
Zhenyu Hou, Xiao Liu, Yukuo Cen +4
cs.LGarXiv:2205.10803v32022ShapeR: Robust Conditional 3D Shape Generation from Casual Captures
Yawar Siddiqui, Duncan Frost, Samir Aroudj +9
cs.CVcs.LGarXiv:2601.11514v12026The Expressive Power of Neural Networks: A View from the Width
Zhou Lu, Hongming Pu, Feicheng Wang +2
cs.LGarXiv:1709.02540v32017Spectral Signatures in Backdoor Attacks
Brandon Tran, Jerry Li, Aleksander Madry
cs.LGcs.CRstat.MLarXiv:1811.00636v12018Memory Caching: RNNs with Growing Memory
Ali Behrouz, Zeman Li, Yuan Deng +3
cs.LGcs.AIarXiv:2602.24281v12026A Survey on Knowledge Graph-Based Recommender Systems
Qingyu Guo, Fuzhen Zhuang, Chuan Qin +4
cs.IRcs.LGstat.MLarXiv:2003.00911v12020Molecular Transformer - A Model for Uncertainty-Calibrated Chemical Reaction Prediction
Philippe Schwaller, Teodoro Laino, Théophile Gaudin +3
physics.chem-phcs.LGarXiv:1811.02633v22018On Evaluating Adversarial Robustness
Nicholas Carlini, Anish Athalye, Nicolas Papernot +6
cs.LGcs.CRstat.MLarXiv:1902.06705v22019VTAM: Video-Tactile-Action Models for Complex Physical Interaction Beyond VLAs
Haoran Yuan, Weigang Yi, Zhenyu Zhang +9
cs.ROcs.AIcs.CVarXiv:2603.23481v12026MTEB: Massive Text Embedding Benchmark
Niklas Muennighoff, Nouamane Tazi, Loïc Magne +1
cs.CLcs.IRcs.LGarXiv:2210.07316v32022