Machine Learning

Papers filed under cs.LG on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.

Search paper metadata (including unsummarized papers)

19,081 to 19,140 of 20,454

  1. Breaking the Bubble: Asynchronous Pipeline Parallel Training with Bounded Weight Inconsistency

    Itay Elam, Eliron Rahimi, Avi Mendelson +1

    cs.LGarXiv:2606.07881v12026
  2. Recurrent Neural Networks for Multivariate Time Series with Missing Values

    Zhengping Che, Sanjay Purushotham, Kyunghyun Cho +2

    cs.LGcs.NEstat.MLarXiv:1606.01865v22016
  3. Chiaroscuro Attention: Spending Compute in the Dark

    Prateek Kumar Sikdar

    cs.CLcs.AIcs.LGarXiv:2606.08327v22026
  4. Generating Sentences from a Continuous Space

    Samuel R. Bowman, Luke Vilnis, Oriol Vinyals +3

    cs.LGcs.CLarXiv:1511.06349v42015
  5. Learn to Explain: Multimodal Reasoning via Thought Chains for Science Question Answering

    Pan Lu, Swaroop Mishra, Tony Xia +6

    cs.CLcs.AIcs.CVarXiv:2209.09513v22022
  6. Big Self-Supervised Models are Strong Semi-Supervised Learners

    Ting Chen, Simon Kornblith, Kevin Swersky +2

    cs.LGcs.CVstat.MLarXiv:2006.10029v22020
  7. Certified Adversarial Robustness via Randomized Smoothing

    Jeremy M Cohen, Elan Rosenfeld, J. Zico Kolter

    cs.LGstat.MLarXiv:1902.02918v22019
  8. Common Voice: A Massively-Multilingual Speech Corpus

    Rosana Ardila, Megan Branson, Kelly Davis +7

    cs.CLcs.LGarXiv:1912.06670v22019
  9. Understanding Contrastive Representation Learning through Alignment and Uniformity on the Hypersphere

    Tongzhou Wang, Phillip Isola

    cs.LGcs.CVstat.MLarXiv:2005.10242v102020
  10. Methods for Interpreting and Understanding Deep Neural Networks

    Grégoire Montavon, Wojciech Samek, Klaus-Robert Müller

    cs.LGstat.MLarXiv:1706.07979v12017
  11. When Behavioral Safety Evaluation Fails: A Representation-Level Perspective

    Enyi Jiang, Anders Gjølbye, Yibo Jacky Zhang +1

    cs.LGcs.AIcs.CLarXiv:2606.08044v22026
  12. On the Spectral Bias of Neural Networks

    Nasim Rahaman, Aristide Baratin, Devansh Arpit +5

    stat.MLcs.LGarXiv:1806.08734v32018
  13. Hierarchical Graph Representation Learning with Differentiable Pooling

    Rex Ying, Jiaxuan You, Christopher Morris +3

    cs.LGcs.NEcs.SIarXiv:1806.08804v42018
  14. Phase Marginalization for Patch-Grid Instability in Vision Transformers

    Oğuzhan Ercan

    cs.CVcs.LGarXiv:2606.08132v12026
  15. A Comparative Analysis of XGBoost

    Candice Bentéjac, Anna Csörgő, Gonzalo Martínez-Muñoz

    cs.LGstat.MLarXiv:1911.01914v12019
  16. Revisiting Semi-Supervised Learning with Graph Embeddings

    Zhilin Yang, William W. Cohen, Ruslan Salakhutdinov

    cs.LGarXiv:1603.08861v22016
  17. Sanity Checks for Saliency Maps

    Julius Adebayo, Justin Gilmer, Michael Muelly +3

    cs.CVcs.LGstat.MLarXiv:1810.03292v32018
  18. PIPE-Cypher: Automatic Enterprise Benchmark Generation for Text-to-Cypher Systems

    Suraj Ranganath, Anish Raghavendra

    cs.LGcs.AIcs.DBarXiv:2606.08481v12026
  19. Beyond the Imitation Game: Quantifying and extrapolating the capabilities of language models

    Aarohi Srivastava, Abhinav Rastogi, Abhishek Rao +448

    cs.CLcs.AIcs.CYarXiv:2206.04615v32022
  20. Imbalanced-learn: A Python Toolbox to Tackle the Curse of Imbalanced Datasets in Machine Learning

    Guillaume Lemaitre, Fernando Nogueira, Christos K. Aridas

    cs.LGarXiv:1609.06570v12016
  21. SeqGAN: Sequence Generative Adversarial Nets with Policy Gradient

    Lantao Yu, Weinan Zhang, Jun Wang +1

    cs.LGcs.AIarXiv:1609.05473v62016
  22. End-to-End Context Compression at Scale

    Ang Li, Sean McLeish, Haozhe Chen +12

    cs.CLcs.AIcs.LGarXiv:2606.09659v12026
  23. One pixel attack for fooling deep neural networks

    Jiawei Su, Danilo Vasconcellos Vargas, Sakurai Kouichi

    cs.LGcs.CVstat.MLarXiv:1710.08864v72017
  24. Late-Layer Fusion is Enough: Dual-Path Vision Token Routing for Multimodal Large Language Models under Visual Saturation

    Siyuan Liu, Jinyang Wu

    cs.AIcs.CLcs.CVarXiv:2606.09131v12026
  25. MilliVid: Hierarchical Latents for Long-Range Consistency in Video Generation

    Ishaan Preetam Chandratreya, David Charatan, Basile Van Hoorick +4

    cs.CVcs.LGarXiv:2606.09056v12026
  26. CommonsenseQA: A Question Answering Challenge Targeting Commonsense Knowledge

    Alon Talmor, Jonathan Herzig, Nicholas Lourie +1

    cs.CLcs.AIcs.LGarXiv:1811.00937v22018
  27. Image Super-Resolution via Iterative Refinement

    Chitwan Saharia, Jonathan Ho, William Chan +3

    eess.IVcs.CVcs.LGarXiv:2104.07636v22021
  28. MoleculeNet: A Benchmark for Molecular Machine Learning

    Zhenqin Wu, Bharath Ramsundar, Evan N. Feinberg +5

    cs.LGphysics.chem-phstat.MLarXiv:1703.00564v32017
  29. RT-1: Robotics Transformer for Real-World Control at Scale

    Anthony Brohan, Noah Brown, Justice Carbajal +48

    cs.ROcs.AIcs.CLarXiv:2212.06817v22022
  30. Generating Long Sequences with Sparse Transformers

    Rewon Child, Scott Gray, Alec Radford +1

    cs.LGstat.MLarXiv:1904.10509v12019
  31. Fine-Tuning Language Models from Human Preferences

    Daniel M. Ziegler, Nisan Stiennon, Jeffrey Wu +5

    cs.CLcs.LGstat.MLarXiv:1909.08593v22019
  32. Mastering Atari, Go, Chess and Shogi by Planning with a Learned Model

    Julian Schrittwieser, Ioannis Antonoglou, Thomas Hubert +9

    cs.LGstat.MLarXiv:1911.08265v22019
  33. A Review of Uncertainty Quantification in Deep Learning: Techniques, Applications and Challenges

    Moloud Abdar, Farhad Pourpanah, Sadiq Hussain +9

    cs.LGcs.AIcs.CVarXiv:2011.06225v42020
  34. Co-teaching: Robust Training of Deep Neural Networks with Extremely Noisy Labels

    Bo Han, Quanming Yao, Xingrui Yu +5

    cs.LGstat.MLarXiv:1804.06872v32018
  35. Multimodal Unsupervised Image-to-Image Translation

    Xun Huang, Ming-Yu Liu, Serge Belongie +1

    cs.CVcs.LGstat.MLarXiv:1804.04732v22018
  36. Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems

    Sergey Levine, Aviral Kumar, George Tucker +1

    cs.LGcs.AIstat.MLarXiv:2005.01643v32020
  37. Attention-based Deep Multiple Instance Learning

    Maximilian Ilse, Jakub M. Tomczak, Max Welling

    cs.LGstat.MLarXiv:1802.04712v42018
  38. Unsupervised Learning of Video Representations using LSTMs

    Nitish Srivastava, Elman Mansimov, Ruslan Salakhutdinov

    cs.LGcs.CVcs.NEarXiv:1502.04681v32015
  39. On First-Order Meta-Learning Algorithms

    Alex Nichol, Joshua Achiam, John Schulman

    cs.LGarXiv:1803.02999v32018
  40. Training Region-based Object Detectors with Online Hard Example Mining

    Abhinav Shrivastava, Abhinav Gupta, Ross Girshick

    cs.CVcs.LGarXiv:1604.03540v12016
  41. NICE: Non-linear Independent Components Estimation

    Laurent Dinh, David Krueger, Yoshua Bengio

    cs.LGarXiv:1410.8516v62014
  42. A Critical Review of Recurrent Neural Networks for Sequence Learning

    Zachary C. Lipton, John Berkowitz, Charles Elkan

    cs.LGcs.NEarXiv:1506.00019v42015
  43. Domain-Specific Language Model Pretraining for Biomedical Natural Language Processing

    Yu Gu, Robert Tinn, Hao Cheng +6

    cs.CLcs.LGarXiv:2007.15779v62020
  44. Rainbow: Combining Improvements in Deep Reinforcement Learning

    Matteo Hessel, Joseph Modayil, Hado van Hasselt +7

    cs.AIcs.LGarXiv:1710.02298v12017
  45. Online Learning for Matrix Factorization and Sparse Coding

    Julien Mairal, Francis Bach, Jean Ponce +1

    stat.MLcs.LGmath.OCarXiv:0908.0050v22009
  46. SmoothGrad: removing noise by adding noise

    Daniel Smilkov, Nikhil Thorat, Been Kim +2

    cs.LGcs.CVstat.MLarXiv:1706.03825v12017
  47. Unsupervised Data Augmentation for Consistency Training

    Qizhe Xie, Zihang Dai, Eduard Hovy +2

    cs.LGcs.AIcs.CLarXiv:1904.12848v62019
  48. Deep Networks with Stochastic Depth

    Gao Huang, Yu Sun, Zhuang Liu +2

    cs.LGcs.CVcs.NEarXiv:1603.09382v32016
  49. Decoupled Weight Decay Regularization

    Ilya Loshchilov, Frank Hutter

    cs.LGcs.NEmath.OCarXiv:1711.05101v32017
  50. A Tutorial on Principal Component Analysis

    Jonathon Shlens

    cs.LGstat.MLarXiv:1404.1100v12014
  51. Meta-Learning in Neural Networks: A Survey

    Timothy Hospedales, Antreas Antoniou, Paul Micaelli +1

    cs.LGstat.MLarXiv:2004.05439v22020
  52. UNet 3+: A Full-Scale Connected UNet for Medical Image Segmentation

    Huimin Huang, Lanfen Lin, Ruofeng Tong +6

    eess.IVcs.CVcs.LGarXiv:2004.08790v12020
  53. COVID-Net: A Tailored Deep Convolutional Neural Network Design for Detection of COVID-19 Cases from Chest X-Ray Images

    Linda Wang, Alexander Wong

    eess.IVcs.CVcs.LGarXiv:2003.09871v42020
  54. Probabilistic Circuits as Reasoning Machines in Artificial Intelligence (Part I)

    Robert Peharz

    cs.AIcs.LGmath.PRarXiv:2608.16565v12026
  55. Lean Refactor: Multi-Objective Controllable Proof Optimization via Agentic Strategy Search

    Jialin Lu, Soonho Kong, Rodrigo Stehling +4

    cs.LOcs.AIcs.CLarXiv:2605.20244v22026
  56. Reasoning Arena: Trace Tournaments When Verifiable Rewards Fall Short

    Han Zhou, Adam X. Yang, Laurence Aitchison +2

    cs.LGcs.AIcs.CLarXiv:2606.09380v12026
  57. BrainSurgery: Reproducible and Reliable Declarative Weight Manipulations for Model Editing and Upcycling

    Gianluca Barmina, Annemette Broch Pirchert, Andrea Blasi Núñez +2

    cs.LGcs.CLarXiv:2606.09707v12026
  58. Interpreting and Steering a Text-to-Speech Language Model with Sparse Autoencoders

    Nikita Koriagin, Georgii Aparin, Nikita Balagansky +1

    cs.LGcs.AIcs.CLarXiv:2606.10029v12026
  59. Your Language Model is Its Own Critic: Reinforcement Learning with Value Estimation from Actor's Internal States

    Yunho Choi, Jongwon Lim, Woojin Ahn +3

    cs.LGcs.AIcs.CLarXiv:2605.07579v22026
  60. Attention-Based Models for Speech Recognition

    Jan Chorowski, Dzmitry Bahdanau, Dmitriy Serdyuk +2

    cs.CLcs.LGcs.NEarXiv:1506.07503v12015