Machine Learning

Papers filed under cs.LG on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.

Search paper metadata (including unsummarized papers)

18,841 to 18,900 of 20,199

  1. Revisiting Semi-Supervised Learning with Graph Embeddings

    Zhilin Yang, William W. Cohen, Ruslan Salakhutdinov

    cs.LGarXiv:1603.08861v22016
  2. Sanity Checks for Saliency Maps

    Julius Adebayo, Justin Gilmer, Michael Muelly +3

    cs.CVcs.LGstat.MLarXiv:1810.03292v32018
  3. PIPE-Cypher: Automatic Enterprise Benchmark Generation for Text-to-Cypher Systems

    Suraj Ranganath, Anish Raghavendra

    cs.LGcs.AIcs.DBarXiv:2606.08481v12026
  4. Beyond the Imitation Game: Quantifying and extrapolating the capabilities of language models

    Aarohi Srivastava, Abhinav Rastogi, Abhishek Rao +448

    cs.CLcs.AIcs.CYarXiv:2206.04615v32022
  5. Imbalanced-learn: A Python Toolbox to Tackle the Curse of Imbalanced Datasets in Machine Learning

    Guillaume Lemaitre, Fernando Nogueira, Christos K. Aridas

    cs.LGarXiv:1609.06570v12016
  6. SeqGAN: Sequence Generative Adversarial Nets with Policy Gradient

    Lantao Yu, Weinan Zhang, Jun Wang +1

    cs.LGcs.AIarXiv:1609.05473v62016
  7. End-to-End Context Compression at Scale

    Ang Li, Sean McLeish, Haozhe Chen +12

    cs.CLcs.AIcs.LGarXiv:2606.09659v12026
  8. One pixel attack for fooling deep neural networks

    Jiawei Su, Danilo Vasconcellos Vargas, Sakurai Kouichi

    cs.LGcs.CVstat.MLarXiv:1710.08864v72017
  9. Late-Layer Fusion is Enough: Dual-Path Vision Token Routing for Multimodal Large Language Models under Visual Saturation

    Siyuan Liu, Jinyang Wu

    cs.AIcs.CLcs.CVarXiv:2606.09131v12026
  10. MilliVid: Hierarchical Latents for Long-Range Consistency in Video Generation

    Ishaan Preetam Chandratreya, David Charatan, Basile Van Hoorick +4

    cs.CVcs.LGarXiv:2606.09056v12026
  11. CommonsenseQA: A Question Answering Challenge Targeting Commonsense Knowledge

    Alon Talmor, Jonathan Herzig, Nicholas Lourie +1

    cs.CLcs.AIcs.LGarXiv:1811.00937v22018
  12. Image Super-Resolution via Iterative Refinement

    Chitwan Saharia, Jonathan Ho, William Chan +3

    eess.IVcs.CVcs.LGarXiv:2104.07636v22021
  13. MoleculeNet: A Benchmark for Molecular Machine Learning

    Zhenqin Wu, Bharath Ramsundar, Evan N. Feinberg +5

    cs.LGphysics.chem-phstat.MLarXiv:1703.00564v32017
  14. RT-1: Robotics Transformer for Real-World Control at Scale

    Anthony Brohan, Noah Brown, Justice Carbajal +48

    cs.ROcs.AIcs.CLarXiv:2212.06817v22022
  15. Generating Long Sequences with Sparse Transformers

    Rewon Child, Scott Gray, Alec Radford +1

    cs.LGstat.MLarXiv:1904.10509v12019
  16. Fine-Tuning Language Models from Human Preferences

    Daniel M. Ziegler, Nisan Stiennon, Jeffrey Wu +5

    cs.CLcs.LGstat.MLarXiv:1909.08593v22019
  17. Mastering Atari, Go, Chess and Shogi by Planning with a Learned Model

    Julian Schrittwieser, Ioannis Antonoglou, Thomas Hubert +9

    cs.LGstat.MLarXiv:1911.08265v22019
  18. A Review of Uncertainty Quantification in Deep Learning: Techniques, Applications and Challenges

    Moloud Abdar, Farhad Pourpanah, Sadiq Hussain +9

    cs.LGcs.AIcs.CVarXiv:2011.06225v42020
  19. Co-teaching: Robust Training of Deep Neural Networks with Extremely Noisy Labels

    Bo Han, Quanming Yao, Xingrui Yu +5

    cs.LGstat.MLarXiv:1804.06872v32018
  20. Multimodal Unsupervised Image-to-Image Translation

    Xun Huang, Ming-Yu Liu, Serge Belongie +1

    cs.CVcs.LGstat.MLarXiv:1804.04732v22018
  21. Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems

    Sergey Levine, Aviral Kumar, George Tucker +1

    cs.LGcs.AIstat.MLarXiv:2005.01643v32020
  22. Attention-based Deep Multiple Instance Learning

    Maximilian Ilse, Jakub M. Tomczak, Max Welling

    cs.LGstat.MLarXiv:1802.04712v42018
  23. Unsupervised Learning of Video Representations using LSTMs

    Nitish Srivastava, Elman Mansimov, Ruslan Salakhutdinov

    cs.LGcs.CVcs.NEarXiv:1502.04681v32015
  24. On First-Order Meta-Learning Algorithms

    Alex Nichol, Joshua Achiam, John Schulman

    cs.LGarXiv:1803.02999v32018
  25. Training Region-based Object Detectors with Online Hard Example Mining

    Abhinav Shrivastava, Abhinav Gupta, Ross Girshick

    cs.CVcs.LGarXiv:1604.03540v12016
  26. NICE: Non-linear Independent Components Estimation

    Laurent Dinh, David Krueger, Yoshua Bengio

    cs.LGarXiv:1410.8516v62014
  27. A Critical Review of Recurrent Neural Networks for Sequence Learning

    Zachary C. Lipton, John Berkowitz, Charles Elkan

    cs.LGcs.NEarXiv:1506.00019v42015
  28. Domain-Specific Language Model Pretraining for Biomedical Natural Language Processing

    Yu Gu, Robert Tinn, Hao Cheng +6

    cs.CLcs.LGarXiv:2007.15779v62020
  29. Rainbow: Combining Improvements in Deep Reinforcement Learning

    Matteo Hessel, Joseph Modayil, Hado van Hasselt +7

    cs.AIcs.LGarXiv:1710.02298v12017
  30. Online Learning for Matrix Factorization and Sparse Coding

    Julien Mairal, Francis Bach, Jean Ponce +1

    stat.MLcs.LGmath.OCarXiv:0908.0050v22009
  31. SmoothGrad: removing noise by adding noise

    Daniel Smilkov, Nikhil Thorat, Been Kim +2

    cs.LGcs.CVstat.MLarXiv:1706.03825v12017
  32. Unsupervised Data Augmentation for Consistency Training

    Qizhe Xie, Zihang Dai, Eduard Hovy +2

    cs.LGcs.AIcs.CLarXiv:1904.12848v62019
  33. Deep Networks with Stochastic Depth

    Gao Huang, Yu Sun, Zhuang Liu +2

    cs.LGcs.CVcs.NEarXiv:1603.09382v32016
  34. Decoupled Weight Decay Regularization

    Ilya Loshchilov, Frank Hutter

    cs.LGcs.NEmath.OCarXiv:1711.05101v32017
  35. A Tutorial on Principal Component Analysis

    Jonathon Shlens

    cs.LGstat.MLarXiv:1404.1100v12014
  36. Meta-Learning in Neural Networks: A Survey

    Timothy Hospedales, Antreas Antoniou, Paul Micaelli +1

    cs.LGstat.MLarXiv:2004.05439v22020
  37. UNet 3+: A Full-Scale Connected UNet for Medical Image Segmentation

    Huimin Huang, Lanfen Lin, Ruofeng Tong +6

    eess.IVcs.CVcs.LGarXiv:2004.08790v12020
  38. COVID-Net: A Tailored Deep Convolutional Neural Network Design for Detection of COVID-19 Cases from Chest X-Ray Images

    Linda Wang, Alexander Wong

    eess.IVcs.CVcs.LGarXiv:2003.09871v42020
  39. Probabilistic Circuits as Reasoning Machines in Artificial Intelligence (Part I)

    Robert Peharz

    cs.AIcs.LGmath.PRarXiv:2608.16565v12026
    Summaries:한국어
  40. Lean Refactor: Multi-Objective Controllable Proof Optimization via Agentic Strategy Search

    Jialin Lu, Soonho Kong, Rodrigo Stehling +4

    cs.LOcs.AIcs.CLarXiv:2605.20244v22026
  41. Reasoning Arena: Trace Tournaments When Verifiable Rewards Fall Short

    Han Zhou, Adam X. Yang, Laurence Aitchison +2

    cs.LGcs.AIcs.CLarXiv:2606.09380v12026
  42. BrainSurgery: Reproducible and Reliable Declarative Weight Manipulations for Model Editing and Upcycling

    Gianluca Barmina, Annemette Broch Pirchert, Andrea Blasi Núñez +2

    cs.LGcs.CLarXiv:2606.09707v12026
  43. Interpreting and Steering a Text-to-Speech Language Model with Sparse Autoencoders

    Nikita Koriagin, Georgii Aparin, Nikita Balagansky +1

    cs.LGcs.AIcs.CLarXiv:2606.10029v12026
  44. Your Language Model is Its Own Critic: Reinforcement Learning with Value Estimation from Actor's Internal States

    Yunho Choi, Jongwon Lim, Woojin Ahn +3

    cs.LGcs.AIcs.CLarXiv:2605.07579v22026
  45. Attention-Based Models for Speech Recognition

    Jan Chorowski, Dzmitry Bahdanau, Dmitriy Serdyuk +2

    cs.CLcs.LGcs.NEarXiv:1506.07503v12015
  46. Hindsight Experience Replay

    Marcin Andrychowicz, Filip Wolski, Alex Ray +7

    cs.LGcs.AIcs.NEarXiv:1707.01495v32017
  47. TRIAGE: Dialectical Reasoning for Explainable Risk Prediction on Irregularly Sampled Medical Time Series with LLMs

    Hyeongwon Jang, Gyouk Chu, Changhun Kim +3

    cs.LGcs.AIcs.CLarXiv:2606.09030v12026
  48. Deeply-Recursive Convolutional Network for Image Super-Resolution

    Jiwon Kim, Jung Kwon Lee, Kyoung Mu Lee

    cs.CVcs.LGarXiv:1511.04491v22015
  49. Event-based Vision: A Survey

    Guillermo Gallego, Tobi Delbruck, Garrick Orchard +8

    cs.CVcs.AIcs.LGarXiv:1904.08405v32019
  50. HiFi-GAN: Generative Adversarial Networks for Efficient and High Fidelity Speech Synthesis

    Jungil Kong, Jaehyeon Kim, Jaekyoung Bae

    cs.SDcs.LGeess.ASarXiv:2010.05646v22020
  51. Deep Transfer Learning with Joint Adaptation Networks

    Mingsheng Long, Han Zhu, Jianmin Wang +1

    cs.LGstat.MLarXiv:1605.06636v22016
  52. Video Diffusion Models

    Jonathan Ho, Tim Salimans, Alexey Gritsenko +3

    cs.CVcs.AIcs.LGarXiv:2204.03458v22022
  53. Conditional Image Generation with PixelCNN Decoders

    Aaron van den Oord, Nal Kalchbrenner, Oriol Vinyals +3

    cs.CVcs.LGarXiv:1606.05328v22016
  54. End-to-end Sequence Labeling via Bi-directional LSTM-CNNs-CRF

    Xuezhe Ma, Eduard Hovy

    cs.LGcs.CLstat.MLarXiv:1603.01354v52016
  55. Self-training with Noisy Student improves ImageNet classification

    Qizhe Xie, Minh-Thang Luong, Eduard Hovy +1

    cs.LGcs.CVstat.MLarXiv:1911.04252v42019
  56. Asymptotic Equivalence of Bayes Cross Validation and Widely Applicable Information Criterion in Singular Learning Theory

    Sumio Watanabe

    cs.LGarXiv:1004.2316v22010
  57. Universal adversarial perturbations

    Seyed-Mohsen Moosavi-Dezfooli, Alhussein Fawzi, Omar Fawzi +1

    cs.CVcs.AIcs.LGarXiv:1610.08401v32016
  58. Learning Efficient Convolutional Networks through Network Slimming

    Zhuang Liu, Jianguo Li, Zhiqiang Shen +3

    cs.CVcs.AIcs.LGarXiv:1708.06519v12017
  59. An Image is Worth One Word: Personalizing Text-to-Image Generation using Textual Inversion

    Rinon Gal, Yuval Alaluf, Yuval Atzmon +4

    cs.CVcs.CLcs.GRarXiv:2208.01618v12022
  60. Pixel Recurrent Neural Networks

    Aaron van den Oord, Nal Kalchbrenner, Koray Kavukcuoglu

    cs.CVcs.LGcs.NEarXiv:1601.06759v32016