Machine Learning

Papers filed under cs.LG on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.

Search paper metadata (including unsummarized papers)

17,281 to 17,340 of 20,205

  1. Temporal Pattern Attention for Multivariate Time Series Forecasting

    Shun-Yao Shih, Fan-Keng Sun, Hung-yi Lee

    cs.LGcs.CLstat.MLarXiv:1809.04206v32018
  2. Offline Reinforcement Learning as One Big Sequence Modeling Problem

    Michael Janner, Qiyang Li, Sergey Levine

    cs.LGcs.AIarXiv:2106.02039v42021
  3. Hypergraph Convolution and Hypergraph Attention

    Song Bai, Feihu Zhang, Philip H. S. Torr

    cs.LGcs.CVstat.MLarXiv:1901.08150v22019
  4. Scaling Vision Transformers to 22 Billion Parameters

    Mostafa Dehghani, Josip Djolonga, Basil Mustafa +39

    cs.CVcs.AIcs.LGarXiv:2302.05442v12023
  5. Classification of COVID-19 in chest X-ray images using DeTraC deep convolutional neural network

    Asmaa Abbas, Mohammed M. Abdelsamea, Mohamed Medhat Gaber

    eess.IVcs.CVcs.LGarXiv:2003.13815v32020
  6. Differentiable Convex Optimization Layers

    Akshay Agrawal, Brandon Amos, Shane Barratt +3

    cs.LGmath.OCstat.MLarXiv:1910.12430v12019
  7. Instruction Tuning for Large Language Models: A Survey

    Shengyu Zhang, Linfeng Dong, Xiaoya Li +8

    cs.CLcs.AIcs.LGarXiv:2308.10792v102023
  8. A Survey on Deep Semi-supervised Learning

    Xiangli Yang, Zixing Song, Irwin King +1

    cs.LGarXiv:2103.00550v22021
  9. Sanity Checks for Sparse Autoencoders: Do SAEs Beat Random Baselines?

    Anton Korznikov, Andrey Galichin, Alexey Dontsov +3

    cs.LGarXiv:2602.14111v12026
  10. The Creation and Detection of Deepfakes: A Survey

    Yisroel Mirsky, Wenke Lee

    cs.CVcs.LGeess.IVarXiv:2004.11138v32020
  11. Human Motion Trajectory Prediction: A Survey

    Andrey Rudenko, Luigi Palmieri, Michael Herman +3

    cs.ROcs.CVcs.LGarXiv:1905.06113v32019
  12. Deep Hidden Physics Models: Deep Learning of Nonlinear Partial Differential Equations

    Maziar Raissi

    stat.MLcs.LGmath.AParXiv:1801.06637v12018
  13. Explainability for Large Language Models: A Survey

    Haiyan Zhao, Hanjie Chen, Fan Yang +6

    cs.CLcs.AIcs.LGarXiv:2309.01029v32023
  14. Bayesian Nonparametric Federated Learning of Neural Networks

    Mikhail Yurochkin, Mayank Agarwal, Soumya Ghosh +3

    stat.MLcs.LGarXiv:1905.12022v12019
  15. PyOD: A Python Toolbox for Scalable Outlier Detection

    Yue Zhao, Zain Nasrullah, Zheng Li

    cs.LGcs.IRstat.MLarXiv:1901.01588v22019
  16. Beyond Inferring Class Representatives: User-Level Privacy Leakage From Federated Learning

    Zhibo Wang, Mengkai Song, Zhifei Zhang +3

    cs.LGcs.CRcs.CVarXiv:1812.00535v32018
  17. Explanation of Machine Learning Models Using Shapley Additive Explanation and Application for Real Data in Hospital

    Yasunobu Nohara, Koutarou Matsumoto, Hidehisa Soejima +1

    cs.LGstat.MLarXiv:2112.11071v22021
  18. Training Spiking Neural Networks Using Lessons From Deep Learning

    Jason K. Eshraghian, Max Ward, Emre Neftci +6

    cs.NEcs.ETcs.LGarXiv:2109.12894v62021
  19. InCoder: A Generative Model for Code Infilling and Synthesis

    Daniel Fried, Armen Aghajanyan, Jessy Lin +7

    cs.SEcs.CLcs.LGarXiv:2204.05999v32022
  20. X-Coder: Advancing Competitive Programming with Fully Synthetic Tasks, Solutions, and Tests

    Jie Wu, Haoling Li, Xin Zhang +7

    cs.CLcs.LGarXiv:2601.06953v22026
  21. Apodex 1.1: Scaling Agentic Intelligence for Complex Work

    Apodex Team, B. An, B. Li +68

    cs.AIcs.CLcs.LGarXiv:2608.23283v12026
  22. Towards Learning Universal, Regional, and Local Hydrological Behaviors via Machine-Learning Applied to Large-Sample Datasets

    Frederik Kratzert, Daniel Klotz, Guy Shalev +3

    cs.LGstat.MLarXiv:1907.08456v22019
  23. Reasoning Models Generate Societies of Thought

    Junsol Kim, Shiyang Lai, Nino Scherrer +2

    cs.CLcs.CYcs.LGarXiv:2601.10825v12026
  24. Is Q-learning Provably Efficient?

    Chi Jin, Zeyuan Allen-Zhu, Sebastien Bubeck +1

    cs.LGcs.AImath.OCarXiv:1807.03765v12018
  25. SimVLA: A Simple VLA Baseline for Robotic Manipulation

    Yuankai Luo, Woping Chen, Tong Liang +2

    cs.ROcs.LGarXiv:2602.18224v12026
  26. Style Tokens: Unsupervised Style Modeling, Control and Transfer in End-to-End Speech Synthesis

    Yuxuan Wang, Daisy Stanton, Yu Zhang +7

    cs.CLcs.LGcs.SDarXiv:1803.09017v12018
  27. Long Range Arena: A Benchmark for Efficient Transformers

    Yi Tay, Mostafa Dehghani, Samira Abnar +7

    cs.LGcs.AIcs.CLarXiv:2011.04006v12020
  28. Maximum Likelihood Training of Score-Based Diffusion Models

    Yang Song, Conor Durkan, Iain Murray +1

    stat.MLcs.LGarXiv:2101.09258v42021
  29. End-to-end Autonomous Driving: Challenges and Frontiers

    Li Chen, Penghao Wu, Kashyap Chitta +3

    cs.ROcs.AIcs.CVarXiv:2306.16927v32023
  30. Sharp Minima Can Generalize For Deep Nets

    Laurent Dinh, Razvan Pascanu, Samy Bengio +1

    cs.LGarXiv:1703.04933v22017
  31. Non-Autoregressive Neural Machine Translation

    Jiatao Gu, James Bradbury, Caiming Xiong +2

    cs.CLcs.LGarXiv:1711.02281v22017
  32. The Truthfulness Spectrum Hypothesis

    Zhuofan Josh Ying, Shauli Ravfogel, Nikolaus Kriegeskorte +1

    cs.LGarXiv:2602.20273v12026
  33. Interpolation Consistency Training for Semi-Supervised Learning

    Vikas Verma, Kenji Kawaguchi, Alex Lamb +4

    stat.MLcs.AIcs.LGarXiv:1903.03825v52019
  34. From Scale to Speed: Adaptive Test-Time Scaling for Image Editing

    Xiangyan Qu, Zhenlong Yuan, Jing Tang +9

    cs.CVcs.AIcs.LGarXiv:2603.00141v32026
  35. UI-Voyager: A Self-Evolving GUI Agent Learning via Failed Experience

    Zichuan Lin, Feiyu Liu, Yijun Yang +9

    cs.LGcs.AIcs.CVarXiv:2603.24533v12026
  36. Image Generation from Scene Graphs

    Justin Johnson, Agrim Gupta, Li Fei-Fei

    cs.CVcs.LGarXiv:1804.01622v12018
  37. A Survey of Modern Deep Learning based Object Detection Models

    Syed Sahil Abbas Zaidi, Mohammad Samar Ansari, Asra Aslam +3

    cs.CVcs.LGeess.IVarXiv:2104.11892v22021
  38. Auto-Keras: An Efficient Neural Architecture Search System

    Haifeng Jin, Qingquan Song, Xia Hu

    cs.LGcs.AIstat.MLarXiv:1806.10282v32018
  39. Co-occurrence Feature Learning for Skeleton based Action Recognition using Regularized Deep LSTM Networks

    Wentao Zhu, Cuiling Lan, Junliang Xing +4

    cs.CVcs.LGarXiv:1603.07772v12016
  40. GraphLab: A New Framework for Parallel Machine Learning

    Yucheng Low, Joseph Gonzalez, Aapo Kyrola +3

    cs.LGcs.DCarXiv:1006.4990v12010
  41. Large Scale Interactive Motion Forecasting for Autonomous Driving : The Waymo Open Motion Dataset

    Scott Ettinger, Shuyang Cheng, Benjamin Caine +15

    cs.CVcs.LGcs.ROarXiv:2104.10133v12021
  42. ArenaRL: Scaling RL for Open-Ended Agents via Tournament-based Relative Ranking

    Qiang Zhang, Boli Chen, Fanrui Zhang +14

    cs.LGcs.AIarXiv:2601.06487v32026
  43. Learning to Represent Programs with Graphs

    Miltiadis Allamanis, Marc Brockschmidt, Mahmoud Khademi

    cs.LGcs.AIcs.PLarXiv:1711.00740v32017
  44. Compressive Transformers for Long-Range Sequence Modelling

    Jack W. Rae, Anna Potapenko, Siddhant M. Jayakumar +1

    cs.LGstat.MLarXiv:1911.05507v12019
  45. Efficient Neural Audio Synthesis

    Nal Kalchbrenner, Erich Elsen, Karen Simonyan +7

    cs.SDcs.LGeess.ASarXiv:1802.08435v22018
  46. Deep Learning using Linear Support Vector Machines

    Yichuan Tang

    cs.LGstat.MLarXiv:1306.0239v42013
  47. Grammar as a Foreign Language

    Oriol Vinyals, Lukasz Kaiser, Terry Koo +3

    cs.CLcs.LGstat.MLarXiv:1412.7449v32014
  48. K-BERT: Enabling Language Representation with Knowledge Graph

    Weijie Liu, Peng Zhou, Zhe Zhao +4

    cs.CLcs.LGarXiv:1909.07606v12019
  49. Can LLMs Clean Up Your Mess? A Survey of Application-Ready Data Preparation with LLMs

    Wei Zhou, Jun Zhou, Haoyu Wang +16

    cs.DBcs.AIcs.CLarXiv:2601.17058v12026
  50. Small Generalizable Prompt Predictive Models Can Steer Efficient RL Post-Training of Large Reasoning Models

    Yun Qu, Qi Wang, Yixiu Mao +8

    cs.AIcs.LGarXiv:2602.01970v22026
  51. Learning to Repair Lean Proofs from Compiler Feedback

    Evan Wang, Simon Chess, Daniel Lee +4

    cs.LGarXiv:2602.02990v22026
  52. Graph Structure Learning for Robust Graph Neural Networks

    Wei Jin, Yao Ma, Xiaorui Liu +3

    cs.LGcs.CRcs.SIarXiv:2005.10203v32020
  53. LatentMem: Customizing Latent Memory for Multi-Agent Systems

    Muxin Fu, Xiangyuan Xue, Yafu Li +5

    cs.CLcs.LGcs.MAarXiv:2602.03036v22026
  54. On the Entropy Dynamics in Reinforcement Fine-Tuning of Large Language Models

    Shumin Wang, Yuexiang Xie, Wenhao Zhang +4

    cs.LGcs.AIarXiv:2602.03392v12026
  55. Compressed Sensing using Generative Models

    Ashish Bora, Ajil Jalal, Eric Price +1

    stat.MLcs.ITcs.LGarXiv:1703.03208v12017
  56. RegionCLIP: Region-based Language-Image Pretraining

    Yiwu Zhong, Jianwei Yang, Pengchuan Zhang +8

    cs.CVcs.AIcs.LGarXiv:2112.09106v12021
  57. Safe, Multi-Agent, Reinforcement Learning for Autonomous Driving

    Shai Shalev-Shwartz, Shaked Shammah, Amnon Shashua

    cs.AIcs.LGstat.MLarXiv:1610.03295v12016
  58. RoboPocket: Improve Robot Policies Instantly with Your Phone

    Junjie Fang, Wendi Chen, Han Xue +7

    cs.ROcs.AIcs.LGarXiv:2603.05504v22026
  59. Learned in Translation: Contextualized Word Vectors

    Bryan McCann, James Bradbury, Caiming Xiong +1

    cs.CLcs.AIcs.LGarXiv:1708.00107v22017
  60. DSPy: Compiling Declarative Language Model Calls into Self-Improving Pipelines

    Omar Khattab, Arnav Singhvi, Paridhi Maheshwari +10

    cs.CLcs.AIcs.IRarXiv:2310.03714v12023