Machine Learning

Papers filed under cs.LG on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.

Search paper metadata (including unsummarized papers)

8,461 to 8,520 of 20,221

  1. Molecular graph generation with Graph Neural Networks

    Pietro Bongini, Monica Bianchini, Franco Scarselli

    stat.MLcs.LGq-bio.BMarXiv:2012.07397v22020
  2. A Review on Methods and Applications in Multimodal Deep Learning

    Jabeen Summaira, Xi Li, Amin Muhammad Shoib +1

    cs.LGcs.MMarXiv:2202.09195v12022
  3. Illuminating Generalization in Deep Reinforcement Learning through Procedural Level Generation

    Niels Justesen, Ruben Rodriguez Torrado, Philip Bontrager +3

    cs.LGcs.AIstat.MLarXiv:1806.10729v52018
  4. DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning

    DeepSeek-AI, Daya Guo, Dejian Yang +197

    cs.CLcs.AIcs.LGarXiv:2501.12948v22025
  5. Theoretical Issues in Deep Networks: Approximation, Optimization and Generalization

    Tomaso Poggio, Andrzej Banburski, Qianli Liao

    cs.LGstat.MLarXiv:1908.09375v12019
  6. Automated Machine Learning: State-of-The-Art and Open Challenges

    Radwa Elshawi, Mohamed Maher, Sherif Sakr

    cs.LGstat.MLarXiv:1906.02287v22019
  7. Deep Factors for Forecasting

    Yuyang Wang, Alex Smola, Danielle C. Maddix +3

    stat.MLcs.LGarXiv:1905.12417v12019
  8. Generative Diffusion Surrogates with Analytical Variance Schedule

    Patrick Reichherzer, Gianluca Gregori, David N. Hosking +1

    cs.LGastro-ph.IMphysics.plasm-pharXiv:2609.01705v12026
  9. LoRA Fine-tuning Efficiently Undoes Safety Training in Llama 2-Chat 70B

    Simon Lermen, Charlie Rogers-Smith, Jeffrey Ladish

    cs.LGcs.AIarXiv:2310.20624v22023
  10. Knapsack based Optimal Policies for Budget-Limited Multi-Armed Bandits

    Long Tran-Thanh, Archie Chapman, Alex Rogers +1

    cs.AIcs.LGarXiv:1204.1909v12012
  11. Linking Points With Labels in 3D: A Review of Point Cloud Semantic Segmentation

    Yuxing Xie, Jiaojiao Tian, Xiao Xiang Zhu

    cs.CVcs.LGeess.IVarXiv:1908.08854v32019
  12. Deep Reinforcement Learning and Permissioned Blockchain for Content Caching in Vehicular Edge Computing and Networks

    Yueyue Dai, Du Xu, Ke Zhang +2

    cs.CRcs.LGarXiv:2011.08449v22020
  13. Fast Sampling of Diffusion Models via Operator Learning

    Hongkai Zheng, Weili Nie, Arash Vahdat +2

    cs.LGcs.CVarXiv:2211.13449v32022
  14. DINOv3

    Oriane Siméoni, Huy V. Vo, Maximilian Seitzer +23

    cs.CVcs.LGarXiv:2508.10104v12025
  15. Time Travel in LLMs: Tracing Data Contamination in Large Language Models

    Shahriar Golchin, Mihai Surdeanu

    cs.CLcs.AIcs.CRarXiv:2308.08493v32023
  16. $π_{0.5}$: a Vision-Language-Action Model with Open-World Generalization

    Physical Intelligence, Kevin Black, Noah Brown +33

    cs.LGcs.ROarXiv:2504.16054v12025
  17. MIDR: Enrichment-Augmented Indexing for Multimodal Document Retrieval

    Debanjan Mahata, Atharva Tendle, Daniel Preotiuc-Pietro +2

    cs.IRcs.AIcs.CLarXiv:2609.01316v12026
  18. Rubrics as Rewards: Reinforcement Learning Beyond Verifiable Domains

    Anisha Gunjal, Anthony Wang, Elaine Lau +4

    cs.LGcs.AIcs.CLarXiv:2507.17746v22025
  19. DAPO: An Open-Source LLM Reinforcement Learning System at Scale

    Qiying Yu, Zheng Zhang, Ruofei Zhu +32

    cs.LGcs.CLarXiv:2503.14476v22025
  20. Distributed Optimization with Arbitrary Local Solvers

    Chenxin Ma, Jakub Konečný, Martin Jaggi +4

    cs.LGmath.OCarXiv:1512.04039v22015
  21. Designing Fair AI for Managing Employees in Organizations: A Review, Critique, and Design Agenda

    Lionel P. Robert, Casey Pierce, Liz Morris +2

    cs.HCcs.AIcs.CYarXiv:2002.09054v12020
  22. Unsupervised Predictive Memory in a Goal-Directed Agent

    Greg Wayne, Chia-Chun Hung, David Amos +21

    cs.LGstat.MLarXiv:1803.10760v12018
  23. StoryBuddy: A Human-AI Collaborative Chatbot for Parent-Child Interactive Storytelling with Flexible Parental Involvement

    Zheng Zhang, Ying Xu, Yanhao Wang +6

    cs.HCcs.AIcs.CLarXiv:2202.06205v22022
  24. EEG-Inception: An Accurate and Robust End-to-End Neural Network for EEG-based Motor Imagery Classification

    Ce Zhang, Young-Keun Kim, Azim Eskandarian

    eess.SPcs.HCcs.LGarXiv:2101.10932v32021
  25. MLI: An API for Distributed Machine Learning

    Evan R. Sparks, Ameet Talwalkar, Virginia Smith +6

    cs.LGcs.DCstat.MLarXiv:1310.5426v22013
  26. Sampling Permutations for Shapley Value Estimation

    Rory Mitchell, Joshua Cooper, Eibe Frank +1

    stat.MLcs.LGmath.COarXiv:2104.12199v22021
  27. Evaluating Explainability for Graph Neural Networks

    Chirag Agarwal, Owen Queen, Himabindu Lakkaraju +1

    cs.LGcs.AIarXiv:2208.09339v22022
  28. Joint Line Segmentation and Transcription for End-to-End Handwritten Paragraph Recognition

    Théodore Bluche

    cs.CVcs.LGcs.NEarXiv:1604.08352v12016
  29. HarmoFL: Harmonizing Local and Global Drifts in Federated Learning on Heterogeneous Medical Images

    Meirui Jiang, Zirui Wang, Qi Dou

    eess.IVcs.AIcs.CVarXiv:2112.10775v32021
  30. Transferring Subspaces Between Subjects in Brain-Computer Interfacing

    Wojciech Samek, Frank C. Meinecke, Klaus-Robert Müller

    stat.MLcs.HCcs.LGarXiv:1209.4115v22012
  31. Multi-view Graph Contrastive Representation Learning for Drug-Drug Interaction Prediction

    Yingheng Wang, Yaosen Min, Xin Chen +1

    cs.LGcs.AIarXiv:2010.11711v32020
  32. Node Selection Toward Faster Convergence for Federated Learning on Non-IID Data

    Hongda Wu, Ping Wang

    cs.LGcs.AIarXiv:2105.07066v32021
  33. Deep Learning for Hybrid 5G Services in Mobile Edge Computing Systems: Learn from a Digital Twin

    Rui Dong, Changyang She, Wibowo Hardjawana +2

    eess.SPcs.ITcs.LGarXiv:1907.01523v12019
  34. Label Sanitization against Label Flipping Poisoning Attacks

    Andrea Paudice, Luis Muñoz-González, Emil C. Lupu

    stat.MLcs.CRcs.LGarXiv:1803.00992v22018
  35. Compositional Exemplars for In-context Learning

    Jiacheng Ye, Zhiyong Wu, Jiangtao Feng +2

    cs.CLcs.AIcs.LGarXiv:2302.05698v32023
  36. Defending Neural Backdoors via Generative Distribution Modeling

    Ximing Qiao, Yukun Yang, Hai Li

    cs.LGstat.MLarXiv:1910.04749v22019
  37. Deciphering antibody affinity maturation with language models and weakly supervised learning

    Jeffrey A. Ruffolo, Jeffrey J. Gray, Jeremias Sulam

    q-bio.BMcs.LGarXiv:2112.07782v12021
  38. Enhancing Person-Job Fit for Talent Recruitment: An Ability-aware Neural Network Approach

    Chuan Qin, Hengshu Zhu, Tong Xu +4

    cs.AIcs.CLcs.LGarXiv:1812.08947v12018
  39. Interpolation-Prediction Networks for Irregularly Sampled Time Series

    Satya Narayan Shukla, Benjamin M. Marlin

    cs.LGstat.MLarXiv:1909.07782v12019
  40. Let Confidence Change, Not the Prediction: Prediction-Preserving Repair for Post-hoc Calibration

    Daehwan Kim, Haejun Chung, Ikbeom Jang

    cs.LGcs.CVarXiv:2609.01072v22026
  41. SimMTM: A Simple Pre-Training Framework for Masked Time-Series Modeling

    Jiaxiang Dong, Haixu Wu, Haoran Zhang +3

    cs.LGarXiv:2302.00861v42023
  42. Barzilai-Borwein Step Size for Stochastic Gradient Descent

    Conghui Tan, Shiqian Ma, Yu-Hong Dai +1

    math.OCcs.LGstat.MLarXiv:1605.04131v22016
  43. Adversarial Reprogramming of Neural Networks

    Gamaleldin F. Elsayed, Ian Goodfellow, Jascha Sohl-Dickstein

    cs.LGcs.CRcs.CVarXiv:1806.11146v22018
  44. MAMO: Memory-Augmented Meta-Optimization for Cold-start Recommendation

    Manqing Dong, Feng Yuan, Lina Yao +2

    cs.IRcs.LGstat.MLarXiv:2007.03183v12020
  45. Compile by Training: Turning Natural-Language Specifications into Local Neural Functions

    Yuntian Deng, Pengyu Nie, Stuart Shieber

    cs.CLcs.AIcs.LGarXiv:2609.04199v12026
  46. ConCare: Personalized Clinical Feature Embedding via Capturing the Healthcare Context

    Liantao Ma, Chaohe Zhang, Yasha Wang +6

    cs.LGstat.MLarXiv:1911.12216v12019
  47. Compositional generalization through meta sequence-to-sequence learning

    Brenden M. Lake

    cs.CLcs.AIcs.LGarXiv:1906.05381v22019
  48. Self-supervised Learning on Graphs: Deep Insights and New Direction

    Wei Jin, Tyler Derr, Haochen Liu +4

    cs.LGstat.MLarXiv:2006.10141v12020
  49. $QD$-Learning: A Collaborative Distributed Strategy for Multi-Agent Reinforcement Learning Through Consensus + Innovations

    Soummya Kar, Jose' M. F. Moura, H. Vincent Poor

    stat.MLcs.LGcs.MAarXiv:1205.0047v22012
  50. SciREX: A Challenge Dataset for Document-Level Information Extraction

    Sarthak Jain, Madeleine van Zuylen, Hannaneh Hajishirzi +1

    cs.CLcs.IRcs.LGarXiv:2005.00512v12020
  51. Behavior Generation with Latent Actions

    Seungjae Lee, Yibin Wang, Haritheja Etukuru +3

    cs.LGcs.AIcs.ROarXiv:2403.03181v22024
  52. RAVE: A variational autoencoder for fast and high-quality neural audio synthesis

    Antoine Caillon, Philippe Esling

    cs.LGcs.SDeess.ASarXiv:2111.05011v22021
  53. Factoring nonnegative matrices with linear programs

    Victor Bittorf, Benjamin Recht, Christopher Re +1

    math.OCcs.LGstat.MLarXiv:1206.1270v22012
  54. Maximum-Entropy Adversarial Data Augmentation for Improved Generalization and Robustness

    Long Zhao, Ting Liu, Xi Peng +1

    cs.LGcs.CVarXiv:2010.08001v22020
  55. Kernel Instrumental Variable Regression

    Rahul Singh, Maneesh Sahani, Arthur Gretton

    cs.LGecon.EMmath.FAarXiv:1906.00232v62019
  56. Deep Learning Based Regression and Multi-class Models for Acute Oral Toxicity Prediction with Automatic Chemical Feature Extraction

    Youjun Xu, Jianfeng Pei, Luhua Lai

    stat.MLcs.LGq-bio.QMarXiv:1704.04718v32017
  57. On the Convergence Rate of Training Recurrent Neural Networks

    Zeyuan Allen-Zhu, Yuanzhi Li, Zhao Song

    cs.LGcs.DScs.NEarXiv:1810.12065v42018
  58. Generalized Product of Experts for Automatic and Principled Fusion of Gaussian Process Predictions

    Yanshuai Cao, David J. Fleet

    cs.LGcs.AIstat.MLarXiv:1410.7827v22014
  59. Zeus: Understanding and Optimizing GPU Energy Consumption of DNN Training

    Jie You, Jae-Won Chung, Mosharaf Chowdhury

    cs.LGcs.AIcs.DCarXiv:2208.06102v22022
  60. FeTrIL: Feature Translation for Exemplar-Free Class-Incremental Learning

    Grégoire Petit, Adrian Popescu, Hugo Schindler +2

    cs.CVcs.AIcs.LGarXiv:2211.13131v22022