Machine Learning
Papers filed under cs.LG on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.
Search paper metadata (including unsummarized papers)
8,461 to 8,520 of 20,221
Molecular graph generation with Graph Neural Networks
Pietro Bongini, Monica Bianchini, Franco Scarselli
stat.MLcs.LGq-bio.BMarXiv:2012.07397v22020A Review on Methods and Applications in Multimodal Deep Learning
Jabeen Summaira, Xi Li, Amin Muhammad Shoib +1
cs.LGcs.MMarXiv:2202.09195v12022Illuminating Generalization in Deep Reinforcement Learning through Procedural Level Generation
Niels Justesen, Ruben Rodriguez Torrado, Philip Bontrager +3
cs.LGcs.AIstat.MLarXiv:1806.10729v52018DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning
DeepSeek-AI, Daya Guo, Dejian Yang +197
cs.CLcs.AIcs.LGarXiv:2501.12948v22025Theoretical Issues in Deep Networks: Approximation, Optimization and Generalization
Tomaso Poggio, Andrzej Banburski, Qianli Liao
cs.LGstat.MLarXiv:1908.09375v12019Automated Machine Learning: State-of-The-Art and Open Challenges
Radwa Elshawi, Mohamed Maher, Sherif Sakr
cs.LGstat.MLarXiv:1906.02287v22019Deep Factors for Forecasting
Yuyang Wang, Alex Smola, Danielle C. Maddix +3
stat.MLcs.LGarXiv:1905.12417v12019Generative Diffusion Surrogates with Analytical Variance Schedule
Patrick Reichherzer, Gianluca Gregori, David N. Hosking +1
cs.LGastro-ph.IMphysics.plasm-pharXiv:2609.01705v12026LoRA Fine-tuning Efficiently Undoes Safety Training in Llama 2-Chat 70B
Simon Lermen, Charlie Rogers-Smith, Jeffrey Ladish
cs.LGcs.AIarXiv:2310.20624v22023Knapsack based Optimal Policies for Budget-Limited Multi-Armed Bandits
Long Tran-Thanh, Archie Chapman, Alex Rogers +1
cs.AIcs.LGarXiv:1204.1909v12012Linking Points With Labels in 3D: A Review of Point Cloud Semantic Segmentation
Yuxing Xie, Jiaojiao Tian, Xiao Xiang Zhu
cs.CVcs.LGeess.IVarXiv:1908.08854v32019Deep Reinforcement Learning and Permissioned Blockchain for Content Caching in Vehicular Edge Computing and Networks
Yueyue Dai, Du Xu, Ke Zhang +2
cs.CRcs.LGarXiv:2011.08449v22020Fast Sampling of Diffusion Models via Operator Learning
Hongkai Zheng, Weili Nie, Arash Vahdat +2
cs.LGcs.CVarXiv:2211.13449v32022DINOv3
Oriane Siméoni, Huy V. Vo, Maximilian Seitzer +23
cs.CVcs.LGarXiv:2508.10104v12025Time Travel in LLMs: Tracing Data Contamination in Large Language Models
Shahriar Golchin, Mihai Surdeanu
cs.CLcs.AIcs.CRarXiv:2308.08493v32023$π_{0.5}$: a Vision-Language-Action Model with Open-World Generalization
Physical Intelligence, Kevin Black, Noah Brown +33
cs.LGcs.ROarXiv:2504.16054v12025MIDR: Enrichment-Augmented Indexing for Multimodal Document Retrieval
Debanjan Mahata, Atharva Tendle, Daniel Preotiuc-Pietro +2
cs.IRcs.AIcs.CLarXiv:2609.01316v12026Rubrics as Rewards: Reinforcement Learning Beyond Verifiable Domains
Anisha Gunjal, Anthony Wang, Elaine Lau +4
cs.LGcs.AIcs.CLarXiv:2507.17746v22025DAPO: An Open-Source LLM Reinforcement Learning System at Scale
Qiying Yu, Zheng Zhang, Ruofei Zhu +32
cs.LGcs.CLarXiv:2503.14476v22025Distributed Optimization with Arbitrary Local Solvers
Chenxin Ma, Jakub Konečný, Martin Jaggi +4
cs.LGmath.OCarXiv:1512.04039v22015Designing Fair AI for Managing Employees in Organizations: A Review, Critique, and Design Agenda
Lionel P. Robert, Casey Pierce, Liz Morris +2
cs.HCcs.AIcs.CYarXiv:2002.09054v12020Unsupervised Predictive Memory in a Goal-Directed Agent
Greg Wayne, Chia-Chun Hung, David Amos +21
cs.LGstat.MLarXiv:1803.10760v12018StoryBuddy: A Human-AI Collaborative Chatbot for Parent-Child Interactive Storytelling with Flexible Parental Involvement
Zheng Zhang, Ying Xu, Yanhao Wang +6
cs.HCcs.AIcs.CLarXiv:2202.06205v22022EEG-Inception: An Accurate and Robust End-to-End Neural Network for EEG-based Motor Imagery Classification
Ce Zhang, Young-Keun Kim, Azim Eskandarian
eess.SPcs.HCcs.LGarXiv:2101.10932v32021MLI: An API for Distributed Machine Learning
Evan R. Sparks, Ameet Talwalkar, Virginia Smith +6
cs.LGcs.DCstat.MLarXiv:1310.5426v22013Sampling Permutations for Shapley Value Estimation
Rory Mitchell, Joshua Cooper, Eibe Frank +1
stat.MLcs.LGmath.COarXiv:2104.12199v22021Evaluating Explainability for Graph Neural Networks
Chirag Agarwal, Owen Queen, Himabindu Lakkaraju +1
cs.LGcs.AIarXiv:2208.09339v22022Joint Line Segmentation and Transcription for End-to-End Handwritten Paragraph Recognition
Théodore Bluche
cs.CVcs.LGcs.NEarXiv:1604.08352v12016HarmoFL: Harmonizing Local and Global Drifts in Federated Learning on Heterogeneous Medical Images
Meirui Jiang, Zirui Wang, Qi Dou
eess.IVcs.AIcs.CVarXiv:2112.10775v32021Transferring Subspaces Between Subjects in Brain-Computer Interfacing
Wojciech Samek, Frank C. Meinecke, Klaus-Robert Müller
stat.MLcs.HCcs.LGarXiv:1209.4115v22012Multi-view Graph Contrastive Representation Learning for Drug-Drug Interaction Prediction
Yingheng Wang, Yaosen Min, Xin Chen +1
cs.LGcs.AIarXiv:2010.11711v32020Node Selection Toward Faster Convergence for Federated Learning on Non-IID Data
Hongda Wu, Ping Wang
cs.LGcs.AIarXiv:2105.07066v32021Deep Learning for Hybrid 5G Services in Mobile Edge Computing Systems: Learn from a Digital Twin
Rui Dong, Changyang She, Wibowo Hardjawana +2
eess.SPcs.ITcs.LGarXiv:1907.01523v12019Label Sanitization against Label Flipping Poisoning Attacks
Andrea Paudice, Luis Muñoz-González, Emil C. Lupu
stat.MLcs.CRcs.LGarXiv:1803.00992v22018Compositional Exemplars for In-context Learning
Jiacheng Ye, Zhiyong Wu, Jiangtao Feng +2
cs.CLcs.AIcs.LGarXiv:2302.05698v32023Defending Neural Backdoors via Generative Distribution Modeling
Ximing Qiao, Yukun Yang, Hai Li
cs.LGstat.MLarXiv:1910.04749v22019Deciphering antibody affinity maturation with language models and weakly supervised learning
Jeffrey A. Ruffolo, Jeffrey J. Gray, Jeremias Sulam
q-bio.BMcs.LGarXiv:2112.07782v12021Enhancing Person-Job Fit for Talent Recruitment: An Ability-aware Neural Network Approach
Chuan Qin, Hengshu Zhu, Tong Xu +4
cs.AIcs.CLcs.LGarXiv:1812.08947v12018Interpolation-Prediction Networks for Irregularly Sampled Time Series
Satya Narayan Shukla, Benjamin M. Marlin
cs.LGstat.MLarXiv:1909.07782v12019Let Confidence Change, Not the Prediction: Prediction-Preserving Repair for Post-hoc Calibration
Daehwan Kim, Haejun Chung, Ikbeom Jang
cs.LGcs.CVarXiv:2609.01072v22026SimMTM: A Simple Pre-Training Framework for Masked Time-Series Modeling
Jiaxiang Dong, Haixu Wu, Haoran Zhang +3
cs.LGarXiv:2302.00861v42023Barzilai-Borwein Step Size for Stochastic Gradient Descent
Conghui Tan, Shiqian Ma, Yu-Hong Dai +1
math.OCcs.LGstat.MLarXiv:1605.04131v22016Adversarial Reprogramming of Neural Networks
Gamaleldin F. Elsayed, Ian Goodfellow, Jascha Sohl-Dickstein
cs.LGcs.CRcs.CVarXiv:1806.11146v22018MAMO: Memory-Augmented Meta-Optimization for Cold-start Recommendation
Manqing Dong, Feng Yuan, Lina Yao +2
cs.IRcs.LGstat.MLarXiv:2007.03183v12020Compile by Training: Turning Natural-Language Specifications into Local Neural Functions
Yuntian Deng, Pengyu Nie, Stuart Shieber
cs.CLcs.AIcs.LGarXiv:2609.04199v12026ConCare: Personalized Clinical Feature Embedding via Capturing the Healthcare Context
Liantao Ma, Chaohe Zhang, Yasha Wang +6
cs.LGstat.MLarXiv:1911.12216v12019Compositional generalization through meta sequence-to-sequence learning
Brenden M. Lake
cs.CLcs.AIcs.LGarXiv:1906.05381v22019Self-supervised Learning on Graphs: Deep Insights and New Direction
Wei Jin, Tyler Derr, Haochen Liu +4
cs.LGstat.MLarXiv:2006.10141v12020$QD$-Learning: A Collaborative Distributed Strategy for Multi-Agent Reinforcement Learning Through Consensus + Innovations
Soummya Kar, Jose' M. F. Moura, H. Vincent Poor
stat.MLcs.LGcs.MAarXiv:1205.0047v22012SciREX: A Challenge Dataset for Document-Level Information Extraction
Sarthak Jain, Madeleine van Zuylen, Hannaneh Hajishirzi +1
cs.CLcs.IRcs.LGarXiv:2005.00512v12020Behavior Generation with Latent Actions
Seungjae Lee, Yibin Wang, Haritheja Etukuru +3
cs.LGcs.AIcs.ROarXiv:2403.03181v22024RAVE: A variational autoencoder for fast and high-quality neural audio synthesis
Antoine Caillon, Philippe Esling
cs.LGcs.SDeess.ASarXiv:2111.05011v22021Factoring nonnegative matrices with linear programs
Victor Bittorf, Benjamin Recht, Christopher Re +1
math.OCcs.LGstat.MLarXiv:1206.1270v22012Maximum-Entropy Adversarial Data Augmentation for Improved Generalization and Robustness
Long Zhao, Ting Liu, Xi Peng +1
cs.LGcs.CVarXiv:2010.08001v22020Kernel Instrumental Variable Regression
Rahul Singh, Maneesh Sahani, Arthur Gretton
cs.LGecon.EMmath.FAarXiv:1906.00232v62019Deep Learning Based Regression and Multi-class Models for Acute Oral Toxicity Prediction with Automatic Chemical Feature Extraction
Youjun Xu, Jianfeng Pei, Luhua Lai
stat.MLcs.LGq-bio.QMarXiv:1704.04718v32017On the Convergence Rate of Training Recurrent Neural Networks
Zeyuan Allen-Zhu, Yuanzhi Li, Zhao Song
cs.LGcs.DScs.NEarXiv:1810.12065v42018Generalized Product of Experts for Automatic and Principled Fusion of Gaussian Process Predictions
Yanshuai Cao, David J. Fleet
cs.LGcs.AIstat.MLarXiv:1410.7827v22014Zeus: Understanding and Optimizing GPU Energy Consumption of DNN Training
Jie You, Jae-Won Chung, Mosharaf Chowdhury
cs.LGcs.AIcs.DCarXiv:2208.06102v22022FeTrIL: Feature Translation for Exemplar-Free Class-Incremental Learning
Grégoire Petit, Adrian Popescu, Hugo Schindler +2
cs.CVcs.AIcs.LGarXiv:2211.13131v22022