Machine Learning
Papers filed under cs.LG on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.
Search paper metadata (including unsummarized papers)
6,481 to 6,540 of 20,205
Adaptive Gradient Descent without Descent
Yura Malitsky, Konstantin Mishchenko
math.OCcs.LGmath.NAarXiv:1910.09529v22019Interactive Post-Training for Vision-Language-Action Models
Shuhan Tan, Kairan Dou, Yue Zhao +1
cs.LGcs.AIcs.CVarXiv:2505.17016v12025Self-supervised Video Object Segmentation by Motion Grouping
Charig Yang, Hala Lamdouar, Erika Lu +2
cs.CVcs.LGarXiv:2104.07658v2202114 Examples of How LLMs Can Transform Materials Science and Chemistry: A Reflection on a Large Language Model Hackathon
Kevin Maik Jablonka, Qianxiang Ai, Alexander Al-Feghali +50
cond-mat.mtrl-scics.LGphysics.chem-pharXiv:2306.06283v42023Gated Graph Recurrent Neural Networks
Luana Ruiz, Fernando Gama, Alejandro Ribeiro
eess.SPcs.LGarXiv:2002.01038v22020A Large Scale Event-based Detection Dataset for Automotive
Pierre de Tournemire, Davide Nitti, Etienne Perot +2
cs.CVcs.LGcs.ROarXiv:2001.08499v32020Provably Efficient Safe Exploration via Primal-Dual Policy Optimization
Dongsheng Ding, Xiaohan Wei, Zhuoran Yang +2
cs.LGmath.OCstat.MLarXiv:2003.00534v22020LayoutTransformer: Layout Generation and Completion with Self-attention
Kamal Gupta, Justin Lazarow, Alessandro Achille +3
cs.CVcs.LGarXiv:2006.14615v22020OS-Harm: A Benchmark for Measuring Safety of Computer Use Agents
Thomas Kuntz, Agatha Duzan, Hao Zhao +4
cs.SEcs.LGarXiv:2506.14866v22025Are You Thinking What I am Thinking? : Examining Conceptual Separation in Neural Architectures
Jaee Ponde, Roshni Agarwal, Subhashis Banerjee
cs.LGcs.AIarXiv:2609.00764v12026Beyond Periodicity: Towards a Unifying Framework for Activations in Coordinate-MLPs
Sameera Ramasinghe, Simon Lucey
cs.LGarXiv:2111.15135v22021DriveTransformer: Unified Transformer for Scalable End-to-End Autonomous Driving
Xiaosong Jia, Junqi You, Zhiyuan Zhang +1
cs.LGcs.CVcs.ROarXiv:2503.07656v22025Unifying Conformal Language Tasks with In-Context Ensembles
Xiao Shi Huang, Chen-Yuan Lin, Bruce Kuwahara +2
cs.CLcs.LGstat.MLarXiv:2609.03005v12026The Geometry of Ignorance: LLMs Know When to Temper Bayesian Priors
Toni J. B. Liu, Jiajun Bao, Yizhou Liu +4
cs.LGcs.AIcs.CLarXiv:2609.02959v12026Reinforcement Learning in Economics and Finance
Arthur Charpentier, Romuald Elie, Carl Remlinger
econ.THcs.LGq-fin.CParXiv:2003.10014v12020Machine Learning Methods for Cancer Classification Using Gene Expression Data: A Review
Fadi Alharbi, Aleksandar Vakanski
cs.LGarXiv:2301.12222v12023Orthogonal Ensembles and Tested Explanations for Performer-Independent Body-Motion Emotion Recognition
Naoto Nishida, Yoshio Ishiguro
cs.CVcs.HCcs.LGarXiv:2609.02510v12026SoK: Where Do Flow Labels Come From? Auditing Label Provenance in Encrypted Traffic Benchmarks
Sizhe Huang, Shujie Yang
cs.NIcs.LGarXiv:2609.02140v12026Learning From Labeled And Unlabeled Data: An Empirical Study Across Techniques And Domains
N. V. Chawla, Grigoris Karakoulas
cs.LGarXiv:1109.2047v12011LLMs Accelerate Annotation for Medical Information Extraction
Akshay Goel, Almog Gueta, Omry Gilon +10
cs.CLcs.AIcs.LGarXiv:2312.02296v12023Learning Deep Networks from Noisy Labels with Dropout Regularization
Ishan Jindal, Matthew Nokleby, Xuewen Chen
cs.CVcs.LGstat.MLarXiv:1705.03419v12017AMO: Adaptive Motion Optimization for Hyper-Dexterous Humanoid Whole-Body Control
Jialong Li, Xuxin Cheng, Tianshu Huang +3
cs.ROcs.AIcs.LGarXiv:2505.03738v12025A Survey on Deep Learning for Neuroimaging-based Brain Disorder Analysis
Li Zhang, Mingliang Wang, Mingxia Liu +1
eess.IVcs.CVcs.LGarXiv:2005.04573v12020Hardware-Accelerated Instance Segmentation for Resource-Constrained Space Robotics with Criticality Analysis
Siddhant Shete, Hilmi Dogu Kücüker, Udo Frese +1
cs.ROcs.ARcs.CVarXiv:2609.02219v12026MCP Safety Audit: LLMs with the Model Context Protocol Allow Major Security Exploits
Brandon Radosevich, John Halloran
cs.CRcs.AIcs.LGarXiv:2504.03767v22025Random vector functional link network: recent developments, applications, and future directions
A. K. Malik, Ruobin Gao, M. A. Ganaie +2
cs.NEcs.LGcs.ROarXiv:2203.11316v22022Test-Time Training Done Right
Tianyuan Zhang, Sai Bi, Yicong Hong +6
cs.LGcs.CLcs.CVarXiv:2505.23884v12025Source Distribution Estimation by Posterior Averaging
Trung-Dung Hoang, Lisa M. Koch
cs.LGarXiv:2609.02622v12026FlexTok: Resampling Images into 1D Token Sequences of Flexible Length
Roman Bachmann, Jesse Allardice, David Mizrahi +6
cs.CVcs.LGarXiv:2502.13967v22025Scalable Direction-Following TTS via Voice Impression-Guided Pseudo Triplet Construction
Kenichi Fujita, Yusuke Ijima
cs.SDcs.CLcs.LGarXiv:2609.02623v12026Breadth Beats Depth: Improving GCG-Based Jailbreak Optimization with Breadth-Oriented Suffix Search
Shiliang Xiao, Jingsong Wei, Yuzhi Liang +3
cs.CLcs.LGarXiv:2609.02172v12026Personalized HeartSteps: A Reinforcement Learning Algorithm for Optimizing Physical Activity
Peng Liao, Kristjan Greenewald, Predrag Klasnja +1
cs.LGcs.AIarXiv:1909.03539v12019Defining and Characterizing Reward Hacking
Joar Skalse, Nikolaus H. R. Howe, Dmitrii Krasheninnikov +1
cs.LGstat.MLarXiv:2209.13085v22022Anomaly Detection of Time Series with Smoothness-Inducing Sequential Variational Auto-Encoder
Longyuan Li, Junchi Yan, Haiyang Wang +1
cs.LGcs.AIarXiv:2102.01331v12021Reasoning with Sampling: Your Base Model is Smarter Than You Think
Aayush Karan, Yilun Du
cs.LGcs.AIcs.CLarXiv:2510.14901v12025Skill-SD: Skill-Conditioned Self-Distillation for Multi-turn LLM Agents
Hao Wang, Guozhi Wang, Han Xiao +8
cs.LGcs.AIcs.CLarXiv:2604.10674v12026Counterfactual Explainable Recommendation
Juntao Tan, Shuyuan Xu, Yingqiang Ge +3
cs.IRcs.LGarXiv:2108.10539v32021Second-order Non-local Attention Networks for Person Re-identification
Bryan, Xia, Yuan Gong +2
cs.CVcs.AIcs.LGarXiv:1909.00295v12019Technology Readiness Levels for AI & ML
Alexander Lavin, Gregory Renard
cs.SEcs.AIcs.LGarXiv:2006.12497v32020TransPolymer: a Transformer-based language model for polymer property predictions
Changwen Xu, Yuyang Wang, Amir Barati Farimani
cs.LGphysics.chem-pharXiv:2209.01307v42022More Agents Is All You Need
Junyou Li, Qin Zhang, Yangbin Yu +2
cs.CLcs.AIcs.LGarXiv:2402.05120v22024SEAL: Reinforcing Global Safety in Mixture-of-Experts through Shared Expert ALignment
Qingyu Meng, Yiwei Zha, Jiahuan Pei +3
cs.LGcs.AIcs.CRarXiv:2609.02293v12026Untangling the Mechanisms of Misleading Context in Medical Question Answering
Robin Linzmayer, Noémie Elhadad
cs.CLcs.AIcs.LGarXiv:2609.02754v12026Recurrent Neural Network Attention Mechanisms for Interpretable System Log Anomaly Detection
Andy Brown, Aaron Tuor, Brian Hutchinson +1
cs.LGcs.NEstat.MLarXiv:1803.04967v12018Towards One-for-All Robustness Across a Continuum of Threat Levels
Zhichao Hou, Xiaorui Liu
cs.LGcs.AIarXiv:2609.02440v12026Audio-Reasoner: Improving Reasoning Capability in Large Audio Language Models
Zhifei Xie, Mingbao Lin, Zihang Liu +3
cs.SDcs.AIcs.CLarXiv:2503.02318v22025DeepOPF: A Feasibility-Optimized Deep Neural Network Approach for AC Optimal Power Flow Problems
Xiang Pan, Minghua Chen, Tianyu Zhao +1
eess.SYcs.LGarXiv:2007.01002v62020Right Question is Already Half the Answer: Fully Unsupervised LLM Reasoning Incentivization
Qingyang Zhang, Haitao Wu, Changqing Zhang +2
cs.LGarXiv:2504.05812v32025ResearchRubrics: A Benchmark of Prompts and Rubrics For Evaluating Deep Research Agents
Manasi Sharma, Chen Bo Calvin Zhang, Chaithanya Bandi +13
cs.AIcs.CLcs.LGarXiv:2511.07685v12025On the Power and Limitations of Random Features for Understanding Neural Networks
Gilad Yehudai, Ohad Shamir
cs.LGcs.NEstat.MLarXiv:1904.00687v42019Unifying Graph Convolutional Neural Networks and Label Propagation
Hongwei Wang, Jure Leskovec
cs.LGstat.MLarXiv:2002.06755v12020What matters for Representation Alignment: Global Information or Spatial Structure?
Jaskirat Singh, Xingjian Leng, Zongze Wu +4
cs.CVcs.AIcs.GRarXiv:2512.10794v12025R1-Zero's "Aha Moment" in Visual Reasoning on a 2B Non-SFT Model
Hengguang Zhou, Xirui Li, Ruochen Wang +3
cs.AIcs.CVcs.LGarXiv:2503.05132v22025Fine-Tuning Large Neural Language Models for Biomedical Natural Language Processing
Robert Tinn, Hao Cheng, Yu Gu +5
cs.CLcs.LGarXiv:2112.07869v12021KodCode: A Diverse, Challenging, and Verifiable Synthetic Dataset for Coding
Zhangchen Xu, Yang Liu, Yueqin Yin +2
cs.LGcs.AIcs.CLarXiv:2503.02951v22025Low-Rank Modeling and Its Applications in Image Analysis
Xiaowei Zhou, Can Yang, Hongyu Zhao +1
cs.CVcs.LGstat.MLarXiv:1401.3409v32014The Art of Scaling Reinforcement Learning Compute for LLMs
Devvrit Khatri, Lovish Madaan, Rishabh Tiwari +6
cs.LGcs.AIarXiv:2510.13786v12025Network Morphism
Tao Wei, Changhu Wang, Yong Rui +1
cs.LGcs.CVcs.NEarXiv:1603.01670v22016TinyOL: TinyML with Online-Learning on Microcontrollers
Haoyu Ren, Darko Anicic, Thomas Runkler
cs.LGcs.DCeess.SYarXiv:2103.08295v32021Memory Injection Attacks on LLM Agents via Query-Only Interaction
Shen Dong, Shaochen Xu, Pengfei He +5
cs.LGarXiv:2503.03704v52025