Machine Learning

Papers filed under cs.LG on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.

Search paper metadata (including unsummarized papers)

6,481 to 6,540 of 20,205

  1. Adaptive Gradient Descent without Descent

    Yura Malitsky, Konstantin Mishchenko

    math.OCcs.LGmath.NAarXiv:1910.09529v22019
  2. Interactive Post-Training for Vision-Language-Action Models

    Shuhan Tan, Kairan Dou, Yue Zhao +1

    cs.LGcs.AIcs.CVarXiv:2505.17016v12025
  3. Self-supervised Video Object Segmentation by Motion Grouping

    Charig Yang, Hala Lamdouar, Erika Lu +2

    cs.CVcs.LGarXiv:2104.07658v22021
  4. 14 Examples of How LLMs Can Transform Materials Science and Chemistry: A Reflection on a Large Language Model Hackathon

    Kevin Maik Jablonka, Qianxiang Ai, Alexander Al-Feghali +50

    cond-mat.mtrl-scics.LGphysics.chem-pharXiv:2306.06283v42023
  5. Gated Graph Recurrent Neural Networks

    Luana Ruiz, Fernando Gama, Alejandro Ribeiro

    eess.SPcs.LGarXiv:2002.01038v22020
  6. A Large Scale Event-based Detection Dataset for Automotive

    Pierre de Tournemire, Davide Nitti, Etienne Perot +2

    cs.CVcs.LGcs.ROarXiv:2001.08499v32020
  7. Provably Efficient Safe Exploration via Primal-Dual Policy Optimization

    Dongsheng Ding, Xiaohan Wei, Zhuoran Yang +2

    cs.LGmath.OCstat.MLarXiv:2003.00534v22020
  8. LayoutTransformer: Layout Generation and Completion with Self-attention

    Kamal Gupta, Justin Lazarow, Alessandro Achille +3

    cs.CVcs.LGarXiv:2006.14615v22020
  9. OS-Harm: A Benchmark for Measuring Safety of Computer Use Agents

    Thomas Kuntz, Agatha Duzan, Hao Zhao +4

    cs.SEcs.LGarXiv:2506.14866v22025
  10. Are You Thinking What I am Thinking? : Examining Conceptual Separation in Neural Architectures

    Jaee Ponde, Roshni Agarwal, Subhashis Banerjee

    cs.LGcs.AIarXiv:2609.00764v12026
  11. Beyond Periodicity: Towards a Unifying Framework for Activations in Coordinate-MLPs

    Sameera Ramasinghe, Simon Lucey

    cs.LGarXiv:2111.15135v22021
  12. DriveTransformer: Unified Transformer for Scalable End-to-End Autonomous Driving

    Xiaosong Jia, Junqi You, Zhiyuan Zhang +1

    cs.LGcs.CVcs.ROarXiv:2503.07656v22025
  13. Unifying Conformal Language Tasks with In-Context Ensembles

    Xiao Shi Huang, Chen-Yuan Lin, Bruce Kuwahara +2

    cs.CLcs.LGstat.MLarXiv:2609.03005v12026
  14. The Geometry of Ignorance: LLMs Know When to Temper Bayesian Priors

    Toni J. B. Liu, Jiajun Bao, Yizhou Liu +4

    cs.LGcs.AIcs.CLarXiv:2609.02959v12026
  15. Reinforcement Learning in Economics and Finance

    Arthur Charpentier, Romuald Elie, Carl Remlinger

    econ.THcs.LGq-fin.CParXiv:2003.10014v12020
  16. Machine Learning Methods for Cancer Classification Using Gene Expression Data: A Review

    Fadi Alharbi, Aleksandar Vakanski

    cs.LGarXiv:2301.12222v12023
  17. Orthogonal Ensembles and Tested Explanations for Performer-Independent Body-Motion Emotion Recognition

    Naoto Nishida, Yoshio Ishiguro

    cs.CVcs.HCcs.LGarXiv:2609.02510v12026
  18. SoK: Where Do Flow Labels Come From? Auditing Label Provenance in Encrypted Traffic Benchmarks

    Sizhe Huang, Shujie Yang

    cs.NIcs.LGarXiv:2609.02140v12026
  19. Learning From Labeled And Unlabeled Data: An Empirical Study Across Techniques And Domains

    N. V. Chawla, Grigoris Karakoulas

    cs.LGarXiv:1109.2047v12011
  20. LLMs Accelerate Annotation for Medical Information Extraction

    Akshay Goel, Almog Gueta, Omry Gilon +10

    cs.CLcs.AIcs.LGarXiv:2312.02296v12023
  21. Learning Deep Networks from Noisy Labels with Dropout Regularization

    Ishan Jindal, Matthew Nokleby, Xuewen Chen

    cs.CVcs.LGstat.MLarXiv:1705.03419v12017
  22. AMO: Adaptive Motion Optimization for Hyper-Dexterous Humanoid Whole-Body Control

    Jialong Li, Xuxin Cheng, Tianshu Huang +3

    cs.ROcs.AIcs.LGarXiv:2505.03738v12025
  23. A Survey on Deep Learning for Neuroimaging-based Brain Disorder Analysis

    Li Zhang, Mingliang Wang, Mingxia Liu +1

    eess.IVcs.CVcs.LGarXiv:2005.04573v12020
  24. Hardware-Accelerated Instance Segmentation for Resource-Constrained Space Robotics with Criticality Analysis

    Siddhant Shete, Hilmi Dogu Kücüker, Udo Frese +1

    cs.ROcs.ARcs.CVarXiv:2609.02219v12026
  25. MCP Safety Audit: LLMs with the Model Context Protocol Allow Major Security Exploits

    Brandon Radosevich, John Halloran

    cs.CRcs.AIcs.LGarXiv:2504.03767v22025
  26. Random vector functional link network: recent developments, applications, and future directions

    A. K. Malik, Ruobin Gao, M. A. Ganaie +2

    cs.NEcs.LGcs.ROarXiv:2203.11316v22022
  27. Test-Time Training Done Right

    Tianyuan Zhang, Sai Bi, Yicong Hong +6

    cs.LGcs.CLcs.CVarXiv:2505.23884v12025
  28. Source Distribution Estimation by Posterior Averaging

    Trung-Dung Hoang, Lisa M. Koch

    cs.LGarXiv:2609.02622v12026
  29. FlexTok: Resampling Images into 1D Token Sequences of Flexible Length

    Roman Bachmann, Jesse Allardice, David Mizrahi +6

    cs.CVcs.LGarXiv:2502.13967v22025
  30. Scalable Direction-Following TTS via Voice Impression-Guided Pseudo Triplet Construction

    Kenichi Fujita, Yusuke Ijima

    cs.SDcs.CLcs.LGarXiv:2609.02623v12026
  31. Breadth Beats Depth: Improving GCG-Based Jailbreak Optimization with Breadth-Oriented Suffix Search

    Shiliang Xiao, Jingsong Wei, Yuzhi Liang +3

    cs.CLcs.LGarXiv:2609.02172v12026
  32. Personalized HeartSteps: A Reinforcement Learning Algorithm for Optimizing Physical Activity

    Peng Liao, Kristjan Greenewald, Predrag Klasnja +1

    cs.LGcs.AIarXiv:1909.03539v12019
  33. Defining and Characterizing Reward Hacking

    Joar Skalse, Nikolaus H. R. Howe, Dmitrii Krasheninnikov +1

    cs.LGstat.MLarXiv:2209.13085v22022
  34. Anomaly Detection of Time Series with Smoothness-Inducing Sequential Variational Auto-Encoder

    Longyuan Li, Junchi Yan, Haiyang Wang +1

    cs.LGcs.AIarXiv:2102.01331v12021
  35. Reasoning with Sampling: Your Base Model is Smarter Than You Think

    Aayush Karan, Yilun Du

    cs.LGcs.AIcs.CLarXiv:2510.14901v12025
  36. Skill-SD: Skill-Conditioned Self-Distillation for Multi-turn LLM Agents

    Hao Wang, Guozhi Wang, Han Xiao +8

    cs.LGcs.AIcs.CLarXiv:2604.10674v12026
  37. Counterfactual Explainable Recommendation

    Juntao Tan, Shuyuan Xu, Yingqiang Ge +3

    cs.IRcs.LGarXiv:2108.10539v32021
  38. Second-order Non-local Attention Networks for Person Re-identification

    Bryan, Xia, Yuan Gong +2

    cs.CVcs.AIcs.LGarXiv:1909.00295v12019
  39. Technology Readiness Levels for AI & ML

    Alexander Lavin, Gregory Renard

    cs.SEcs.AIcs.LGarXiv:2006.12497v32020
  40. TransPolymer: a Transformer-based language model for polymer property predictions

    Changwen Xu, Yuyang Wang, Amir Barati Farimani

    cs.LGphysics.chem-pharXiv:2209.01307v42022
  41. More Agents Is All You Need

    Junyou Li, Qin Zhang, Yangbin Yu +2

    cs.CLcs.AIcs.LGarXiv:2402.05120v22024
  42. SEAL: Reinforcing Global Safety in Mixture-of-Experts through Shared Expert ALignment

    Qingyu Meng, Yiwei Zha, Jiahuan Pei +3

    cs.LGcs.AIcs.CRarXiv:2609.02293v12026
  43. Untangling the Mechanisms of Misleading Context in Medical Question Answering

    Robin Linzmayer, Noémie Elhadad

    cs.CLcs.AIcs.LGarXiv:2609.02754v12026
  44. Recurrent Neural Network Attention Mechanisms for Interpretable System Log Anomaly Detection

    Andy Brown, Aaron Tuor, Brian Hutchinson +1

    cs.LGcs.NEstat.MLarXiv:1803.04967v12018
  45. Towards One-for-All Robustness Across a Continuum of Threat Levels

    Zhichao Hou, Xiaorui Liu

    cs.LGcs.AIarXiv:2609.02440v12026
  46. Audio-Reasoner: Improving Reasoning Capability in Large Audio Language Models

    Zhifei Xie, Mingbao Lin, Zihang Liu +3

    cs.SDcs.AIcs.CLarXiv:2503.02318v22025
  47. DeepOPF: A Feasibility-Optimized Deep Neural Network Approach for AC Optimal Power Flow Problems

    Xiang Pan, Minghua Chen, Tianyu Zhao +1

    eess.SYcs.LGarXiv:2007.01002v62020
  48. Right Question is Already Half the Answer: Fully Unsupervised LLM Reasoning Incentivization

    Qingyang Zhang, Haitao Wu, Changqing Zhang +2

    cs.LGarXiv:2504.05812v32025
  49. ResearchRubrics: A Benchmark of Prompts and Rubrics For Evaluating Deep Research Agents

    Manasi Sharma, Chen Bo Calvin Zhang, Chaithanya Bandi +13

    cs.AIcs.CLcs.LGarXiv:2511.07685v12025
  50. On the Power and Limitations of Random Features for Understanding Neural Networks

    Gilad Yehudai, Ohad Shamir

    cs.LGcs.NEstat.MLarXiv:1904.00687v42019
  51. Unifying Graph Convolutional Neural Networks and Label Propagation

    Hongwei Wang, Jure Leskovec

    cs.LGstat.MLarXiv:2002.06755v12020
  52. What matters for Representation Alignment: Global Information or Spatial Structure?

    Jaskirat Singh, Xingjian Leng, Zongze Wu +4

    cs.CVcs.AIcs.GRarXiv:2512.10794v12025
  53. R1-Zero's "Aha Moment" in Visual Reasoning on a 2B Non-SFT Model

    Hengguang Zhou, Xirui Li, Ruochen Wang +3

    cs.AIcs.CVcs.LGarXiv:2503.05132v22025
  54. Fine-Tuning Large Neural Language Models for Biomedical Natural Language Processing

    Robert Tinn, Hao Cheng, Yu Gu +5

    cs.CLcs.LGarXiv:2112.07869v12021
  55. KodCode: A Diverse, Challenging, and Verifiable Synthetic Dataset for Coding

    Zhangchen Xu, Yang Liu, Yueqin Yin +2

    cs.LGcs.AIcs.CLarXiv:2503.02951v22025
  56. Low-Rank Modeling and Its Applications in Image Analysis

    Xiaowei Zhou, Can Yang, Hongyu Zhao +1

    cs.CVcs.LGstat.MLarXiv:1401.3409v32014
  57. The Art of Scaling Reinforcement Learning Compute for LLMs

    Devvrit Khatri, Lovish Madaan, Rishabh Tiwari +6

    cs.LGcs.AIarXiv:2510.13786v12025
  58. Network Morphism

    Tao Wei, Changhu Wang, Yong Rui +1

    cs.LGcs.CVcs.NEarXiv:1603.01670v22016
  59. TinyOL: TinyML with Online-Learning on Microcontrollers

    Haoyu Ren, Darko Anicic, Thomas Runkler

    cs.LGcs.DCeess.SYarXiv:2103.08295v32021
  60. Memory Injection Attacks on LLM Agents via Query-Only Interaction

    Shen Dong, Shaochen Xu, Pengfei He +5

    cs.LGarXiv:2503.03704v52025