Machine Learning

Papers filed under cs.LG on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.

Search paper metadata (including unsummarized papers)

4,921 to 4,980 of 20,223

  1. Unsafer in Many Turns: Benchmarking and Defending Multi-Turn Safety Risks in Tool-Using Agents

    Xu Li, Simon Yu, Minzhou Pan +5

    cs.CRcs.AIcs.CLarXiv:2602.13379v22026
  2. A Comprehensive Review of Multi-Agent Reinforcement Learning in Video Games

    Zhengyang Li, Qijin Ji, Xinghong Ling +1

    cs.LGarXiv:2509.03682v12025
  3. Vision Mamba: A Comprehensive Survey and Taxonomy

    Xiao Liu, Chenxu Zhang, Lei Zhang

    cs.CVcs.AIcs.CLarXiv:2405.04404v12024
  4. OCTID: Optical Coherence Tomography Image Database

    Peyman Gholami, Priyanka Roy, Mohana Kuppuswamy Parthasarathy +1

    cs.CVcs.LGarXiv:1812.07056v22018
  5. Audiobox: Unified Audio Generation with Natural Language Prompts

    Apoorv Vyas, Bowen Shi, Matthew Le +21

    cs.SDcs.LGeess.ASarXiv:2312.15821v12023
  6. Free-rider Attacks on Model Aggregation in Federated Learning

    Yann Fraboni, Richard Vidal, Marco Lorenzi

    cs.LGstat.MLarXiv:2006.11901v52020
  7. Out of the BLEU: how should we assess quality of the Code Generation models?

    Mikhail Evtikhiev, Egor Bogomolov, Yaroslav Sokolov +1

    cs.SEcs.LGarXiv:2208.03133v22022
  8. QuestBench: Can LLMs ask the right question to acquire information in reasoning tasks?

    Belinda Z. Li, Been Kim, Zi Wang

    cs.AIcs.CLcs.LGarXiv:2503.22674v22025
  9. DIAL: Decoupling Intent and Action via Latent World Modeling for End-to-End VLA

    Yi Chen, Yuying Ge, Hui Zhou +3

    cs.ROcs.AIcs.CVarXiv:2603.29844v22026
  10. A Mechanistic Analysis of Looped Reasoning Language Models

    Hugh Blayney, Álvaro Arroyo, Johan Obando-Ceron +4

    cs.LGcs.AIarXiv:2604.11791v12026
  11. UAV-VLA: Vision-Language-Action System for Large Scale Aerial Mission Generation

    Oleg Sautenkov, Yasheerah Yaqoot, Artem Lykov +7

    cs.ROcs.AIcs.CVarXiv:2501.05014v22025
  12. Naive Bayes and Text Classification I - Introduction and Theory

    Sebastian Raschka

    cs.LGarXiv:1410.5329v42014
  13. Federated Learning for Cyber Physical Systems: A Comprehensive Survey

    Minh K. Quan, Pubudu N. Pathirana, Mayuri Wijayasundara +5

    cs.LGcs.AIcs.CRarXiv:2505.04873v12025
  14. Fast Task Inference with Variational Intrinsic Successor Features

    Steven Hansen, Will Dabney, Andre Barreto +3

    cs.LGcs.AIstat.MLarXiv:1906.05030v22019
  15. Exploit Bounding Box Annotations for Multi-label Object Recognition

    Hao Yang, Joey Tianyi Zhou, Yu Zhang +3

    cs.CVcs.LGarXiv:1504.05843v22015
  16. Extremely Fast Decision Tree

    Chaitanya Manapragada, Geoff Webb, Mahsa Salehi

    cs.LGstat.MLarXiv:1802.08780v12018
  17. Wasserstein Learning of Deep Generative Point Process Models

    Shuai Xiao, Mehrdad Farajtabar, Xiaojing Ye +3

    cs.LGstat.MLarXiv:1705.08051v12017
  18. Code to Think, Think to Code: A Survey on Code-Enhanced Reasoning and Reasoning-Driven Code Intelligence in LLMs

    Dayu Yang, Tianyang Liu, Daoan Zhang +8

    cs.CLcs.AIcs.LGarXiv:2502.19411v12025
  19. Learning multiview 3D point cloud registration

    Zan Gojcic, Caifa Zhou, Jan D. Wegner +2

    cs.CVcs.LGarXiv:2001.05119v22020
  20. Trace and Pace: Controllable Pedestrian Animation via Guided Trajectory Diffusion

    Davis Rempe, Zhengyi Luo, Xue Bin Peng +5

    cs.CVcs.GRcs.LGarXiv:2304.01893v12023
  21. Sym-NCO: Leveraging Symmetricity for Neural Combinatorial Optimization

    Minsu Kim, Junyoung Park, Jinkyoo Park

    cs.LGstat.MLarXiv:2205.13209v22022
  22. Provable Inductive Matrix Completion

    Prateek Jain, Inderjit S. Dhillon

    cs.LGcs.ITstat.MLarXiv:1306.0626v12013
  23. A Survey on Diffusion Language Models

    Tianyi Li, Mingda Chen, Bowei Guo +1

    cs.CLcs.AIcs.LGarXiv:2508.10875v32025
  24. Hashing with binary autoencoders

    Miguel Á. Carreira-Perpiñán, Ramin Raziperchikolaei

    cs.LGcs.CVmath.OCarXiv:1501.00756v12015
  25. Trajectory Prediction for Autonomous Driving: Progress, Limitations, and Future Directions

    Nadya Abdel Madjid, Abdulrahman Ahmad, Murad Mebrahtu +7

    cs.ROcs.AIcs.CVarXiv:2503.03262v32025
  26. Multi-modal Graph Learning for Disease Prediction

    Shuai Zheng, Zhenfeng Zhu, Zhizhe Liu +4

    cs.LGcs.AIcs.CVarXiv:2203.05880v12022
  27. NAS evaluation is frustratingly hard

    Antoine Yang, Pedro M. Esperança, Fabio M. Carlucci

    cs.LGcs.CVstat.MLarXiv:1912.12522v32019
  28. Timeline: A Dynamic Hierarchical Dirichlet Process Model for Recovering Birth/Death and Evolution of Topics in Text Stream

    Amr Ahmed, Eric P. Xing

    cs.IRcs.LGstat.MLarXiv:1203.3463v12012
  29. Solving Linear Inverse Problems Using GAN Priors: An Algorithm with Provable Guarantees

    Viraj Shah, Chinmay Hegde

    stat.MLcs.LGarXiv:1802.08406v12018
  30. CLadder: Assessing Causal Reasoning in Language Models

    Zhijing Jin, Yuen Chen, Felix Leeb +8

    cs.CLcs.AIcs.LGarXiv:2312.04350v32023
  31. MATCHA: Speeding Up Decentralized SGD via Matching Decomposition Sampling

    Jianyu Wang, Anit Kumar Sahu, Zhouyi Yang +2

    cs.LGeess.SYmath.OCarXiv:1905.09435v32019
  32. Odysseys: Benchmarking Web Agents on Realistic Long Horizon Tasks

    Lawrence Keunho Jang, Jing Yu Koh, Daniel Fried +1

    cs.LGcs.CLarXiv:2604.24964v12026
  33. TimeKAN: KAN-based Frequency Decomposition Learning Architecture for Long-term Time Series Forecasting

    Songtao Huang, Zhen Zhao, Can Li +1

    cs.LGcs.AIarXiv:2502.06910v22025
  34. CoDEx: A Comprehensive Knowledge Graph Completion Benchmark

    Tara Safavi, Danai Koutra

    cs.CLcs.AIcs.IRarXiv:2009.07810v22020
  35. Learning Getting-Up Policies for Real-World Humanoid Robots

    Xialin He, Runpei Dong, Zixuan Chen +1

    cs.ROcs.LGarXiv:2502.12152v22025
  36. Score-Based Causal Discovery of Latent Variable Causal Models

    Ignavier Ng, Xinshuai Dong, Haoyue Dai +3

    cs.LGstat.MLarXiv:2605.20396v12026
  37. The Optimal Sample Complexity of PAC Learning

    Steve Hanneke

    cs.LGstat.MLarXiv:1507.00473v42015
  38. SPEED+: Next-Generation Dataset for Spacecraft Pose Estimation across Domain Gap

    Tae Ha Park, Marcus Märtens, Gurvan Lecuyer +2

    cs.CVcs.LGarXiv:2110.03101v22021
  39. Implicit Graph Neural Networks

    Fangda Gu, Heng Chang, Wenwu Zhu +2

    cs.LGstat.MLarXiv:2009.06211v32020
  40. Neural Transformation Learning for Deep Anomaly Detection Beyond Images

    Chen Qiu, Timo Pfrommer, Marius Kloft +2

    cs.LGcs.AIarXiv:2103.16440v42021
  41. Aligning Language Models from User Interactions

    Thomas Kleine Buening, Jonas Hübotter, Barna Pásztor +3

    cs.CLcs.AIcs.LGarXiv:2603.12273v12026
  42. On the generalization of language models from in-context learning and finetuning: a controlled study

    Andrew K. Lampinen, Arslan Chaudhry, Stephanie C. Y. Chan +7

    cs.CLcs.AIcs.LGarXiv:2505.00661v32025
  43. COOT: Cooperative Hierarchical Transformer for Video-Text Representation Learning

    Simon Ging, Mohammadreza Zolfaghari, Hamed Pirsiavash +1

    cs.CVcs.AIcs.CLarXiv:2011.00597v12020
  44. TransformerFusion: Monocular RGB Scene Reconstruction using Transformers

    Aljaž Božič, Pablo Palafox, Justus Thies +2

    cs.CVcs.GRcs.LGarXiv:2107.02191v12021
  45. Large Language Models for Automated Data Science: Introducing CAAFE for Context-Aware Automated Feature Engineering

    Noah Hollmann, Samuel Müller, Frank Hutter

    cs.AIcs.LGarXiv:2305.03403v52023
  46. ReasonFlux: Hierarchical LLM Reasoning via Scaling Thought Templates

    Ling Yang, Zhaochen Yu, Bin Cui +1

    cs.CLcs.AIcs.LGarXiv:2502.06772v22025
  47. An Open-Source, Event-Driven Pipeline for Cryptocurrency Market Data: Ingestion, Forecasting, and On-Chain Fraud Detection

    Basil Sajid Shaikh, Melrick Mascarenhas, Nuzhat Faiz Shaikh

    cs.AIcs.LGarXiv:2608.29973v12026
  48. Temporal Straightening for Latent Planning

    Ying Wang, Oumayma Bounou, Gaoyue Zhou +4

    cs.LGarXiv:2603.12231v32026
  49. DisPFL: Towards Communication-Efficient Personalized Federated Learning via Decentralized Sparse Training

    Rong Dai, Li Shen, Fengxiang He +2

    cs.LGarXiv:2206.00187v12022
  50. Understanding and Improving Recurrent Networks for Human Activity Recognition by Continuous Attention

    Ming Zeng, Haoxiang Gao, Tong Yu +4

    cs.LGcs.AIstat.MLarXiv:1810.04038v12018
  51. Few-shot Text Classification with Distributional Signatures

    Yujia Bao, Menghua Wu, Shiyu Chang +1

    cs.CLcs.LGarXiv:1908.06039v32019
  52. (Certified!!) Adversarial Robustness for Free!

    Nicholas Carlini, Florian Tramer, Krishnamurthy Dj Dvijotham +3

    cs.LGcs.CRarXiv:2206.10550v22022
  53. A Quantum Variational Approach to Prototypical Recurrent Unit

    Mahyar Sadeghi Garjan, Tommaso Cesari, Michel Barbeau

    cs.LGarXiv:2609.04354v12026
  54. CRPO: A New Approach for Safe Reinforcement Learning with Convergence Guarantee

    Tengyu Xu, Yingbin Liang, Guanghui Lan

    cs.LGstat.MLarXiv:2011.05869v32020
  55. Are Diffusion Models Vulnerable to Membership Inference Attacks?

    Jinhao Duan, Fei Kong, Shiqi Wang +2

    cs.CVcs.AIcs.CRarXiv:2302.01316v22023
  56. Efficient Multi-view Clustering via Unified and Discrete Bipartite Graph Learning

    Si-Guo Fang, Dong Huang, Xiao-Sha Cai +3

    cs.LGcs.AIarXiv:2209.04187v22022
  57. Set You Straight: Auto-Steering Denoising Trajectories to Sidestep Unwanted Concepts

    Leyang Li, Shilin Lu, Yan Ren +1

    cs.CVcs.AIcs.CRarXiv:2504.12782v12025
  58. Machine Learning-Aided Operations and Communications of Unmanned Aerial Vehicles: A Contemporary Survey

    Harrison Kurunathan, Hailong Huang, Kai Li +2

    cs.ROcs.CVcs.LGarXiv:2211.04324v12022
  59. Physics Informed Neural Networks for Control Oriented Thermal Modeling of Buildings

    Gargya Gokhale, Bert Claessens, Chris Develder

    eess.SPcs.LGeess.SYarXiv:2111.12066v22021
  60. DALL-E-Bot: Introducing Web-Scale Diffusion Models to Robotics

    Ivan Kapelyukh, Vitalis Vosylius, Edward Johns

    cs.ROcs.CVcs.LGarXiv:2210.02438v32022