Machine Learning

Papers filed under cs.LG on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.

Search paper metadata (including unsummarized papers)

7,141 to 7,200 of 20,199

  1. Simple Contrastive Graph Clustering

    Yue Liu, Xihong Yang, Sihang Zhou +1

    cs.LGcs.AIarXiv:2205.07865v32022
  2. Flexible Entropy Control in RLVR with a Gradient-Preserving Perspective

    Kun Chen, Peng Shi, Fanfan Liu +4

    cs.LGcs.AIcs.CLarXiv:2602.09782v22026
  3. Trust The Typical

    Debargha Ganguly, Sreehari Sankar, Biyao Zhang +8

    cs.CLcs.AIcs.DCarXiv:2602.04581v12026
  4. Does "AI" stand for augmenting inequality in the era of covid-19 healthcare?

    David Leslie, Anjali Mazumder, Aidan Peppin +2

    cs.CYcs.LGarXiv:2105.07844v12021
  5. Agentic AI: A Comprehensive Survey of Architectures, Applications, and Future Directions

    Mohamad Abou Ali, Fadi Dornaika

    cs.AIcs.LGarXiv:2510.25445v12025
  6. Context Learning for Multi-Agent Discussion

    Xingyuan Hua, Sheng Yue, Xinyi Li +3

    cs.AIcs.LGcs.MAarXiv:2602.02350v32026
  7. Steering LLMs via Scalable Interactive Oversight

    Enyu Zhou, Zhiheng Xi, Long Ma +9

    cs.AIcs.LGarXiv:2602.04210v22026
  8. Alternating Reinforcement Learning for Rubric-Based Reward Modeling in Non-Verifiable LLM Post-Training

    Ran Xu, Tianci Liu, Zihan Dong +6

    cs.CLcs.LGarXiv:2602.01511v22026
  9. Multi-agent Architecture Search via Agentic Supernet

    Guibin Zhang, Luyang Niu, Junfeng Fang +3

    cs.LGcs.CLcs.MAarXiv:2502.04180v22025
  10. FineInstructions: Scaling Synthetic Instructions to Pre-Training Scale

    Ajay Patel, Colin Raffel, Chris Callison-Burch

    cs.CLcs.LGarXiv:2601.22146v32026
  11. InT: Self-Proposed Interventions Enable Credit Assignment in LLM Reasoning

    Matthew Y. R. Yang, Hao Bai, Ian Wu +3

    cs.LGcs.AIcs.CLarXiv:2601.14209v12026
  12. Introduction to Machine Learning

    Laurent Younes

    stat.MLcs.LGarXiv:2409.02668v22024
  13. LLM Agents Grounded in Self-Reports Enable General-Purpose Simulation of Individuals

    Joon Sung Park, Carolyn Q. Zou, Jonne Kamphorst +8

    cs.AIcs.HCcs.LGarXiv:2411.10109v32024
  14. Taming Visually Guided Sound Generation

    Vladimir Iashin, Esa Rahtu

    cs.CVcs.AIcs.LGarXiv:2110.08791v12021
  15. From PINNs to PIKANs: Recent Advances in Physics-Informed Machine Learning

    Juan Diego Toscano, Vivek Oommen, Alan John Varghese +4

    cs.LGcs.AIphysics.comp-pharXiv:2410.13228v22024
  16. A Survey on Diffusion Models for Inverse Problems

    Giannis Daras, Hyungjin Chung, Chieh-Hsin Lai +5

    cs.LGcs.AIcs.CVarXiv:2410.00083v12024
  17. Simple and Effective Masked Diffusion Language Models

    Subham Sekhar Sahoo, Marianne Arriola, Yair Schiff +5

    cs.CLcs.AIcs.LGarXiv:2406.07524v22024
  18. RewardBench: Evaluating Reward Models for Language Modeling

    Nathan Lambert, Valentina Pyatkin, Jacob Morrison +9

    cs.LGarXiv:2403.13787v22024
  19. Keeping it Simple: Language Models can learn Complex Molecular Distributions

    Daniel Flam-Shepherd, Kevin Zhu, Alán Aspuru-Guzik

    cs.LGcs.AIq-bio.QMarXiv:2112.03041v12021
  20. Momentum Improves Normalized SGD

    Ashok Cutkosky, Harsh Mehta

    cs.LGmath.OCstat.MLarXiv:2002.03305v22020
  21. Hiding Among the Clones: A Simple and Nearly Optimal Analysis of Privacy Amplification by Shuffling

    Vitaly Feldman, Audra McMillan, Kunal Talwar

    cs.LGcs.CRcs.DSarXiv:2012.12803v32020
  22. Empowering Edge Intelligence: A Comprehensive Survey on On-Device AI Models

    Xubin Wang, Zhiqing Tang, Jianxiong Guo +4

    cs.AIcs.LGcs.NIarXiv:2503.06027v22025
  23. Photorealistic Video Generation with Diffusion Models

    Agrim Gupta, Lijun Yu, Kihyuk Sohn +6

    cs.CVcs.AIcs.LGarXiv:2312.06662v12023
  24. Generative Adversarial Networks: A Survey Towards Private and Secure Applications

    Zhipeng Cai, Zuobin Xiong, Honghui Xu +3

    cs.LGcs.CRarXiv:2106.03785v12021
  25. Towards Understanding Sycophancy in Language Models

    Mrinank Sharma, Meg Tong, Tomasz Korbak +16

    cs.CLcs.AIcs.LGarXiv:2310.13548v42023
  26. Multivariate Time Series Forecasting with Dynamic Graph Neural ODEs

    Ming Jin, Yu Zheng, Yuan-Fang Li +3

    cs.LGarXiv:2202.08408v22022
  27. T2I-R1: Reinforcing Image Generation with Collaborative Semantic-level and Token-level CoT

    Dongzhi Jiang, Ziyu Guo, Renrui Zhang +6

    cs.CVcs.AIcs.CLarXiv:2505.00703v22025
  28. Deep Multiagent Reinforcement Learning: Challenges and Directions

    Annie Wong, Thomas Bäck, Anna V. Kononova +1

    cs.LGcs.AIcs.MAarXiv:2106.15691v22021
  29. Identifying the Risks of LM Agents with an LM-Emulated Sandbox

    Yangjun Ruan, Honghua Dong, Andrew Wang +6

    cs.AIcs.CLcs.LGarXiv:2309.15817v22023
  30. Consistency Trajectory Models: Learning Probability Flow ODE Trajectory of Diffusion

    Dongjun Kim, Chieh-Hsin Lai, Wei-Hsiang Liao +6

    cs.LGcs.AIcs.CVarXiv:2310.02279v32023
  31. Consciousness in Artificial Intelligence: Insights from the Science of Consciousness

    Patrick Butlin, Robert Long, Eric Elmoznino +16

    cs.AIcs.CYcs.LGarXiv:2308.08708v32023
  32. Self-Consuming Generative Models Go MAD

    Sina Alemohammad, Josue Casco-Rodriguez, Lorenzo Luzi +5

    cs.LGcs.AIcs.CVarXiv:2307.01850v12023
  33. Stochastic Interpolants: A Unifying Framework for Flows and Diffusions

    Michael S. Albergo, Nicholas M. Boffi, Eric Vanden-Eijnden

    cs.LGcond-mat.dis-nnmath.PRarXiv:2303.08797v42023
  34. Consistency Models

    Yang Song, Prafulla Dhariwal, Mark Chen +1

    cs.LGcs.CVstat.MLarXiv:2303.01469v22023
  35. How to DP-fy ML: A Practical Guide to Machine Learning with Differential Privacy

    Natalia Ponomareva, Hussein Hazimeh, Alex Kurakin +6

    cs.LGcs.CRstat.MLarXiv:2303.00654v32023
  36. ChatGPT Makes Medicine Easy to Swallow: An Exploratory Case Study on Simplified Radiology Reports

    Katharina Jeblick, Balthasar Schachtner, Jakob Dexl +8

    cs.CLcs.LGarXiv:2212.14882v12022
  37. Dataless Knowledge Fusion by Merging Weights of Language Models

    Xisen Jin, Xiang Ren, Daniel Preotiuc-Pietro +1

    cs.CLcs.LGarXiv:2212.09849v62022
  38. Explainable Artificial Intelligence (XAI) from a user perspective- A synthesis of prior literature and problematizing avenues for future research

    AKM Bahalul Haque, A. K. M. Najmul Islam, Patrick Mikalef

    cs.AIcs.LGarXiv:2211.15343v12022
  39. ReAct: Synergizing Reasoning and Acting in Language Models

    Shunyu Yao, Jeffrey Zhao, Dian Yu +4

    cs.CLcs.AIcs.LGarXiv:2210.03629v32022
  40. Pass@k Training for Adaptively Balancing Exploration and Exploitation of Large Reasoning Models

    Zhipeng Chen, Xiaobo Qin, Youbin Wu +4

    cs.LGcs.AIcs.CLarXiv:2508.10751v12025
  41. Diffusion Models: A Comprehensive Survey of Methods and Applications

    Ling Yang, Zhilong Zhang, Yang Song +6

    cs.LGcs.AIcs.CVarXiv:2209.00796v152022
  42. Music Source Separation with Band-split RNN

    Yi Luo, Jianwei Yu

    eess.AScs.LGcs.SDarXiv:2209.15174v12022
  43. A Comprehensive Review of Digital Twin -- Part 1: Modeling and Twinning Enabling Technologies

    Adam Thelen, Xiaoge Zhang, Olga Fink +7

    cs.CEcs.AIcs.LGarXiv:2208.14197v22022
  44. Graphs, Convolutions, and Neural Networks: From Graph Filters to Graph Neural Networks

    Fernando Gama, Elvin Isufi, Geert Leus +1

    cs.LGeess.SYstat.MLarXiv:2003.03777v52020
  45. State of the Art on Diffusion Models for Visual Computing

    Ryan Po, Wang Yifan, Vladislav Golyanik +15

    cs.AIcs.CVcs.GRarXiv:2310.07204v12023
  46. Streaming 4D Visual Geometry Transformer

    Dong Zhuo, Wenzhao Zheng, Jiahe Guo +3

    cs.CVcs.AIcs.LGarXiv:2507.11539v22025
  47. IBM Federated Learning: an Enterprise Framework White Paper V0.1

    Heiko Ludwig, Nathalie Baracaldo, Gegi Thomas +21

    cs.LGcs.CRcs.DCarXiv:2007.10987v12020
  48. Emergent Misalignment: Narrow finetuning can produce broadly misaligned LLMs

    Jan Betley, Daniel Tan, Niels Warncke +5

    cs.CLcs.AIcs.CRarXiv:2502.17424v72025
  49. Diffusion-Based Planning for Autonomous Driving with Flexible Guidance

    Yinan Zheng, Ruiming Liang, Kexin Zheng +8

    cs.ROcs.AIcs.LGarXiv:2501.15564v22025
  50. Personalized Speech recognition on mobile devices

    Ian McGraw, Rohit Prabhavalkar, Raziel Alvarez +8

    cs.CLcs.LGcs.SDarXiv:1603.03185v22016
  51. Large Language Models are Competitive Near Cold-start Recommenders for Language- and Item-based Preferences

    Scott Sanner, Krisztian Balog, Filip Radlinski +2

    cs.IRcs.LGarXiv:2307.14225v12023
  52. Latent-Space No-Arbitrage Geometry of Generative Models for Implied Volatility Surfaces

    Jing Wang, Shuaiqiang Liu, Cornelis Vuik

    q-fin.CPcs.AIcs.LGarXiv:2609.00332v12026
  53. Equiformer: Equivariant Graph Attention Transformer for 3D Atomistic Graphs

    Yi-Lun Liao, Tess Smidt

    cs.LGcs.AIphysics.comp-pharXiv:2206.11990v22022
  54. GLIPv2: Unifying Localization and Vision-Language Understanding

    Haotian Zhang, Pengchuan Zhang, Xiaowei Hu +7

    cs.CVcs.AIcs.CLarXiv:2206.05836v22022
  55. Optimizing Quantum Error Correction Codes with Reinforcement Learning

    Hendrik Poulsen Nautrup, Nicolas Delfosse, Vedran Dunjko +2

    quant-phcs.AIcs.LGarXiv:1812.08451v52018
  56. FlashAttention: Fast and Memory-Efficient Exact Attention with IO-Awareness

    Tri Dao, Daniel Y. Fu, Stefano Ermon +2

    cs.LGarXiv:2205.14135v22022
    Summaries:한국어
  57. ThinkAct: Vision-Language-Action Reasoning via Reinforced Visual Latent Planning

    Chi-Pin Huang, Yueh-Hua Wu, Min-Hung Chen +2

    cs.CVcs.AIcs.LGarXiv:2507.16815v22025
  58. Deep Multi-modal Fusion of Image and Non-image Data in Disease Diagnosis and Prognosis: A Review

    Can Cui, Haichun Yang, Yaohong Wang +6

    cs.LGcs.AIcs.CVarXiv:2203.15588v32022
  59. TWIST: Teleoperated Whole-Body Imitation System

    Yanjie Ze, Zixuan Chen, João Pedro Araújo +4

    cs.ROcs.CVcs.LGarXiv:2505.02833v12025
  60. FedDC: Federated Learning with Non-IID Data via Local Drift Decoupling and Correction

    Liang Gao, Huazhu Fu, Li Li +3

    cs.LGcs.AIarXiv:2203.11751v12022