Machine Learning
Papers filed under cs.LG on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.
Search paper metadata (including unsummarized papers)
7,141 to 7,200 of 20,199
Simple Contrastive Graph Clustering
Yue Liu, Xihong Yang, Sihang Zhou +1
cs.LGcs.AIarXiv:2205.07865v32022Flexible Entropy Control in RLVR with a Gradient-Preserving Perspective
Kun Chen, Peng Shi, Fanfan Liu +4
cs.LGcs.AIcs.CLarXiv:2602.09782v22026Trust The Typical
Debargha Ganguly, Sreehari Sankar, Biyao Zhang +8
cs.CLcs.AIcs.DCarXiv:2602.04581v12026Does "AI" stand for augmenting inequality in the era of covid-19 healthcare?
David Leslie, Anjali Mazumder, Aidan Peppin +2
cs.CYcs.LGarXiv:2105.07844v12021Agentic AI: A Comprehensive Survey of Architectures, Applications, and Future Directions
Mohamad Abou Ali, Fadi Dornaika
cs.AIcs.LGarXiv:2510.25445v12025Context Learning for Multi-Agent Discussion
Xingyuan Hua, Sheng Yue, Xinyi Li +3
cs.AIcs.LGcs.MAarXiv:2602.02350v32026Steering LLMs via Scalable Interactive Oversight
Enyu Zhou, Zhiheng Xi, Long Ma +9
cs.AIcs.LGarXiv:2602.04210v22026Alternating Reinforcement Learning for Rubric-Based Reward Modeling in Non-Verifiable LLM Post-Training
Ran Xu, Tianci Liu, Zihan Dong +6
cs.CLcs.LGarXiv:2602.01511v22026Multi-agent Architecture Search via Agentic Supernet
Guibin Zhang, Luyang Niu, Junfeng Fang +3
cs.LGcs.CLcs.MAarXiv:2502.04180v22025FineInstructions: Scaling Synthetic Instructions to Pre-Training Scale
Ajay Patel, Colin Raffel, Chris Callison-Burch
cs.CLcs.LGarXiv:2601.22146v32026InT: Self-Proposed Interventions Enable Credit Assignment in LLM Reasoning
Matthew Y. R. Yang, Hao Bai, Ian Wu +3
cs.LGcs.AIcs.CLarXiv:2601.14209v12026Introduction to Machine Learning
Laurent Younes
stat.MLcs.LGarXiv:2409.02668v22024LLM Agents Grounded in Self-Reports Enable General-Purpose Simulation of Individuals
Joon Sung Park, Carolyn Q. Zou, Jonne Kamphorst +8
cs.AIcs.HCcs.LGarXiv:2411.10109v32024Taming Visually Guided Sound Generation
Vladimir Iashin, Esa Rahtu
cs.CVcs.AIcs.LGarXiv:2110.08791v12021From PINNs to PIKANs: Recent Advances in Physics-Informed Machine Learning
Juan Diego Toscano, Vivek Oommen, Alan John Varghese +4
cs.LGcs.AIphysics.comp-pharXiv:2410.13228v22024A Survey on Diffusion Models for Inverse Problems
Giannis Daras, Hyungjin Chung, Chieh-Hsin Lai +5
cs.LGcs.AIcs.CVarXiv:2410.00083v12024Simple and Effective Masked Diffusion Language Models
Subham Sekhar Sahoo, Marianne Arriola, Yair Schiff +5
cs.CLcs.AIcs.LGarXiv:2406.07524v22024RewardBench: Evaluating Reward Models for Language Modeling
Nathan Lambert, Valentina Pyatkin, Jacob Morrison +9
cs.LGarXiv:2403.13787v22024Keeping it Simple: Language Models can learn Complex Molecular Distributions
Daniel Flam-Shepherd, Kevin Zhu, Alán Aspuru-Guzik
cs.LGcs.AIq-bio.QMarXiv:2112.03041v12021Momentum Improves Normalized SGD
Ashok Cutkosky, Harsh Mehta
cs.LGmath.OCstat.MLarXiv:2002.03305v22020Hiding Among the Clones: A Simple and Nearly Optimal Analysis of Privacy Amplification by Shuffling
Vitaly Feldman, Audra McMillan, Kunal Talwar
cs.LGcs.CRcs.DSarXiv:2012.12803v32020Empowering Edge Intelligence: A Comprehensive Survey on On-Device AI Models
Xubin Wang, Zhiqing Tang, Jianxiong Guo +4
cs.AIcs.LGcs.NIarXiv:2503.06027v22025Photorealistic Video Generation with Diffusion Models
Agrim Gupta, Lijun Yu, Kihyuk Sohn +6
cs.CVcs.AIcs.LGarXiv:2312.06662v12023Generative Adversarial Networks: A Survey Towards Private and Secure Applications
Zhipeng Cai, Zuobin Xiong, Honghui Xu +3
cs.LGcs.CRarXiv:2106.03785v12021Towards Understanding Sycophancy in Language Models
Mrinank Sharma, Meg Tong, Tomasz Korbak +16
cs.CLcs.AIcs.LGarXiv:2310.13548v42023Multivariate Time Series Forecasting with Dynamic Graph Neural ODEs
Ming Jin, Yu Zheng, Yuan-Fang Li +3
cs.LGarXiv:2202.08408v22022T2I-R1: Reinforcing Image Generation with Collaborative Semantic-level and Token-level CoT
Dongzhi Jiang, Ziyu Guo, Renrui Zhang +6
cs.CVcs.AIcs.CLarXiv:2505.00703v22025Deep Multiagent Reinforcement Learning: Challenges and Directions
Annie Wong, Thomas Bäck, Anna V. Kononova +1
cs.LGcs.AIcs.MAarXiv:2106.15691v22021Identifying the Risks of LM Agents with an LM-Emulated Sandbox
Yangjun Ruan, Honghua Dong, Andrew Wang +6
cs.AIcs.CLcs.LGarXiv:2309.15817v22023Consistency Trajectory Models: Learning Probability Flow ODE Trajectory of Diffusion
Dongjun Kim, Chieh-Hsin Lai, Wei-Hsiang Liao +6
cs.LGcs.AIcs.CVarXiv:2310.02279v32023Consciousness in Artificial Intelligence: Insights from the Science of Consciousness
Patrick Butlin, Robert Long, Eric Elmoznino +16
cs.AIcs.CYcs.LGarXiv:2308.08708v32023Self-Consuming Generative Models Go MAD
Sina Alemohammad, Josue Casco-Rodriguez, Lorenzo Luzi +5
cs.LGcs.AIcs.CVarXiv:2307.01850v12023Stochastic Interpolants: A Unifying Framework for Flows and Diffusions
Michael S. Albergo, Nicholas M. Boffi, Eric Vanden-Eijnden
cs.LGcond-mat.dis-nnmath.PRarXiv:2303.08797v42023Consistency Models
Yang Song, Prafulla Dhariwal, Mark Chen +1
cs.LGcs.CVstat.MLarXiv:2303.01469v22023How to DP-fy ML: A Practical Guide to Machine Learning with Differential Privacy
Natalia Ponomareva, Hussein Hazimeh, Alex Kurakin +6
cs.LGcs.CRstat.MLarXiv:2303.00654v32023ChatGPT Makes Medicine Easy to Swallow: An Exploratory Case Study on Simplified Radiology Reports
Katharina Jeblick, Balthasar Schachtner, Jakob Dexl +8
cs.CLcs.LGarXiv:2212.14882v12022Dataless Knowledge Fusion by Merging Weights of Language Models
Xisen Jin, Xiang Ren, Daniel Preotiuc-Pietro +1
cs.CLcs.LGarXiv:2212.09849v62022Explainable Artificial Intelligence (XAI) from a user perspective- A synthesis of prior literature and problematizing avenues for future research
AKM Bahalul Haque, A. K. M. Najmul Islam, Patrick Mikalef
cs.AIcs.LGarXiv:2211.15343v12022ReAct: Synergizing Reasoning and Acting in Language Models
Shunyu Yao, Jeffrey Zhao, Dian Yu +4
cs.CLcs.AIcs.LGarXiv:2210.03629v32022Pass@k Training for Adaptively Balancing Exploration and Exploitation of Large Reasoning Models
Zhipeng Chen, Xiaobo Qin, Youbin Wu +4
cs.LGcs.AIcs.CLarXiv:2508.10751v12025Diffusion Models: A Comprehensive Survey of Methods and Applications
Ling Yang, Zhilong Zhang, Yang Song +6
cs.LGcs.AIcs.CVarXiv:2209.00796v152022Music Source Separation with Band-split RNN
Yi Luo, Jianwei Yu
eess.AScs.LGcs.SDarXiv:2209.15174v12022A Comprehensive Review of Digital Twin -- Part 1: Modeling and Twinning Enabling Technologies
Adam Thelen, Xiaoge Zhang, Olga Fink +7
cs.CEcs.AIcs.LGarXiv:2208.14197v22022Graphs, Convolutions, and Neural Networks: From Graph Filters to Graph Neural Networks
Fernando Gama, Elvin Isufi, Geert Leus +1
cs.LGeess.SYstat.MLarXiv:2003.03777v52020State of the Art on Diffusion Models for Visual Computing
Ryan Po, Wang Yifan, Vladislav Golyanik +15
cs.AIcs.CVcs.GRarXiv:2310.07204v12023Streaming 4D Visual Geometry Transformer
Dong Zhuo, Wenzhao Zheng, Jiahe Guo +3
cs.CVcs.AIcs.LGarXiv:2507.11539v22025IBM Federated Learning: an Enterprise Framework White Paper V0.1
Heiko Ludwig, Nathalie Baracaldo, Gegi Thomas +21
cs.LGcs.CRcs.DCarXiv:2007.10987v12020Emergent Misalignment: Narrow finetuning can produce broadly misaligned LLMs
Jan Betley, Daniel Tan, Niels Warncke +5
cs.CLcs.AIcs.CRarXiv:2502.17424v72025Diffusion-Based Planning for Autonomous Driving with Flexible Guidance
Yinan Zheng, Ruiming Liang, Kexin Zheng +8
cs.ROcs.AIcs.LGarXiv:2501.15564v22025Personalized Speech recognition on mobile devices
Ian McGraw, Rohit Prabhavalkar, Raziel Alvarez +8
cs.CLcs.LGcs.SDarXiv:1603.03185v22016Large Language Models are Competitive Near Cold-start Recommenders for Language- and Item-based Preferences
Scott Sanner, Krisztian Balog, Filip Radlinski +2
cs.IRcs.LGarXiv:2307.14225v12023Latent-Space No-Arbitrage Geometry of Generative Models for Implied Volatility Surfaces
Jing Wang, Shuaiqiang Liu, Cornelis Vuik
q-fin.CPcs.AIcs.LGarXiv:2609.00332v12026Equiformer: Equivariant Graph Attention Transformer for 3D Atomistic Graphs
Yi-Lun Liao, Tess Smidt
cs.LGcs.AIphysics.comp-pharXiv:2206.11990v22022GLIPv2: Unifying Localization and Vision-Language Understanding
Haotian Zhang, Pengchuan Zhang, Xiaowei Hu +7
cs.CVcs.AIcs.CLarXiv:2206.05836v22022Optimizing Quantum Error Correction Codes with Reinforcement Learning
Hendrik Poulsen Nautrup, Nicolas Delfosse, Vedran Dunjko +2
quant-phcs.AIcs.LGarXiv:1812.08451v52018FlashAttention: Fast and Memory-Efficient Exact Attention with IO-Awareness
Tri Dao, Daniel Y. Fu, Stefano Ermon +2
cs.LGarXiv:2205.14135v22022Summaries:한국어ThinkAct: Vision-Language-Action Reasoning via Reinforced Visual Latent Planning
Chi-Pin Huang, Yueh-Hua Wu, Min-Hung Chen +2
cs.CVcs.AIcs.LGarXiv:2507.16815v22025Deep Multi-modal Fusion of Image and Non-image Data in Disease Diagnosis and Prognosis: A Review
Can Cui, Haichun Yang, Yaohong Wang +6
cs.LGcs.AIcs.CVarXiv:2203.15588v32022TWIST: Teleoperated Whole-Body Imitation System
Yanjie Ze, Zixuan Chen, João Pedro Araújo +4
cs.ROcs.CVcs.LGarXiv:2505.02833v12025FedDC: Federated Learning with Non-IID Data via Local Drift Decoupling and Correction
Liang Gao, Huazhu Fu, Li Li +3
cs.LGcs.AIarXiv:2203.11751v12022