Machine Learning
Papers filed under cs.LG on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.
Search paper metadata (including unsummarized papers)
4,921 to 4,980 of 20,223
Unsafer in Many Turns: Benchmarking and Defending Multi-Turn Safety Risks in Tool-Using Agents
Xu Li, Simon Yu, Minzhou Pan +5
cs.CRcs.AIcs.CLarXiv:2602.13379v22026A Comprehensive Review of Multi-Agent Reinforcement Learning in Video Games
Zhengyang Li, Qijin Ji, Xinghong Ling +1
cs.LGarXiv:2509.03682v12025Vision Mamba: A Comprehensive Survey and Taxonomy
Xiao Liu, Chenxu Zhang, Lei Zhang
cs.CVcs.AIcs.CLarXiv:2405.04404v12024OCTID: Optical Coherence Tomography Image Database
Peyman Gholami, Priyanka Roy, Mohana Kuppuswamy Parthasarathy +1
cs.CVcs.LGarXiv:1812.07056v22018Audiobox: Unified Audio Generation with Natural Language Prompts
Apoorv Vyas, Bowen Shi, Matthew Le +21
cs.SDcs.LGeess.ASarXiv:2312.15821v12023Free-rider Attacks on Model Aggregation in Federated Learning
Yann Fraboni, Richard Vidal, Marco Lorenzi
cs.LGstat.MLarXiv:2006.11901v52020Out of the BLEU: how should we assess quality of the Code Generation models?
Mikhail Evtikhiev, Egor Bogomolov, Yaroslav Sokolov +1
cs.SEcs.LGarXiv:2208.03133v22022QuestBench: Can LLMs ask the right question to acquire information in reasoning tasks?
Belinda Z. Li, Been Kim, Zi Wang
cs.AIcs.CLcs.LGarXiv:2503.22674v22025DIAL: Decoupling Intent and Action via Latent World Modeling for End-to-End VLA
Yi Chen, Yuying Ge, Hui Zhou +3
cs.ROcs.AIcs.CVarXiv:2603.29844v22026A Mechanistic Analysis of Looped Reasoning Language Models
Hugh Blayney, Álvaro Arroyo, Johan Obando-Ceron +4
cs.LGcs.AIarXiv:2604.11791v12026UAV-VLA: Vision-Language-Action System for Large Scale Aerial Mission Generation
Oleg Sautenkov, Yasheerah Yaqoot, Artem Lykov +7
cs.ROcs.AIcs.CVarXiv:2501.05014v22025Naive Bayes and Text Classification I - Introduction and Theory
Sebastian Raschka
cs.LGarXiv:1410.5329v42014Federated Learning for Cyber Physical Systems: A Comprehensive Survey
Minh K. Quan, Pubudu N. Pathirana, Mayuri Wijayasundara +5
cs.LGcs.AIcs.CRarXiv:2505.04873v12025Fast Task Inference with Variational Intrinsic Successor Features
Steven Hansen, Will Dabney, Andre Barreto +3
cs.LGcs.AIstat.MLarXiv:1906.05030v22019Exploit Bounding Box Annotations for Multi-label Object Recognition
Hao Yang, Joey Tianyi Zhou, Yu Zhang +3
cs.CVcs.LGarXiv:1504.05843v22015Extremely Fast Decision Tree
Chaitanya Manapragada, Geoff Webb, Mahsa Salehi
cs.LGstat.MLarXiv:1802.08780v12018Wasserstein Learning of Deep Generative Point Process Models
Shuai Xiao, Mehrdad Farajtabar, Xiaojing Ye +3
cs.LGstat.MLarXiv:1705.08051v12017Code to Think, Think to Code: A Survey on Code-Enhanced Reasoning and Reasoning-Driven Code Intelligence in LLMs
Dayu Yang, Tianyang Liu, Daoan Zhang +8
cs.CLcs.AIcs.LGarXiv:2502.19411v12025Learning multiview 3D point cloud registration
Zan Gojcic, Caifa Zhou, Jan D. Wegner +2
cs.CVcs.LGarXiv:2001.05119v22020Trace and Pace: Controllable Pedestrian Animation via Guided Trajectory Diffusion
Davis Rempe, Zhengyi Luo, Xue Bin Peng +5
cs.CVcs.GRcs.LGarXiv:2304.01893v12023Sym-NCO: Leveraging Symmetricity for Neural Combinatorial Optimization
Minsu Kim, Junyoung Park, Jinkyoo Park
cs.LGstat.MLarXiv:2205.13209v22022Provable Inductive Matrix Completion
Prateek Jain, Inderjit S. Dhillon
cs.LGcs.ITstat.MLarXiv:1306.0626v12013A Survey on Diffusion Language Models
Tianyi Li, Mingda Chen, Bowei Guo +1
cs.CLcs.AIcs.LGarXiv:2508.10875v32025Hashing with binary autoencoders
Miguel Á. Carreira-Perpiñán, Ramin Raziperchikolaei
cs.LGcs.CVmath.OCarXiv:1501.00756v12015Trajectory Prediction for Autonomous Driving: Progress, Limitations, and Future Directions
Nadya Abdel Madjid, Abdulrahman Ahmad, Murad Mebrahtu +7
cs.ROcs.AIcs.CVarXiv:2503.03262v32025Multi-modal Graph Learning for Disease Prediction
Shuai Zheng, Zhenfeng Zhu, Zhizhe Liu +4
cs.LGcs.AIcs.CVarXiv:2203.05880v12022NAS evaluation is frustratingly hard
Antoine Yang, Pedro M. Esperança, Fabio M. Carlucci
cs.LGcs.CVstat.MLarXiv:1912.12522v32019Timeline: A Dynamic Hierarchical Dirichlet Process Model for Recovering Birth/Death and Evolution of Topics in Text Stream
Amr Ahmed, Eric P. Xing
cs.IRcs.LGstat.MLarXiv:1203.3463v12012Solving Linear Inverse Problems Using GAN Priors: An Algorithm with Provable Guarantees
Viraj Shah, Chinmay Hegde
stat.MLcs.LGarXiv:1802.08406v12018CLadder: Assessing Causal Reasoning in Language Models
Zhijing Jin, Yuen Chen, Felix Leeb +8
cs.CLcs.AIcs.LGarXiv:2312.04350v32023MATCHA: Speeding Up Decentralized SGD via Matching Decomposition Sampling
Jianyu Wang, Anit Kumar Sahu, Zhouyi Yang +2
cs.LGeess.SYmath.OCarXiv:1905.09435v32019Odysseys: Benchmarking Web Agents on Realistic Long Horizon Tasks
Lawrence Keunho Jang, Jing Yu Koh, Daniel Fried +1
cs.LGcs.CLarXiv:2604.24964v12026TimeKAN: KAN-based Frequency Decomposition Learning Architecture for Long-term Time Series Forecasting
Songtao Huang, Zhen Zhao, Can Li +1
cs.LGcs.AIarXiv:2502.06910v22025CoDEx: A Comprehensive Knowledge Graph Completion Benchmark
Tara Safavi, Danai Koutra
cs.CLcs.AIcs.IRarXiv:2009.07810v22020Learning Getting-Up Policies for Real-World Humanoid Robots
Xialin He, Runpei Dong, Zixuan Chen +1
cs.ROcs.LGarXiv:2502.12152v22025Score-Based Causal Discovery of Latent Variable Causal Models
Ignavier Ng, Xinshuai Dong, Haoyue Dai +3
cs.LGstat.MLarXiv:2605.20396v12026The Optimal Sample Complexity of PAC Learning
Steve Hanneke
cs.LGstat.MLarXiv:1507.00473v42015SPEED+: Next-Generation Dataset for Spacecraft Pose Estimation across Domain Gap
Tae Ha Park, Marcus Märtens, Gurvan Lecuyer +2
cs.CVcs.LGarXiv:2110.03101v22021Implicit Graph Neural Networks
Fangda Gu, Heng Chang, Wenwu Zhu +2
cs.LGstat.MLarXiv:2009.06211v32020Neural Transformation Learning for Deep Anomaly Detection Beyond Images
Chen Qiu, Timo Pfrommer, Marius Kloft +2
cs.LGcs.AIarXiv:2103.16440v42021Aligning Language Models from User Interactions
Thomas Kleine Buening, Jonas Hübotter, Barna Pásztor +3
cs.CLcs.AIcs.LGarXiv:2603.12273v12026On the generalization of language models from in-context learning and finetuning: a controlled study
Andrew K. Lampinen, Arslan Chaudhry, Stephanie C. Y. Chan +7
cs.CLcs.AIcs.LGarXiv:2505.00661v32025COOT: Cooperative Hierarchical Transformer for Video-Text Representation Learning
Simon Ging, Mohammadreza Zolfaghari, Hamed Pirsiavash +1
cs.CVcs.AIcs.CLarXiv:2011.00597v12020TransformerFusion: Monocular RGB Scene Reconstruction using Transformers
Aljaž Božič, Pablo Palafox, Justus Thies +2
cs.CVcs.GRcs.LGarXiv:2107.02191v12021Large Language Models for Automated Data Science: Introducing CAAFE for Context-Aware Automated Feature Engineering
Noah Hollmann, Samuel Müller, Frank Hutter
cs.AIcs.LGarXiv:2305.03403v52023ReasonFlux: Hierarchical LLM Reasoning via Scaling Thought Templates
Ling Yang, Zhaochen Yu, Bin Cui +1
cs.CLcs.AIcs.LGarXiv:2502.06772v22025An Open-Source, Event-Driven Pipeline for Cryptocurrency Market Data: Ingestion, Forecasting, and On-Chain Fraud Detection
Basil Sajid Shaikh, Melrick Mascarenhas, Nuzhat Faiz Shaikh
cs.AIcs.LGarXiv:2608.29973v12026Temporal Straightening for Latent Planning
Ying Wang, Oumayma Bounou, Gaoyue Zhou +4
cs.LGarXiv:2603.12231v32026DisPFL: Towards Communication-Efficient Personalized Federated Learning via Decentralized Sparse Training
Rong Dai, Li Shen, Fengxiang He +2
cs.LGarXiv:2206.00187v12022Understanding and Improving Recurrent Networks for Human Activity Recognition by Continuous Attention
Ming Zeng, Haoxiang Gao, Tong Yu +4
cs.LGcs.AIstat.MLarXiv:1810.04038v12018Few-shot Text Classification with Distributional Signatures
Yujia Bao, Menghua Wu, Shiyu Chang +1
cs.CLcs.LGarXiv:1908.06039v32019(Certified!!) Adversarial Robustness for Free!
Nicholas Carlini, Florian Tramer, Krishnamurthy Dj Dvijotham +3
cs.LGcs.CRarXiv:2206.10550v22022A Quantum Variational Approach to Prototypical Recurrent Unit
Mahyar Sadeghi Garjan, Tommaso Cesari, Michel Barbeau
cs.LGarXiv:2609.04354v12026CRPO: A New Approach for Safe Reinforcement Learning with Convergence Guarantee
Tengyu Xu, Yingbin Liang, Guanghui Lan
cs.LGstat.MLarXiv:2011.05869v32020Are Diffusion Models Vulnerable to Membership Inference Attacks?
Jinhao Duan, Fei Kong, Shiqi Wang +2
cs.CVcs.AIcs.CRarXiv:2302.01316v22023Efficient Multi-view Clustering via Unified and Discrete Bipartite Graph Learning
Si-Guo Fang, Dong Huang, Xiao-Sha Cai +3
cs.LGcs.AIarXiv:2209.04187v22022Set You Straight: Auto-Steering Denoising Trajectories to Sidestep Unwanted Concepts
Leyang Li, Shilin Lu, Yan Ren +1
cs.CVcs.AIcs.CRarXiv:2504.12782v12025Machine Learning-Aided Operations and Communications of Unmanned Aerial Vehicles: A Contemporary Survey
Harrison Kurunathan, Hailong Huang, Kai Li +2
cs.ROcs.CVcs.LGarXiv:2211.04324v12022Physics Informed Neural Networks for Control Oriented Thermal Modeling of Buildings
Gargya Gokhale, Bert Claessens, Chris Develder
eess.SPcs.LGeess.SYarXiv:2111.12066v22021DALL-E-Bot: Introducing Web-Scale Diffusion Models to Robotics
Ivan Kapelyukh, Vitalis Vosylius, Edward Johns
cs.ROcs.CVcs.LGarXiv:2210.02438v32022