Artificial Intelligence
Papers filed under cs.AI on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.
Search paper metadata (including unsummarized papers)
3,781 to 3,840 of 15,222
LiveMathematicianBench: A Live Benchmark for Mathematician-Level Reasoning with Proof Sketches
Linyang He, Qiyao Yu, Hanze Dong +5
cs.CLcs.AIcs.LGarXiv:2604.01754v12026Federated Learning for Healthcare Domain - Pipeline, Applications and Challenges
Madhura Joshi, Ankit Pal, Malaikannan Sankarasubbu
cs.LGcs.AIcs.CRarXiv:2211.07893v22022Controlled LLM Training on Spectral Sphere
Tian Xie, Haoming Luo, Haoyu Tang +9
cs.LGcs.AIarXiv:2601.08393v32026Cold-Start Recommendation towards the Era of Large Language Models (LLMs): A Comprehensive Survey and Roadmap
Weizhi Zhang, Yuanchen Bei, Liangwei Yang +15
cs.IRcs.AIarXiv:2501.01945v22025trajectory-judge: What Outcome-Only LLM Judges Miss on Agent Trajectories
Hadi Mohammadi
cs.CLcs.AIcs.SEarXiv:2609.00038v12026Towards Agentic RAG with Deep Reasoning: A Survey of RAG-Reasoning Systems in LLMs
Yangning Li, Weizhi Zhang, Yuyao Yang +17
cs.CLcs.AIarXiv:2507.09477v22025A Survey on LoRA of Large Language Models
Yuren Mao, Yuhang Ge, Yijiang Fan +4
cs.LGcs.AIcs.CLarXiv:2407.11046v42024Cross-Relational Preference Learning for Better LLM Instruction Following
Runsheng Li, Kai Sun, Bin Shi +1
cs.AIarXiv:2608.29352v12026Learning Combinatorial Optimization on Graphs: A Survey with Applications to Networking
Natalia Vesselinova, Rebecca Steinert, Daniel F. Perez-Ramirez +1
cs.LGcs.AIstat.MLarXiv:2005.11081v22020Security Concerns for Large Language Models: A Survey
Miles Q. Li, Benjamin C. M. Fung
cs.CRcs.AIarXiv:2505.18889v52025Don't Trust ChatGPT when Your Question is not in English: A Study of Multilingual Abilities and Types of LLMs
Xiang Zhang, Senyu Li, Bradley Hauer +2
cs.CLcs.AIarXiv:2305.16339v22023RoboClaw: An Agentic Framework for Scalable Long-Horizon Robotic Tasks
Ruiying Li, Yunlang Zhou, YuYao Zhu +15
cs.ROcs.AIarXiv:2603.11558v32026AutoEval: Autonomous Evaluation of Generalist Robot Manipulation Policies in the Real World
Zhiyuan Zhou, Pranav Atreya, You Liang Tan +2
cs.ROcs.AIarXiv:2503.24278v22025Towards Human-Bot Collaborative Software Architecting with ChatGPT
Aakash Ahmad, Muhammad Waseem, Peng Liang +3
cs.SEcs.AIarXiv:2302.14600v12023RLHF Deciphered: A Critical Analysis of Reinforcement Learning from Human Feedback for LLMs
Shreyas Chaudhari, Pranjal Aggarwal, Vishvak Murahari +5
cs.LGcs.AIcs.CLarXiv:2404.08555v22024Computational Depth Measurement in Thermographic Video: Overcoming Spatial Overfitting via Spatio-Temporal Decoupling
Zain Ul Abidin, Habeeban Memon, Junaid Ahmed
cs.AIcs.CVarXiv:2608.29223v12026AutoAgent: A Fully-Automated and Zero-Code Framework for LLM Agents
Jiabin Tang, Tianyu Fan, Chao Huang
cs.AIcs.CLarXiv:2502.05957v32025Recent Advances in Zero-shot Recognition
Yanwei Fu, Tao Xiang, Yu-Gang Jiang +3
cs.CVcs.AIcs.LGarXiv:1710.04837v12017Foundation Models in Autonomous Driving: A Survey on Scenario Generation and Scenario Analysis
Yuan Gao, Mattia Piccinini, Yuchen Zhang +12
cs.ROcs.AIarXiv:2506.11526v42025SkillGen: Verified Inference-Time Agent Skill Synthesis
Yuchen Ma, Yue Huang, Han Bao +5
cs.LGcs.AIcs.MAarXiv:2605.10999v12026$τ^τ$-Bench: An Environment for End-To-End, Realistic Agent Construction
Quan Shi, Keshav Dhandhania, Karthik Narasimhan +1
cs.AIarXiv:2609.04611v12026A Survey of Learning Causality with Data: Problems and Methods
Ruocheng Guo, Lu Cheng, Jundong Li +2
cs.AIstat.MEarXiv:1809.09337v42018SafeVLA: Towards Safety Alignment of Vision-Language-Action Model via Constrained Learning
Borong Zhang, Yuhao Zhang, Jiaming Ji +5
cs.ROcs.AIarXiv:2503.03480v42025Plan Explicability and Predictability for Robot Task Planning
Yu Zhang, Sarath Sreedharan, Anagha Kulkarni +3
cs.AIcs.ROarXiv:1511.08158v22015Deception Abilities Emerged in Large Language Models
Thilo Hagendorff
cs.CLcs.AIcs.LGarXiv:2307.16513v22023Accelerating LLM Inference with Staged Speculative Decoding
Benjamin Spector, Chris Re
cs.AIcs.CLarXiv:2308.04623v12023Mind the Value-Action Gap: Do LLMs Act in Alignment with Their Values?
Hua Shen, Nicholas Clark, Tanushree Mitra
cs.HCcs.AIcs.CLarXiv:2501.15463v42025DataSciBench: An LLM Agent Benchmark for Data Science
Dan Zhang, Sining Zhoubian, Min Cai +7
cs.CLcs.AIcs.LGarXiv:2502.13897v12025Tiny Machine Learning: Progress and Futures
Ji Lin, Ligeng Zhu, Wei-Ming Chen +2
cs.LGcs.AIcs.CVarXiv:2403.19076v22024Enhancing Multimodal Emotion Recognition via Multi-Feature Encoding and Attention-Based Fusion
Xu Lin, Ke Wang, Hui Kang +1
cs.CVcs.AIcs.LGarXiv:2609.04690v12026Fast and Effective On-policy Distillation from Reasoning Prefixes
Dongxu Zhang, Zhichao Yang, Sepehr Janghorbani +6
cs.LGcs.AIarXiv:2602.15260v12026When Are Teacher Tokens Reliable? Position-Weighted On-Policy Self-Distillation for Reasoning
Xiaogeng Liu, Xinyan Wang, Yingzi Ma +2
cs.LGcs.AIarXiv:2605.21606v12026Towards Efficient Generative Large Language Model Serving: A Survey from Algorithms to Systems
Xupeng Miao, Gabriele Oliaro, Zhihao Zhang +4
cs.LGcs.AIcs.DCarXiv:2312.15234v22023Sim-and-Real Co-Training: A Simple Recipe for Vision-Based Robotic Manipulation
Abhiram Maddukuri, Zhenyu Jiang, Lawrence Yunliang Chen +12
cs.ROcs.AIcs.LGarXiv:2503.24361v22025Refuse without Refusal: A Structural Analysis of Safety-Tuning Responses for Reducing False Refusals in Language Models
Minji Kim, Hyounghun Kim
cs.CLcs.AIarXiv:2609.04714v12026Frontier AI Regulation: Managing Emerging Risks to Public Safety
Markus Anderljung, Joslyn Barnhart, Anton Korinek +21
cs.CYcs.AIarXiv:2307.03718v42023Why We Care About Understanding: Competence through Predictive Compression
Matthieu Queloz, Pierre Beckmann
cs.AIcs.CLarXiv:2609.04962v12026Gym-Anything: Turn any Software into an Agent Environment
Pranjal Aggarwal, Graham Neubig, Sean Welleck
cs.LGcs.AIarXiv:2604.06126v12026REVEAL: Retrieval-Augmented Visual-Language Pre-Training with Multi-Source Multimodal Knowledge Memory
Ziniu Hu, Ahmet Iscen, Chen Sun +6
cs.CVcs.AIarXiv:2212.05221v22022Dynamic Heterogeneous Graph Representation Learning: A Survey
Huan Liu, Pengfei Jiao, Jie Yin +2
cs.LGcs.AIcs.SIarXiv:2609.04779v12026Projecting Assumptions: The Duality Between Sparse Autoencoders and Concept Geometry
Sai Sumedh R. Hindupur, Ekdeep Singh Lubana, Thomas Fel +1
cs.LGcs.AIarXiv:2503.01822v22025Breaking the Protocol: Security Analysis of the Model Context Protocol Specification and Prompt Injection Vulnerabilities in Tool-Integrated LLM Agents
Narek Maloyan, Dmitry Namiot
cs.CRcs.AIarXiv:2601.17549v12026Leveraging Imperfect Restoration for Data Availability Attack
Yi Huang, Jeremy Styborski, Mingzhi Lyu +2
cs.AIarXiv:2609.04627v12026$τ$-Voice: Benchmarking Full-Duplex Voice Agents on Real-World Domains
Soham Ray, Keshav Dhandhania, Victor Barres +1
cs.SDcs.AIarXiv:2603.13686v12026On the Position Bias of On-Policy Distillation
Yan Xie, Sijie Zhu, Tiansheng Wen +2
cs.LGcs.AIarXiv:2606.22600v32026Inducing Relational Knowledge from BERT
Zied Bouraoui, Jose Camacho-Collados, Steven Schockaert
cs.CLcs.AIarXiv:1911.12753v12019De-skilling, Cognitive Offloading, and Misplaced Responsibilities: Potential Ironies of AI-Assisted Design
Prakash Shukla, Phuong Bui, Sean S Levy +3
cs.HCcs.AIarXiv:2503.03924v12025A Survey of Multimodal Retrieval-Augmented Generation
Lang Mei, Siyu Mo, Zhihan Yang +1
cs.IRcs.AIcs.CLarXiv:2504.08748v12025Which Attention Heads Matter for In-Context Learning?
Kayo Yin, Jacob Steinhardt
cs.LGcs.AIcs.CLarXiv:2502.14010v12025LeVERB: Humanoid Whole-Body Control with Latent Vision-Language Instruction
Haoru Xue, Xiaoyu Huang, Dantong Niu +8
cs.ROcs.AIarXiv:2506.13751v32025MIPaaL: Mixed Integer Program as a Layer
Aaron Ferber, Bryan Wilder, Bistra Dilkina +1
cs.LGcs.AIarXiv:1907.05912v22019Imitation Is Not Enough: Robustifying Imitation with Reinforcement Learning for Challenging Driving Scenarios
Yiren Lu, Justin Fu, George Tucker +9
cs.AIcs.ROarXiv:2212.11419v22022Toward Postural State Classification in Immersive VR with Multimodal Data and Explainability Analysis
Nipa Anjum, Md Irfan Pavel, Robert Gonzalez +4
cs.HCcs.AIcs.LGarXiv:2608.28844v12026EvoTool: Self-Evolving Tool-Use Policy Optimization in LLM Agents via Blame-Aware Mutation and Diversity-Aware Selection
Shuo Yang, Soyeon Caren Han, Xueqi Ma +3
cs.AIarXiv:2603.04900v12026Capability-Stratified Degradation in Ternary Language Models
Anirudh Malik, M Sparsh Mehra, Poojith Devan
cs.AIarXiv:2608.28809v12026Exploring the Responses of Large Language Models to Beginner Programmers' Help Requests
Arto Hellas, Juho Leinonen, Sami Sarsa +3
cs.CYcs.AIcs.CLarXiv:2306.05715v12023ARC-AGI-3: A New Challenge for Frontier Agentic Intelligence
ARC Prize Foundation
cs.AIarXiv:2603.24621v22026Do Massively Pretrained Language Models Make Better Storytellers?
Abigail See, Aneesh Pappu, Rohun Saxena +2
cs.CLcs.AIcs.LGarXiv:1909.10705v12019Efficient Reasoning with Hidden Thinking
Xuan Shen, Yizhou Wang, Yufa Zhou +4
cs.CLcs.AIcs.LGarXiv:2501.19201v22025Who's in Charge? Disempowerment Patterns in Real-World LLM Usage
Mrinank Sharma, Miles McCain, Raymond Douglas +1
cs.CYcs.AIcs.CLarXiv:2601.19062v12026