Artificial Intelligence

Papers filed under cs.AI on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.

Search paper metadata (including unsummarized papers)

3,781 to 3,840 of 15,222

  1. LiveMathematicianBench: A Live Benchmark for Mathematician-Level Reasoning with Proof Sketches

    Linyang He, Qiyao Yu, Hanze Dong +5

    cs.CLcs.AIcs.LGarXiv:2604.01754v12026
  2. Federated Learning for Healthcare Domain - Pipeline, Applications and Challenges

    Madhura Joshi, Ankit Pal, Malaikannan Sankarasubbu

    cs.LGcs.AIcs.CRarXiv:2211.07893v22022
  3. Controlled LLM Training on Spectral Sphere

    Tian Xie, Haoming Luo, Haoyu Tang +9

    cs.LGcs.AIarXiv:2601.08393v32026
  4. Cold-Start Recommendation towards the Era of Large Language Models (LLMs): A Comprehensive Survey and Roadmap

    Weizhi Zhang, Yuanchen Bei, Liangwei Yang +15

    cs.IRcs.AIarXiv:2501.01945v22025
  5. trajectory-judge: What Outcome-Only LLM Judges Miss on Agent Trajectories

    Hadi Mohammadi

    cs.CLcs.AIcs.SEarXiv:2609.00038v12026
  6. Towards Agentic RAG with Deep Reasoning: A Survey of RAG-Reasoning Systems in LLMs

    Yangning Li, Weizhi Zhang, Yuyao Yang +17

    cs.CLcs.AIarXiv:2507.09477v22025
  7. A Survey on LoRA of Large Language Models

    Yuren Mao, Yuhang Ge, Yijiang Fan +4

    cs.LGcs.AIcs.CLarXiv:2407.11046v42024
  8. Cross-Relational Preference Learning for Better LLM Instruction Following

    Runsheng Li, Kai Sun, Bin Shi +1

    cs.AIarXiv:2608.29352v12026
  9. Learning Combinatorial Optimization on Graphs: A Survey with Applications to Networking

    Natalia Vesselinova, Rebecca Steinert, Daniel F. Perez-Ramirez +1

    cs.LGcs.AIstat.MLarXiv:2005.11081v22020
  10. Security Concerns for Large Language Models: A Survey

    Miles Q. Li, Benjamin C. M. Fung

    cs.CRcs.AIarXiv:2505.18889v52025
  11. Don't Trust ChatGPT when Your Question is not in English: A Study of Multilingual Abilities and Types of LLMs

    Xiang Zhang, Senyu Li, Bradley Hauer +2

    cs.CLcs.AIarXiv:2305.16339v22023
  12. RoboClaw: An Agentic Framework for Scalable Long-Horizon Robotic Tasks

    Ruiying Li, Yunlang Zhou, YuYao Zhu +15

    cs.ROcs.AIarXiv:2603.11558v32026
  13. AutoEval: Autonomous Evaluation of Generalist Robot Manipulation Policies in the Real World

    Zhiyuan Zhou, Pranav Atreya, You Liang Tan +2

    cs.ROcs.AIarXiv:2503.24278v22025
  14. Towards Human-Bot Collaborative Software Architecting with ChatGPT

    Aakash Ahmad, Muhammad Waseem, Peng Liang +3

    cs.SEcs.AIarXiv:2302.14600v12023
  15. RLHF Deciphered: A Critical Analysis of Reinforcement Learning from Human Feedback for LLMs

    Shreyas Chaudhari, Pranjal Aggarwal, Vishvak Murahari +5

    cs.LGcs.AIcs.CLarXiv:2404.08555v22024
  16. Computational Depth Measurement in Thermographic Video: Overcoming Spatial Overfitting via Spatio-Temporal Decoupling

    Zain Ul Abidin, Habeeban Memon, Junaid Ahmed

    cs.AIcs.CVarXiv:2608.29223v12026
  17. AutoAgent: A Fully-Automated and Zero-Code Framework for LLM Agents

    Jiabin Tang, Tianyu Fan, Chao Huang

    cs.AIcs.CLarXiv:2502.05957v32025
  18. Recent Advances in Zero-shot Recognition

    Yanwei Fu, Tao Xiang, Yu-Gang Jiang +3

    cs.CVcs.AIcs.LGarXiv:1710.04837v12017
  19. Foundation Models in Autonomous Driving: A Survey on Scenario Generation and Scenario Analysis

    Yuan Gao, Mattia Piccinini, Yuchen Zhang +12

    cs.ROcs.AIarXiv:2506.11526v42025
  20. SkillGen: Verified Inference-Time Agent Skill Synthesis

    Yuchen Ma, Yue Huang, Han Bao +5

    cs.LGcs.AIcs.MAarXiv:2605.10999v12026
  21. $τ^τ$-Bench: An Environment for End-To-End, Realistic Agent Construction

    Quan Shi, Keshav Dhandhania, Karthik Narasimhan +1

    cs.AIarXiv:2609.04611v12026
  22. A Survey of Learning Causality with Data: Problems and Methods

    Ruocheng Guo, Lu Cheng, Jundong Li +2

    cs.AIstat.MEarXiv:1809.09337v42018
  23. SafeVLA: Towards Safety Alignment of Vision-Language-Action Model via Constrained Learning

    Borong Zhang, Yuhao Zhang, Jiaming Ji +5

    cs.ROcs.AIarXiv:2503.03480v42025
  24. Plan Explicability and Predictability for Robot Task Planning

    Yu Zhang, Sarath Sreedharan, Anagha Kulkarni +3

    cs.AIcs.ROarXiv:1511.08158v22015
  25. Deception Abilities Emerged in Large Language Models

    Thilo Hagendorff

    cs.CLcs.AIcs.LGarXiv:2307.16513v22023
  26. Accelerating LLM Inference with Staged Speculative Decoding

    Benjamin Spector, Chris Re

    cs.AIcs.CLarXiv:2308.04623v12023
  27. Mind the Value-Action Gap: Do LLMs Act in Alignment with Their Values?

    Hua Shen, Nicholas Clark, Tanushree Mitra

    cs.HCcs.AIcs.CLarXiv:2501.15463v42025
  28. DataSciBench: An LLM Agent Benchmark for Data Science

    Dan Zhang, Sining Zhoubian, Min Cai +7

    cs.CLcs.AIcs.LGarXiv:2502.13897v12025
  29. Tiny Machine Learning: Progress and Futures

    Ji Lin, Ligeng Zhu, Wei-Ming Chen +2

    cs.LGcs.AIcs.CVarXiv:2403.19076v22024
  30. Enhancing Multimodal Emotion Recognition via Multi-Feature Encoding and Attention-Based Fusion

    Xu Lin, Ke Wang, Hui Kang +1

    cs.CVcs.AIcs.LGarXiv:2609.04690v12026
  31. Fast and Effective On-policy Distillation from Reasoning Prefixes

    Dongxu Zhang, Zhichao Yang, Sepehr Janghorbani +6

    cs.LGcs.AIarXiv:2602.15260v12026
  32. When Are Teacher Tokens Reliable? Position-Weighted On-Policy Self-Distillation for Reasoning

    Xiaogeng Liu, Xinyan Wang, Yingzi Ma +2

    cs.LGcs.AIarXiv:2605.21606v12026
  33. Towards Efficient Generative Large Language Model Serving: A Survey from Algorithms to Systems

    Xupeng Miao, Gabriele Oliaro, Zhihao Zhang +4

    cs.LGcs.AIcs.DCarXiv:2312.15234v22023
  34. Sim-and-Real Co-Training: A Simple Recipe for Vision-Based Robotic Manipulation

    Abhiram Maddukuri, Zhenyu Jiang, Lawrence Yunliang Chen +12

    cs.ROcs.AIcs.LGarXiv:2503.24361v22025
  35. Refuse without Refusal: A Structural Analysis of Safety-Tuning Responses for Reducing False Refusals in Language Models

    Minji Kim, Hyounghun Kim

    cs.CLcs.AIarXiv:2609.04714v12026
  36. Frontier AI Regulation: Managing Emerging Risks to Public Safety

    Markus Anderljung, Joslyn Barnhart, Anton Korinek +21

    cs.CYcs.AIarXiv:2307.03718v42023
  37. Why We Care About Understanding: Competence through Predictive Compression

    Matthieu Queloz, Pierre Beckmann

    cs.AIcs.CLarXiv:2609.04962v12026
  38. Gym-Anything: Turn any Software into an Agent Environment

    Pranjal Aggarwal, Graham Neubig, Sean Welleck

    cs.LGcs.AIarXiv:2604.06126v12026
  39. REVEAL: Retrieval-Augmented Visual-Language Pre-Training with Multi-Source Multimodal Knowledge Memory

    Ziniu Hu, Ahmet Iscen, Chen Sun +6

    cs.CVcs.AIarXiv:2212.05221v22022
  40. Dynamic Heterogeneous Graph Representation Learning: A Survey

    Huan Liu, Pengfei Jiao, Jie Yin +2

    cs.LGcs.AIcs.SIarXiv:2609.04779v12026
  41. Projecting Assumptions: The Duality Between Sparse Autoencoders and Concept Geometry

    Sai Sumedh R. Hindupur, Ekdeep Singh Lubana, Thomas Fel +1

    cs.LGcs.AIarXiv:2503.01822v22025
  42. Breaking the Protocol: Security Analysis of the Model Context Protocol Specification and Prompt Injection Vulnerabilities in Tool-Integrated LLM Agents

    Narek Maloyan, Dmitry Namiot

    cs.CRcs.AIarXiv:2601.17549v12026
  43. Leveraging Imperfect Restoration for Data Availability Attack

    Yi Huang, Jeremy Styborski, Mingzhi Lyu +2

    cs.AIarXiv:2609.04627v12026
  44. $τ$-Voice: Benchmarking Full-Duplex Voice Agents on Real-World Domains

    Soham Ray, Keshav Dhandhania, Victor Barres +1

    cs.SDcs.AIarXiv:2603.13686v12026
  45. On the Position Bias of On-Policy Distillation

    Yan Xie, Sijie Zhu, Tiansheng Wen +2

    cs.LGcs.AIarXiv:2606.22600v32026
  46. Inducing Relational Knowledge from BERT

    Zied Bouraoui, Jose Camacho-Collados, Steven Schockaert

    cs.CLcs.AIarXiv:1911.12753v12019
  47. De-skilling, Cognitive Offloading, and Misplaced Responsibilities: Potential Ironies of AI-Assisted Design

    Prakash Shukla, Phuong Bui, Sean S Levy +3

    cs.HCcs.AIarXiv:2503.03924v12025
  48. A Survey of Multimodal Retrieval-Augmented Generation

    Lang Mei, Siyu Mo, Zhihan Yang +1

    cs.IRcs.AIcs.CLarXiv:2504.08748v12025
  49. Which Attention Heads Matter for In-Context Learning?

    Kayo Yin, Jacob Steinhardt

    cs.LGcs.AIcs.CLarXiv:2502.14010v12025
  50. LeVERB: Humanoid Whole-Body Control with Latent Vision-Language Instruction

    Haoru Xue, Xiaoyu Huang, Dantong Niu +8

    cs.ROcs.AIarXiv:2506.13751v32025
  51. MIPaaL: Mixed Integer Program as a Layer

    Aaron Ferber, Bryan Wilder, Bistra Dilkina +1

    cs.LGcs.AIarXiv:1907.05912v22019
  52. Imitation Is Not Enough: Robustifying Imitation with Reinforcement Learning for Challenging Driving Scenarios

    Yiren Lu, Justin Fu, George Tucker +9

    cs.AIcs.ROarXiv:2212.11419v22022
  53. Toward Postural State Classification in Immersive VR with Multimodal Data and Explainability Analysis

    Nipa Anjum, Md Irfan Pavel, Robert Gonzalez +4

    cs.HCcs.AIcs.LGarXiv:2608.28844v12026
  54. EvoTool: Self-Evolving Tool-Use Policy Optimization in LLM Agents via Blame-Aware Mutation and Diversity-Aware Selection

    Shuo Yang, Soyeon Caren Han, Xueqi Ma +3

    cs.AIarXiv:2603.04900v12026
  55. Capability-Stratified Degradation in Ternary Language Models

    Anirudh Malik, M Sparsh Mehra, Poojith Devan

    cs.AIarXiv:2608.28809v12026
  56. Exploring the Responses of Large Language Models to Beginner Programmers' Help Requests

    Arto Hellas, Juho Leinonen, Sami Sarsa +3

    cs.CYcs.AIcs.CLarXiv:2306.05715v12023
  57. ARC-AGI-3: A New Challenge for Frontier Agentic Intelligence

    ARC Prize Foundation

    cs.AIarXiv:2603.24621v22026
  58. Do Massively Pretrained Language Models Make Better Storytellers?

    Abigail See, Aneesh Pappu, Rohun Saxena +2

    cs.CLcs.AIcs.LGarXiv:1909.10705v12019
  59. Efficient Reasoning with Hidden Thinking

    Xuan Shen, Yizhou Wang, Yufa Zhou +4

    cs.CLcs.AIcs.LGarXiv:2501.19201v22025
  60. Who's in Charge? Disempowerment Patterns in Real-World LLM Usage

    Mrinank Sharma, Miles McCain, Raymond Douglas +1

    cs.CYcs.AIcs.CLarXiv:2601.19062v12026