Artificial Intelligence

Papers filed under cs.AI on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.

Search paper metadata (including unsummarized papers)

1 to 60 of 15,209

  1. Linear attention is (maybe) all you need (to understand transformer optimization)

    Kwangjun Ahn, Xiang Cheng, Minhak Song +3

    cs.LGcs.AImath.OCarXiv:2310.01082v22023
  2. HABERTOR: An Efficient and Effective Deep Hatespeech Detector

    Thanh Tran, Yifan Hu, Changwei Hu +4

    cs.CLcs.AIcs.IRarXiv:2010.08865v12020
  3. Detecting Data Contamination from Reinforcement Learning Post-training for Large Language Models

    Yongding Tao, Tian Wang, Yihong Dong +4

    cs.CLcs.AIcs.LGarXiv:2510.09259v22025
  4. LLMs as Scalable, General-Purpose Simulators For Evolving Digital Agent Training

    Yiming Wang, Da Yin, Yuedong Cui +8

    cs.CLcs.AIcs.LGarXiv:2510.14969v12025
  5. Agentic Knowledgeable Self-awareness

    Shuofei Qiao, Zhisong Qiu, Baochang Ren +8

    cs.CLcs.AIcs.CVarXiv:2504.03553v22025
  6. UniTraj: Learning a Universal Trajectory Foundation Model from Billion-Scale Worldwide Traces

    Yuanshao Zhu, James Jianqiao Yu, Xiangyu Zhao +4

    cs.ETcs.AIcs.LGarXiv:2411.03859v32024
  7. Finding Generalizable Evidence by Learning to Convince Q&A Models

    Ethan Perez, Siddharth Karamcheti, Rob Fergus +3

    cs.CLcs.AIcs.IRarXiv:1909.05863v12019
  8. Keyword search is all you need: Achieving RAG-Level Performance without vector databases using agentic tool use

    Shreyas Subramanian, Adewale Akinfaderin, Yanyan Zhang +4

    cs.IRcs.AIarXiv:2602.23368v12025
  9. Plan-over-Graph: Towards Parallelable LLM Agent Schedule

    Shiqi Zhang, Xinbei Ma, Zouying Cao +2

    cs.AIarXiv:2502.14563v12025
  10. The Evolution of Thought: Tracking LLM Overthinking via Reasoning Dynamics Analysis

    Zihao Wei, Liang Pang, Jiahao Liu +7

    cs.CLcs.AIarXiv:2508.17627v22025
  11. Diving Deep into Modes of Fact Hallucinations in Dialogue Systems

    Souvik Das, Sougata Saha, Rohini K. Srihari

    cs.CLcs.AIarXiv:2301.04449v12023
  12. TeLoGraF: Temporal Logic Planning via Graph-encoded Flow Matching

    Yue Meng, Chuchu Fan

    cs.ROcs.AIcs.FLarXiv:2505.00562v12025
  13. Is Behavior Cloning All You Need? Understanding Horizon in Imitation Learning

    Dylan J. Foster, Adam Block, Dipendra Misra

    cs.LGcs.AImath.STarXiv:2407.15007v22024
  14. Seeing Beyond Words: Self-Supervised Visual Learning for Multimodal Large Language Models

    Davide Caffagni, Sara Sarto, Marcella Cornia +5

    cs.CVcs.AIcs.CLarXiv:2512.15885v12025
  15. Predictive Concept Decoders: Training Scalable End-to-End Interpretability Assistants

    Vincent Huang, Dami Choi, Daniel D. Johnson +2

    cs.AIcs.CLcs.LGarXiv:2512.15712v12025
  16. Building Better Activation Oracles

    Jan Bauer, Celeste De Schamphelaere, Adam Karvonen +2

    cs.LGcs.AIarXiv:2606.02609v22026
  17. Activation Oracles: Training and Evaluating LLMs as General-Purpose Activation Explainers

    Adam Karvonen, James Chua, Clément Dumas +8

    cs.CLcs.AIcs.LGarXiv:2512.15674v22025
  18. Generating Images Part by Part with Composite Generative Adversarial Networks

    Hanock Kwak, Byoung-Tak Zhang

    cs.AIcs.CVcs.LGarXiv:1607.05387v22016
  19. Fun-ASR Technical Report

    Keyu An, Yanni Chen, Zhigao Chen +35

    cs.CLcs.AIcs.SDarXiv:2509.12508v42025
  20. Interscript: A dataset for interactive learning of scripts through error feedback

    Niket Tandon, Aman Madaan, Peter Clark +2

    cs.AIarXiv:2112.07867v22021
  21. Self-Improving Language Models for Evolutionary Program Synthesis: A Case Study on ARC-AGI

    Julien Pourcel, Cédric Colas, Pierre-Yves Oudeyer

    cs.LGcs.AIcs.NEarXiv:2507.14172v22025
  22. Rethinking On-Policy Self-Distillation for Thinking Models

    Simran Kaur, Narutatsu Ri, Yinghui He +2

    cs.AIcs.LGarXiv:2607.05184v12026
  23. SubtleMemory: A Benchmark for Fine-Grained Relational Memory Discrimination in Long-Horizon AI Agents

    Wenxuan Wang, Haoyu Sun, Fukuan Hou +4

    cs.AIcs.CLarXiv:2606.05761v22026
  24. SimWorld Studio: Automatic Environment Generation with Evolving Coding Agent for Embodied Agent Learning

    Haoqiang Kang, Xiaokang Ye, Yuhan Liu +5

    cs.AIarXiv:2605.09423v22026
  25. The Amazing Agent Race: Strong Tool Users, Weak Navigators

    Zae Myung Kim, Dongseok Lee, Jaehyung Kim +2

    cs.AIcs.CLcs.LGarXiv:2604.10261v22026
  26. Strategic Navigation or Stochastic Search? How Agents and Humans Reason Over Document Collections

    Łukasz Borchmann, Jordy Van Landeghem, Michał Turski +12

    cs.CLcs.AIarXiv:2603.12180v22026
  27. VisPhyWorld: Probing Physical Reasoning via Code-Driven Video Reconstruction

    Jiarong Liang, Max Ku, Ka-Hei Hui +2

    cs.CVcs.AIarXiv:2602.13294v32026
  28. Autonomous Agents on Blockchains: Standards, Execution Models, and Trust Boundaries

    Saad Alqithami

    cs.AIcs.MAarXiv:2601.04583v12026
  29. Artificial Intelligence Index Report 2025

    Nestor Maslej, Loredana Fattorini, Raymond Perrault +20

    cs.AIarXiv:2504.07139v32025
  30. Generative AI Act II: Test Time Scaling Drives Cognition Engineering

    Shijie Xia, Yiwei Qin, Xuefeng Li +11

    cs.CLcs.AIarXiv:2504.13828v32025
  31. Contrastive Instruction Tuning

    Tianyi Lorena Yan, Fei Wang, James Y. Huang +5

    cs.CLcs.AIcs.LGarXiv:2402.11138v22024
  32. Habitat-Web: Learning Embodied Object-Search Strategies from Human Demonstrations at Scale

    Ram Ramrakhya, Eric Undersander, Dhruv Batra +1

    cs.AIcs.CVcs.ROarXiv:2204.03514v22022
  33. Hybrid Transformer with Multi-level Fusion for Multimodal Knowledge Graph Completion

    Xiang Chen, Ningyu Zhang, Lei Li +6

    cs.CLcs.AIcs.CVarXiv:2205.02357v52022
  34. MetaSpatial: Reinforcing 3D Spatial Reasoning in VLMs for the Metaverse

    Zhenyu Pan, Han Liu

    cs.CVcs.AIarXiv:2503.18470v22025
  35. From Attribution Maps to Human-Understandable Explanations through Concept Relevance Propagation

    Reduan Achtibat, Maximilian Dreyer, Ilona Eisenbraun +4

    cs.LGcs.AIarXiv:2206.03208v22022
  36. Comprehending and Ordering Semantics for Image Captioning

    Yehao Li, Yingwei Pan, Ting Yao +1

    cs.CVcs.AIcs.CLarXiv:2206.06930v12022
  37. Can large language models reason about medical questions?

    Valentin Liévin, Christoffer Egeberg Hother, Andreas Geert Motzfeldt +1

    cs.CLcs.AIcs.LGarXiv:2207.08143v42022
  38. iTool: Reinforced Fine-Tuning with Dynamic Deficiency Calibration for Advanced Tool Use

    Yirong Zeng, Xiao Ding, Yuxian Wang +8

    cs.CLcs.AIcs.LGarXiv:2501.09766v52025
  39. Federated Learning on Non-IID Graphs via Structural Knowledge Sharing

    Yue Tan, Yixin Liu, Guodong Long +3

    cs.LGcs.AIcs.DCarXiv:2211.13009v12022
  40. Reinforcement Fine-Tuning Naturally Mitigates Forgetting in Continual Post-Training

    Song Lai, Haohan Zhao, Rong Feng +9

    cs.LGcs.AIcs.CLarXiv:2507.05386v62025
  41. Learning Performance-Improving Code Edits

    Alexander Shypula, Aman Madaan, Yimeng Zeng +7

    cs.SEcs.AIcs.LGarXiv:2302.07867v52023
  42. Can Pre-trained Vision and Language Models Answer Visual Information-Seeking Questions?

    Yang Chen, Hexiang Hu, Yi Luan +4

    cs.CVcs.AIcs.CLarXiv:2302.11713v52023
  43. Learning by Distilling Context

    Charlie Snell, Dan Klein, Ruiqi Zhong

    cs.CLcs.AIarXiv:2209.15189v12022
  44. GPT-4 Technical Report

    OpenAI, Josh Achiam, Steven Adler +278

    cs.CLcs.AIarXiv:2303.08774v62023
  45. LLaMA-Adapter: Efficient Fine-tuning of Language Models with Zero-init Attention

    Renrui Zhang, Jiaming Han, Chris Liu +7

    cs.CVcs.AIcs.CLarXiv:2303.16199v32023
  46. LLaMA-Adapter V2: Parameter-Efficient Visual Instruction Model

    Peng Gao, Jiaming Han, Renrui Zhang +9

    cs.CVcs.AIcs.CLarXiv:2304.15010v12023
  47. RS5M and GeoRSCLIP: A Large Scale Vision-Language Dataset and A Large Vision-Language Model for Remote Sensing

    Zilun Zhang, Tiancheng Zhao, Yulong Guo +1

    cs.CVcs.AIcs.CLarXiv:2306.11300v52023
  48. Matching Patients to Clinical Trials with Large Language Models

    Qiao Jin, Zifeng Wang, Charalampos S. Floudas +7

    cs.CLcs.AIarXiv:2307.15051v52023
  49. Catastrophic Jailbreak of Open-source LLMs via Exploiting Generation

    Yangsibo Huang, Samyak Gupta, Mengzhou Xia +2

    cs.CLcs.AIcs.CRarXiv:2310.06987v12023
  50. Linear Representations of Sentiment in Large Language Models

    Curt Tigges, Oskar John Hollinsworth, Atticus Geiger +1

    cs.LGcs.AIcs.CLarXiv:2310.15154v12023
  51. Gibbs Sampling with People

    Peter M. C. Harrison, Raja Marjieh, Federico Adolfi +5

    q-bio.NCcs.AIcs.CVarXiv:2008.02595v22020
  52. Large Language Models for Robotics: A Survey

    Fanlong Zeng, Wensheng Gan, Zezheng Huai +5

    cs.ROcs.AIarXiv:2311.07226v22023
  53. Emu Video: Factorizing Text-to-Video Generation by Explicit Image Conditioning

    Rohit Girdhar, Mannat Singh, Andrew Brown +7

    cs.CVcs.AIcs.GRarXiv:2311.10709v22023
  54. HalluciDoctor: Mitigating Hallucinatory Toxicity in Visual Instruction Data

    Qifan Yu, Juncheng Li, Longhui Wei +6

    cs.CVcs.AIarXiv:2311.13614v22023
  55. The AI Assessment Scale (AIAS): A Framework for Ethical Integration of Generative AI in Educational Assessment

    Mike Perkins, Leon Furze, Jasper Roe +1

    cs.AIarXiv:2312.07086v22023
  56. Large Legal Fictions: Profiling Legal Hallucinations in Large Language Models

    Matthew Dahl, Varun Magesh, Mirac Suzgun +1

    cs.CLcs.AIcs.CYarXiv:2401.01301v22024
  57. FlightLLM: Efficient Large Language Model Inference with a Complete Mapping Flow on FPGAs

    Shulin Zeng, Jun Liu, Guohao Dai +14

    cs.ARcs.AIarXiv:2401.03868v22024
  58. One Polluted Page Is Enough: Evaluating Web Content Pollution in LLM Recommenders

    Minghao Luo, Liang Chen

    cs.CLcs.AIarXiv:2606.13610v22026
  59. SkillAdaptor: Self-Adapting Skills for LLM Agents from Trajectories

    Zhuoyun Yu, Xin Xie, Wuguannan Yao +4

    cs.CLcs.AIcs.LGarXiv:2606.01311v12026
  60. Memory-Bound but Not Bandwidth-Limited: The Physical AI Inference Gap in Batch-1 LLM Decode

    Josef Chen

    cs.ARcs.AIcs.DCarXiv:2605.30571v12026