Artificial Intelligence
Papers filed under cs.AI on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.
Search paper metadata (including unsummarized papers)
1 to 60 of 15,209
Linear attention is (maybe) all you need (to understand transformer optimization)
Kwangjun Ahn, Xiang Cheng, Minhak Song +3
cs.LGcs.AImath.OCarXiv:2310.01082v22023HABERTOR: An Efficient and Effective Deep Hatespeech Detector
Thanh Tran, Yifan Hu, Changwei Hu +4
cs.CLcs.AIcs.IRarXiv:2010.08865v12020Detecting Data Contamination from Reinforcement Learning Post-training for Large Language Models
Yongding Tao, Tian Wang, Yihong Dong +4
cs.CLcs.AIcs.LGarXiv:2510.09259v22025LLMs as Scalable, General-Purpose Simulators For Evolving Digital Agent Training
Yiming Wang, Da Yin, Yuedong Cui +8
cs.CLcs.AIcs.LGarXiv:2510.14969v12025Agentic Knowledgeable Self-awareness
Shuofei Qiao, Zhisong Qiu, Baochang Ren +8
cs.CLcs.AIcs.CVarXiv:2504.03553v22025UniTraj: Learning a Universal Trajectory Foundation Model from Billion-Scale Worldwide Traces
Yuanshao Zhu, James Jianqiao Yu, Xiangyu Zhao +4
cs.ETcs.AIcs.LGarXiv:2411.03859v32024Finding Generalizable Evidence by Learning to Convince Q&A Models
Ethan Perez, Siddharth Karamcheti, Rob Fergus +3
cs.CLcs.AIcs.IRarXiv:1909.05863v12019Keyword search is all you need: Achieving RAG-Level Performance without vector databases using agentic tool use
Shreyas Subramanian, Adewale Akinfaderin, Yanyan Zhang +4
cs.IRcs.AIarXiv:2602.23368v12025Plan-over-Graph: Towards Parallelable LLM Agent Schedule
Shiqi Zhang, Xinbei Ma, Zouying Cao +2
cs.AIarXiv:2502.14563v12025The Evolution of Thought: Tracking LLM Overthinking via Reasoning Dynamics Analysis
Zihao Wei, Liang Pang, Jiahao Liu +7
cs.CLcs.AIarXiv:2508.17627v22025Diving Deep into Modes of Fact Hallucinations in Dialogue Systems
Souvik Das, Sougata Saha, Rohini K. Srihari
cs.CLcs.AIarXiv:2301.04449v12023TeLoGraF: Temporal Logic Planning via Graph-encoded Flow Matching
Yue Meng, Chuchu Fan
cs.ROcs.AIcs.FLarXiv:2505.00562v12025Is Behavior Cloning All You Need? Understanding Horizon in Imitation Learning
Dylan J. Foster, Adam Block, Dipendra Misra
cs.LGcs.AImath.STarXiv:2407.15007v22024Seeing Beyond Words: Self-Supervised Visual Learning for Multimodal Large Language Models
Davide Caffagni, Sara Sarto, Marcella Cornia +5
cs.CVcs.AIcs.CLarXiv:2512.15885v12025Predictive Concept Decoders: Training Scalable End-to-End Interpretability Assistants
Vincent Huang, Dami Choi, Daniel D. Johnson +2
cs.AIcs.CLcs.LGarXiv:2512.15712v12025Building Better Activation Oracles
Jan Bauer, Celeste De Schamphelaere, Adam Karvonen +2
cs.LGcs.AIarXiv:2606.02609v22026Activation Oracles: Training and Evaluating LLMs as General-Purpose Activation Explainers
Adam Karvonen, James Chua, Clément Dumas +8
cs.CLcs.AIcs.LGarXiv:2512.15674v22025Generating Images Part by Part with Composite Generative Adversarial Networks
Hanock Kwak, Byoung-Tak Zhang
cs.AIcs.CVcs.LGarXiv:1607.05387v22016Fun-ASR Technical Report
Keyu An, Yanni Chen, Zhigao Chen +35
cs.CLcs.AIcs.SDarXiv:2509.12508v42025Interscript: A dataset for interactive learning of scripts through error feedback
Niket Tandon, Aman Madaan, Peter Clark +2
cs.AIarXiv:2112.07867v22021Self-Improving Language Models for Evolutionary Program Synthesis: A Case Study on ARC-AGI
Julien Pourcel, Cédric Colas, Pierre-Yves Oudeyer
cs.LGcs.AIcs.NEarXiv:2507.14172v22025Rethinking On-Policy Self-Distillation for Thinking Models
Simran Kaur, Narutatsu Ri, Yinghui He +2
cs.AIcs.LGarXiv:2607.05184v12026SubtleMemory: A Benchmark for Fine-Grained Relational Memory Discrimination in Long-Horizon AI Agents
Wenxuan Wang, Haoyu Sun, Fukuan Hou +4
cs.AIcs.CLarXiv:2606.05761v22026SimWorld Studio: Automatic Environment Generation with Evolving Coding Agent for Embodied Agent Learning
Haoqiang Kang, Xiaokang Ye, Yuhan Liu +5
cs.AIarXiv:2605.09423v22026The Amazing Agent Race: Strong Tool Users, Weak Navigators
Zae Myung Kim, Dongseok Lee, Jaehyung Kim +2
cs.AIcs.CLcs.LGarXiv:2604.10261v22026Strategic Navigation or Stochastic Search? How Agents and Humans Reason Over Document Collections
Łukasz Borchmann, Jordy Van Landeghem, Michał Turski +12
cs.CLcs.AIarXiv:2603.12180v22026VisPhyWorld: Probing Physical Reasoning via Code-Driven Video Reconstruction
Jiarong Liang, Max Ku, Ka-Hei Hui +2
cs.CVcs.AIarXiv:2602.13294v32026Autonomous Agents on Blockchains: Standards, Execution Models, and Trust Boundaries
Saad Alqithami
cs.AIcs.MAarXiv:2601.04583v12026Artificial Intelligence Index Report 2025
Nestor Maslej, Loredana Fattorini, Raymond Perrault +20
cs.AIarXiv:2504.07139v32025Generative AI Act II: Test Time Scaling Drives Cognition Engineering
Shijie Xia, Yiwei Qin, Xuefeng Li +11
cs.CLcs.AIarXiv:2504.13828v32025Contrastive Instruction Tuning
Tianyi Lorena Yan, Fei Wang, James Y. Huang +5
cs.CLcs.AIcs.LGarXiv:2402.11138v22024Habitat-Web: Learning Embodied Object-Search Strategies from Human Demonstrations at Scale
Ram Ramrakhya, Eric Undersander, Dhruv Batra +1
cs.AIcs.CVcs.ROarXiv:2204.03514v22022Hybrid Transformer with Multi-level Fusion for Multimodal Knowledge Graph Completion
Xiang Chen, Ningyu Zhang, Lei Li +6
cs.CLcs.AIcs.CVarXiv:2205.02357v52022MetaSpatial: Reinforcing 3D Spatial Reasoning in VLMs for the Metaverse
Zhenyu Pan, Han Liu
cs.CVcs.AIarXiv:2503.18470v22025From Attribution Maps to Human-Understandable Explanations through Concept Relevance Propagation
Reduan Achtibat, Maximilian Dreyer, Ilona Eisenbraun +4
cs.LGcs.AIarXiv:2206.03208v22022Comprehending and Ordering Semantics for Image Captioning
Yehao Li, Yingwei Pan, Ting Yao +1
cs.CVcs.AIcs.CLarXiv:2206.06930v12022Can large language models reason about medical questions?
Valentin Liévin, Christoffer Egeberg Hother, Andreas Geert Motzfeldt +1
cs.CLcs.AIcs.LGarXiv:2207.08143v42022iTool: Reinforced Fine-Tuning with Dynamic Deficiency Calibration for Advanced Tool Use
Yirong Zeng, Xiao Ding, Yuxian Wang +8
cs.CLcs.AIcs.LGarXiv:2501.09766v52025Federated Learning on Non-IID Graphs via Structural Knowledge Sharing
Yue Tan, Yixin Liu, Guodong Long +3
cs.LGcs.AIcs.DCarXiv:2211.13009v12022Reinforcement Fine-Tuning Naturally Mitigates Forgetting in Continual Post-Training
Song Lai, Haohan Zhao, Rong Feng +9
cs.LGcs.AIcs.CLarXiv:2507.05386v62025Learning Performance-Improving Code Edits
Alexander Shypula, Aman Madaan, Yimeng Zeng +7
cs.SEcs.AIcs.LGarXiv:2302.07867v52023Can Pre-trained Vision and Language Models Answer Visual Information-Seeking Questions?
Yang Chen, Hexiang Hu, Yi Luan +4
cs.CVcs.AIcs.CLarXiv:2302.11713v52023Learning by Distilling Context
Charlie Snell, Dan Klein, Ruiqi Zhong
cs.CLcs.AIarXiv:2209.15189v12022GPT-4 Technical Report
OpenAI, Josh Achiam, Steven Adler +278
cs.CLcs.AIarXiv:2303.08774v62023LLaMA-Adapter: Efficient Fine-tuning of Language Models with Zero-init Attention
Renrui Zhang, Jiaming Han, Chris Liu +7
cs.CVcs.AIcs.CLarXiv:2303.16199v32023LLaMA-Adapter V2: Parameter-Efficient Visual Instruction Model
Peng Gao, Jiaming Han, Renrui Zhang +9
cs.CVcs.AIcs.CLarXiv:2304.15010v12023RS5M and GeoRSCLIP: A Large Scale Vision-Language Dataset and A Large Vision-Language Model for Remote Sensing
Zilun Zhang, Tiancheng Zhao, Yulong Guo +1
cs.CVcs.AIcs.CLarXiv:2306.11300v52023Matching Patients to Clinical Trials with Large Language Models
Qiao Jin, Zifeng Wang, Charalampos S. Floudas +7
cs.CLcs.AIarXiv:2307.15051v52023Catastrophic Jailbreak of Open-source LLMs via Exploiting Generation
Yangsibo Huang, Samyak Gupta, Mengzhou Xia +2
cs.CLcs.AIcs.CRarXiv:2310.06987v12023Linear Representations of Sentiment in Large Language Models
Curt Tigges, Oskar John Hollinsworth, Atticus Geiger +1
cs.LGcs.AIcs.CLarXiv:2310.15154v12023Gibbs Sampling with People
Peter M. C. Harrison, Raja Marjieh, Federico Adolfi +5
q-bio.NCcs.AIcs.CVarXiv:2008.02595v22020Large Language Models for Robotics: A Survey
Fanlong Zeng, Wensheng Gan, Zezheng Huai +5
cs.ROcs.AIarXiv:2311.07226v22023Emu Video: Factorizing Text-to-Video Generation by Explicit Image Conditioning
Rohit Girdhar, Mannat Singh, Andrew Brown +7
cs.CVcs.AIcs.GRarXiv:2311.10709v22023HalluciDoctor: Mitigating Hallucinatory Toxicity in Visual Instruction Data
Qifan Yu, Juncheng Li, Longhui Wei +6
cs.CVcs.AIarXiv:2311.13614v22023The AI Assessment Scale (AIAS): A Framework for Ethical Integration of Generative AI in Educational Assessment
Mike Perkins, Leon Furze, Jasper Roe +1
cs.AIarXiv:2312.07086v22023Large Legal Fictions: Profiling Legal Hallucinations in Large Language Models
Matthew Dahl, Varun Magesh, Mirac Suzgun +1
cs.CLcs.AIcs.CYarXiv:2401.01301v22024FlightLLM: Efficient Large Language Model Inference with a Complete Mapping Flow on FPGAs
Shulin Zeng, Jun Liu, Guohao Dai +14
cs.ARcs.AIarXiv:2401.03868v22024One Polluted Page Is Enough: Evaluating Web Content Pollution in LLM Recommenders
Minghao Luo, Liang Chen
cs.CLcs.AIarXiv:2606.13610v22026SkillAdaptor: Self-Adapting Skills for LLM Agents from Trajectories
Zhuoyun Yu, Xin Xie, Wuguannan Yao +4
cs.CLcs.AIcs.LGarXiv:2606.01311v12026Memory-Bound but Not Bandwidth-Limited: The Physical AI Inference Gap in Batch-1 LLM Decode
Josef Chen
cs.ARcs.AIcs.DCarXiv:2605.30571v12026