Artificial Intelligence

Papers filed under cs.AI on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.

Search paper metadata (including unsummarized papers)

5,701 to 5,760 of 15,455

  1. AdaptThink: Reasoning Models Can Learn When to Think

    Jiajie Zhang, Nianyi Lin, Lei Hou +2

    cs.CLcs.AIcs.LGarXiv:2505.13417v12025
  2. Transfiver: Human-AI Co-Inference through a Shared Editable State

    Minji Park, Seunghyun Yoon, Hyuk Lim

    cs.AIcs.CLcs.HCarXiv:2609.03797v12026
  3. From Reusing to Forecasting: Accelerating Diffusion Models with TaylorSeers

    Jiacheng Liu, Chang Zou, Yuanhuiyi Lyu +2

    cs.CVcs.AIarXiv:2503.06923v22025
  4. MCP-Universe: Benchmarking Large Language Models with Real-World Model Context Protocol Servers

    Ziyang Luo, Zhiqi Shen, Wenzhuo Yang +7

    cs.AIcs.CLarXiv:2508.14704v12025
  5. LLaVA-Mini: Efficient Image and Video Large Multimodal Models with One Vision Token

    Shaolei Zhang, Qingkai Fang, Zhe Yang +1

    cs.CVcs.AIcs.CLarXiv:2501.03895v22025
  6. PLAS: Latent Action Space for Offline Reinforcement Learning

    Wenxuan Zhou, Sujay Bajracharya, David Held

    cs.ROcs.AIcs.LGarXiv:2011.07213v12020
  7. Benchmarking Cognitive Biases in Large Language Models as Evaluators

    Ryan Koo, Minhwa Lee, Vipul Raheja +3

    cs.CLcs.AIcs.LGarXiv:2309.17012v32023
  8. WorldScore: A Unified Evaluation Benchmark for World Generation

    Haoyi Duan, Hong-Xing Yu, Sirui Chen +2

    cs.GRcs.AIcs.CVarXiv:2504.00983v22025
  9. PhysBench: Benchmarking and Enhancing Vision-Language Models for Physical World Understanding

    Wei Chow, Jiageng Mao, Boyi Li +3

    cs.CVcs.AIcs.CLarXiv:2501.16411v22025
  10. Air-Ground Collaborative Vision-and-Language Navigation via Shared Bird's-Eye Maps

    Shuning Zhang, Liang Li, Yunheng Wang +3

    cs.ROcs.AIarXiv:2609.03483v12026
  11. Contrastive Behavioral Similarity Embeddings for Generalization in Reinforcement Learning

    Rishabh Agarwal, Marlos C. Machado, Pablo Samuel Castro +1

    cs.LGcs.AIstat.MLarXiv:2101.05265v22021
  12. Reasoning Models Know When They're Right: Probing Hidden States for Self-Verification

    Anqi Zhang, Yulin Chen, Jane Pan +4

    cs.AIcs.CLarXiv:2504.05419v12025
  13. EEG-to-Report: An Annotation and Feature-Text Framework for Training Language Models on Clinical EEG

    Xuan-The Tran, Le Trung Kien Nguyen

    cs.AIarXiv:2608.26153v12026
  14. ICON Decomposition: Multivariate Concept-Level Explanations of Deep Representations for Model Auditing

    Roshan Prakash Rane, Marco Simnacher, Manuel Pfeuffer +7

    cs.LGcs.AIcs.CVarXiv:2608.26083v12026
  15. Which Economic Tasks are Performed with AI? Evidence from Millions of Claude Conversations

    Kunal Handa, Alex Tamkin, Miles McCain +12

    cs.CYcs.AIcs.CLarXiv:2503.04761v12025
  16. PANDA - Prototype-Anchored Alignment for Partially Unpaired Multimodal Learning, with Applications to Alzheimers MRI and TCGA Pathology

    Sheethal Bhat, Mahfuzur Rahman Chowdhury, Paula Andrea Perez-Toro +4

    cs.CVcs.AIarXiv:2608.25970v12026
  17. PromptArmor: Simple yet Effective Prompt Injection Defenses

    Tianneng Shi, Kaijie Zhu, Zhun Wang +13

    cs.CRcs.AIarXiv:2507.15219v12025
  18. WorldSense: Evaluating Real-world Omnimodal Understanding for Multimodal LLMs

    Jack Hong, Shilin Yan, Jiayin Cai +3

    cs.CVcs.AIarXiv:2502.04326v32025
  19. Scientific production in the era of Large Language Models

    Keigo Kusumegi, Xinyu Yang, Paul Ginsparg +3

    cs.DLcs.AIcs.CYarXiv:2601.13187v12026
  20. Entity-level Factual Consistency of Abstractive Text Summarization

    Feng Nan, Ramesh Nallapati, Zhiguo Wang +5

    cs.CLcs.AIarXiv:2102.09130v12021
  21. GRIT: Teaching MLLMs to Think with Images

    Yue Fan, Xuehai He, Diji Yang +6

    cs.CVcs.AIcs.CLarXiv:2505.15879v22025
  22. Time-R1: Post-Training Large Vision Language Model for Temporal Video Grounding

    Ye Wang, Ziheng Wang, Boshen Xu +14

    cs.CVcs.AIcs.CLarXiv:2503.13377v32025
  23. LLMs for Academic Workflows: An Evaluation of Literature Reviews Generated with Short and Long Context Windows of LLMs

    Muhammad Ali Chaudhry, Xinyuan Hao, Haifa Alwahaby

    cs.AIcs.HCcs.IRarXiv:2608.26145v12026
  24. Knowledge Base Completion: Baselines Strike Back

    Rudolf Kadlec, Ondrej Bajgar, Jan Kleindienst

    cs.LGcs.AIarXiv:1705.10744v12017
  25. Multi-Agent Risks from Advanced AI

    Lewis Hammond, Alan Chan, Jesse Clifton +41

    cs.MAcs.AIcs.CYarXiv:2502.14143v12025
  26. DAST: Difficulty-Adaptive Slow-Thinking for Large Reasoning Models

    Yi Shen, Jian Zhang, Jieyun Huang +7

    cs.LGcs.AIarXiv:2503.04472v32025
  27. ConRFT: A Reinforced Fine-tuning Method for VLA Models via Consistency Policy

    Yuhui Chen, Shuai Tian, Shugao Liu +3

    cs.ROcs.AIarXiv:2502.05450v22025
  28. Emotion Concepts and their Function in a Large Language Model

    Nicholas Sofroniew, Isaac Kauvar, William Saunders +13

    cs.AIcs.CLarXiv:2604.07729v12026
  29. Goedel-Prover: A Frontier Model for Open-Source Automated Theorem Proving

    Yong Lin, Shange Tang, Bohan Lyu +8

    cs.LGcs.AIarXiv:2502.07640v32025
  30. Hunyuan3D 2.5: Towards High-Fidelity 3D Assets Generation with Ultimate Details

    Zeqiang Lai, Yunfei Zhao, Haolin Liu +23

    cs.CVcs.AIarXiv:2506.16504v12025
  31. Skywork Open Reasoner 1 Technical Report

    Jujie He, Jiacai Liu, Chris Yuhao Liu +14

    cs.LGcs.AIcs.CLarXiv:2505.22312v22025
  32. Neural Analysis and Synthesis: Reconstructing Speech from Self-Supervised Representations

    Hyeong-Seok Choi, Juheon Lee, Wansoo Kim +3

    cs.SDcs.AIeess.ASarXiv:2110.14513v22021
  33. ProRL: Prolonged Reinforcement Learning Expands Reasoning Boundaries in Large Language Models

    Mingjie Liu, Shizhe Diao, Ximing Lu +5

    cs.CLcs.AIarXiv:2505.24864v12025
  34. AutoSkill: Experience-Driven Lifelong Learning via Skill Self-Evolution

    Yutao Yang, Junsong Li, Qianjun Pan +9

    cs.AIarXiv:2603.01145v22026
  35. DeepEyesV2: Toward Agentic Multimodal Model

    Jack Hong, Chenxiao Zhao, ChengLin Zhu +3

    cs.CVcs.AIarXiv:2511.05271v42025
  36. Learning from Videos for 3D World: Enhancing MLLMs with 3D Vision Geometry Priors

    Duo Zheng, Shijia Huang, Yanyang Li +1

    cs.CVcs.AIarXiv:2505.24625v32025
  37. Conditional Memory via Scalable Lookup: A New Axis of Sparsity for Large Language Models

    Xin Cheng, Rui Tian, Wangding Zeng +18

    cs.CLcs.AIarXiv:2601.07372v22026
  38. SciFive: a text-to-text transformer model for biomedical literature

    Long N. Phan, James T. Anibal, Hieu Tran +4

    cs.CLcs.AIcs.LGarXiv:2106.03598v12021
  39. Entailment as Few-Shot Learner

    Sinong Wang, Han Fang, Madian Khabsa +2

    cs.CLcs.AIarXiv:2104.14690v12021
  40. RWKV-7 "Goose" with Expressive Dynamic State Evolution

    Bo Peng, Ruichong Zhang, Daniel Goldstein +15

    cs.CLcs.AIcs.LGarXiv:2503.14456v22025
  41. Multi-SWE-bench: A Multilingual Benchmark for Issue Resolving

    Daoguang Zan, Zhirong Huang, Wei Liu +16

    cs.SEcs.AIcs.CLarXiv:2504.02605v12025
  42. Trae Agent: An LLM-based Agent for Software Engineering with Test-time Scaling

    Trae Research Team, Pengfei Gao, Zhao Tian +12

    cs.SEcs.AIarXiv:2507.23370v12025
  43. CulturalMenuBench: Probing the Knowledge-Application Gap in Multimodal Culinary Reasoning

    Bo Zeng, Linfeng Gao, Peiqin Lin +9

    cs.AIarXiv:2609.03526v12026
  44. DataSentinel: A Game-Theoretic Detection of Prompt Injection Attacks

    Yupei Liu, Yuqi Jia, Jinyuan Jia +2

    cs.CRcs.AIarXiv:2504.11358v42025
  45. From Language to Programs: Bridging Reinforcement Learning and Maximum Marginal Likelihood

    Kelvin Guu, Panupong Pasupat, Evan Zheran Liu +1

    cs.AIcs.LGstat.MLarXiv:1704.07926v12017
  46. Towards Large Reasoning Models: A Survey of Reinforced Reasoning with Large Language Models

    Fengli Xu, Qianyue Hao, Zefang Zong +17

    cs.AIcs.CLarXiv:2501.09686v32025
  47. Autonomous Drone Racing with Deep Reinforcement Learning

    Yunlong Song, Mats Steinweg, Elia Kaufmann +1

    cs.ROcs.AIarXiv:2103.08624v22021
  48. dLLM-Cache: Accelerating Diffusion Large Language Models with Adaptive Caching

    Zhiyuan Liu, Yicun Yang, Yaojie Zhang +6

    cs.LGcs.AIcs.CLarXiv:2506.06295v32025
  49. Multiple decision trees

    Suk Wah Kwok, Chris Carter

    cs.LGcs.AIstat.MLarXiv:1304.2363v12013
  50. Agent Skills for Large Language Models: Architecture, Acquisition, Security, and the Path Forward

    Renjun Xu, Yang Yan

    cs.MAcs.AIarXiv:2602.12430v42026
  51. Toward Transparent AI: A Survey on Interpreting the Inner Structures of Deep Neural Networks

    Tilman Räuker, Anson Ho, Stephen Casper +1

    cs.LGcs.AIcs.CLarXiv:2207.13243v62022
  52. Composite Monte Carlo Decision Making under High Uncertainty of Novel Coronavirus Epidemic Using Hybridized Deep Learning and Fuzzy Rule Induction

    Simon James Fong, Gloria Li, Nilanjan Dey +2

    cs.AIq-bio.PEarXiv:2003.09868v12020
  53. WASP: Benchmarking Web Agent Security Against Prompt Injection Attacks

    Ivan Evtimov, Arman Zharmagambetov, Aaron Grattafiori +2

    cs.CRcs.AIarXiv:2504.18575v32025
  54. Cosmos-Reason1: From Physical Common Sense To Embodied Reasoning

    NVIDIA, :, Alisson Azzolini +51

    cs.AIcs.CVcs.LGarXiv:2503.15558v32025
  55. On the Effects of Data Scale on UI Control Agents

    Wei Li, William Bishop, Alice Li +4

    cs.AIcs.LGarXiv:2406.03679v62024
  56. Nested Learning: The Illusion of Deep Learning Architectures

    Ali Behrouz, Meisam Razaviyayn, Peilin Zhong +1

    cs.LGcs.AIarXiv:2512.24695v12025
  57. Evaluation and Benchmarking of LLM Agents: A Survey

    Mahmoud Mohammadi, Yipeng Li, Jane Lo +1

    cs.LGcs.AIarXiv:2507.21504v12025
  58. mHC: Manifold-Constrained Hyper-Connections

    Zhenda Xie, Yixuan Wei, Huanqi Cao +17

    cs.CLcs.AIcs.LGarXiv:2512.24880v22025
  59. The Diffusion Duality

    Subham Sekhar Sahoo, Justin Deschenaux, Aaron Gokaslan +3

    cs.LGcs.AIcs.CLarXiv:2506.10892v32025
  60. A survey of agent interoperability protocols: Model Context Protocol (MCP), Agent Communication Protocol (ACP), Agent-to-Agent Protocol (A2A), and Agent Network Protocol (ANP)

    Abul Ehtesham, Aditi Singh, Gaurav Kumar Gupta +1

    cs.AIarXiv:2505.02279v22025