Artificial Intelligence

Papers filed under cs.AI on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.

Search paper metadata (including unsummarized papers)

15,421 to 15,437 of 15,437

  1. Are You Sure You're Sure? On the Impact of Instruction Tuning on Confidence and Lexical Diversity

    Irina Proskurina, Mayank Kumar, Oyindolapo O. Komolafe

    cs.CLcs.AIarXiv:2608.13430v12026
  2. Forecast Collapse in Time-Series Foundation Models

    Shu Wan, Miles Ma, Hank Zhu +4

    cs.LGcs.AIcs.CEarXiv:2608.14106v12026
  3. Activity Frames: Deterministic Screen-Activity Compilation for Agent Memory and Replay

    Nossa Iyamu

    cs.AIarXiv:2608.05784v12026
  4. MobileMem: Learning from a Year of Mobile Experiences

    Xinle Deng, Yida Xue, Xiangyuan Ru +14

    cs.AIcs.CLcs.LGarXiv:2608.13606v12026
  5. Long-Horizon-Terminal-Bench: Testing the Limits of Agents on Long-Horizon Terminal Tasks with Dense Reward-Based Grading

    Zongxia Li, Zhongzhi Li, Yucheng Shi +10

    cs.AIarXiv:2607.08964v22026
  6. SPIEval: Evaluating Large Language Models as Mobile Assistants over Scattered Personal Information

    Junjie Ye, Zhuohui Sheng, Shaofan Liu +12

    cs.CLcs.AIarXiv:2608.10692v12026
  7. SkillOpt-Lite: Better and Faster Agent Self-evolution via One Line of Vibe

    Yifei Shen, Bo Li, Xinjie Zhang

    cs.SEcs.AIcs.LGarXiv:2607.03451v12026
  8. $A^2E$ : An End-to-End Agent Auditing Engine

    Haoning Wang, Mingxun Zhang, Chenyue Yu +4

    cs.AIarXiv:2608.07346v22026
  9. Gemma 4 Technical Report

    Gemma Team, Sherif El Abd, Vaibhav Aggarwal +320

    cs.CLcs.AIarXiv:2607.02770v22026
  10. ComBodied Agents: a New Paradigm of Human-Centric Agentic AI

    Qianggang Ding, Xingyao Wang, Rui Feng +19

    cs.AIarXiv:2608.10915v22026
  11. Persistent Recursive Worlds Enable Autonomous Software Evolution

    Beichen Huang, Zhenyu Liang, Bowen Zheng +1

    cs.SEcs.AIcs.MAarXiv:2608.10450v22026
  12. AgenticDataBench: A Comprehensive Benchmark for Data Agents

    Zhaoyan Sun, Shan Zhong, Daizhou Wen +10

    cs.DBcs.AIcs.CLarXiv:2607.01647v12026
  13. The State-Prediction Separation Hypothesis

    Giovanni Monea, Nathan Godey, Kianté Brantley +1

    cs.CLcs.AIcs.LGarXiv:2607.01218v12026
  14. SkillCoach: Self-Evolving Rubrics for Evaluating and Enhancing Agentic Skill-Use

    Jiayin Zhu, Kelong Mao, Yudong Guo +4

    cs.AIcs.CLarXiv:2607.01874v12026
  15. AutoMem: Automated Learning of Memory as a Cognitive Skill

    Shengguang Wu, Hao Zhu, Yuhui Zhang +2

    cs.AIcs.CLcs.MAarXiv:2607.01224v12026
  16. Cross-Domain Generalization Failure in Lightweight Intrusion Detection Models for IIoT Networks

    MD Azizul Hakim, Md Shihab Uddin, Talha Ibne Anis

    cs.CRcs.AIarXiv:2607.00553v12026
  17. SWE-Together: Evaluating Coding Agents in Interactive User Sessions

    Yifan Wu, Zhuokai Zhao, Songlin Li +8

    cs.SEcs.AIarXiv:2606.29957v12026