Artificial Intelligence

Papers filed under cs.AI on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.

Search paper metadata (including unsummarized papers)

5,401 to 5,460 of 15,404

  1. The Polar Express: Optimal Matrix Sign Methods and Their Application to the Muon Algorithm

    Noah Amsel, David Persson, Christopher Musco +1

    cs.LGcs.AIcs.CLarXiv:2505.16932v52025
  2. Autonomous discovery of new structure-plausibility laws for explainable and rapid crystal diagnosis and screening

    Zhilong Song, Lixue Cheng

    cond-mat.mtrl-scics.AIarXiv:2609.01209v12026
  3. SoK: When Safe Agents Fail Together: The Security of Multi Agent LLM Systems

    Rui Yang, Junjie Xu, Zhengyu Liu +4

    cs.CRcs.AIarXiv:2609.00595v12026
  4. Ultralytics YOLO Evolution: An Overview of YOLO26, YOLO11, YOLOv8 and YOLOv5 Object Detectors for Computer Vision and Pattern Recognition

    Ranjan Sapkota, Manoj Karkee

    cs.CVcs.AIarXiv:2510.09653v32025
  5. Programming Refusal with Conditional Activation Steering

    Bruce W. Lee, Inkit Padhi, Karthikeyan Natesan Ramamurthy +4

    cs.LGcs.AIcs.CLarXiv:2409.05907v32024
  6. Vision-OPD: Learning to See Fine Details for Multimodal LLMs via On-Policy Self-Distillation

    Qianhao Yuan, Jie Lou, Xing Yu +4

    cs.CVcs.AIcs.CLarXiv:2605.18740v42026
  7. GOOD: A Graph Out-of-Distribution Benchmark

    Shurui Gui, Xiner Li, Limei Wang +1

    cs.LGcs.AIarXiv:2206.08452v22022
  8. Deep Exemplar-based Video Colorization

    Bo Zhang, Mingming He, Jing Liao +4

    cs.CVcs.AIcs.LGarXiv:1906.09909v12019
  9. Act Only When It Pays: Efficient Reinforcement Learning for LLM Reasoning via Selective Rollouts

    Haizhong Zheng, Yang Zhou, Brian R. Bartoldson +4

    cs.AIcs.LGarXiv:2506.02177v12025
  10. TurboQuant: Online Vector Quantization with Near-optimal Distortion Rate

    Amir Zandieh, Majid Daliri, Majid Hadian +1

    cs.LGcs.AIcs.DBarXiv:2504.19874v12025
  11. GTA1: GUI Test-time Scaling Agent

    Yan Yang, Dongxu Li, Yutong Dai +12

    cs.AIarXiv:2507.05791v52025
  12. Calibration is the Bottleneck: An Action-Class Diagnostic of Multi-Turn Tool-Calling

    Kangjia Zhao, Jiajun Li, Haozhan Shen +8

    cs.CLcs.AIarXiv:2609.00949v12026
  13. In-Context Neurofeedback: Can LLMs Control Their Internal Representations through Privileged Access?

    Koshiro Aoki, Ryota Takatsuki, Gouki Minegishi +2

    cs.AIarXiv:2609.00904v12026
  14. Machine Mental Imagery: Empower Multimodal Reasoning with Latent Visual Tokens

    Zeyuan Yang, Xueyang Yu, Delin Chen +2

    cs.CVcs.AIarXiv:2506.17218v12025
  15. A Minimalist Approach to LLM Reasoning: from Rejection Sampling to Reinforce

    Wei Xiong, Jiarui Yao, Yuhui Xu +8

    cs.LGcs.AIcs.CLarXiv:2504.11343v22025
  16. dLLM: Simple Diffusion Language Modeling

    Zhanhui Zhou, Lingjie Chen, Hanghang Tong +1

    cs.CLcs.AIcs.LGarXiv:2602.22661v12026
  17. DR Tulu: Reinforcement Learning with Evolving Rubrics for Deep Research

    Rulin Shao, Akari Asai, Shannon Zejiang Shen +18

    cs.CLcs.AIcs.LGarXiv:2511.19399v32025
  18. FG-CLIP: Fine-Grained Visual and Textual Alignment

    Chunyu Xie, Bin Wang, Fanjing Kong +5

    cs.CVcs.AIarXiv:2505.05071v32025
  19. DNC-IMM: Early Lane-Change Intention Recognition via Neural Calibration Based on Driving Context Information

    Woong-Chan Byun, Seung-Hyun Kong

    cs.ROcs.AIarXiv:2609.01120v12026
  20. Learning to Win by Reading Manuals in a Monte-Carlo Framework

    S. R. K. Branavan, David Silver, Regina Barzilay

    cs.CLcs.AIcs.LGarXiv:1401.5390v12014
  21. Does Math Reasoning Improve General LLM Capabilities? Understanding Transferability of LLM Reasoning

    Maggie Huan, Yuetai Li, Tuney Zheng +6

    cs.AIcs.CLarXiv:2507.00432v22025
  22. Comprehensive Taxonomies of Nature- and Bio-inspired Optimization: Inspiration versus Algorithmic Behavior, Critical Analysis and Recommendations (from 2020 to 2024)

    Daniel Molina, Javier Poyatos, Javier Del Ser +3

    cs.AIarXiv:2002.08136v52020
  23. Measuring what Matters: Construct Validity in Large Language Model Benchmarks

    Andrew M. Bean, Ryan Othniel Kearns, Angelika Romanou +39

    cs.CLcs.AIarXiv:2511.04703v12025
  24. Large Models for Time Series and Spatio-Temporal Data: A Survey and Outlook

    Ming Jin, Yaxuan Kong, Yuxuan Liang +13

    cs.LGcs.AIarXiv:2310.10196v32023
  25. NORA: A Small Open-Sourced Generalist Vision Language Action Model for Embodied Tasks

    Chia-Yu Hung, Qi Sun, Pengfei Hong +5

    cs.ROcs.AIcs.CVarXiv:2504.19854v12025
  26. Exploration in Deep Reinforcement Learning: From Single-Agent to Multiagent Domain

    Jianye Hao, Tianpei Yang, Hongyao Tang +5

    cs.AIcs.LGcs.MAarXiv:2109.06668v62021
  27. Phi-4-reasoning Technical Report

    Marah Abdin, Sahaj Agarwal, Ahmed Awadallah +20

    cs.AIcs.CLarXiv:2504.21318v12025
  28. InfiGUI-R1: Advancing Multimodal GUI Agents from Reactive Actors to Deliberative Reasoners

    Yuhang Liu, Pengxiang Li, Congkai Xie +5

    cs.AIcs.CLarXiv:2504.14239v12025
  29. EvoSkill: Automated Skill Discovery for Multi-Agent Systems

    Salaheddin Alzubi, Noah Provenzano, Jaydon Bingham +2

    cs.AIcs.MAarXiv:2603.02766v12026
  30. DecodingTrust-Agent Platform (DTap): A Controllable and Interactive Red-Teaming Platform for AI Agents

    Zhaorun Chen, Xun Liu, Haibo Tong +14

    cs.AIarXiv:2605.04808v12026
  31. Natural Emergent Misalignment from Reward Hacking in Production RL

    Monte MacDiarmid, Benjamin Wright, Jonathan Uesato +19

    cs.AIcs.SEarXiv:2511.18397v12025
  32. Diffusion LLMs Can Do Faster-Than-AR Inference via Discrete Diffusion Forcing

    Xu Wang, Chenkai Xu, Yijie Jin +3

    cs.LGcs.AIarXiv:2508.09192v12025
  33. PEARL: Path-Entity Aligned Relational Learning with Contextual Subgraphs for Inductive Knowledge Graph Completion

    Yunchi Yang, Longlong Li, Cunquan Qu

    cs.AIarXiv:2609.02216v12026
  34. Agent Gym: A Framework for Continuous Evaluation and Evolution of LLM Agents Through Human-in-the-Loop Feedback

    Pouya Ghiasnezhad Omran, Michael Zimmermann, Duncan Cambridge +2

    cs.AIarXiv:2608.15591v12026
  35. Thinkless: LLM Learns When to Think

    Gongfan Fang, Xinyin Ma, Xinchao Wang

    cs.CLcs.AIarXiv:2505.13379v22025
  36. From LLM Reasoning to Autonomous AI Agents: A Comprehensive Review

    Mohamed Amine Ferrag, Norbert Tihanyi, Merouane Debbah

    cs.AIcs.LGarXiv:2504.19678v22025
  37. Visual prompt engineering for video models

    Robert Geirhos, Yuxuan Li, Thaddäus Wiedemer +7

    cs.CVcs.AIarXiv:2607.25537v12026
  38. WOD-E2E: Waymo Open Dataset for End-to-End Driving in Challenging Long-tail Scenarios

    Runsheng Xu, Hubert Lin, Wonseok Jeon +11

    cs.CVcs.AIarXiv:2510.26125v32025
  39. Rethinking Classifier-Free Guidance in On-Policy Diffusion Distillation

    Bingnan Li, Haozhe Wang, Haozhong Xiong +5

    cs.CVcs.AIcs.LGarXiv:2607.24731v22026
  40. SpatialCLI: Learning to Reason With Spatial Tools, Then Without Them

    Yang Zhou, Zixuan Huang, Sunzhu Li +10

    cs.AIarXiv:2607.27703v22026
  41. Where LLM Agents Fail and How They can Learn From Failures

    Kunlun Zhu, Zijia Liu, Bingxuan Li +15

    cs.AIarXiv:2509.25370v12025
  42. Securing AI Agents with Information-Flow Control

    Manuel Costa, Boris Köpf, Aashish Kolluri +6

    cs.CRcs.AIarXiv:2505.23643v22025
  43. Why human-AI relationships need socioaffective alignment

    Hannah Rose Kirk, Iason Gabriel, Chris Summerfield +2

    cs.HCcs.AIarXiv:2502.02528v12025
  44. The Next Decade in AI: Four Steps Towards Robust Artificial Intelligence

    Gary Marcus

    cs.AIcs.LGarXiv:2002.06177v32020
  45. Are Language Models Actually Useful for Time Series Forecasting?

    Mingtian Tan, Mike A. Merrill, Vinayak Gupta +2

    cs.LGcs.AIarXiv:2406.16964v22024
  46. CEDAR: Automata as Verifiable Interfaces for Language-Guided Embodied Action

    Lekai Chen, Alvaro Velasquez, Ashutosh Trivedi

    cs.AIcs.CLcs.FLarXiv:2608.27797v12026
  47. Performance Foundations of Parallel & Distributed Reasoning Language Models

    Maciej Besta, Leonard Schmidt, Lara Nonino +7

    cs.LGcs.AIcs.DCarXiv:2608.27046v12026
  48. A Lightweight Multimodal Vision-Language Framework for Early-Stage Anatomical Green Fruit Classification in Commercial Orchards

    Ranjan Sapkota, William Bu, Chen Chen +2

    cs.CVcs.AIarXiv:2608.24935v12026
  49. Six misconceptions about large language models: A minimal model and diagnostic taxonomy

    Zhicheng Lin

    cs.CYcs.AIarXiv:2608.20421v12026
  50. Which Negatives Matter? Ask Your Text Encoder: Adaptive Similarity Margins for Dense-Caption Retrieval

    Haoyue Liu, Ye Chen, Zhichao Wang +1

    cs.AIcs.LGarXiv:2608.18521v12026
  51. Beyond Suspicious Steps: Ontological Trust in Long-Horizon Agents

    An He, Yao Wang, Haibin Zhang

    cs.AIarXiv:2608.17718v12026
  52. When Do Explanations Help In-Context Learning? A Comparative Study of Natural Language Explanation Types and Faithfulness

    Mahdi Dhaini, Adam Dejl, Juraj Vladika +3

    cs.CLcs.AIarXiv:2608.16627v12026
  53. Intent-Driven Situation Tracking for User-Centric Multi-Turn Agents

    Meiling Tao, Yiling Tao, Peng Wang

    cs.AIarXiv:2608.15755v12026
  54. Engineering Reliable Coding Agents: Evaluating and Operating the System Around the Model

    Stephanie Jarmak

    cs.SEcs.AIarXiv:2608.13867v12026
  55. When the Algorithm Becomes the Brand Crisis: A Sociotechnical Theory of Distributed Responsibility and Accountable Transparency

    Mohammad Saleh Torkestani, Taha Mansouri

    cs.AIarXiv:2609.00510v12026
  56. Metacognition in LLMs: Foundations, Progress, and Opportunities

    Gabrielle Kaili-May Liu, Areeb Gani, Jacqueline Lu +3

    cs.CLcs.AIarXiv:2607.11881v12026
  57. Automated Vulnerability Injection in Smart Contracts Using Large Language Models

    Luca Migliaccio, Roberto Natella, Naghmeh Ivaki +2

    cs.SEcs.AIcs.CRarXiv:2609.02624v12026
    Summaries:한국어
  58. CordisBench: Can Language Models Reason About Component Lifecycles in Dynamic Agent Harnesses?

    Damien Sileo, Dimitri Kachler

    cs.CLcs.AIarXiv:2609.01600v12026
  59. The Hitchhiker's Guide to Agentic AI: From Foundations to Systems

    Haggai Roitman

    cs.AIcs.CLcs.IRarXiv:2606.24937v22026
    Summaries:한국어
  60. Agentic Large Language Models, a survey

    Aske Plaat, Max van Duijn, Niki van Stein +3

    cs.AIcs.CLcs.LGarXiv:2503.23037v32025