Artificial Intelligence
Papers filed under cs.AI on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.
Search paper metadata (including unsummarized papers)
5,401 to 5,460 of 15,404
The Polar Express: Optimal Matrix Sign Methods and Their Application to the Muon Algorithm
Noah Amsel, David Persson, Christopher Musco +1
cs.LGcs.AIcs.CLarXiv:2505.16932v52025Autonomous discovery of new structure-plausibility laws for explainable and rapid crystal diagnosis and screening
Zhilong Song, Lixue Cheng
cond-mat.mtrl-scics.AIarXiv:2609.01209v12026SoK: When Safe Agents Fail Together: The Security of Multi Agent LLM Systems
Rui Yang, Junjie Xu, Zhengyu Liu +4
cs.CRcs.AIarXiv:2609.00595v12026Ultralytics YOLO Evolution: An Overview of YOLO26, YOLO11, YOLOv8 and YOLOv5 Object Detectors for Computer Vision and Pattern Recognition
Ranjan Sapkota, Manoj Karkee
cs.CVcs.AIarXiv:2510.09653v32025Programming Refusal with Conditional Activation Steering
Bruce W. Lee, Inkit Padhi, Karthikeyan Natesan Ramamurthy +4
cs.LGcs.AIcs.CLarXiv:2409.05907v32024Vision-OPD: Learning to See Fine Details for Multimodal LLMs via On-Policy Self-Distillation
Qianhao Yuan, Jie Lou, Xing Yu +4
cs.CVcs.AIcs.CLarXiv:2605.18740v42026GOOD: A Graph Out-of-Distribution Benchmark
Shurui Gui, Xiner Li, Limei Wang +1
cs.LGcs.AIarXiv:2206.08452v22022Deep Exemplar-based Video Colorization
Bo Zhang, Mingming He, Jing Liao +4
cs.CVcs.AIcs.LGarXiv:1906.09909v12019Act Only When It Pays: Efficient Reinforcement Learning for LLM Reasoning via Selective Rollouts
Haizhong Zheng, Yang Zhou, Brian R. Bartoldson +4
cs.AIcs.LGarXiv:2506.02177v12025TurboQuant: Online Vector Quantization with Near-optimal Distortion Rate
Amir Zandieh, Majid Daliri, Majid Hadian +1
cs.LGcs.AIcs.DBarXiv:2504.19874v12025GTA1: GUI Test-time Scaling Agent
Yan Yang, Dongxu Li, Yutong Dai +12
cs.AIarXiv:2507.05791v52025Calibration is the Bottleneck: An Action-Class Diagnostic of Multi-Turn Tool-Calling
Kangjia Zhao, Jiajun Li, Haozhan Shen +8
cs.CLcs.AIarXiv:2609.00949v12026In-Context Neurofeedback: Can LLMs Control Their Internal Representations through Privileged Access?
Koshiro Aoki, Ryota Takatsuki, Gouki Minegishi +2
cs.AIarXiv:2609.00904v12026Machine Mental Imagery: Empower Multimodal Reasoning with Latent Visual Tokens
Zeyuan Yang, Xueyang Yu, Delin Chen +2
cs.CVcs.AIarXiv:2506.17218v12025A Minimalist Approach to LLM Reasoning: from Rejection Sampling to Reinforce
Wei Xiong, Jiarui Yao, Yuhui Xu +8
cs.LGcs.AIcs.CLarXiv:2504.11343v22025dLLM: Simple Diffusion Language Modeling
Zhanhui Zhou, Lingjie Chen, Hanghang Tong +1
cs.CLcs.AIcs.LGarXiv:2602.22661v12026DR Tulu: Reinforcement Learning with Evolving Rubrics for Deep Research
Rulin Shao, Akari Asai, Shannon Zejiang Shen +18
cs.CLcs.AIcs.LGarXiv:2511.19399v32025FG-CLIP: Fine-Grained Visual and Textual Alignment
Chunyu Xie, Bin Wang, Fanjing Kong +5
cs.CVcs.AIarXiv:2505.05071v32025DNC-IMM: Early Lane-Change Intention Recognition via Neural Calibration Based on Driving Context Information
Woong-Chan Byun, Seung-Hyun Kong
cs.ROcs.AIarXiv:2609.01120v12026Learning to Win by Reading Manuals in a Monte-Carlo Framework
S. R. K. Branavan, David Silver, Regina Barzilay
cs.CLcs.AIcs.LGarXiv:1401.5390v12014Does Math Reasoning Improve General LLM Capabilities? Understanding Transferability of LLM Reasoning
Maggie Huan, Yuetai Li, Tuney Zheng +6
cs.AIcs.CLarXiv:2507.00432v22025Comprehensive Taxonomies of Nature- and Bio-inspired Optimization: Inspiration versus Algorithmic Behavior, Critical Analysis and Recommendations (from 2020 to 2024)
Daniel Molina, Javier Poyatos, Javier Del Ser +3
cs.AIarXiv:2002.08136v52020Measuring what Matters: Construct Validity in Large Language Model Benchmarks
Andrew M. Bean, Ryan Othniel Kearns, Angelika Romanou +39
cs.CLcs.AIarXiv:2511.04703v12025Large Models for Time Series and Spatio-Temporal Data: A Survey and Outlook
Ming Jin, Yaxuan Kong, Yuxuan Liang +13
cs.LGcs.AIarXiv:2310.10196v32023NORA: A Small Open-Sourced Generalist Vision Language Action Model for Embodied Tasks
Chia-Yu Hung, Qi Sun, Pengfei Hong +5
cs.ROcs.AIcs.CVarXiv:2504.19854v12025Exploration in Deep Reinforcement Learning: From Single-Agent to Multiagent Domain
Jianye Hao, Tianpei Yang, Hongyao Tang +5
cs.AIcs.LGcs.MAarXiv:2109.06668v62021Phi-4-reasoning Technical Report
Marah Abdin, Sahaj Agarwal, Ahmed Awadallah +20
cs.AIcs.CLarXiv:2504.21318v12025InfiGUI-R1: Advancing Multimodal GUI Agents from Reactive Actors to Deliberative Reasoners
Yuhang Liu, Pengxiang Li, Congkai Xie +5
cs.AIcs.CLarXiv:2504.14239v12025EvoSkill: Automated Skill Discovery for Multi-Agent Systems
Salaheddin Alzubi, Noah Provenzano, Jaydon Bingham +2
cs.AIcs.MAarXiv:2603.02766v12026DecodingTrust-Agent Platform (DTap): A Controllable and Interactive Red-Teaming Platform for AI Agents
Zhaorun Chen, Xun Liu, Haibo Tong +14
cs.AIarXiv:2605.04808v12026Natural Emergent Misalignment from Reward Hacking in Production RL
Monte MacDiarmid, Benjamin Wright, Jonathan Uesato +19
cs.AIcs.SEarXiv:2511.18397v12025Diffusion LLMs Can Do Faster-Than-AR Inference via Discrete Diffusion Forcing
Xu Wang, Chenkai Xu, Yijie Jin +3
cs.LGcs.AIarXiv:2508.09192v12025PEARL: Path-Entity Aligned Relational Learning with Contextual Subgraphs for Inductive Knowledge Graph Completion
Yunchi Yang, Longlong Li, Cunquan Qu
cs.AIarXiv:2609.02216v12026Agent Gym: A Framework for Continuous Evaluation and Evolution of LLM Agents Through Human-in-the-Loop Feedback
Pouya Ghiasnezhad Omran, Michael Zimmermann, Duncan Cambridge +2
cs.AIarXiv:2608.15591v12026Thinkless: LLM Learns When to Think
Gongfan Fang, Xinyin Ma, Xinchao Wang
cs.CLcs.AIarXiv:2505.13379v22025From LLM Reasoning to Autonomous AI Agents: A Comprehensive Review
Mohamed Amine Ferrag, Norbert Tihanyi, Merouane Debbah
cs.AIcs.LGarXiv:2504.19678v22025Visual prompt engineering for video models
Robert Geirhos, Yuxuan Li, Thaddäus Wiedemer +7
cs.CVcs.AIarXiv:2607.25537v12026WOD-E2E: Waymo Open Dataset for End-to-End Driving in Challenging Long-tail Scenarios
Runsheng Xu, Hubert Lin, Wonseok Jeon +11
cs.CVcs.AIarXiv:2510.26125v32025Rethinking Classifier-Free Guidance in On-Policy Diffusion Distillation
Bingnan Li, Haozhe Wang, Haozhong Xiong +5
cs.CVcs.AIcs.LGarXiv:2607.24731v22026SpatialCLI: Learning to Reason With Spatial Tools, Then Without Them
Yang Zhou, Zixuan Huang, Sunzhu Li +10
cs.AIarXiv:2607.27703v22026Where LLM Agents Fail and How They can Learn From Failures
Kunlun Zhu, Zijia Liu, Bingxuan Li +15
cs.AIarXiv:2509.25370v12025Securing AI Agents with Information-Flow Control
Manuel Costa, Boris Köpf, Aashish Kolluri +6
cs.CRcs.AIarXiv:2505.23643v22025Why human-AI relationships need socioaffective alignment
Hannah Rose Kirk, Iason Gabriel, Chris Summerfield +2
cs.HCcs.AIarXiv:2502.02528v12025The Next Decade in AI: Four Steps Towards Robust Artificial Intelligence
Gary Marcus
cs.AIcs.LGarXiv:2002.06177v32020Are Language Models Actually Useful for Time Series Forecasting?
Mingtian Tan, Mike A. Merrill, Vinayak Gupta +2
cs.LGcs.AIarXiv:2406.16964v22024CEDAR: Automata as Verifiable Interfaces for Language-Guided Embodied Action
Lekai Chen, Alvaro Velasquez, Ashutosh Trivedi
cs.AIcs.CLcs.FLarXiv:2608.27797v12026Performance Foundations of Parallel & Distributed Reasoning Language Models
Maciej Besta, Leonard Schmidt, Lara Nonino +7
cs.LGcs.AIcs.DCarXiv:2608.27046v12026A Lightweight Multimodal Vision-Language Framework for Early-Stage Anatomical Green Fruit Classification in Commercial Orchards
Ranjan Sapkota, William Bu, Chen Chen +2
cs.CVcs.AIarXiv:2608.24935v12026Six misconceptions about large language models: A minimal model and diagnostic taxonomy
Zhicheng Lin
cs.CYcs.AIarXiv:2608.20421v12026Which Negatives Matter? Ask Your Text Encoder: Adaptive Similarity Margins for Dense-Caption Retrieval
Haoyue Liu, Ye Chen, Zhichao Wang +1
cs.AIcs.LGarXiv:2608.18521v12026Beyond Suspicious Steps: Ontological Trust in Long-Horizon Agents
An He, Yao Wang, Haibin Zhang
cs.AIarXiv:2608.17718v12026When Do Explanations Help In-Context Learning? A Comparative Study of Natural Language Explanation Types and Faithfulness
Mahdi Dhaini, Adam Dejl, Juraj Vladika +3
cs.CLcs.AIarXiv:2608.16627v12026Intent-Driven Situation Tracking for User-Centric Multi-Turn Agents
Meiling Tao, Yiling Tao, Peng Wang
cs.AIarXiv:2608.15755v12026Engineering Reliable Coding Agents: Evaluating and Operating the System Around the Model
Stephanie Jarmak
cs.SEcs.AIarXiv:2608.13867v12026When the Algorithm Becomes the Brand Crisis: A Sociotechnical Theory of Distributed Responsibility and Accountable Transparency
Mohammad Saleh Torkestani, Taha Mansouri
cs.AIarXiv:2609.00510v12026Metacognition in LLMs: Foundations, Progress, and Opportunities
Gabrielle Kaili-May Liu, Areeb Gani, Jacqueline Lu +3
cs.CLcs.AIarXiv:2607.11881v12026Automated Vulnerability Injection in Smart Contracts Using Large Language Models
Luca Migliaccio, Roberto Natella, Naghmeh Ivaki +2
cs.SEcs.AIcs.CRarXiv:2609.02624v12026Summaries:한국어CordisBench: Can Language Models Reason About Component Lifecycles in Dynamic Agent Harnesses?
Damien Sileo, Dimitri Kachler
cs.CLcs.AIarXiv:2609.01600v12026The Hitchhiker's Guide to Agentic AI: From Foundations to Systems
Haggai Roitman
cs.AIcs.CLcs.IRarXiv:2606.24937v22026Summaries:한국어Agentic Large Language Models, a survey
Aske Plaat, Max van Duijn, Niki van Stein +3
cs.AIcs.CLcs.LGarXiv:2503.23037v32025