Artificial Intelligence
Papers filed under cs.AI on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.
Search paper metadata (including unsummarized papers)
15,001 to 15,060 of 15,326
Towards Autonomous and Auditable Medical Imaging Model Development
Shengyuan Liu, Jia-Xuan Jiang, Boyun Zheng +8
cs.CVcs.AIarXiv:2607.10522v12026LightMem-Ego: Your AI Memory for Everyday Life
Yijun Chen, Boyi Xiao, Yixian Zhao +10
cs.CLcs.AIcs.CVarXiv:2607.11487v12026Are LLMs Ready for Scientific Discovery? A Capability-Oriented Benchmark for AI Scientists
Chuhan Shi, Xiaoquan Ren, Sicheng Song +3
cs.AIarXiv:2607.11079v12026See like a Robot: Robot-Centric Pointmaps for Vision-Language-Action Models
Byungkun Lee, Dongyoon Hwang, Dongjin Kim +3
cs.ROcs.AIarXiv:2607.11498v12026SVR-R1: Bootstrapping Multi-modal Reasoning with Self-verification in Reinforcement Learning
Mingyuan Wu, Jingcheng Yang, Shengyi Qian +11
cs.AIarXiv:2607.10966v12026ABot-N1: Toward a General Visual Language Navigation Foundation Model
Ruiyan Gong, Yingnan Guo, Junjun Hu +43
cs.CVcs.AIcs.ROarXiv:2607.10383v32026RAGU: A Multi-Step GraphRAG Engine with a Compact Domain-Adapted LLM
Mikhail Komarov, Ivan Bondarenko, Stanislav Shtuka +5
cs.CLcs.AIarXiv:2607.11683v12026From Controlled to the Wild: Evaluation of Pentesting Agents for the Real-World
Pedro Conde, Henrique Branquinho, Valerio Mazzone +3
cs.AIcs.CRarXiv:2605.10834v32026PalmClaw: A Native On-Device Agent Framework for Mobile Phones
Hongru Cai, Yongqi Li, Ran Wei +1
cs.CLcs.AIarXiv:2607.13027v12026Hallo4D: Multi-Modal Hallucination Mitigation for Consistent Spatio-Temporal Generation
Hongbo Wang, Huaibo Huang, Jie Cao +3
cs.CVcs.AIarXiv:2607.12752v22026Chat2Scenic: An Iterative RAG-Based Framework for Scenario Generation in Autonomous Driving
Yuan Gao, Wenting Miao, Mattia Piccinini +3
cs.AIcs.ROarXiv:2607.14387v12026Smarter and Cheaper at Once: Byte-Exact KV-Cache Grafting Turns a Frozen Small Model into a Verified-Knowledge Flywheel
Sietse Schelpe
cs.CLcs.AIcs.LGarXiv:2607.14431v12026SearchOS-V1: Towards Robust Open-Domain Information-Seeking Agent Collaboration
Yuyao Zhang, Junjie Gao, Zhengxian Wu +11
cs.AIcs.IRarXiv:2607.15257v12026Understanding Reasoning from Pretraining to Post-Training
Jingyan Shen, Ang Li, Salman Rahman +4
cs.LGcs.AIcs.CLarXiv:2607.16097v22026Beyond Entropy: Correctness-Aware Advantage Shaping via Contrastive Policy Optimization
Weiwen Xu, Jia Liu, Hou Pong Chan +4
cs.LGcs.AIcs.CLarXiv:2607.14614v12026When Does Muon Help Agentic Reinforcement Learning?
Kai Ruan, Jinghao Lin, Zihe Huang +4
cs.LGcs.AIarXiv:2607.16169v42026DSWorld: A Data Science World Model for Efficient Autonomous Agents
Zherui Yang, Fan Liu, Hao Liu
cs.AIarXiv:2607.15901v12026Beyond Success Rate: Cost-Aware Evaluation of Offensive and Defensive Security Agents
Paul Kassianik, Blaine Nelson, Yaron Singer
cs.CRcs.AIarXiv:2607.15263v32026ResearchStudio-Idea: An Evidence-Grounded Research-Ideation Skill Suite from ML Conference Outcomes
Qihao Zhao, Yangyu Huang, Yalun Dai +8
cs.AIarXiv:2607.04439v12026PixCon: Clean-Positive Contrastive Learning for Foundation-Model Semi-Supervised Segmentation
Ebenezer Tarubinga
cs.CVcs.AIcs.LGarXiv:2607.03068v12026RL-Index: Reinforcement Learning for Retrieval Index Reasoning
Yongjia Lei, Nedim Lipka, Zhisheng Qi +7
cs.IRcs.AIcs.LGarXiv:2606.16316v22026When Lower Privileges Suffice: Investigating Over-Privileged Tool Selection in LLM Agents
Kaiyue Yang, Yuyan Bu, Jingwei Yi +5
cs.SEcs.AIcs.CLarXiv:2606.20023v22026Demystifying Training-Time Augmentation for Data-Constrained Language Model Pretraining
Michael K. Chen, Xikun Zhang, Fan Bai +2
cs.LGcs.AIcs.CLarXiv:2606.16246v22026Improving Text-to-Music Generation with Human Preference Rewards
Yonghyun Kim, Junwon Lee, Haiwen Xia +2
cs.SDcs.AIcs.LGarXiv:2606.21670v12026BioInsight: Multi-Agent Orchestration for Interactive Biomedical Knowledge Discovery
Jieyi Wang, Bingxuan Li, Nanyi Jiang +9
cs.AIarXiv:2606.20997v22026TheoremGraph: Bridging Formal and Informal Mathematics
Simon Kurgan, Evan Wang, Eric Leonen +6
cs.IRcs.AImath.HOarXiv:2606.25363v12026ProMSA:Progressive Multimodal Search Agents for Knowledge-Based Visual Question Answering
ZhengXian Wu, Hangrui Xu, Kai Shi +8
cs.CVcs.AIarXiv:2606.27974v12026ResearchStudio-Reel: Automate the Last Mile of Research from Paper to Poster, Video, and Blog
Lingao Xiao, Yalun Dai, Yangyu Huang +17
cs.CVcs.AIcs.HCarXiv:2607.04438v22026Remember When It Matters: Proactive Memory Agent for Long-Horizon Agents
Yifan Wu, Lizhu Zhang, Yuhang Zhou +5
cs.AIcs.CLarXiv:2607.08716v12026Towards Mechanistically Understanding Why Memorized Knowledge Fails to Generalize in Large Language Model Finetuning
Lu Dai, Ziyang Rao, Yili Wang +3
cs.AIcs.CLarXiv:2607.08393v12026A Sovereign, Open-Source Foundation Model for German and English
Soofi-Team, :, Benedikt Droste +30
cs.CLcs.AIcs.LGarXiv:2607.09424v32026OvisOCR2 Technical Report
Shiyin Lu, Yinglun Li, Yu Xia +10
cs.CVcs.AIarXiv:2607.13639v12026RxBrain: Embodied Cognition Foundation Model with Joint Language-Visual Reasoning and Imagination
Haotian Liang, Mingkang Chen, Yufei Huang +27
cs.AIcs.ROarXiv:2607.14187v12026Vesta: A Generalist Embodied Reasoning Model
Johan Bjorck, Zhiqi Li, Yunze Man +29
cs.ROcs.AIarXiv:2606.20905v12026Look Light, Think Heavy: What Multimodal Chain-of-Thought Reasoning Can and Cannot Do
Zhuoran Jin, Kejian Zhu, Hongbang Yuan +5
cs.CLcs.AIcs.CVarXiv:2606.22565v12026IV-CoT: Implicit Visual Chain-of-Thought for Structure-Aware Text-to-Image Generation
Zixuan Li, Haokun Lin, Yicheng Xiao +10
cs.CVcs.AIarXiv:2606.24849v12026Managing Procedural Memory in LLM Agents: Control, Adaptation, and Evaluation
Julia Belikova, Rauf Parchiev, Evgeny Egorov +4
cs.AIcs.CLcs.SEarXiv:2606.23127v12026TUA-Bench: A Benchmark for General-Purpose Terminal-Use Agents
Shoufa Chen, Luyuan Wang, Xuan Yang +7
cs.SEcs.AIarXiv:2606.28480v12026Cognitive Episodes in LLM Reasoning Traces Enable Interpretable Human Item Difficulty Prediction
Chenguang Wang, Ming Li, Xinyue Zeng +4
cs.CLcs.AIcs.CYarXiv:2606.28186v32026Dockerless: Environment-Free Program Verifier for Coding Agents
Wenhao Zeng, Yuling Shi, Xiaodong Gu +10
cs.SEcs.AIarXiv:2606.28436v12026SPEAR: A Simulator for Photorealistic Embodied AI Research
Mike Roberts, Renhan Wang, Rushikesh Zawar +10
cs.CVcs.AIcs.GRarXiv:2607.06701v12026Wan-Streamer v0.2: Higher Resolution, Same Latency
Lianghua Huang, Zhi-Fan Wu, Yupeng Shi +23
cs.CVcs.AIcs.GRarXiv:2607.04443v32026Building to the Test: Coding Agents Deliver What You Check, Not What You Requested
Yanuo Ma, Ben Kereopa-Yorke, Ben Schultz
cs.SEcs.AIarXiv:2606.28430v12026Measuring the Gap Between Human and LLM Research Ideas
Ziyu Chen, Yilun Zhao, Arman Cohan
cs.CLcs.AIarXiv:2607.01233v12026When More Sampling Hurts: The Modal Ceiling and Correlation Ceiling of Test-Time Scaling
Yong Yi Bay, Kathleen A. Yearick
cs.LGcs.AIcs.CLarXiv:2606.28661v12026Parameter-Efficient Quantum-Inspired Fast Weight Programmers for Traffic-Matrix Forecasting
Kuo-Chung Peng, Jiun-Cheng Jiang, Chun-Hua Lin +3
quant-phcs.AIcs.LGarXiv:2606.27821v12026ENTRAP-VL: A Taxonomic Probe for Dual Contextual Entrainment in Vision-Language Models
Karan Goyal, Afreen Hossain, Debojyoti Das +1
cs.CVcs.AIcs.CLarXiv:2607.20092v12026Measuring Cross-Task Behavioral Consistency in Language Model Agents
Amritesh Banerjee, Pranil Raichura
cs.AIarXiv:2608.13598v12026No Universal Signal Predicts Sample-Level LLM Regression under Version Updates
Jia Sheng, Yiwei Lu
cs.AIcs.CLcs.LGarXiv:2608.13607v12026Optimal Power Allocation and AI Receiver Design for Superimposed DMRS and Data Transmission
Sha Hu, Zhongwang Fu
cs.ITcs.AIarXiv:2608.13809v12026Does ISO-Grounded NFR Specification Improve LLM Code Generation? A Comparison of Rich and Structured Interventions against a Natural-Language Baseline
Joào Pedro Monteiro Pereira, Vinicius Cardoso Garcia
cs.SEcs.AIcs.LGarXiv:2608.13742v12026SDO: Subspace Deconflicting Operator for Multi-Adapter Composition
Zhongsheng Wang, Zhedong Lin, Qian Liu +2
cs.AIarXiv:2608.13820v12026Explanation Multiplicity: Circuit-Level Interpretability Evidence Does Not Survive Defensible Analytic Variation
Ajay Pravin Mahale
cs.AIarXiv:2608.13754v12026Ontology-Grounded Project Memory for Coding Agents
James Adam
cs.AIcs.SEarXiv:2608.13662v12026A Calibrated Test of Internal Action Maps: State Signals Without Global Affine Closure
Dekun Yang
cs.AIcs.CLcs.LGarXiv:2608.13626v12026Computational Humor with Multimodal LLMs: Methods, Datasets, Evaluation, and Challenges
Tuo Liang, Zhe Hu, Disheng Liu +2
cs.CLcs.AIcs.MMarXiv:2607.19011v12026Self-State Attacks on Self-Hosted AI Agents: How Far Can OS Defenses Go?
Yimeng Chen, Nathanaël Denis, Roberto Di Pietro +1
cs.CRcs.AIcs.CLarXiv:2607.17986v12026Differentiable Logic Gate Networks for Low-Latency EEG Classification on Edge Devices
Shyamal Y. Dharia, Stephen D. Smith, Camilo E. Valderrama
cs.LGcs.AIarXiv:2607.18149v12026AI Tour Meeting: Group Travel Planning by LLM Agents
Daisuke Kikuta
cs.AIcs.CLcs.MAarXiv:2607.18806v12026Train the Model, Not the Reader: Decodability Supervision for Verifiable Activation Explanations
Hiskias Dingeto
cs.AIcs.CLarXiv:2607.20379v12026