Artificial Intelligence
Papers filed under cs.AI on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.
Search paper metadata (including unsummarized papers)
541 to 600 of 15,291
Not All Prompts Are Equal: Exploration-Guided Prompt Scaffolding for Multimodal Reinforcement Post-Training
Yuanhao Yue, Qianli Ma, Chengyu Wang +3
cs.LGcs.AIcs.CLarXiv:2609.15051v12026SkillMOO: Multi-Objective Optimization of Agent Skills for Software Engineering
Jingzhi Gong, Ruizhen Gu, Zhiwei Fei +7
cs.SEcs.AIarXiv:2604.09297v32026Improving Coherence and Consistency in Neural Sequence Models with Dual-System, Neuro-Symbolic Reasoning
Maxwell Nye, Michael Henry Tessler, Joshua B. Tenenbaum +1
cs.AIcs.CLcs.LGarXiv:2107.02794v22021"Death" of a Chatbot: Investigating and Designing Toward Psychologically Safe Endings for Human-AI Relationships
Rachel Poonsiriwong, Chayapatr Archiwaranguprok, Pat Pataranutaporn
cs.HCcs.AIarXiv:2602.07193v22026EVA: Aligning Video World Models with Executable Robot Actions via Inverse Dynamics Rewards
Ruixiang Wang, Qingming Liu, Yueci Deng +3
cs.ROcs.AIarXiv:2603.17808v22026Pick Your Poison: Learning to Select Poison Sets for Stronger LLM Backdoor Attacks
Aashiq Muhamed, Mona T. Diab, Virginia Smith +2
cs.LGcs.AIcs.CLarXiv:2609.15029v12026MemoRAG: Boosting Long Context Processing with Global Memory-Enhanced Retrieval Augmentation
Hongjin Qian, Zheng Liu, Peitian Zhang +4
cs.CLcs.AIarXiv:2409.05591v32024How Lossless Is Lossless Speculative Decoding? The Role of Numerical Precision in Orthrus
Ilya Koziev, Leonid Sinev, Ivan Oseledets
cs.CLcs.AIarXiv:2609.15504v12026Beyond Quacking: Deep Integration of Language Models and RAG into DuckDB
Anas Dorbani, Sunny Yasser, Jimmy Lin +1
cs.DBcs.AIcs.IRarXiv:2504.01157v12025HazardAuditor: From Executable Threats to Safer Computer-Use Agents
Yunhao Feng, Ruixiao Lin, Ming Wen +5
cs.AIarXiv:2609.15134v12026CORE: Simple and Effective Session-based Recommendation within Consistent Representation Space
Yupeng Hou, Binbin Hu, Zhiqiang Zhang +1
cs.IRcs.AIarXiv:2204.11067v12022AdaRubric: Task-Adaptive Rubrics for Reliable LLM Agent Evaluation and Reward Learning
Liang Ding
cs.AIcs.CLarXiv:2603.21362v32026PhysStream: Streaming Physics-Grounded Video Generation with Structured Scene Memory and Fine-Grained Motion Control
Chuhao Chen, Peter Wonka, Chaoyang Wang +4
cs.CVcs.AIcs.GRarXiv:2609.17521v12026FlexiTac: A Low-Cost, Open-Source, Scalable Tactile Sensing Solution for Robotic Systems
Binghao Huang, Yunzhu Li
cs.ROcs.AIcs.LGarXiv:2604.28156v12026KaiNinja: Extending Native 3D Generators to the Part Level
Ruihan Yu, Lian Fu, Muyao Niu +9
cs.GRcs.AIcs.CVarXiv:2609.15659v22026How Coding Agents Fail Their Users: A Large-Scale Analysis of Developer-Agent Misalignment in 20,574 Real-World Sessions
Ningzhi Tang, Chaoran Chen, Gelei Xu +5
cs.SEcs.AIcs.HCarXiv:2605.29442v22026MS2: Multi-Document Summarization of Medical Studies
Jay DeYoung, Iz Beltagy, Madeleine van Zuylen +2
cs.CLcs.AIcs.LGarXiv:2104.06486v32021Maximum Entropy Gain Exploration for Long Horizon Multi-goal Reinforcement Learning
Silviu Pitis, Harris Chan, Stephen Zhao +2
cs.LGcs.AIcs.ROarXiv:2007.02832v12020Reflective Planning: Vision-Language Models for Multi-Stage Long-Horizon Robotic Manipulation
Yunhai Feng, Jiaming Han, Zhuoran Yang +3
cs.ROcs.AIcs.LGarXiv:2502.16707v12025TJ4DRadSet: A 4D Radar Dataset for Autonomous Driving
Lianqing Zheng, Zhixiong Ma, Xichan Zhu +9
cs.CVcs.AIarXiv:2204.13483v32022RSIAgent: Autonomous Exploration for Recursive Self-improvement in New Environments
Sibo Zhu, Shicheng Fan, Xinyue Wang +3
cs.AIcs.CLcs.CVarXiv:2609.15364v12026Labels Are Not Endpoints: Treatment Leakage and Construct Validity in MCP Agent Security Evaluation
Rana Muhammad Ahmed, Sabahat Abbas
cs.CRcs.AIarXiv:2608.12880v12026Atria Dawn: The Dawn of Agentic Superintelligence
Honglin Guo, Tao Gui, Yicheng Chen +140
cs.AIarXiv:2609.15818v12026Meta-World+: An Improved, Standardized, RL Benchmark
Reginald McLean, Evangelos Chatzaroulas, Luc McCutcheon +9
cs.AIcs.LGarXiv:2505.11289v22025AI for Games in the Foundation Model Era
Meng Luo, Yanlin Li, Hao Li +7
cs.AIarXiv:2609.16679v12026Estimating Uncertain Spatial Relationships in Robotics
Randall Smith, Matthew Self, Peter Cheeseman
cs.AIcs.ROarXiv:1304.3111v22013ScienceBuddy: Recursive-in-Recursive Self-Improvement for Interactive Scientific Agents
Shuhan Xue, Jianyuan Zhong, Ziyuan Nan +10
cs.AIcs.CLarXiv:2609.17523v12026Wukong: Towards a Scaling Law for Large-Scale Recommendation
Buyun Zhang, Liang Luo, Yuxin Chen +12
cs.LGcs.AIarXiv:2403.02545v42024CMGAN: Conformer-Based Metric-GAN for Monaural Speech Enhancement
Sherif Abdulatif, Ruizhe Cao, Bin Yang
cs.SDcs.AIcs.LGarXiv:2209.11112v32022Bee: A High-Quality Corpus and Full-Stack Suite to Unlock Advanced Fully Open MLLMs
Yi Zhang, Bolin Ni, Xin-Sheng Chen +7
cs.CVcs.AIarXiv:2510.13795v42025Benchmark Radar: A Living Database and Search Engine for AI Benchmarks and Evaluation
Koutian Wu, Junjie Zhou, Ergan Shang +5
cs.AIcs.IRarXiv:2609.11115v22026Agentic Memory Enhanced Recursive Reasoning for Root Cause Localization in Microservices
Lingzhe Zhang, Tong Jia, Yunpeng Zhai +5
cs.SEcs.AIarXiv:2601.02732v12026AttnLRP: Attention-Aware Layer-Wise Relevance Propagation for Transformers
Reduan Achtibat, Sayed Mohammad Vakilzadeh Hatefi, Maximilian Dreyer +4
cs.CLcs.AIcs.CVarXiv:2402.05602v22024Ghostbuster: Detecting Text Ghostwritten by Large Language Models
Vivek Verma, Eve Fleisig, Nicholas Tomlin +1
cs.CLcs.AIarXiv:2305.15047v32023Task-Embedded Control Networks for Few-Shot Imitation Learning
Stephen James, Michael Bloesch, Andrew J. Davison
cs.ROcs.AIcs.CVarXiv:1810.03237v12018Verifier-free Test-Time Sampling for Vision-Language-Action Models
Suhyeok Jang, Dongyoung Kim, Changyeon Kim +2
cs.ROcs.AIcs.LGarXiv:2510.05681v22025RLAD: Training LLMs to Discover Abstractions for Solving Reasoning Problems
Yuxiao Qu, Anikait Singh, Yoonho Lee +4
cs.AIcs.CLcs.LGarXiv:2510.02263v12025Nemotron-Math: Efficient Long-Context Distillation of Mathematical Reasoning from Multi-Mode Supervision
Wei Du, Shubham Toshniwal, Branislav Kisacanin +7
cs.AIarXiv:2512.15489v12025AnimeCeleb: Large-Scale Animation CelebHeads Dataset for Head Reenactment
Kangyeol Kim, Sunghyun Park, Jaeseong Lee +3
cs.AIcs.CVarXiv:2111.07640v22021TrustJudge: Inconsistencies of LLM-as-a-Judge and How to Alleviate Them
Yidong Wang, Yunze Song, Tingyuan Zhu +11
cs.AIcs.CLarXiv:2509.21117v22025Local Updates, Global Learning (LUGL): Playing Games with non-incremental Learners
David Milec, Spyridon Samothrakis, Michael Fairbank +1
cs.LGcs.AIarXiv:2609.03660v12026ClaimReceipt: Verifying Evidence Sufficiency and Coverage in Agent Evaluations
Peiying Zhu, Sidi Chang
cs.AIcs.CRcs.MAarXiv:2609.01992v12026Spec2Twin-Chain: Orchestrating Bi-Level Optimization with LLMs for Blockchain Digital Twin Construction
Haoting Zhang, Haoxian Chen, Jiayuan Sheng +4
cs.AIarXiv:2608.30050v12026ScienceFlow: A long-horizon agent for ML research, scientific discovery and beyond
Mingming Zhao, Jiqian Dong, Kangping Xu +16
cs.AIarXiv:2608.14354v12026AutoWorldModel-Bench: A State-Centric Benchmark for Automated World-Model Research
Marjan Moodi, Xuankang Zhu, Fernando De Mesentier Silva +2
cs.AIarXiv:2608.11216v12026From RLVR to RLSVR: Task Transformation Induces Self-Verifiable Rewards for Open-Ended LLM Self-Improvement
Qinsi Wang, Jing Shi, Huazheng Wang +8
cs.AIarXiv:2607.23802v22026KWBench: Measuring Unprompted Problem Recognition in Knowledge Work
Ankit Maloo
cs.AIcs.GTarXiv:2604.15760v12026Look Where It Matters: High-Resolution Crops Retrieval for Efficient VLMs
Nimrod Shabtay, Moshe Kimhi, Artem Spector +5
cs.CVcs.AIarXiv:2603.16932v12026LLM Maybe LongLM: Self-Extend LLM Context Window Without Tuning
Hongye Jin, Xiaotian Han, Jingfeng Yang +5
cs.CLcs.AIcs.LGarXiv:2401.01325v32024Mamba: Linear-Time Sequence Modeling with Selective State Spaces
Albert Gu, Tri Dao
cs.LGcs.AIarXiv:2312.00752v22023Summaries:한국어Symbolic Discovery of Optimization Algorithms
Xiangning Chen, Chen Liang, Da Huang +9
cs.LGcs.AIcs.CLarXiv:2302.06675v42023Bias Out-of-the-Box: An Empirical Analysis of Intersectional Occupational Biases in Popular Generative Language Models
Hannah Kirk, Yennie Jun, Haider Iqbal +5
cs.CLcs.AIarXiv:2102.04130v32021Transfer Learning for Named-Entity Recognition with Neural Networks
Ji Young Lee, Franck Dernoncourt, Peter Szolovits
cs.CLcs.AIcs.NEarXiv:1705.06273v12017Domain Generalization using Causal Matching
Divyat Mahajan, Shruti Tople, Amit Sharma
cs.LGcs.AIstat.MLarXiv:2006.07500v32020Ming-Flash-Omni: A Sparse, Unified Architecture for Multimodal Perception and Generation
Inclusion AI, :, Bowen Ma +73
cs.CVcs.AIarXiv:2510.24821v32025When Modalities Conflict: How Unimodal Reasoning Uncertainty Governs Preference Dynamics in MLLMs
Zhuoran Zhang, Tengyue Wang, Xilin Gong +4
cs.AIarXiv:2511.02243v12025Multi-Agent Actor-Critic with Hierarchical Graph Attention Network
Heechang Ryu, Hayong Shin, Jinkyoo Park
cs.LGcs.AIcs.MAarXiv:1909.12557v22019Shaping capabilities with token-level data filtering
Neil Rathi, Alec Radford
cs.LGcs.AIcs.CLarXiv:2601.21571v22026Reinforcement Learning via Self-Distillation
Jonas Hübotter, Frederike Lübeck, Lejs Behric +8
cs.LGcs.AIarXiv:2601.20802v22026Agentic Reasoning for Large Language Models
Tianxin Wei, Ting-Wei Li, Zhining Liu +26
cs.AIcs.CLarXiv:2601.12538v12026