Artificial Intelligence

Papers filed under cs.AI on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.

Search paper metadata (including unsummarized papers)

541 to 600 of 15,291

  1. Not All Prompts Are Equal: Exploration-Guided Prompt Scaffolding for Multimodal Reinforcement Post-Training

    Yuanhao Yue, Qianli Ma, Chengyu Wang +3

    cs.LGcs.AIcs.CLarXiv:2609.15051v12026
  2. SkillMOO: Multi-Objective Optimization of Agent Skills for Software Engineering

    Jingzhi Gong, Ruizhen Gu, Zhiwei Fei +7

    cs.SEcs.AIarXiv:2604.09297v32026
  3. Improving Coherence and Consistency in Neural Sequence Models with Dual-System, Neuro-Symbolic Reasoning

    Maxwell Nye, Michael Henry Tessler, Joshua B. Tenenbaum +1

    cs.AIcs.CLcs.LGarXiv:2107.02794v22021
  4. "Death" of a Chatbot: Investigating and Designing Toward Psychologically Safe Endings for Human-AI Relationships

    Rachel Poonsiriwong, Chayapatr Archiwaranguprok, Pat Pataranutaporn

    cs.HCcs.AIarXiv:2602.07193v22026
  5. EVA: Aligning Video World Models with Executable Robot Actions via Inverse Dynamics Rewards

    Ruixiang Wang, Qingming Liu, Yueci Deng +3

    cs.ROcs.AIarXiv:2603.17808v22026
  6. Pick Your Poison: Learning to Select Poison Sets for Stronger LLM Backdoor Attacks

    Aashiq Muhamed, Mona T. Diab, Virginia Smith +2

    cs.LGcs.AIcs.CLarXiv:2609.15029v12026
  7. MemoRAG: Boosting Long Context Processing with Global Memory-Enhanced Retrieval Augmentation

    Hongjin Qian, Zheng Liu, Peitian Zhang +4

    cs.CLcs.AIarXiv:2409.05591v32024
  8. How Lossless Is Lossless Speculative Decoding? The Role of Numerical Precision in Orthrus

    Ilya Koziev, Leonid Sinev, Ivan Oseledets

    cs.CLcs.AIarXiv:2609.15504v12026
  9. Beyond Quacking: Deep Integration of Language Models and RAG into DuckDB

    Anas Dorbani, Sunny Yasser, Jimmy Lin +1

    cs.DBcs.AIcs.IRarXiv:2504.01157v12025
  10. HazardAuditor: From Executable Threats to Safer Computer-Use Agents

    Yunhao Feng, Ruixiao Lin, Ming Wen +5

    cs.AIarXiv:2609.15134v12026
  11. CORE: Simple and Effective Session-based Recommendation within Consistent Representation Space

    Yupeng Hou, Binbin Hu, Zhiqiang Zhang +1

    cs.IRcs.AIarXiv:2204.11067v12022
  12. AdaRubric: Task-Adaptive Rubrics for Reliable LLM Agent Evaluation and Reward Learning

    Liang Ding

    cs.AIcs.CLarXiv:2603.21362v32026
  13. PhysStream: Streaming Physics-Grounded Video Generation with Structured Scene Memory and Fine-Grained Motion Control

    Chuhao Chen, Peter Wonka, Chaoyang Wang +4

    cs.CVcs.AIcs.GRarXiv:2609.17521v12026
  14. FlexiTac: A Low-Cost, Open-Source, Scalable Tactile Sensing Solution for Robotic Systems

    Binghao Huang, Yunzhu Li

    cs.ROcs.AIcs.LGarXiv:2604.28156v12026
  15. KaiNinja: Extending Native 3D Generators to the Part Level

    Ruihan Yu, Lian Fu, Muyao Niu +9

    cs.GRcs.AIcs.CVarXiv:2609.15659v22026
  16. How Coding Agents Fail Their Users: A Large-Scale Analysis of Developer-Agent Misalignment in 20,574 Real-World Sessions

    Ningzhi Tang, Chaoran Chen, Gelei Xu +5

    cs.SEcs.AIcs.HCarXiv:2605.29442v22026
  17. MS2: Multi-Document Summarization of Medical Studies

    Jay DeYoung, Iz Beltagy, Madeleine van Zuylen +2

    cs.CLcs.AIcs.LGarXiv:2104.06486v32021
  18. Maximum Entropy Gain Exploration for Long Horizon Multi-goal Reinforcement Learning

    Silviu Pitis, Harris Chan, Stephen Zhao +2

    cs.LGcs.AIcs.ROarXiv:2007.02832v12020
  19. Reflective Planning: Vision-Language Models for Multi-Stage Long-Horizon Robotic Manipulation

    Yunhai Feng, Jiaming Han, Zhuoran Yang +3

    cs.ROcs.AIcs.LGarXiv:2502.16707v12025
  20. TJ4DRadSet: A 4D Radar Dataset for Autonomous Driving

    Lianqing Zheng, Zhixiong Ma, Xichan Zhu +9

    cs.CVcs.AIarXiv:2204.13483v32022
  21. RSIAgent: Autonomous Exploration for Recursive Self-improvement in New Environments

    Sibo Zhu, Shicheng Fan, Xinyue Wang +3

    cs.AIcs.CLcs.CVarXiv:2609.15364v12026
  22. Labels Are Not Endpoints: Treatment Leakage and Construct Validity in MCP Agent Security Evaluation

    Rana Muhammad Ahmed, Sabahat Abbas

    cs.CRcs.AIarXiv:2608.12880v12026
  23. Atria Dawn: The Dawn of Agentic Superintelligence

    Honglin Guo, Tao Gui, Yicheng Chen +140

    cs.AIarXiv:2609.15818v12026
  24. Meta-World+: An Improved, Standardized, RL Benchmark

    Reginald McLean, Evangelos Chatzaroulas, Luc McCutcheon +9

    cs.AIcs.LGarXiv:2505.11289v22025
  25. AI for Games in the Foundation Model Era

    Meng Luo, Yanlin Li, Hao Li +7

    cs.AIarXiv:2609.16679v12026
  26. Estimating Uncertain Spatial Relationships in Robotics

    Randall Smith, Matthew Self, Peter Cheeseman

    cs.AIcs.ROarXiv:1304.3111v22013
  27. ScienceBuddy: Recursive-in-Recursive Self-Improvement for Interactive Scientific Agents

    Shuhan Xue, Jianyuan Zhong, Ziyuan Nan +10

    cs.AIcs.CLarXiv:2609.17523v12026
  28. Wukong: Towards a Scaling Law for Large-Scale Recommendation

    Buyun Zhang, Liang Luo, Yuxin Chen +12

    cs.LGcs.AIarXiv:2403.02545v42024
  29. CMGAN: Conformer-Based Metric-GAN for Monaural Speech Enhancement

    Sherif Abdulatif, Ruizhe Cao, Bin Yang

    cs.SDcs.AIcs.LGarXiv:2209.11112v32022
  30. Bee: A High-Quality Corpus and Full-Stack Suite to Unlock Advanced Fully Open MLLMs

    Yi Zhang, Bolin Ni, Xin-Sheng Chen +7

    cs.CVcs.AIarXiv:2510.13795v42025
  31. Benchmark Radar: A Living Database and Search Engine for AI Benchmarks and Evaluation

    Koutian Wu, Junjie Zhou, Ergan Shang +5

    cs.AIcs.IRarXiv:2609.11115v22026
  32. Agentic Memory Enhanced Recursive Reasoning for Root Cause Localization in Microservices

    Lingzhe Zhang, Tong Jia, Yunpeng Zhai +5

    cs.SEcs.AIarXiv:2601.02732v12026
  33. AttnLRP: Attention-Aware Layer-Wise Relevance Propagation for Transformers

    Reduan Achtibat, Sayed Mohammad Vakilzadeh Hatefi, Maximilian Dreyer +4

    cs.CLcs.AIcs.CVarXiv:2402.05602v22024
  34. Ghostbuster: Detecting Text Ghostwritten by Large Language Models

    Vivek Verma, Eve Fleisig, Nicholas Tomlin +1

    cs.CLcs.AIarXiv:2305.15047v32023
  35. Task-Embedded Control Networks for Few-Shot Imitation Learning

    Stephen James, Michael Bloesch, Andrew J. Davison

    cs.ROcs.AIcs.CVarXiv:1810.03237v12018
  36. Verifier-free Test-Time Sampling for Vision-Language-Action Models

    Suhyeok Jang, Dongyoung Kim, Changyeon Kim +2

    cs.ROcs.AIcs.LGarXiv:2510.05681v22025
  37. RLAD: Training LLMs to Discover Abstractions for Solving Reasoning Problems

    Yuxiao Qu, Anikait Singh, Yoonho Lee +4

    cs.AIcs.CLcs.LGarXiv:2510.02263v12025
  38. Nemotron-Math: Efficient Long-Context Distillation of Mathematical Reasoning from Multi-Mode Supervision

    Wei Du, Shubham Toshniwal, Branislav Kisacanin +7

    cs.AIarXiv:2512.15489v12025
  39. AnimeCeleb: Large-Scale Animation CelebHeads Dataset for Head Reenactment

    Kangyeol Kim, Sunghyun Park, Jaeseong Lee +3

    cs.AIcs.CVarXiv:2111.07640v22021
  40. TrustJudge: Inconsistencies of LLM-as-a-Judge and How to Alleviate Them

    Yidong Wang, Yunze Song, Tingyuan Zhu +11

    cs.AIcs.CLarXiv:2509.21117v22025
  41. Local Updates, Global Learning (LUGL): Playing Games with non-incremental Learners

    David Milec, Spyridon Samothrakis, Michael Fairbank +1

    cs.LGcs.AIarXiv:2609.03660v12026
  42. ClaimReceipt: Verifying Evidence Sufficiency and Coverage in Agent Evaluations

    Peiying Zhu, Sidi Chang

    cs.AIcs.CRcs.MAarXiv:2609.01992v12026
  43. Spec2Twin-Chain: Orchestrating Bi-Level Optimization with LLMs for Blockchain Digital Twin Construction

    Haoting Zhang, Haoxian Chen, Jiayuan Sheng +4

    cs.AIarXiv:2608.30050v12026
  44. ScienceFlow: A long-horizon agent for ML research, scientific discovery and beyond

    Mingming Zhao, Jiqian Dong, Kangping Xu +16

    cs.AIarXiv:2608.14354v12026
  45. AutoWorldModel-Bench: A State-Centric Benchmark for Automated World-Model Research

    Marjan Moodi, Xuankang Zhu, Fernando De Mesentier Silva +2

    cs.AIarXiv:2608.11216v12026
  46. From RLVR to RLSVR: Task Transformation Induces Self-Verifiable Rewards for Open-Ended LLM Self-Improvement

    Qinsi Wang, Jing Shi, Huazheng Wang +8

    cs.AIarXiv:2607.23802v22026
  47. KWBench: Measuring Unprompted Problem Recognition in Knowledge Work

    Ankit Maloo

    cs.AIcs.GTarXiv:2604.15760v12026
  48. Look Where It Matters: High-Resolution Crops Retrieval for Efficient VLMs

    Nimrod Shabtay, Moshe Kimhi, Artem Spector +5

    cs.CVcs.AIarXiv:2603.16932v12026
  49. LLM Maybe LongLM: Self-Extend LLM Context Window Without Tuning

    Hongye Jin, Xiaotian Han, Jingfeng Yang +5

    cs.CLcs.AIcs.LGarXiv:2401.01325v32024
  50. Mamba: Linear-Time Sequence Modeling with Selective State Spaces

    Albert Gu, Tri Dao

    cs.LGcs.AIarXiv:2312.00752v22023
    Summaries:한국어
  51. Symbolic Discovery of Optimization Algorithms

    Xiangning Chen, Chen Liang, Da Huang +9

    cs.LGcs.AIcs.CLarXiv:2302.06675v42023
  52. Bias Out-of-the-Box: An Empirical Analysis of Intersectional Occupational Biases in Popular Generative Language Models

    Hannah Kirk, Yennie Jun, Haider Iqbal +5

    cs.CLcs.AIarXiv:2102.04130v32021
  53. Transfer Learning for Named-Entity Recognition with Neural Networks

    Ji Young Lee, Franck Dernoncourt, Peter Szolovits

    cs.CLcs.AIcs.NEarXiv:1705.06273v12017
  54. Domain Generalization using Causal Matching

    Divyat Mahajan, Shruti Tople, Amit Sharma

    cs.LGcs.AIstat.MLarXiv:2006.07500v32020
  55. Ming-Flash-Omni: A Sparse, Unified Architecture for Multimodal Perception and Generation

    Inclusion AI, :, Bowen Ma +73

    cs.CVcs.AIarXiv:2510.24821v32025
  56. When Modalities Conflict: How Unimodal Reasoning Uncertainty Governs Preference Dynamics in MLLMs

    Zhuoran Zhang, Tengyue Wang, Xilin Gong +4

    cs.AIarXiv:2511.02243v12025
  57. Multi-Agent Actor-Critic with Hierarchical Graph Attention Network

    Heechang Ryu, Hayong Shin, Jinkyoo Park

    cs.LGcs.AIcs.MAarXiv:1909.12557v22019
  58. Shaping capabilities with token-level data filtering

    Neil Rathi, Alec Radford

    cs.LGcs.AIcs.CLarXiv:2601.21571v22026
  59. Reinforcement Learning via Self-Distillation

    Jonas Hübotter, Frederike Lübeck, Lejs Behric +8

    cs.LGcs.AIarXiv:2601.20802v22026
  60. Agentic Reasoning for Large Language Models

    Tianxin Wei, Ting-Wei Li, Zhining Liu +26

    cs.AIcs.CLarXiv:2601.12538v12026