Artificial Intelligence

Papers filed under cs.AI on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.

Search paper metadata (including unsummarized papers)

3,901 to 3,960 of 15,209

  1. Ask Before You Optimize: Dynamic Pre-Formulation Clarification for Interactive Optimization

    Sihan Ge, Yichen Lin, Chenyu Zhou +3

    math.OCcs.AIarXiv:2609.05258v12026
  2. DEMix Layers: Disentangling Domains for Modular Language Modeling

    Suchin Gururangan, Mike Lewis, Ari Holtzman +2

    cs.CLcs.AIarXiv:2108.05036v22021
  3. Mitigating Over-Optimization in PRM-Guided Search in Mathematical Reasoning by Optimizing the Guide

    Taejong Joo, Diego Klabjan

    cs.AIcs.LGarXiv:2608.30051v12026
  4. POPE: Learning to Reason on Hard Problems via Privileged On-Policy Exploration

    Yuxiao Qu, Amrith Setlur, Virginia Smith +2

    cs.LGcs.AIcs.CLarXiv:2601.18779v12026
  5. A fast PC algorithm for high dimensional causal discovery with multi-core PCs

    Thuc Duy Le, Tao Hoang, Jiuyong Li +2

    cs.AIarXiv:1502.02454v32015
  6. Humanoid Manipulation Interface: Humanoid Whole-Body Manipulation from Robot-Free Demonstrations

    Ruiqian Nai, Boyuan Zheng, Junming Zhao +8

    cs.ROcs.AIcs.LGarXiv:2602.06643v22026
  7. PhysReason: A Comprehensive Benchmark towards Physics-Based Reasoning

    Xinyu Zhang, Yuxuan Dong, Yanrui Wu +6

    cs.AIarXiv:2502.12054v22025
  8. Pure Vision Language Action (VLA) Models: A Comprehensive Survey

    Dapeng Zhang, Jing Sun, Chenghui Hu +5

    cs.ROcs.AIarXiv:2509.19012v32025
  9. Frequency-Aligned Knowledge Distillation for Lightweight Spatiotemporal Forecasting

    Yuqi Li, Chuanguang Yang, Hansheng Zeng +5

    cs.LGcs.AIcs.CVarXiv:2507.02939v22025
  10. StreamMem: Query-Agnostic KV Cache Memory for Streaming Video Understanding

    Yanlai Yang, Zhuokai Zhao, Satya Narayan Shukla +4

    cs.CVcs.AIarXiv:2508.15717v12025
  11. Stop Summation: Min-Form Credit Assignment Is All Process Reward Model Needs for Reasoning

    Jie Cheng, Gang Xiong, Ruixi Qiao +5

    cs.AIcs.LGarXiv:2504.15275v32025
  12. ReWOO: Decoupling Reasoning from Observations for Efficient Augmented Language Models

    Binfeng Xu, Zhiyuan Peng, Bowen Lei +3

    cs.CLcs.AIarXiv:2305.18323v12023
  13. Focused Transformer: Contrastive Training for Context Scaling

    Szymon Tworkowski, Konrad Staniszewski, Mikołaj Pacek +3

    cs.CLcs.AIcs.LGarXiv:2307.03170v22023
  14. Satori: Reinforcement Learning with Chain-of-Action-Thought Enhances LLM Reasoning via Autoregressive Search

    Maohao Shen, Guangtao Zeng, Zhenting Qi +7

    cs.CLcs.AIarXiv:2502.02508v32025
  15. Fully Autonomous AI Agents Should Not be Developed

    Margaret Mitchell, Avijit Ghosh, Alexandra Sasha Luccioni +1

    cs.AIarXiv:2502.02649v32025
  16. Chemception: A Deep Neural Network with Minimal Chemistry Knowledge Matches the Performance of Expert-developed QSAR/QSPR Models

    Garrett B. Goh, Charles Siegel, Abhinav Vishnu +2

    stat.MLcs.AIcs.CEarXiv:1706.06689v12017
  17. Iris: Climbing to the Search Frontier

    Ziyuan Liu, Hengqi Liu, Zichuan Wang +6

    cs.AIarXiv:2609.04304v12026
  18. Error Detection for PET/CT Radiology Reports: Domain-Specific vs Large Language Models

    Hermione Warr, Harry Anthony, Lilli J Freischem +3

    cs.LGcs.AIarXiv:2608.30021v12026
  19. On the Instance Hardness as a Decision Criterion in TinyML Systems

    Tobiasz Puslecki, Krzysztof Walkowiak

    cs.AIcs.LGarXiv:2608.29913v12026
  20. Walk the Talk? Measuring the Faithfulness of Large Language Model Explanations

    Katie Matton, Robert Osazuwa Ness, John Guttag +1

    cs.CLcs.AIcs.LGarXiv:2504.14150v22025
  21. On Vanishing Gradients, Over-Smoothing, and Over-Squashing in GNNs: Bridging Recurrent and Graph Learning

    Álvaro Arroyo, Alessio Gravina, Benjamin Gutteridge +5

    cs.LGcs.AIarXiv:2502.10818v22025
  22. Text-to-Image Diffusion Models are Zero-Shot Classifiers

    Kevin Clark, Priyank Jaini

    cs.CVcs.AIcs.LGarXiv:2303.15233v22023
  23. GIScience in the Era of Artificial Intelligence: A Research Agenda Towards Autonomous GIS

    Zhenlong Li, Huan Ning, Song Gao +13

    cs.AIcs.ETcs.SEarXiv:2503.23633v52025
  24. Hierarchical Memory for High-Efficiency Long-Term Reasoning in LLM Agents

    Haoran Sun, Shaoning Zeng

    cs.CLcs.AIarXiv:2507.22925v12025
  25. Formal Concept Analysis with Three Types of Negation

    Zhenghua Pan

    cs.AIarXiv:2608.29311v12026
  26. MotionGPT: Finetuned LLMs Are General-Purpose Motion Generators

    Yaqi Zhang, Di Huang, Bin Liu +7

    cs.CVcs.AIarXiv:2306.10900v22023
  27. MMPCBench: Benchmarking Multimodal Large Language Models on Proactive Critique of Flawed Inputs

    Jinzhe Li, Gengxu Li, Jinnan Li +2

    cs.AIarXiv:2608.29286v12026
  28. On the Arbitrary-Oriented Object Detection: Classification based Approaches Revisited

    Xue Yang, Junchi Yan

    cs.CVcs.AIarXiv:2003.05597v42020
  29. The MASK Benchmark: Disentangling Honesty From Accuracy in AI Systems

    Richard Ren, Arunim Agarwal, Mantas Mazeika +13

    cs.LGcs.AIcs.CLarXiv:2503.03750v32025
  30. GNN-RAG: Graph Neural Retrieval for Large Language Model Reasoning

    Costas Mavromatis, George Karypis

    cs.CLcs.AIcs.LGarXiv:2405.20139v12024
  31. Multi-Objective Deep Reinforcement Learning

    Hossam Mossalam, Yannis M. Assael, Diederik M. Roijers +1

    cs.AIarXiv:1610.02707v12016
  32. The Curse of Depth in Large Language Models

    Wenfang Sun, Xinyuan Song, Pengxiang Li +3

    cs.LGcs.AIarXiv:2502.05795v62025
  33. VICRegL: Self-Supervised Learning of Local Visual Features

    Adrien Bardes, Jean Ponce, Yann LeCun

    cs.CVcs.AIcs.LGarXiv:2210.01571v12022
  34. SP-VLA: A Joint Model Scheduling and Token Pruning Approach for VLA Model Acceleration

    Ye Li, Yuan Meng, Zewen Sun +7

    cs.CVcs.AIarXiv:2506.12723v32025
  35. OPSDL: On-Policy Self-Distillation for Long-Context Language Models

    Xinsen Zhang, Zhenkai Ding, Tianjun Pan +4

    cs.CLcs.AIarXiv:2604.17535v12026
  36. TAAL: Mitigating Early Beam Pruning in Generative Recommendation via Temporal Autoregressive Alignment

    Lianjie Li, Zhiying Tu, Dianhui Chu +1

    cs.IRcs.AIarXiv:2608.29179v12026
  37. HM-RAG: Hierarchical Multi-Agent Multimodal Retrieval Augmented Generation

    Pei Liu, Xin Liu, Ruoyu Yao +4

    cs.CLcs.AIarXiv:2504.12330v12025
  38. MedAgentBench: A Realistic Virtual EHR Environment to Benchmark Medical LLM Agents

    Yixing Jiang, Kameron C. Black, Gloria Geng +4

    cs.LGcs.AIcs.MAarXiv:2501.14654v22025
  39. SkillTrojan: Backdoor Attacks on Skill-Based Agent Systems

    Yunhao Feng, Yifan Ding, Yingshui Tan +6

    cs.CRcs.AIarXiv:2604.06811v22026
  40. Assisting in Writing Wikipedia-like Articles From Scratch with Large Language Models

    Yijia Shao, Yucheng Jiang, Theodore A. Kanell +3

    cs.CLcs.AIarXiv:2402.14207v22024
  41. Auditing and Mitigating Privacy Leakage in Cloud-Edge Collaborative Decoding

    Kejia Zhang, Tianyuan Zou, Zixuan GU +1

    cs.CRcs.AIarXiv:2608.29111v12026
  42. Interpreting CLIP with Hierarchical Sparse Autoencoders

    Vladimir Zaigrajew, Hubert Baniecki, Przemyslaw Biecek

    cs.CVcs.AIcs.LGarXiv:2502.20578v22025
  43. Agent-R: Training Language Model Agents to Reflect via Iterative Self-Training

    Siyu Yuan, Zehui Chen, Zhiheng Xi +3

    cs.AIarXiv:2501.11425v32025
  44. A User Simulator for Task-Completion Dialogues

    Xiujun Li, Zachary C. Lipton, Bhuwan Dhingra +3

    cs.LGcs.AIcs.CLarXiv:1612.05688v32016
  45. DeepStory: Video Story QA by Deep Embedded Memory Networks

    Kyung-Min Kim, Min-Oh Heo, Seong-Ho Choi +1

    cs.CVcs.AIcs.CLarXiv:1707.00836v12017
  46. RAGDiffusion++: From Macro-Retrieval to Micro-Fidelity Alignment for Garment Generation

    Yuhan Li, Xianfeng Tan, Fangao Zeng +6

    cs.CVcs.AIarXiv:2608.29280v12026
  47. APIFlow-Bench: Measuring Whether Agents Survive Long, Dependent API Workflows

    Zelin Wan, Arash Nourian, Xiaoxiao Li +2

    cs.AIcs.LGcs.SEarXiv:2608.29128v12026
  48. LEGO-Puzzles: How Good Are MLLMs at Multi-Step Spatial Reasoning?

    Kexian Tang, Junyao Gao, Yanhong Zeng +6

    cs.AIarXiv:2503.19990v42025
  49. D2A: A Dataset Built for AI-Based Vulnerability Detection Methods Using Differential Analysis

    Yunhui Zheng, Saurabh Pujar, Burn Lewis +6

    cs.SEcs.AIcs.LGarXiv:2102.07995v12021
  50. VideoRAG: Retrieval-Augmented Generation over Video Corpus

    Soyeong Jeong, Kangsan Kim, Jinheon Baek +1

    cs.CVcs.AIcs.CLarXiv:2501.05874v32025
  51. A Survey on Sparse Autoencoders: Interpreting the Internal Mechanisms of Large Language Models

    Dong Shu, Xuansheng Wu, Haiyan Zhao +4

    cs.LGcs.AIcs.CLarXiv:2503.05613v32025
  52. Wavelet-Assisted Multi-Frequency Attention Network for Pansharpening

    Jie Huang, Rui Huang, Jinghao Xu +3

    eess.IVcs.AIcs.CVarXiv:2502.04903v12025
  53. A Survey on Large Language Models for Mathematical Reasoning

    Peng-Yuan Wang, Tian-Shuo Liu, Chenyang Wang +8

    cs.AIcs.CLarXiv:2506.08446v12025
  54. Integrating LLMs with ITS: Recent Advances, Potentials, Challenges, and Future Directions

    Doaa Mahmud, Hadeel Hajmohamed, Shamma Almentheri +4

    eess.SYcs.AIcs.ETarXiv:2501.04437v12025
  55. Language Models Use Trigonometry to Do Addition

    Subhash Kantamneni, Max Tegmark

    cs.AIcs.CLcs.LGarXiv:2502.00873v12025
  56. AI Literacy in K-12 and Higher Education in the Wake of Generative AI: An Integrative Review

    Xingjian Gu, Barbara J. Ericson

    cs.CYcs.AIarXiv:2503.00079v32025
  57. Kernel Language Entropy: Fine-grained Uncertainty Quantification for LLMs from Semantic Similarities

    Alexander Nikitin, Jannik Kossen, Yarin Gal +1

    cs.LGcs.AIcs.CLarXiv:2405.20003v12024
  58. Formal Policy Enforcement for Real-World Agentic Systems

    Nils Palumbo, Sarthak Choudhary, Jihye Choi +3

    cs.CRcs.AIcs.MAarXiv:2602.16708v32026
  59. Neural Language Modeling by Jointly Learning Syntax and Lexicon

    Yikang Shen, Zhouhan Lin, Chin-Wei Huang +1

    cs.CLcs.AIarXiv:1711.02013v22017
  60. MinTL: Minimalist Transfer Learning for Task-Oriented Dialogue Systems

    Zhaojiang Lin, Andrea Madotto, Genta Indra Winata +1

    cs.CLcs.AIarXiv:2009.12005v22020