Artificial Intelligence
Papers filed under cs.AI on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.
Search paper metadata (including unsummarized papers)
3,901 to 3,960 of 15,209
Ask Before You Optimize: Dynamic Pre-Formulation Clarification for Interactive Optimization
Sihan Ge, Yichen Lin, Chenyu Zhou +3
math.OCcs.AIarXiv:2609.05258v12026DEMix Layers: Disentangling Domains for Modular Language Modeling
Suchin Gururangan, Mike Lewis, Ari Holtzman +2
cs.CLcs.AIarXiv:2108.05036v22021Mitigating Over-Optimization in PRM-Guided Search in Mathematical Reasoning by Optimizing the Guide
Taejong Joo, Diego Klabjan
cs.AIcs.LGarXiv:2608.30051v12026POPE: Learning to Reason on Hard Problems via Privileged On-Policy Exploration
Yuxiao Qu, Amrith Setlur, Virginia Smith +2
cs.LGcs.AIcs.CLarXiv:2601.18779v12026A fast PC algorithm for high dimensional causal discovery with multi-core PCs
Thuc Duy Le, Tao Hoang, Jiuyong Li +2
cs.AIarXiv:1502.02454v32015Humanoid Manipulation Interface: Humanoid Whole-Body Manipulation from Robot-Free Demonstrations
Ruiqian Nai, Boyuan Zheng, Junming Zhao +8
cs.ROcs.AIcs.LGarXiv:2602.06643v22026PhysReason: A Comprehensive Benchmark towards Physics-Based Reasoning
Xinyu Zhang, Yuxuan Dong, Yanrui Wu +6
cs.AIarXiv:2502.12054v22025Pure Vision Language Action (VLA) Models: A Comprehensive Survey
Dapeng Zhang, Jing Sun, Chenghui Hu +5
cs.ROcs.AIarXiv:2509.19012v32025Frequency-Aligned Knowledge Distillation for Lightweight Spatiotemporal Forecasting
Yuqi Li, Chuanguang Yang, Hansheng Zeng +5
cs.LGcs.AIcs.CVarXiv:2507.02939v22025StreamMem: Query-Agnostic KV Cache Memory for Streaming Video Understanding
Yanlai Yang, Zhuokai Zhao, Satya Narayan Shukla +4
cs.CVcs.AIarXiv:2508.15717v12025Stop Summation: Min-Form Credit Assignment Is All Process Reward Model Needs for Reasoning
Jie Cheng, Gang Xiong, Ruixi Qiao +5
cs.AIcs.LGarXiv:2504.15275v32025ReWOO: Decoupling Reasoning from Observations for Efficient Augmented Language Models
Binfeng Xu, Zhiyuan Peng, Bowen Lei +3
cs.CLcs.AIarXiv:2305.18323v12023Focused Transformer: Contrastive Training for Context Scaling
Szymon Tworkowski, Konrad Staniszewski, Mikołaj Pacek +3
cs.CLcs.AIcs.LGarXiv:2307.03170v22023Satori: Reinforcement Learning with Chain-of-Action-Thought Enhances LLM Reasoning via Autoregressive Search
Maohao Shen, Guangtao Zeng, Zhenting Qi +7
cs.CLcs.AIarXiv:2502.02508v32025Fully Autonomous AI Agents Should Not be Developed
Margaret Mitchell, Avijit Ghosh, Alexandra Sasha Luccioni +1
cs.AIarXiv:2502.02649v32025Chemception: A Deep Neural Network with Minimal Chemistry Knowledge Matches the Performance of Expert-developed QSAR/QSPR Models
Garrett B. Goh, Charles Siegel, Abhinav Vishnu +2
stat.MLcs.AIcs.CEarXiv:1706.06689v12017Iris: Climbing to the Search Frontier
Ziyuan Liu, Hengqi Liu, Zichuan Wang +6
cs.AIarXiv:2609.04304v12026Error Detection for PET/CT Radiology Reports: Domain-Specific vs Large Language Models
Hermione Warr, Harry Anthony, Lilli J Freischem +3
cs.LGcs.AIarXiv:2608.30021v12026On the Instance Hardness as a Decision Criterion in TinyML Systems
Tobiasz Puslecki, Krzysztof Walkowiak
cs.AIcs.LGarXiv:2608.29913v12026Walk the Talk? Measuring the Faithfulness of Large Language Model Explanations
Katie Matton, Robert Osazuwa Ness, John Guttag +1
cs.CLcs.AIcs.LGarXiv:2504.14150v22025On Vanishing Gradients, Over-Smoothing, and Over-Squashing in GNNs: Bridging Recurrent and Graph Learning
Álvaro Arroyo, Alessio Gravina, Benjamin Gutteridge +5
cs.LGcs.AIarXiv:2502.10818v22025Text-to-Image Diffusion Models are Zero-Shot Classifiers
Kevin Clark, Priyank Jaini
cs.CVcs.AIcs.LGarXiv:2303.15233v22023GIScience in the Era of Artificial Intelligence: A Research Agenda Towards Autonomous GIS
Zhenlong Li, Huan Ning, Song Gao +13
cs.AIcs.ETcs.SEarXiv:2503.23633v52025Hierarchical Memory for High-Efficiency Long-Term Reasoning in LLM Agents
Haoran Sun, Shaoning Zeng
cs.CLcs.AIarXiv:2507.22925v12025Formal Concept Analysis with Three Types of Negation
Zhenghua Pan
cs.AIarXiv:2608.29311v12026MotionGPT: Finetuned LLMs Are General-Purpose Motion Generators
Yaqi Zhang, Di Huang, Bin Liu +7
cs.CVcs.AIarXiv:2306.10900v22023MMPCBench: Benchmarking Multimodal Large Language Models on Proactive Critique of Flawed Inputs
Jinzhe Li, Gengxu Li, Jinnan Li +2
cs.AIarXiv:2608.29286v12026On the Arbitrary-Oriented Object Detection: Classification based Approaches Revisited
Xue Yang, Junchi Yan
cs.CVcs.AIarXiv:2003.05597v42020The MASK Benchmark: Disentangling Honesty From Accuracy in AI Systems
Richard Ren, Arunim Agarwal, Mantas Mazeika +13
cs.LGcs.AIcs.CLarXiv:2503.03750v32025GNN-RAG: Graph Neural Retrieval for Large Language Model Reasoning
Costas Mavromatis, George Karypis
cs.CLcs.AIcs.LGarXiv:2405.20139v12024Multi-Objective Deep Reinforcement Learning
Hossam Mossalam, Yannis M. Assael, Diederik M. Roijers +1
cs.AIarXiv:1610.02707v12016The Curse of Depth in Large Language Models
Wenfang Sun, Xinyuan Song, Pengxiang Li +3
cs.LGcs.AIarXiv:2502.05795v62025VICRegL: Self-Supervised Learning of Local Visual Features
Adrien Bardes, Jean Ponce, Yann LeCun
cs.CVcs.AIcs.LGarXiv:2210.01571v12022SP-VLA: A Joint Model Scheduling and Token Pruning Approach for VLA Model Acceleration
Ye Li, Yuan Meng, Zewen Sun +7
cs.CVcs.AIarXiv:2506.12723v32025OPSDL: On-Policy Self-Distillation for Long-Context Language Models
Xinsen Zhang, Zhenkai Ding, Tianjun Pan +4
cs.CLcs.AIarXiv:2604.17535v12026TAAL: Mitigating Early Beam Pruning in Generative Recommendation via Temporal Autoregressive Alignment
Lianjie Li, Zhiying Tu, Dianhui Chu +1
cs.IRcs.AIarXiv:2608.29179v12026HM-RAG: Hierarchical Multi-Agent Multimodal Retrieval Augmented Generation
Pei Liu, Xin Liu, Ruoyu Yao +4
cs.CLcs.AIarXiv:2504.12330v12025MedAgentBench: A Realistic Virtual EHR Environment to Benchmark Medical LLM Agents
Yixing Jiang, Kameron C. Black, Gloria Geng +4
cs.LGcs.AIcs.MAarXiv:2501.14654v22025SkillTrojan: Backdoor Attacks on Skill-Based Agent Systems
Yunhao Feng, Yifan Ding, Yingshui Tan +6
cs.CRcs.AIarXiv:2604.06811v22026Assisting in Writing Wikipedia-like Articles From Scratch with Large Language Models
Yijia Shao, Yucheng Jiang, Theodore A. Kanell +3
cs.CLcs.AIarXiv:2402.14207v22024Auditing and Mitigating Privacy Leakage in Cloud-Edge Collaborative Decoding
Kejia Zhang, Tianyuan Zou, Zixuan GU +1
cs.CRcs.AIarXiv:2608.29111v12026Interpreting CLIP with Hierarchical Sparse Autoencoders
Vladimir Zaigrajew, Hubert Baniecki, Przemyslaw Biecek
cs.CVcs.AIcs.LGarXiv:2502.20578v22025Agent-R: Training Language Model Agents to Reflect via Iterative Self-Training
Siyu Yuan, Zehui Chen, Zhiheng Xi +3
cs.AIarXiv:2501.11425v32025A User Simulator for Task-Completion Dialogues
Xiujun Li, Zachary C. Lipton, Bhuwan Dhingra +3
cs.LGcs.AIcs.CLarXiv:1612.05688v32016DeepStory: Video Story QA by Deep Embedded Memory Networks
Kyung-Min Kim, Min-Oh Heo, Seong-Ho Choi +1
cs.CVcs.AIcs.CLarXiv:1707.00836v12017RAGDiffusion++: From Macro-Retrieval to Micro-Fidelity Alignment for Garment Generation
Yuhan Li, Xianfeng Tan, Fangao Zeng +6
cs.CVcs.AIarXiv:2608.29280v12026APIFlow-Bench: Measuring Whether Agents Survive Long, Dependent API Workflows
Zelin Wan, Arash Nourian, Xiaoxiao Li +2
cs.AIcs.LGcs.SEarXiv:2608.29128v12026LEGO-Puzzles: How Good Are MLLMs at Multi-Step Spatial Reasoning?
Kexian Tang, Junyao Gao, Yanhong Zeng +6
cs.AIarXiv:2503.19990v42025D2A: A Dataset Built for AI-Based Vulnerability Detection Methods Using Differential Analysis
Yunhui Zheng, Saurabh Pujar, Burn Lewis +6
cs.SEcs.AIcs.LGarXiv:2102.07995v12021VideoRAG: Retrieval-Augmented Generation over Video Corpus
Soyeong Jeong, Kangsan Kim, Jinheon Baek +1
cs.CVcs.AIcs.CLarXiv:2501.05874v32025A Survey on Sparse Autoencoders: Interpreting the Internal Mechanisms of Large Language Models
Dong Shu, Xuansheng Wu, Haiyan Zhao +4
cs.LGcs.AIcs.CLarXiv:2503.05613v32025Wavelet-Assisted Multi-Frequency Attention Network for Pansharpening
Jie Huang, Rui Huang, Jinghao Xu +3
eess.IVcs.AIcs.CVarXiv:2502.04903v12025A Survey on Large Language Models for Mathematical Reasoning
Peng-Yuan Wang, Tian-Shuo Liu, Chenyang Wang +8
cs.AIcs.CLarXiv:2506.08446v12025Integrating LLMs with ITS: Recent Advances, Potentials, Challenges, and Future Directions
Doaa Mahmud, Hadeel Hajmohamed, Shamma Almentheri +4
eess.SYcs.AIcs.ETarXiv:2501.04437v12025Language Models Use Trigonometry to Do Addition
Subhash Kantamneni, Max Tegmark
cs.AIcs.CLcs.LGarXiv:2502.00873v12025AI Literacy in K-12 and Higher Education in the Wake of Generative AI: An Integrative Review
Xingjian Gu, Barbara J. Ericson
cs.CYcs.AIarXiv:2503.00079v32025Kernel Language Entropy: Fine-grained Uncertainty Quantification for LLMs from Semantic Similarities
Alexander Nikitin, Jannik Kossen, Yarin Gal +1
cs.LGcs.AIcs.CLarXiv:2405.20003v12024Formal Policy Enforcement for Real-World Agentic Systems
Nils Palumbo, Sarthak Choudhary, Jihye Choi +3
cs.CRcs.AIcs.MAarXiv:2602.16708v32026Neural Language Modeling by Jointly Learning Syntax and Lexicon
Yikang Shen, Zhouhan Lin, Chin-Wei Huang +1
cs.CLcs.AIarXiv:1711.02013v22017MinTL: Minimalist Transfer Learning for Task-Oriented Dialogue Systems
Zhaojiang Lin, Andrea Madotto, Genta Indra Winata +1
cs.CLcs.AIarXiv:2009.12005v22020