Artificial Intelligence
Papers filed under cs.AI on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.
Search paper metadata (including unsummarized papers)
4,501 to 4,560 of 15,269
EDGE: Error Dependency Graph-Guided Multi-Error Attribution in Multi-Agent LLM Systems
Jun Hou, Priya Pitre, Yi Fang +1
cs.AIarXiv:2609.01360v12026Efficient Large-Scale Multi-Modal Classification
D. Kiela, E. Grave, A. Joulin +1
cs.CLcs.AIcs.CVarXiv:1802.02892v12018Dion: Distributed Orthonormalized Updates
Kwangjun Ahn, Byron Xu, Natalie Abreu +5
cs.LGcs.AImath.OCarXiv:2504.05295v32025Classical Planning in Deep Latent Space: Bridging the Subsymbolic-Symbolic Boundary
Masataro Asai, Alex Fukunaga
cs.AIarXiv:1705.00154v32017A Study of Reinforcement Learning for Neural Machine Translation
Lijun Wu, Fei Tian, Tao Qin +2
cs.LGcs.AIstat.MLarXiv:1808.08866v12018Beyond the Clock: Measuring the Value of Adaptive Revision
Ayushi Chadha
cs.AIarXiv:2609.00874v12026Molecular-driven Foundation Model for Oncologic Pathology
Anurag Vaidya, Andrew Zhang, Guillaume Jaume +15
cs.CVcs.AIarXiv:2501.16652v12025Towards Generalizable Visually Grounded Exploration of Household Devices
Linhao Zheng, Zeming Liu, Wangke Chen +4
cs.AIarXiv:2609.00845v12026When Features Become Instances: Inverted Contrastive Learning for Unsupervised Feature Selection
Utsab Ghosh, Roshni Chakraborty
cs.AIcs.CLarXiv:2609.00782v12026Self-Reports Are Not Verification: Environment-Grounded Auditing of LLM Operators in Evolutionary Search
Enrong Pan, Ryan Zhou, Ting Hu
cs.AIcs.LGcs.NEarXiv:2609.00652v12026Learning to Plan & Reason for Evaluation with Thinking-LLM-as-a-Judge
Swarnadeep Saha, Xian Li, Marjan Ghazvininejad +2
cs.AIcs.CLarXiv:2501.18099v22025Wave Function Backpropagation with Explicit Temporal-Interval Dynamics
Byunggu Yu, Justin Kim
cs.AIarXiv:2609.00503v12026Sekai: A Video Dataset towards World Exploration
Zhen Li, Chuanhao Li, Xiaofeng Mao +17
cs.CVcs.AIarXiv:2506.15675v32025A Scoping Review of AI-Driven Digital Interventions in Mental Health Care: Mapping Applications Across Screening, Support, Monitoring, Prevention, and Clinical Education
Yang Ni, Fanli Jia
cs.CYcs.AIcs.HCarXiv:2603.16204v12026VTool-R1: VLMs Learn to Think with Images via Reinforcement Learning on Multimodal Tool Use
Mingyuan Wu, Jingcheng Yang, Jize Jiang +6
cs.LGcs.AIarXiv:2505.19255v42025A Convenient Category for Higher-Order Probability Theory
Chris Heunen, Ohad Kammar, Sam Staton +1
cs.PLcs.AIcs.LOarXiv:1701.02547v42017Creating General User Models from Computer Use
Omar Shaikh, Shardul Sapkota, Shan Rizvi +4
cs.HCcs.AIcs.CLarXiv:2505.10831v32025HDPO: Hybrid Distillation Policy Optimization via Privileged Self-Distillation
Ken Ding
cs.LGcs.AIarXiv:2603.23871v12026Hallucinated Neural Radiance Fields in the Wild
Xingyu Chen, Qi Zhang, Xiaoyu Li +4
cs.CVcs.AIarXiv:2111.15246v32021Cartridges: Lightweight and general-purpose long context representations via self-study
Sabri Eyuboglu, Ryan Ehrlich, Simran Arora +8
cs.CLcs.AIcs.LGarXiv:2506.06266v32025Detect Before You Attribute: Cascade Failure Attribution for Multi-Agent Systems
Jiayi Zhang, Zexin Wang, Degang Sun +4
cs.AIcs.MAarXiv:2608.29646v12026The Invisible Leash: Why RLVR May or May Not Escape Its Origin
Fang Wu, Weihao Xuan, Ximing Lu +4
cs.LGcs.AIcs.CLarXiv:2507.14843v42025Towards Understanding Camera Motions in Any Video
Zhiqiu Lin, Siyuan Cen, Daniel Jiang +12
cs.CVcs.AIcs.CLarXiv:2504.15376v22025Urban Driver: Learning to Drive from Real-world Demonstrations Using Policy Gradients
Oliver Scheel, Luca Bergamini, Maciej Wołczyk +2
cs.ROcs.AIcs.CVarXiv:2109.13333v12021Text2Reward: Reward Shaping with Language Models for Reinforcement Learning
Tianbao Xie, Siheng Zhao, Chen Henry Wu +5
cs.LGcs.AIcs.CLarXiv:2309.11489v32023Parametrized quantum policies for reinforcement learning
Sofiene Jerbi, Casper Gyurik, Simon C. Marshall +2
quant-phcs.AIcs.LGarXiv:2103.05577v22021MedRAX: Medical Reasoning Agent for Chest X-ray
Adibvafa Fallahpour, Jun Ma, Alif Munim +2
cs.LGcs.AIcs.MAarXiv:2502.02673v22025Overtrained Language Models Are Harder to Fine-Tune
Jacob Mitchell Springer, Sachin Goyal, Kaiyue Wen +5
cs.CLcs.AIarXiv:2503.19206v22025Character-Level Question Answering with Attention
David Golub, Xiaodong He
cs.CLcs.AIcs.LGarXiv:1604.00727v42016LHM: Large Animatable Human Reconstruction Model from a Single Image in Seconds
Lingteng Qiu, Xiaodong Gu, Peihao Li +8
cs.CVcs.AIarXiv:2503.10625v12025Generating 3D Molecules for Target Protein Binding
Meng Liu, Youzhi Luo, Kanji Uchino +2
q-bio.BMcs.AIcs.LGarXiv:2204.09410v22022SS-ESOAP: Self-Scaled Adaptive Preconditioning for Physics-Informed Learning
Guangyuan Wang, Mads Toftrup, Sebastian Loeschcke +2
cs.LGcs.AImath.OCarXiv:2608.29448v12026Beyond Single-Turn: A Survey on Multi-Turn Interactions with Large Language Models
Yubo Li, Xiaobin Shen, Yidi Miao +4
cs.CLcs.AIarXiv:2504.04717v62025AdaMatch: A Unified Approach to Semi-Supervised Learning and Domain Adaptation
David Berthelot, Rebecca Roelofs, Kihyuk Sohn +2
cs.LGcs.AIcs.CVarXiv:2106.04732v22021A Survey on Vision-Language-Action Models for Autonomous Driving
Sicong Jiang, Zilin Huang, Kangan Qian +17
cs.CVcs.AIcs.ROarXiv:2506.24044v12025MediQ: Question-Asking LLMs and a Benchmark for Reliable Interactive Clinical Reasoning
Shuyue Stella Li, Vidhisha Balachandran, Shangbin Feng +4
cs.CLcs.AIarXiv:2406.00922v32024EnvScaler: Scaling Tool-Interactive Environments for LLM Agent via Programmatic Synthesis
Xiaoshuai Song, Haofei Chang, Guanting Dong +3
cs.CLcs.AIcs.LGarXiv:2601.05808v22026Reducing Overestimation Bias in Multi-Agent Domains Using Double Centralized Critics
Johannes Ackermann, Volker Gabler, Takayuki Osa +1
cs.LGcs.AIcs.MAarXiv:1910.01465v22019Diffusion Language Models Know the Answer Before Decoding
Pengxiang Li, Yefan Zhou, Dilxat Muhtar +5
cs.CLcs.AIarXiv:2508.19982v52025Retrieval, Scoring, and Decoding Shape Performance and Stability in LLM-based Conversational Recommendation
Ante Kapetanovic, Tomislav Duricic, Andro Mercep +1
cs.CLcs.AIarXiv:2609.00086v12026Open Benchmarking for Click-Through Rate Prediction
Jieming Zhu, Jinyang Liu, Shuai Yang +2
cs.IRcs.AIarXiv:2009.05794v62020The Answer Is Not the Argument
Will Yeadon, Sergio Juárez, Paul Mackay +5
cs.AIarXiv:2609.00264v12026UAVs Meet LLMs: Overviews and Perspectives Toward Agentic Low-Altitude Mobility
Yonglin Tian, Fei Lin, Yiduo Li +11
cs.ROcs.AIarXiv:2501.02341v22025Textless Speech-to-Speech Translation on Real Data
Ann Lee, Hongyu Gong, Paul-Ambroise Duquenne +8
cs.CLcs.AIcs.LGarXiv:2112.08352v22021Humanoid Policy ~ Human Policy
Ri-Zhao Qiu, Shiqi Yang, Xuxin Cheng +12
cs.ROcs.AIcs.CVarXiv:2503.13441v32025PAGE-RAG: Provenance-Aware Graph Evidence Promotion for Fixed-Budget Multi-hop Retrieval-Augmented Generation
Haokun Deng, Xunkai Li, Hongchao Qin +1
cs.AIarXiv:2608.29753v12026Acting Less is Reasoning More! Teaching Model to Act Efficiently
Hongru Wang, Cheng Qian, Wanjun Zhong +7
cs.AIcs.CLarXiv:2504.14870v22025Mind2Web 2: Evaluating Agentic Search with Agent-as-a-Judge
Boyu Gou, Zanming Huang, Yuting Ning +23
cs.AIcs.CLarXiv:2506.21506v22025Deep-Learned Collision Avoidance Policy for Distributed Multi-Agent Navigation
Pinxin Long, Wenxi Liu, Jia Pan
cs.AIcs.CVcs.ROarXiv:1609.06838v22016MoLe-VLA: Dynamic Layer-skipping Vision Language Action Model via Mixture-of-Layers for Efficient Robot Manipulation
Rongyu Zhang, Menghang Dong, Yuan Zhang +6
cs.ROcs.AIarXiv:2503.20384v22025The Illusion of Diminishing Returns: Measuring Long Horizon Execution in LLMs
Akshit Sinha, Arvindh Arun, Shashwat Goel +2
cs.AIarXiv:2509.09677v32025Deep Learning for Source Code Modeling and Generation: Models, Applications and Challenges
Triet H. M. Le, Hao Chen, M. Ali Babar
cs.SEcs.AIcs.LGarXiv:2002.05442v12020Near Optimal Behavior via Approximate State Abstraction
David Abel, D. Ellis Hershkowitz, Michael L. Littman
cs.LGcs.AIarXiv:1701.04113v12017Router-R1: Teaching LLMs Multi-Round Routing and Aggregation via Reinforcement Learning
Haozhen Zhang, Tao Feng, Jiaxuan You
cs.CLcs.AIcs.LGarXiv:2506.09033v32025Taming OpenClaw: Security Analysis and Mitigation of Autonomous LLM Agent Threats
Xinhao Deng, Yixiang Zhang, Jiaqing Wu +15
cs.CRcs.AIarXiv:2603.11619v12026Efficient Reasoning Models: A Survey
Sicheng Feng, Gongfan Fang, Xinyin Ma +1
cs.CLcs.AIarXiv:2504.10903v22025SealQA: Raising the Bar for Reasoning in Search-Augmented Language Models
Thinh Pham, Nguyen Nguyen, Pratibha Zunjare +3
cs.CLcs.AIcs.LGarXiv:2506.01062v42025Talk Isn't Always Cheap: Understanding Failure Modes in Multi-Agent Debate
Andrea Wynn, Harsh Satija, Gillian Hadfield
cs.CLcs.AIcs.MAarXiv:2509.05396v22025Learning Latent Action World Models In The Wild
Quentin Garrido, Tushar Nagarajan, Basile Terver +3
cs.AIcs.CVarXiv:2601.05230v22026Synthetic Data Generation with Large Language Models for Text Classification: Potential and Limitations
Zhuoyan Li, Hangxiao Zhu, Zhuoran Lu +1
cs.CLcs.AIarXiv:2310.07849v22023