Artificial Intelligence
Papers filed under cs.AI on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.
Search paper metadata (including unsummarized papers)
2,581 to 2,640 of 15,404
ARC-Bench: Closed-Loop Replanning Masks Broken Action Ranking in Frozen JEPA World Models
Zhengshu Zhang, Zhiyuan Li
cs.AIcs.LGcs.ROarXiv:2609.05461v12026TANGO: Humanoid Navigation in Cluttered Environments with a Whole-Body Vision-Language-Action Model
Anqi Li, Yuxin Chen, Zhaobo Li +4
cs.ROcs.AIarXiv:2609.09158v12026Procedural Graphs: Self-Evolving Execution Structures for LLM Agents
Yuxing Lu, Yicheng Chen, Shanchan Wu +1
cs.AIcs.CLcs.MAarXiv:2609.09153v12026AutoFyn Technical Report: Non-Parametric Expert Iteration for Long-Horizon Agents
Adib Hasan, Daniel Schaffield, Akashnil Dutta +1
cs.AIarXiv:2609.05446v12026Reason Through the Latent! Making Latent Visual Reasoning Necessary
Suhyeong Park, Junha Jung, Jaewoo Kang
cs.AIcs.CLcs.CVarXiv:2609.06746v12026Compiling VGDL into Causal Models
Mohit Jiwatode, Bodo Rosenhahn, Alexander Dockhorn
cs.AIarXiv:2609.05459v12026Navigation with Large Language Models: Semantic Guesswork as a Heuristic for Planning
Dhruv Shah, Michael Equi, Blazej Osinski +3
cs.ROcs.AIcs.CLarXiv:2310.10103v12023Damage-Aware Bandit Pruning for Vision and Language Transformers
Salem Ameen, Sunil Vadera
cs.AIcs.LGarXiv:2609.05448v12026Minimizing Energy Consumption Leads to the Emergence of Gaits in Legged Robots
Zipeng Fu, Ashish Kumar, Jitendra Malik +1
cs.ROcs.AIcs.CVarXiv:2111.01674v12021Does Splitting a Triage Decision Across Agents Hide Bias or Help Catch It? A Multi-Agent Simulation Study of LLM-Based Resource Allocation Under Audit Capacity Constraints
Paul-Peter Arslan
cs.AIcs.MAarXiv:2608.06949v12026When Does Memory Help? A Cost-Aware Evaluation of Long-Term Memory in Tool-Using LLM Agents
Shweta Mishra, Shashank Mishra
cs.AIarXiv:2609.05441v12026CriticGen: Generation-Aware Evaluation as Actionable Feedback
Huifang Du, Zecheng Zuo, Sen Wang +3
cs.AIarXiv:2609.05439v12026EgoBody: Human Body Shape and Motion of Interacting People from Head-Mounted Devices
Siwei Zhang, Qianli Ma, Yan Zhang +5
cs.CVcs.AIarXiv:2112.07642v32021Beyond Right and Wrong: Evaluating Second-order Social Reasoning in Large Language Models
Sunny Rai, Jinyi Kuang, Reyhan Jamalova +7
cs.AIcs.CYarXiv:2609.05437v12026BEHAVIOR-1K: A Human-Centered, Embodied AI Benchmark with 1,000 Everyday Activities and Realistic Simulation
Chengshu Li, Ruohan Zhang, Josiah Wong +32
cs.ROcs.AIarXiv:2403.09227v12024Large Language Models in Mental Health: A Systematic Review of Applications, Innovations, and Ethical Challenges
Yisong Chen, Yifan Gao, Sijing Yu +2
cs.AIarXiv:2608.18080v12026Steering Geometry: Validating Human Value Geometry in LLM Steering Space
Mohammad Mahdi Abootorabi, Armin Saghafian, Ali Bazshoushtari +5
cs.CLcs.AIcs.LGarXiv:2609.06289v12026CLIP-Dissect: Automatic Description of Neuron Representations in Deep Vision Networks
Tuomas Oikarinen, Tsui-Wei Weng
cs.CVcs.AIcs.LGarXiv:2204.10965v52022AFRA: Argumentation framework with recursive attacks
Pietro Baroni, Federico Cerutti, Massimiliano Giacomin +1
cs.AIarXiv:1810.04886v12018TimeSHAP: Explaining Recurrent Models through Sequence Perturbations
João Bento, Pedro Saleiro, André F. Cruz +2
cs.LGcs.AIarXiv:2012.00073v22020Black-box Explanation of Object Detectors via Saliency Maps
Vitali Petsiuk, Rajiv Jain, Varun Manjunatha +4
cs.CVcs.AIcs.LGarXiv:2006.03204v22020Environments as Scaffold: Enriching Feedback to Bootstrap Self-Evolving Agents in Long-Horizon Tasks
Hongbang Yuan, Zhuoran Jin, Yixin Cao
cs.LGcs.AIcs.CLarXiv:2609.08404v12026A Unified Physics-Aware Quantum Machine Learning Framework across Power GaN HEMTs and Logic Nanowire FETs: Predicting Unseen Process Splits and Held-Out Geometry Combinations with Lower Error and Tighter Split-to-Split Variability
Rushat Rai, Yun-Yuan Wang, Autsada Kakaen +10
cs.AIarXiv:2609.05251v12026Unsupervised Depth Completion from Visual Inertial Odometry
Alex Wong, Xiaohan Fei, Stephanie Tsuei +1
cs.CVcs.AIcs.LGarXiv:1905.08616v42019Improving Diffusion Inverse Problem Solving with Decoupled Noise Annealing
Bingliang Zhang, Wenda Chu, Julius Berner +3
cs.LGcs.AIcs.CVarXiv:2407.01521v32024Learning State Representations for Query Optimization with Deep Reinforcement Learning
Jennifer Ortiz, Magdalena Balazinska, Johannes Gehrke +1
cs.DBcs.AIcs.LGarXiv:1803.08604v12018Retro*: Learning Retrosynthetic Planning with Neural Guided A* Search
Binghong Chen, Chengtao Li, Hanjun Dai +1
cs.LGcs.AIstat.MLarXiv:2006.15820v12020Kalman Delta Networks: Uncertainty-aware Associative Memory
Ngoc Bui, Tinglin Huang, Rex Ying
cs.LGcs.AIarXiv:2609.07816v12026What LLM Trading Agents Actually Do in Production: A Six-Month, Population-Scale Record from Two Fleets
T. J. Barton, Chris Constantakis, Patti Hauseman +4
cs.AIcs.CEcs.MAarXiv:2609.05663v12026Fine-tuning large language models for domain adaptation: Exploration of training strategies, scaling, model merging and synergistic capabilities
Wei Lu, Rachel K. Luu, Markus J. Buehler
cs.CLcond-mat.mtrl-scics.AIarXiv:2409.03444v12024A survey on intrinsic motivation in reinforcement learning
Arthur Aubret, Laetitia Matignon, Salima Hassas
cs.LGcs.AIarXiv:1908.06976v22019Ethical Challenges and Evolving Strategies in the Integration of Artificial Intelligence into Clinical Practice
Ellison B. Weiner, Irene Dankwa-Mullan, William A. Nelson +1
cs.CYcs.AIarXiv:2412.03576v12024Teach Me to Explain: A Review of Datasets for Explainable Natural Language Processing
Sarah Wiegreffe, Ana Marasović
cs.CLcs.AIcs.LGarXiv:2102.12060v42021BLEU might be Guilty but References are not Innocent
Markus Freitag, David Grangier, Isaac Caswell
cs.CLcs.AIcs.LGarXiv:2004.06063v22020OCRBench v2: An Improved Benchmark for Evaluating Large Multimodal Models on Visual Text Localization and Reasoning
Ling Fu, Zhebin Kuang, Jiajun Song +21
cs.CVcs.AIarXiv:2501.00321v22024Predicting Clinical Events by Combining Static and Dynamic Information Using Recurrent Neural Networks
Cristóbal Esteban, Oliver Staeck, Yinchong Yang +1
cs.LGcs.AIcs.NEarXiv:1602.02685v22016Efficient Algorithms for Outlier-Robust Regression
Adam Klivans, Pravesh K. Kothari, Raghu Meka
cs.LGcs.AIcs.DSarXiv:1803.03241v32018LLaMEA: A Large Language Model Evolutionary Algorithm for Automatically Generating Metaheuristics
Niki van Stein, Thomas Bäck
cs.NEcs.AIarXiv:2405.20132v42024How to Build the Virtual Cell with Artificial Intelligence: Priorities and Opportunities
Charlotte Bunne, Yusuf Roohani, Yanay Rosen +39
q-bio.QMcs.AIcs.LGarXiv:2409.11654v22024Guidelines and Evaluation of Clinical Explainable AI in Medical Image Analysis
Weina Jin, Xiaoxiao Li, Mostafa Fatehi +1
cs.LGcs.AIcs.CVarXiv:2202.10553v32022Seeing Before Synthesizing: VLM-Guided Transition Event Discovery for Weakly-Supervised Dense Video Captioning
Ye-Chan Kim, Seunghee Choi, SeungJu Cha +4
cs.CVcs.AIarXiv:2609.04183v12026Sequential Beats Joint: On the Interplay between On-Policy Distillation and RLVR
Boyan Li, Bingsen Chen, Chenghao Yang +3
cs.CLcs.AIcs.LGarXiv:2609.04108v22026Language Agents with Reinforcement Learning for Strategic Play in the Werewolf Game
Zelai Xu, Chao Yu, Fei Fang +2
cs.AIcs.LGcs.MAarXiv:2310.18940v42023SWE-Gate: Passing Functional Tests Is Not Enough for Software Engineering Agents
Xin He, Yanlin Wang, Mingwei Liu +3
cs.SEcs.AIarXiv:2609.04167v12026A Theoretical Explanation for Perplexing Behaviors of Backpropagation-based Visualizations
Weili Nie, Yang Zhang, Ankit Patel
cs.CVcs.AIarXiv:1805.07039v42018Structured Adversarial Attack: Towards General Implementation and Better Interpretability
Kaidi Xu, Sijia Liu, Pu Zhao +6
cs.LGcs.AIstat.MLarXiv:1808.01664v32018ESPO: Error-Structured Prompt Optimization via Diagnose, Diversify, and Stabilize
Lihao Liu, Peng Tang, Kunwar Yashraj Singh +1
cs.CLcs.AIarXiv:2609.04197v12026MAP Estimation, Linear Programming and Belief Propagation with Convex Free Energies
Yair Weiss, Chen Yanover, Talya Meltzer
cs.AIcs.LGstat.MLarXiv:1206.5286v12012Knowledge Acquisition During Pre-training? Large Language Models Learn Better With Auxiliary Views
Joseph Lee, Yidi Huang, Dokyoon Kim +2
cs.CLcs.AIarXiv:2609.04180v12026A Transformer-Based Model With Self-Distillation for Multimodal Emotion Recognition in Conversations
Hui Ma, Jian Wang, Hongfei Lin +3
cs.AIcs.MMarXiv:2310.20494v12023Bayesian Action Decoder for Deep Multi-Agent Reinforcement Learning
Jakob N. Foerster, Francis Song, Edward Hughes +5
cs.MAcs.AIcs.LGarXiv:1811.01458v32018SENTINEL-RL: Offloading Topological Reasoning from LLM Agents in the Security Operations Center
Uday Vallabhaneni, Cassie L. Cagwin, David J. Wild
cs.CRcs.AIarXiv:2609.04159v12026Deep Network Guided Proof Search
Sarah Loos, Geoffrey Irving, Christian Szegedy +1
cs.AIcs.LGcs.LOarXiv:1701.06972v12017A Low-Cost, Open Platform for End-to-End Autonomous Driving on a Miniature Ackermann Vehicle
Gustavo Claudio Karl Couto, Eric Aislan Antonelo, Gabriel George Zipperer
cs.LGcs.AIcs.ROarXiv:2609.04147v12026Learning to Understand Goal Specifications by Modelling Reward
Dzmitry Bahdanau, Felix Hill, Jan Leike +4
cs.AIcs.LGarXiv:1806.01946v42018Max-value Entropy Search for Multi-Objective Bayesian Optimization with Constraints
Syrine Belakaria, Aryan Deshwal, Janardhan Rao Doppa
cs.LGcs.AIstat.MLarXiv:2009.01721v22020SDRL: Interpretable and Data-efficient Deep Reinforcement Learning Leveraging Symbolic Planning
Daoming Lyu, Fangkai Yang, Bo Liu +1
cs.AIarXiv:1811.00090v42018Federated Unsupervised Representation Learning
Fengda Zhang, Kun Kuang, Zhaoyang You +6
cs.LGcs.AIarXiv:2010.08982v12020Exponential Moving Average Normalization for Self-supervised and Semi-supervised Learning
Zhaowei Cai, Avinash Ravichandran, Subhransu Maji +3
cs.LGcs.AIcs.CVarXiv:2101.08482v22021Representational alignment yields generalizable safety in language models
Lingyu Li, Yan Teng, Yingchun Wang +1
cs.CLcs.AIarXiv:2609.04022v12026