Artificial Intelligence

Papers filed under cs.AI on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.

Search paper metadata (including unsummarized papers)

2,581 to 2,640 of 15,404

  1. ARC-Bench: Closed-Loop Replanning Masks Broken Action Ranking in Frozen JEPA World Models

    Zhengshu Zhang, Zhiyuan Li

    cs.AIcs.LGcs.ROarXiv:2609.05461v12026
  2. TANGO: Humanoid Navigation in Cluttered Environments with a Whole-Body Vision-Language-Action Model

    Anqi Li, Yuxin Chen, Zhaobo Li +4

    cs.ROcs.AIarXiv:2609.09158v12026
  3. Procedural Graphs: Self-Evolving Execution Structures for LLM Agents

    Yuxing Lu, Yicheng Chen, Shanchan Wu +1

    cs.AIcs.CLcs.MAarXiv:2609.09153v12026
  4. AutoFyn Technical Report: Non-Parametric Expert Iteration for Long-Horizon Agents

    Adib Hasan, Daniel Schaffield, Akashnil Dutta +1

    cs.AIarXiv:2609.05446v12026
  5. Reason Through the Latent! Making Latent Visual Reasoning Necessary

    Suhyeong Park, Junha Jung, Jaewoo Kang

    cs.AIcs.CLcs.CVarXiv:2609.06746v12026
  6. Compiling VGDL into Causal Models

    Mohit Jiwatode, Bodo Rosenhahn, Alexander Dockhorn

    cs.AIarXiv:2609.05459v12026
  7. Navigation with Large Language Models: Semantic Guesswork as a Heuristic for Planning

    Dhruv Shah, Michael Equi, Blazej Osinski +3

    cs.ROcs.AIcs.CLarXiv:2310.10103v12023
  8. Damage-Aware Bandit Pruning for Vision and Language Transformers

    Salem Ameen, Sunil Vadera

    cs.AIcs.LGarXiv:2609.05448v12026
  9. Minimizing Energy Consumption Leads to the Emergence of Gaits in Legged Robots

    Zipeng Fu, Ashish Kumar, Jitendra Malik +1

    cs.ROcs.AIcs.CVarXiv:2111.01674v12021
  10. Does Splitting a Triage Decision Across Agents Hide Bias or Help Catch It? A Multi-Agent Simulation Study of LLM-Based Resource Allocation Under Audit Capacity Constraints

    Paul-Peter Arslan

    cs.AIcs.MAarXiv:2608.06949v12026
  11. When Does Memory Help? A Cost-Aware Evaluation of Long-Term Memory in Tool-Using LLM Agents

    Shweta Mishra, Shashank Mishra

    cs.AIarXiv:2609.05441v12026
  12. CriticGen: Generation-Aware Evaluation as Actionable Feedback

    Huifang Du, Zecheng Zuo, Sen Wang +3

    cs.AIarXiv:2609.05439v12026
  13. EgoBody: Human Body Shape and Motion of Interacting People from Head-Mounted Devices

    Siwei Zhang, Qianli Ma, Yan Zhang +5

    cs.CVcs.AIarXiv:2112.07642v32021
  14. Beyond Right and Wrong: Evaluating Second-order Social Reasoning in Large Language Models

    Sunny Rai, Jinyi Kuang, Reyhan Jamalova +7

    cs.AIcs.CYarXiv:2609.05437v12026
  15. BEHAVIOR-1K: A Human-Centered, Embodied AI Benchmark with 1,000 Everyday Activities and Realistic Simulation

    Chengshu Li, Ruohan Zhang, Josiah Wong +32

    cs.ROcs.AIarXiv:2403.09227v12024
  16. Large Language Models in Mental Health: A Systematic Review of Applications, Innovations, and Ethical Challenges

    Yisong Chen, Yifan Gao, Sijing Yu +2

    cs.AIarXiv:2608.18080v12026
  17. Steering Geometry: Validating Human Value Geometry in LLM Steering Space

    Mohammad Mahdi Abootorabi, Armin Saghafian, Ali Bazshoushtari +5

    cs.CLcs.AIcs.LGarXiv:2609.06289v12026
  18. CLIP-Dissect: Automatic Description of Neuron Representations in Deep Vision Networks

    Tuomas Oikarinen, Tsui-Wei Weng

    cs.CVcs.AIcs.LGarXiv:2204.10965v52022
  19. AFRA: Argumentation framework with recursive attacks

    Pietro Baroni, Federico Cerutti, Massimiliano Giacomin +1

    cs.AIarXiv:1810.04886v12018
  20. TimeSHAP: Explaining Recurrent Models through Sequence Perturbations

    João Bento, Pedro Saleiro, André F. Cruz +2

    cs.LGcs.AIarXiv:2012.00073v22020
  21. Black-box Explanation of Object Detectors via Saliency Maps

    Vitali Petsiuk, Rajiv Jain, Varun Manjunatha +4

    cs.CVcs.AIcs.LGarXiv:2006.03204v22020
  22. Environments as Scaffold: Enriching Feedback to Bootstrap Self-Evolving Agents in Long-Horizon Tasks

    Hongbang Yuan, Zhuoran Jin, Yixin Cao

    cs.LGcs.AIcs.CLarXiv:2609.08404v12026
  23. A Unified Physics-Aware Quantum Machine Learning Framework across Power GaN HEMTs and Logic Nanowire FETs: Predicting Unseen Process Splits and Held-Out Geometry Combinations with Lower Error and Tighter Split-to-Split Variability

    Rushat Rai, Yun-Yuan Wang, Autsada Kakaen +10

    cs.AIarXiv:2609.05251v12026
  24. Unsupervised Depth Completion from Visual Inertial Odometry

    Alex Wong, Xiaohan Fei, Stephanie Tsuei +1

    cs.CVcs.AIcs.LGarXiv:1905.08616v42019
  25. Improving Diffusion Inverse Problem Solving with Decoupled Noise Annealing

    Bingliang Zhang, Wenda Chu, Julius Berner +3

    cs.LGcs.AIcs.CVarXiv:2407.01521v32024
  26. Learning State Representations for Query Optimization with Deep Reinforcement Learning

    Jennifer Ortiz, Magdalena Balazinska, Johannes Gehrke +1

    cs.DBcs.AIcs.LGarXiv:1803.08604v12018
  27. Retro*: Learning Retrosynthetic Planning with Neural Guided A* Search

    Binghong Chen, Chengtao Li, Hanjun Dai +1

    cs.LGcs.AIstat.MLarXiv:2006.15820v12020
  28. Kalman Delta Networks: Uncertainty-aware Associative Memory

    Ngoc Bui, Tinglin Huang, Rex Ying

    cs.LGcs.AIarXiv:2609.07816v12026
  29. What LLM Trading Agents Actually Do in Production: A Six-Month, Population-Scale Record from Two Fleets

    T. J. Barton, Chris Constantakis, Patti Hauseman +4

    cs.AIcs.CEcs.MAarXiv:2609.05663v12026
  30. Fine-tuning large language models for domain adaptation: Exploration of training strategies, scaling, model merging and synergistic capabilities

    Wei Lu, Rachel K. Luu, Markus J. Buehler

    cs.CLcond-mat.mtrl-scics.AIarXiv:2409.03444v12024
  31. A survey on intrinsic motivation in reinforcement learning

    Arthur Aubret, Laetitia Matignon, Salima Hassas

    cs.LGcs.AIarXiv:1908.06976v22019
  32. Ethical Challenges and Evolving Strategies in the Integration of Artificial Intelligence into Clinical Practice

    Ellison B. Weiner, Irene Dankwa-Mullan, William A. Nelson +1

    cs.CYcs.AIarXiv:2412.03576v12024
  33. Teach Me to Explain: A Review of Datasets for Explainable Natural Language Processing

    Sarah Wiegreffe, Ana Marasović

    cs.CLcs.AIcs.LGarXiv:2102.12060v42021
  34. BLEU might be Guilty but References are not Innocent

    Markus Freitag, David Grangier, Isaac Caswell

    cs.CLcs.AIcs.LGarXiv:2004.06063v22020
  35. OCRBench v2: An Improved Benchmark for Evaluating Large Multimodal Models on Visual Text Localization and Reasoning

    Ling Fu, Zhebin Kuang, Jiajun Song +21

    cs.CVcs.AIarXiv:2501.00321v22024
  36. Predicting Clinical Events by Combining Static and Dynamic Information Using Recurrent Neural Networks

    Cristóbal Esteban, Oliver Staeck, Yinchong Yang +1

    cs.LGcs.AIcs.NEarXiv:1602.02685v22016
  37. Efficient Algorithms for Outlier-Robust Regression

    Adam Klivans, Pravesh K. Kothari, Raghu Meka

    cs.LGcs.AIcs.DSarXiv:1803.03241v32018
  38. LLaMEA: A Large Language Model Evolutionary Algorithm for Automatically Generating Metaheuristics

    Niki van Stein, Thomas Bäck

    cs.NEcs.AIarXiv:2405.20132v42024
  39. How to Build the Virtual Cell with Artificial Intelligence: Priorities and Opportunities

    Charlotte Bunne, Yusuf Roohani, Yanay Rosen +39

    q-bio.QMcs.AIcs.LGarXiv:2409.11654v22024
  40. Guidelines and Evaluation of Clinical Explainable AI in Medical Image Analysis

    Weina Jin, Xiaoxiao Li, Mostafa Fatehi +1

    cs.LGcs.AIcs.CVarXiv:2202.10553v32022
  41. Seeing Before Synthesizing: VLM-Guided Transition Event Discovery for Weakly-Supervised Dense Video Captioning

    Ye-Chan Kim, Seunghee Choi, SeungJu Cha +4

    cs.CVcs.AIarXiv:2609.04183v12026
  42. Sequential Beats Joint: On the Interplay between On-Policy Distillation and RLVR

    Boyan Li, Bingsen Chen, Chenghao Yang +3

    cs.CLcs.AIcs.LGarXiv:2609.04108v22026
  43. Language Agents with Reinforcement Learning for Strategic Play in the Werewolf Game

    Zelai Xu, Chao Yu, Fei Fang +2

    cs.AIcs.LGcs.MAarXiv:2310.18940v42023
  44. SWE-Gate: Passing Functional Tests Is Not Enough for Software Engineering Agents

    Xin He, Yanlin Wang, Mingwei Liu +3

    cs.SEcs.AIarXiv:2609.04167v12026
  45. A Theoretical Explanation for Perplexing Behaviors of Backpropagation-based Visualizations

    Weili Nie, Yang Zhang, Ankit Patel

    cs.CVcs.AIarXiv:1805.07039v42018
  46. Structured Adversarial Attack: Towards General Implementation and Better Interpretability

    Kaidi Xu, Sijia Liu, Pu Zhao +6

    cs.LGcs.AIstat.MLarXiv:1808.01664v32018
  47. ESPO: Error-Structured Prompt Optimization via Diagnose, Diversify, and Stabilize

    Lihao Liu, Peng Tang, Kunwar Yashraj Singh +1

    cs.CLcs.AIarXiv:2609.04197v12026
  48. MAP Estimation, Linear Programming and Belief Propagation with Convex Free Energies

    Yair Weiss, Chen Yanover, Talya Meltzer

    cs.AIcs.LGstat.MLarXiv:1206.5286v12012
  49. Knowledge Acquisition During Pre-training? Large Language Models Learn Better With Auxiliary Views

    Joseph Lee, Yidi Huang, Dokyoon Kim +2

    cs.CLcs.AIarXiv:2609.04180v12026
  50. A Transformer-Based Model With Self-Distillation for Multimodal Emotion Recognition in Conversations

    Hui Ma, Jian Wang, Hongfei Lin +3

    cs.AIcs.MMarXiv:2310.20494v12023
  51. Bayesian Action Decoder for Deep Multi-Agent Reinforcement Learning

    Jakob N. Foerster, Francis Song, Edward Hughes +5

    cs.MAcs.AIcs.LGarXiv:1811.01458v32018
  52. SENTINEL-RL: Offloading Topological Reasoning from LLM Agents in the Security Operations Center

    Uday Vallabhaneni, Cassie L. Cagwin, David J. Wild

    cs.CRcs.AIarXiv:2609.04159v12026
  53. Deep Network Guided Proof Search

    Sarah Loos, Geoffrey Irving, Christian Szegedy +1

    cs.AIcs.LGcs.LOarXiv:1701.06972v12017
  54. A Low-Cost, Open Platform for End-to-End Autonomous Driving on a Miniature Ackermann Vehicle

    Gustavo Claudio Karl Couto, Eric Aislan Antonelo, Gabriel George Zipperer

    cs.LGcs.AIcs.ROarXiv:2609.04147v12026
  55. Learning to Understand Goal Specifications by Modelling Reward

    Dzmitry Bahdanau, Felix Hill, Jan Leike +4

    cs.AIcs.LGarXiv:1806.01946v42018
  56. Max-value Entropy Search for Multi-Objective Bayesian Optimization with Constraints

    Syrine Belakaria, Aryan Deshwal, Janardhan Rao Doppa

    cs.LGcs.AIstat.MLarXiv:2009.01721v22020
  57. SDRL: Interpretable and Data-efficient Deep Reinforcement Learning Leveraging Symbolic Planning

    Daoming Lyu, Fangkai Yang, Bo Liu +1

    cs.AIarXiv:1811.00090v42018
  58. Federated Unsupervised Representation Learning

    Fengda Zhang, Kun Kuang, Zhaoyang You +6

    cs.LGcs.AIarXiv:2010.08982v12020
  59. Exponential Moving Average Normalization for Self-supervised and Semi-supervised Learning

    Zhaowei Cai, Avinash Ravichandran, Subhransu Maji +3

    cs.LGcs.AIcs.CVarXiv:2101.08482v22021
  60. Representational alignment yields generalizable safety in language models

    Lingyu Li, Yan Teng, Yingchun Wang +1

    cs.CLcs.AIarXiv:2609.04022v12026