Artificial Intelligence

Papers filed under cs.AI on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.

Search paper metadata (including unsummarized papers)

3,421 to 3,480 of 15,389

  1. LLM Reasoning as Trajectories: Step-Specific Representation Geometry and Correctness Signals

    Lihao Sun, Hang Dong, Bo Qiao +3

    cs.CLcs.AIcs.LGarXiv:2604.05655v12026
  2. SQL-Zero: Self-Evolving Text-to-SQL

    Daniel Machado Pedrozo, Julia Soares Dollis, Bryan Lincoln Marques de Oliveira +3

    cs.AIarXiv:2609.04697v12026
  3. Predicting Spatiotemporal Mobile Sensing-Based PM2.5 Concentrations Using Low-Rank Adapted Spatially Attentive Graph Neural Network

    Om Chiddarwar, Priyanka Mandal, Praveen Kumar Chandaliya +1

    cs.AIarXiv:2609.04693v12026
  4. Cornerstones or Stumbling Blocks? Deciphering the Rock Tokens in On-Policy Distillation

    Yuxuan Jiang, Runchao Li, Shubhashis Roy Dipta +2

    cs.CLcs.AIarXiv:2605.09253v42026
  5. Large Language Model Cascades with Mixture of Thoughts Representations for Cost-efficient Reasoning

    Murong Yue, Jie Zhao, Min Zhang +2

    cs.CLcs.AIcs.LGarXiv:2310.03094v32023
  6. SkillSieve: A Hierarchical Triage Framework for Detecting Malicious AI Agent Skills

    Yinghan Hou, Zongyou Yang

    cs.CRcs.AIarXiv:2604.06550v32026
  7. Contextual Agentic Memory is a Memo, Not True Memory

    Binyan Xu, Xilin Dai, Kehuan Zhang

    cs.AIcs.CLarXiv:2604.27707v22026
  8. Multimodal Transfer: A Hierarchical Deep Convolutional Neural Network for Fast Artistic Style Transfer

    Xin Wang, Geoffrey Oxholm, Da Zhang +1

    cs.CVcs.AIarXiv:1612.01895v22016
  9. Automated Algorithm Selection on Continuous Black-Box Problems By Combining Exploratory Landscape Analysis and Machine Learning

    Pascal Kerschke, Heike Trautmann

    stat.MLcs.AIcs.DSarXiv:1711.08921v32017
  10. Continual Graph Memory for Adaptive Recommendation under Intent Drift

    Hao Nguyen Ngoc, Tung Nguyen, Nguyen Thi Hanh +2

    cs.AIarXiv:2609.04651v12026
  11. Long-Range Indoor Navigation with PRM-RL

    Anthony Francis, Aleksandra Faust, Hao-Tien Lewis Chiang +4

    cs.ROcs.AIcs.LGarXiv:1902.09458v22019
  12. A Cost-Aware Agentic Architecture for NL-to-SQL over Nested Enterprise Schemas, with a New Benchmark

    Yoga Sri Varshan Varadharajan, Ajay Yadav, Ritesh Goru +5

    cs.AIarXiv:2609.04641v12026
  13. MantisV2: Closing the Zero-Shot Gap in Time Series Classification with Synthetic Data and Test-Time Strategies

    Vasilii Feofanov, Songkang Wen, Jianfeng Zhang +2

    cs.LGcs.AIarXiv:2602.17868v12026
  14. Benchmarking LLM Tool-Use in the Wild

    Peijie Yu, Wei Liu, Yifan Yang +4

    cs.HCcs.AIcs.CLarXiv:2604.06185v12026
  15. How well are open sourced AI-generated image detection models out-of-the-box: A comprehensive benchmark study

    Simiao Ren, Yuchen Zhou, Xingyu Shen +9

    cs.CVcs.AIarXiv:2602.07814v12026
  16. Behaviorally Grounded User Profiles from the Wild for Personalized Alignment and Multi-Perspective Reasoning

    Yuxuan Li, Victor Zhong, Ehsan Kamalloo

    cs.CLcs.AIarXiv:2609.00014v12026
  17. SiLR: Structure-Preserving Admission and Process Reward for LLM Tool Agents

    Chenyu Zhou, Qiliang Jiang, Shuning Wu +1

    cs.AIcs.LGeess.SYarXiv:2609.04629v12026
  18. Action-to-Action Flow Matching

    Jindou Jia, Gen Li, Xiangyu Chen +5

    cs.ROcs.AIarXiv:2602.07322v22026
  19. Does the Selected Object Reach the Reader? Auditing Identity Handoffs in Grounded Language-Model Pipelines

    Siddharth Vohra, Runmin Jiang, Xiaomo Li +1

    cs.AIcs.CLarXiv:2609.04579v12026
  20. Tell Me About Yourself: Using an AI-Powered Chatbot to Conduct Conversational Surveys with Open-ended Questions

    Ziang Xiao, Michelle X. Zhou, Q. Vera Liao +4

    cs.HCcs.AIarXiv:1905.10700v22019
  21. Unsupervised Time-Series Representation Learning with Iterative Bilinear Temporal-Spectral Fusion

    Ling Yang, Shenda Hong

    cs.LGcs.AIarXiv:2202.04770v32022
  22. MoltNet: Understanding Social Behavior of AI Agents in the Agent-Native MoltBook

    Yi Feng, Chen Huang, Zhibo Man +4

    cs.SIcs.AIarXiv:2602.13458v22026
  23. Reward-Oracle MCTS for Formal Theorem Proving: Sample-Efficient Search and the Need for Kernel-Level Proof Auditing

    Bodla Krishna Vamshi, Haizhao Yang

    cs.AIcs.LGcs.LOarXiv:2608.28639v12026
  24. A Tractable Inference Algorithm for Diagnosing Multiple Diseases

    David Heckerman

    cs.AIarXiv:1304.1511v22013
  25. Better Document-level Sentiment Analysis from RST Discourse Parsing

    Parminder Bhatia, Yangfeng Ji, Jacob Eisenstein

    cs.CLcs.AIarXiv:1509.01599v22015
  26. DexMimicGen: Automated Data Generation for Bimanual Dexterous Manipulation via Imitation Learning

    Zhenyu Jiang, Yuqi Xie, Kevin Lin +5

    cs.ROcs.AIcs.CVarXiv:2410.24185v22024
  27. AgentWorm: Self-Propagating Attacks Across LLM Agent Ecosystems

    Yihao Zhang, Zeming Wei, Xiaokun Luan +7

    cs.CRcs.AIcs.LGarXiv:2603.15727v32026
  28. Measuring Similarity between Artistic and AI Generated Images using Siamese Neural Networks

    Diego Castro Elvira, Navil Pineda Rugerio, Jesús García-Ramírez +2

    cs.CVcs.AIarXiv:2608.28671v12026
  29. OpenMathInstruct-1: A 1.8 Million Math Instruction Tuning Dataset

    Shubham Toshniwal, Ivan Moshkov, Sean Narenthiran +3

    cs.CLcs.AIcs.LGarXiv:2402.10176v22024
  30. A Rational Analysis of the Effects of Sycophantic AI

    Rafael M. Batista, Thomas L. Griffiths

    cs.CYcs.AIcs.HCarXiv:2602.14270v12026
  31. From Extraction to Governed Memory: Multi-Agent Knowledge Graph Construction with Domain-Expert Review

    Pranav Bykampadi, Neel Mokaria, Vishesh Narayan +2

    cs.AIcs.CEcs.DLarXiv:2608.28642v12026
  32. Reuse your FLOPs: Scaling RL on Hard Problems by Conditioning on Very Off-Policy Prefixes

    Amrith Setlur, Zijian Wang, Andrew Cohen +2

    cs.LGcs.AIcs.CLarXiv:2601.18795v22026
  33. VLABench: A Large-Scale Benchmark for Language-Conditioned Robotics Manipulation with Long-Horizon Reasoning Tasks

    Shiduo Zhang, Zhe Xu, Peiju Liu +8

    cs.ROcs.AIcs.CLarXiv:2412.18194v12024
  34. New Inference Rules for Max-SAT

    C. M. Li, F. Manya, J. Planes

    cs.AIarXiv:1111.0040v12011
  35. Autonomous Evolution of EDA Tools: Multi-Agent Self-Evolved ABC

    Cunxi Yu, Haoxing Ren

    cs.ARcs.AIarXiv:2604.15082v12026
  36. Grading Scale Impact on LLM-as-a-Judge: Human-LLM Alignment Is Highest on 0-5 Grading Scale

    Weiyue Li, Minda Zhao, Weixuan Dong +12

    cs.CLcs.AIcs.HCarXiv:2601.03444v12026
  37. Interpolated Policy Gradient: Merging On-Policy and Off-Policy Gradient Estimation for Deep Reinforcement Learning

    Shixiang Gu, Timothy Lillicrap, Zoubin Ghahramani +3

    cs.LGcs.AIcs.ROarXiv:1706.00387v12017
  38. Agent-as-a-Judge

    Runyang You, Hongru Cai, Caiqi Zhang +5

    cs.CLcs.AIarXiv:2601.05111v12026
  39. Shifting Attention to Relevance: Towards the Predictive Uncertainty Quantification of Free-Form Large Language Models

    Jinhao Duan, Hao Cheng, Shiqi Wang +5

    cs.CLcs.AIcs.LGarXiv:2307.01379v32023
  40. How Language Models Choose Sides: Internal Representations of Instruction Hierarchy

    Enrique Balp-Straffon, Chih-Hao Hsu, Rushiraj Gadhvi +3

    cs.AIcs.CLcs.LGarXiv:2608.28648v12026
  41. PLeak: Prompt Leaking Attacks against Large Language Model Applications

    Bo Hui, Haolin Yuan, Neil Gong +2

    cs.CRcs.AIcs.LGarXiv:2405.06823v32024
  42. PlasticineLab: A Soft-Body Manipulation Benchmark with Differentiable Physics

    Zhiao Huang, Yuanming Hu, Tao Du +4

    cs.LGcs.AIcs.CVarXiv:2104.03311v12021
  43. Peer Oversight in Collective Decision Making

    Sarah Mohsen, Pavel Naumov

    cs.GTcs.AIcs.MAarXiv:2608.28754v12026
  44. The reach of a verification tool decides its value: A controlled study of verification surface, artifact quality, and cost in AI coding agents

    Achint Mehta

    cs.SEcs.AIarXiv:2608.28795v12026
  45. Coding Agents are Effective Long-Context Processors

    Weili Cao, Xunjian Yin, Bhuwan Dhingra +1

    cs.CLcs.AIarXiv:2603.20432v12026
  46. A parallel corpus of Python functions and documentation strings for automated code documentation and code generation

    Antonio Valerio Miceli Barone, Rico Sennrich

    cs.CLcs.AIarXiv:1707.02275v12017
  47. The Halt Vector: Internalizing a Causal Steering Intervention for Efficient Reasoning

    Dylan Jayabahu, Tinuade Adeleke

    cs.LGcs.AIcs.CLarXiv:2608.28859v12026
  48. WebXSkill: Skill Learning for Autonomous Web Agents

    Zhaoyang Wang, Qianhui Wu, Xuchao Zhang +12

    cs.AIcs.CLarXiv:2604.13318v22026
  49. Improving Named Entity Recognition by External Context Retrieving and Cooperative Learning

    Xinyu Wang, Yong Jiang, Nguyen Bach +4

    cs.CLcs.AIcs.LGarXiv:2105.03654v32021
  50. Variational Probabilistic Inference and the QMR-DT Network

    T. S. Jaakkola, M. I. Jordan

    cs.AIarXiv:1105.5462v12011
  51. Competitive Coevolution through Evolutionary Complexification

    R. Miikkulainen, K. O. Stanley

    cs.AIarXiv:1107.0037v12011
  52. SkillJect: Effectively Automating Skill-Based Prompt Injection for Skill-Enabled Agents

    Xiaojun Jia, Jie Liao, Simeng Qin +5

    cs.CRcs.AIarXiv:2602.14211v32026
  53. OWL2Vec*: Embedding of OWL Ontologies

    Jiaoyan Chen, Pan Hu, Ernesto Jimenez-Ruiz +3

    cs.AIarXiv:2009.14654v22020
  54. Good Counterfactuals and Where to Find Them: A Case-Based Technique for Generating Counterfactuals for Explainable AI (XAI)

    Mark T. Keane, Barry Smyth

    cs.AIarXiv:2005.13997v12020
  55. Spec-Driven Development:From Code to Contract in the Age of AI Coding Assistants

    Deepak Babu Piskala

    cs.SEcs.AIarXiv:2602.00180v12026
  56. Auditing Reasoning-Trace Memorization Claims after Unlearning with Head-Conditioned Canaries

    Yanhang Li, Zhichao Fan, Zexin Zhuang

    cs.LGcs.AIarXiv:2605.18891v12026
  57. Survey and cross-benchmark comparison of remaining time prediction methods in business process monitoring

    Ilya Verenich, Marlon Dumas, Marcello La Rosa +2

    cs.AIcs.LGarXiv:1805.02896v22018
  58. OctoBench: Benchmarking Scaffold-Aware Instruction Following in Repository-Grounded Agentic Coding

    Deming Ding, Shichun Liu, Enhui Yang +12

    cs.CLcs.AIarXiv:2601.10343v22026
  59. Independent SE(3)-Equivariant Models for End-to-End Rigid Protein Docking

    Octavian-Eugen Ganea, Xinyuan Huang, Charlotte Bunne +4

    cs.AIcs.LGarXiv:2111.07786v22021
  60. Reinforced Generation of Combinatorial Structures: Ramsey Numbers

    Ansh Nagda, Prabhakar Raghavan, Abhradeep Thakurta

    math.COcs.AIcs.CCarXiv:2603.09172v52026