Artificial Intelligence

Papers filed under cs.AI on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.

Search paper metadata (including unsummarized papers)

4,261 to 4,320 of 15,291

  1. Hallucination Mitigation for Large Vision-Language Models via Implicit Feature Stabilization

    Aditi Sarker, Rafi Ibn Sultan, Hui Zhu +2

    cs.CVcs.AIcs.LGarXiv:2608.29924v12026
  2. GPT Models in Construction Industry: Opportunities, Limitations, and a Use Case Validation

    Abdullahi Saka, Ridwan Taiwo, Nurudeen Saka +4

    cs.HCcs.AIcs.CLarXiv:2305.18997v12023
  3. URLB: Unsupervised Reinforcement Learning Benchmark

    Michael Laskin, Denis Yarats, Hao Liu +6

    cs.LGcs.AIcs.ROarXiv:2110.15191v12021
  4. Reachability-Based Capability Confinement for LLM Agents under Indirect Prompt Injection

    Wujie Xiong, Rabimba Karanjai, Yang Lu +2

    cs.CRcs.AIarXiv:2608.30041v12026
  5. Stable On-Policy Distillation through Adaptive Target Reformulation

    Ijun Jang, Jewon Yeom, Juan Yeo +2

    cs.LGcs.AIarXiv:2601.07155v32026
    Summaries:한국어
  6. Ross3D: Reconstructive Visual Instruction Tuning with 3D-Awareness

    Haochen Wang, Yucheng Zhao, Tiancai Wang +3

    cs.CVcs.AIcs.CLarXiv:2504.01901v12025
  7. Generalizable Multi-Agent Planning from Signal Temporal Logic Specifications via Diffusion

    Joe Eappen, Zikang Xiong, Shreyash S. Iyengar +1

    cs.MAcs.AIcs.ROarXiv:2608.29490v12026
  8. Computational Aspects of Multi-Winner Approval Voting

    Haris Aziz, Serge Gaspers, Joachim Gudmundsson +3

    cs.GTcs.AIcs.MAarXiv:1407.3247v12014
  9. Compact Language Models via Pruning and Knowledge Distillation

    Saurav Muralidharan, Sharath Turuvekere Sreenivas, Raviraj Joshi +6

    cs.CLcs.AIcs.LGarXiv:2407.14679v22024
  10. Scalable Neural Decoders for Practical Fault-Tolerant Quantum Computation

    Andi Gu, J. Pablo Bonilla Ataides, Mikhail D. Lukin +1

    quant-phcs.AIcs.LGarXiv:2604.08358v12026
  11. Scalable methods for computing state similarity in deterministic Markov Decision Processes

    Pablo Samuel Castro

    cs.LGcs.AIstat.MLarXiv:1911.09291v12019
  12. The Hidden Life of Tokens: Reducing Hallucination of Large Vision-Language Models via Visual Information Steering

    Zhuowei Li, Haizhou Shi, Yunhe Gao +7

    cs.CVcs.AIcs.LGarXiv:2502.03628v22025
  13. AIOpsLab: A Holistic Framework to Evaluate AI Agents for Enabling Autonomous Clouds

    Yinfang Chen, Manish Shetty, Gagan Somashekar +6

    cs.AIcs.DCcs.MAarXiv:2501.06706v12025
  14. The role of artificial intelligence in achieving the Sustainable Development Goals

    Ricardo Vinuesa, Hossein Azizpour, Iolanda Leite +7

    cs.CYcs.AIarXiv:1905.00501v12019
  15. Agent RL Scaling Law: Agent RL with Spontaneous Code Execution for Mathematical Problem Solving

    Xinji Mai, Haotian Xu, Zhong-Zhi Li +5

    cs.AIarXiv:2505.07773v42025
  16. Mechanistically Eliciting Latent Behaviors in Language Models

    Andrew Mack, Nina Panickssery, Alexander Matt Turner

    cs.LGcs.AIarXiv:2606.29604v12026
  17. LLM-Rec: Personalized Recommendation via Prompting Large Language Models

    Hanjia Lyu, Song Jiang, Hanqing Zeng +7

    cs.CLcs.AIcs.IRarXiv:2307.15780v32023
  18. The Complexity of Bayesian Network Learning: Revisiting the Superstructure

    Robert Ganian, Viktoriia Korchemna

    cs.DScs.AIarXiv:2602.10253v12026
  19. CGFM-Nav: Cognitive Graph-Field Memory for Semantic-Guided Lifelong Multimodal Embodied Navigation

    Yuxiang Xiao, Xibei Chen, Xin Zhou +3

    cs.ROcs.AIarXiv:2608.29114v12026
  20. Large Language Models for Time Series: A Survey

    Xiyuan Zhang, Ranak Roy Chowdhury, Rajesh K. Gupta +1

    cs.LGcs.AIcs.CLarXiv:2402.01801v32024
  21. Scalable Clinical Data Infrastructure and Comparative ML Evaluation for Hospitalisation Risk Prediction in Elderly Patients with Multiple Long-Term Conditions using CPRD

    Asra Aslam, Volodymyr Chapman, Maurice M. O'Connell +9

    cs.LGcs.AIarXiv:2608.29419v12026
  22. GAN Lab: Understanding Complex Deep Generative Models using Interactive Visual Experimentation

    Minsuk Kahng, Nikhil Thorat, Duen Horng Chau +2

    cs.HCcs.AIcs.LGarXiv:1809.01587v12018
  23. An Explainable Coherence Score for Detecting Temporal Inconsistencies in Political News

    Marius Nicusor Pantea, Adrian Groza

    cs.AIarXiv:2608.29175v12026
  24. Spatiotemporal-aware Trend-Seasonality Decomposition Network for Traffic Flow Forecasting

    Lingxiao Cao, Bin Wang, Guiyuan Jiang +2

    cs.LGcs.AIarXiv:2502.12213v12025
  25. Deep Reinforcement Learning for Optimal Portfolio Allocation: A Comparative Study with Mean-Variance Optimization

    Srijan Sood, Kassiani Papasotiriou, Marius Vaiciulis +1

    q-fin.PMcs.AIcs.LGarXiv:2602.17098v12026
  26. A Survey on Knowledge-Oriented Retrieval-Augmented Generation

    Mingyue Cheng, Yucong Luo, Jie Ouyang +9

    cs.CLcs.AIarXiv:2503.10677v32025
  27. LLM-SRBench: A New Benchmark for Scientific Equation Discovery with Large Language Models

    Parshin Shojaee, Ngoc-Hieu Nguyen, Kazem Meidani +3

    cs.CLcs.AIcs.LGarXiv:2504.10415v22025
  28. ATLAS: Learning to Optimally Memorize the Context at Test Time

    Ali Behrouz, Zeman Li, Praneeth Kacham +5

    cs.CLcs.AIarXiv:2505.23735v12025
  29. Feature Level Fusion of Face and Fingerprint Biometrics

    Ajita Rattani, Dakshina Ranjan Kisku, Manuele Bicego +1

    cs.CVcs.AIarXiv:1002.2523v12010
  30. Learning Adaptive Parallel Reasoning with Language Models

    Jiayi Pan, Xiuyu Li, Long Lian +6

    cs.AIcs.CLarXiv:2504.15466v22025
  31. First Proof

    Mohammed Abouzaid, Andrew J. Blumberg, Martin Hairer +8

    cs.AImath.AGmath.COarXiv:2602.05192v22026
  32. Not All Agreement Counts as Corroboration: Provenance-Conserving Multi-View Fusion for Typed Action Admission in Human-Robot Collaboration

    Zekai Jin, Hanrong Zhang, Yihong Tang +3

    cs.ROcs.AIarXiv:2609.01662v12026
  33. Urban Driving with Conditional Imitation Learning

    Jeffrey Hawke, Richard Shen, Corina Gurau +8

    cs.CVcs.AIcs.LGarXiv:1912.00177v22019
  34. A Survey of LLM-based Deep Search Agents: Paradigm, Optimization, Evaluation, and Challenges

    Yunjia Xi, Jianghao Lin, Yongzhao Xiao +7

    cs.IRcs.AIcs.CLarXiv:2508.05668v32025
  35. MCPSecBench: A Systematic Security Benchmark and Playground for Testing Model Context Protocols

    Yixuan Yang, Cuifeng Gao, Daoyuan Wu +3

    cs.CRcs.AIarXiv:2508.13220v32025
  36. Stabilizing MoE Reinforcement Learning by Aligning Training and Inference Routers

    Wenhan Ma, Hailin Zhang, Liang Zhao +4

    cs.CLcs.AIcs.LGarXiv:2510.11370v22025
  37. EmbodiedScan: A Holistic Multi-Modal 3D Perception Suite Towards Embodied AI

    Tai Wang, Xiaohan Mao, Chenming Zhu +11

    cs.CVcs.AIcs.ROarXiv:2312.16170v12023
  38. Eigenanalysis framework for autoregressive neural emulators of multi-scale chaotic dynamics

    Conrad Ainslie, Pedram Hassanzadeh, Michael W. Mahoney +1

    cs.AIcs.LGnlin.CDarXiv:2608.16084v12026
  39. On the Limitations of Representing Functions on Sets

    Edward Wagstaff, Fabian B. Fuchs, Martin Engelcke +2

    cs.LGcs.AIcs.NEarXiv:1901.09006v22019
  40. PostTrainBench: Can LLM Agents Automate LLM Post-Training?

    Ben Rank, Hardik Bhatnagar, Ameya Prabhu +4

    cs.SEcs.AIcs.LGarXiv:2603.08640v22026
  41. Sheared LLaMA: Accelerating Language Model Pre-training via Structured Pruning

    Mengzhou Xia, Tianyu Gao, Zhiyuan Zeng +1

    cs.CLcs.AIcs.LGarXiv:2310.06694v22023
  42. Release Strategies and the Social Impacts of Language Models

    Irene Solaiman, Miles Brundage, Jack Clark +12

    cs.CLcs.AIcs.CYarXiv:1908.09203v22019
  43. Comet: Fine-grained Computation-communication Overlapping for Mixture-of-Experts

    Shulai Zhang, Ningxin Zheng, Haibin Lin +9

    cs.DCcs.AIcs.LGarXiv:2502.19811v32025
  44. Drug Similarity Integration Through Attentive Multi-view Graph Auto-Encoders

    Tengfei Ma, Cao Xiao, Jiayu Zhou +1

    cs.LGcs.AIstat.MLarXiv:1804.10850v12018
  45. Training-Free Hidden-State Refinement for Flow-Matching Image Generators

    Yuanyi Yan, Xinzhe Rao, Canyu Shen +5

    cs.CVcs.AIarXiv:2608.29160v12026
  46. Advancing Mathematics Research with AI-Driven Formal Proof Search

    George Tsoukalas, Anton Kovsharov, Sergey Shirobokov +18

    cs.AIarXiv:2605.22763v22026
  47. Large Language Models and Games: A Survey and Roadmap

    Roberto Gallotta, Graham Todd, Marvin Zammit +4

    cs.CLcs.AIcs.HCarXiv:2402.18659v52024
  48. InfiGUIAgent: A Multimodal Generalist GUI Agent with Native Reasoning and Reflection

    Yuhang Liu, Pengxiang Li, Zishu Wei +7

    cs.AIcs.CLcs.HCarXiv:2501.04575v12025
  49. EvoFlint: An Evolutionary Atlas of Multi-Turn LLM Vulnerabilities

    Feitong Qiao, Liren Peng, Shiming Ren +7

    cs.CLcs.AIcs.CRarXiv:2609.00487v12026
  50. SCROLLS: Standardized CompaRison Over Long Language Sequences

    Uri Shaham, Elad Segal, Maor Ivgi +8

    cs.CLcs.AIcs.LGarXiv:2201.03533v22022
  51. Demystifying Reasoning Dynamics with Mutual Information: Thinking Tokens are Information Peaks in LLM Reasoning

    Chen Qian, Dongrui Liu, Haochen Wen +3

    cs.AIcs.CLarXiv:2506.02867v22025
  52. Write, Execute, Assess: Program Synthesis with a REPL

    Kevin Ellis, Maxwell Nye, Yewen Pu +3

    cs.PLcs.AIcs.LGarXiv:1906.04604v12019
  53. Raw2Drive: Reinforcement Learning with Aligned World Models for End-to-End Autonomous Driving (in CARLA v2)

    Zhenjie Yang, Xiaosong Jia, Qifeng Li +3

    cs.ROcs.AIcs.CVarXiv:2505.16394v22025
  54. A Human-AI Theorem Connecting Spontaneous and Field-Induced Mechanisms of Collective Behavior in One Dimension

    Weiguo Yin

    cond-mat.stat-mechcs.AIcs.HCarXiv:2609.00322v12026
  55. Workload Identification with Physical Side Channels for AI Governance

    Simone Gargiulo, Gabriel Kulp

    cs.CRcs.AIcs.CYarXiv:2609.00309v12026
  56. MoFlow: One-Step Flow Matching for Human Trajectory Forecasting via Implicit Maximum Likelihood Estimation based Distillation

    Yuxiang Fu, Qi Yan, Lele Wang +2

    cs.CVcs.AIcs.LGarXiv:2503.09950v12025
  57. Lost in Simulation: LLM-Simulated Users are Unreliable Proxies for Human Users in Agentic Evaluations

    Preethi Seshadri, Samuel Cahyawijaya, Ayomide Odumakinde +2

    cs.HCcs.AIcs.CYarXiv:2601.17087v22026
  58. The CoT Collection: Improving Zero-shot and Few-shot Learning of Language Models via Chain-of-Thought Fine-Tuning

    Seungone Kim, Se June Joo, Doyoung Kim +4

    cs.CLcs.AIcs.LGarXiv:2305.14045v22023
  59. Voting or Consensus? Decision-Making in Multi-Agent Debate

    Lars Benedikt Kaesberg, Jonas Becker, Jan Philip Wahle +2

    cs.MAcs.AIcs.CLarXiv:2502.19130v42025
  60. Explainable Artificial Intelligence (XAI) on TimeSeries Data: A Survey

    Thomas Rojat, Raphaël Puget, David Filliat +3

    cs.LGcs.AIarXiv:2104.00950v12021