Artificial Intelligence

Papers filed under cs.AI on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.

Search paper metadata (including unsummarized papers)

421 to 480 of 15,235

  1. Efficient Multi-turn RL for GUI Agents via Decoupled Training and Adaptive Data Curation

    Pengxiang Li, Zechen Hu, Zirui Shang +15

    cs.LGcs.AIcs.CVarXiv:2509.23866v12025
  2. Distributed JEPA: A Self-Supervised Framework for Energy Forecasting

    Liana Toderean, Tudor Cioara, Vasilis Michalakopoulos +3

    cs.LGcs.AIarXiv:2609.17029v12026
  3. Interactive Memory Learning for Long-Term Conversations

    Cai Ke, Jiangyue Yan, Han Zhang +5

    cs.AIcs.CLarXiv:2609.17088v12026
  4. ThinkFlow: Self-Evolving Probabilistic Latent Memory for Lifelong Conversational Agents

    Cai Ke, Xin Liu, Han Zhang +6

    cs.AIcs.CLarXiv:2609.17010v12026
  5. Controlled Decoding from Language Models

    Sidharth Mudgal, Jong Lee, Harish Ganapathy +10

    cs.LGcs.AIcs.CLarXiv:2310.17022v32023
  6. WebQA: Multihop and Multimodal QA

    Yingshan Chang, Mridu Narang, Hisami Suzuki +3

    cs.CLcs.AIcs.CVarXiv:2109.00590v42021
  7. CoAdapt: An LLM-based Framework for Adaptive Collaborative Perception in IIoT Robotic Swarms

    Houssam Hajj Hassan, Antonia Maria Masucci, Lynda Zitoune +1

    cs.AIcs.ROarXiv:2609.16852v12026
  8. LimiX-2: A Contextual Mechanism Network Towards General Structured-Data Intelligence

    Xingxuan Zhang, Gang Ren, Hao Yuan +57

    cs.AIarXiv:2609.17488v12026
  9. InRTL: Effective Intra-Inter Interaction Learning for Relational Tables

    Weichen Li, Ken Zhong, Zheng Wang +2

    cs.LGcs.AIarXiv:2609.12712v12026
  10. EHRSHOT: An EHR Benchmark for Few-Shot Evaluation of Foundation Models

    Michael Wornow, Rahul Thapa, Ethan Steinberg +2

    cs.LGcs.AIcs.CLarXiv:2307.02028v32023
  11. SIMS: Scale-Invariant Merit-Function-Based Scalarization for Multi-Task Learning

    Zebin Chen, Fei Xing, Yang Chen +4

    cs.LGcs.AImath.OCarXiv:2609.12599v12026
  12. Let the LLMs Talk: Simulating Human-to-Human Conversational QA via Zero-Shot LLM-to-LLM Interactions

    Zahra Abbasiantaeb, Yifei Yuan, Evangelos Kanoulas +1

    cs.CLcs.AIcs.IRarXiv:2312.02913v12023
  13. Chain of Thoughtlessness? An Analysis of CoT in Planning

    Kaya Stechly, Karthik Valmeekam, Subbarao Kambhampati

    cs.AIarXiv:2405.04776v32024
  14. Diagnosing Faults in Reinforcement Learning Simulators and World Models with Canonical Polynomial Invariants

    Tesfay Zemuy Gebrekidan, Hadush Hailu Gebrerufael

    cs.LGcs.AIarXiv:2609.13194v12026
  15. World Models for Embodied Intelligence: From Plausible to Controllable to Actionable

    Nanjie Yao, Hao Wang, Chong Cheng +10

    cs.ROcs.AIarXiv:2609.16697v12026
  16. On the Importance of Gating: Memorization vs. In-Context Learning in State Space Models

    William L. Tong, Aryo Lotfi, Emmanuel Abbe +6

    cs.LGcs.AIarXiv:2609.16540v12026
  17. Vision-based Human Fall Detection Systems using Deep Learning: A Review

    Ekram Alam, Abu Sufian, Paramartha Dutta +1

    cs.CVcs.AIarXiv:2207.10952v12022
  18. Continual Learning for Traversability Prediction with Uncertainty-Aware Adaptation

    Hojin Lee, Yunho Lee, Daniel A Duecker +1

    cs.ROcs.AIcs.LGarXiv:2609.17141v12026
  19. A Multi-Vehicle Dataset with Camera, LiDAR, and Radar Sensors and Scanned 3D Models for Custom Auto-Annotation using RTK-GNSS

    Philipp Berthold, Bianca Forkel, Mirko Maehlisch

    cs.ROcs.AIcs.CVarXiv:2609.12871v12026
  20. The Router Within: Eliciting Native Skill Routing from a Frozen LLM

    Ruishuo Chen, Xun Wang, Yu Chen +2

    cs.LGcs.AIcs.CLarXiv:2609.15982v12026
  21. RLP: Reinforcement as a Pretraining Objective

    Ali Hatamizadeh, Syeda Nahida Akter, Shrimai Prabhumoye +5

    cs.LGcs.AIcs.CLarXiv:2510.01265v22025
  22. AI Policies: Help or Hindrance? A Software Developer's Perspective

    Samuel Ferino, Rashina Hoda, John Grundy +2

    cs.SEcs.AIarXiv:2609.16496v12026
  23. ResLRP: The Role of Residual Cancellation in Attribution Instability in Vision Transformers

    Jim Berend, Reduan Achtibat, Daniel Schäffer +4

    cs.CVcs.AIcs.LGarXiv:2609.17152v12026
  24. Meta-Prompting: Enhancing Language Models with Task-Agnostic Scaffolding

    Mirac Suzgun, Adam Tauman Kalai

    cs.CLcs.AIcs.HCarXiv:2401.12954v12024
  25. FluxVLA Engine: A One-Stop VLA Engineering Platform for Embodied Intelligence

    Yinhao Li, Weixin Mao, Zihan Lan +21

    cs.ROcs.AIarXiv:2609.17210v12026
  26. LLMs as Master Forgers: Generating Synthetic Time Series Data for Manufacturing

    Mantek Singh, Jeshwanth Challagundla, Prateek Karnal +3

    cs.LGcs.AIarXiv:2609.16155v12026
  27. Decoder Design Matters for ECG Delineation

    Joseph Scharpf, William Han, Chaojing Duan +3

    cs.LGcs.AIarXiv:2609.16489v12026
  28. The MAL Simulator: Cyber Operations Simulation based on Attack & Defense Graphs

    Jakob Nyberg, Sandor Berglund, Andrei Buhaiu +3

    cs.CRcs.AIarXiv:2609.16563v12026
  29. Can Large Language Models Serve as Rational Players in Game Theory? A Systematic Analysis

    Caoyun Fan, Jindou Chen, Yaohui Jin +1

    cs.AIcs.CLcs.GTarXiv:2312.05488v22023
  30. Vision And Text Transformer For Predicting Answerability On Visual Question Answering

    Tung Le, Huy Tien Nguyen, Le Minh Nguyen

    cs.CVcs.AIarXiv:2609.16565v12026
  31. gradSim: Differentiable simulation for system identification and visuomotor control

    Krishna Murthy Jatavallabhula, Miles Macklin, Florian Golemo +11

    cs.CVcs.AIcs.LGarXiv:2104.02646v12021
  32. WebGLM: Towards An Efficient Web-Enhanced Question Answering System with Human Preferences

    Xiao Liu, Hanyu Lai, Hao Yu +6

    cs.CLcs.AIarXiv:2306.07906v12023
  33. A Cyber Range Evaluation of Autonomous Network Incident Response Agents

    Jakob Nyberg, Teodor Sommestad, Andrei Buhaiu +3

    cs.CRcs.AIarXiv:2609.16541v12026
  34. Neuro-Symbolic Hierarchical Intention Anticipation in Human Behavior

    Farnaz Soleimani, Abdelghani Chibani, Yacine Amirat +1

    cs.AIcs.CVcs.HCarXiv:2609.17064v12026
  35. Contrastive Triple Extraction with Generative Transformer

    Hongbin Ye, Ningyu Zhang, Shumin Deng +4

    cs.CLcs.AIcs.DBarXiv:2009.06207v82020
  36. How to estimate carbon footprint when training deep learning models? A guide and review

    Lucia Bouza Heguerte, Aurélie Bugeau, Loïc Lannelongue

    cs.LGcs.AIcs.CYarXiv:2306.08323v22023
  37. Natural-Language to SysMLv2 Translation via Conformance-Driven Iterative Refinement

    Chance LaVoie, Eladio Andujar Lugo, Taylan G. Topcu +1

    cs.SEcs.AIarXiv:2607.14162v12026
  38. FlexKBQA: A Flexible LLM-Powered Framework for Few-Shot Knowledge Base Question Answering

    Zhenyu Li, Sunqi Fan, Yu Gu +5

    cs.CLcs.AIarXiv:2308.12060v32023
  39. The METRIC-framework for assessing data quality for trustworthy AI in medicine: a systematic review

    Daniel Schwabe, Katinka Becker, Martin Seyferth +2

    cs.LGcs.AIarXiv:2402.13635v12024
  40. UI-S1: Advancing GUI Automation via Semi-online Reinforcement Learning

    Zhengxi Lu, Jiabo Ye, Fei Tang +8

    cs.LGcs.AIarXiv:2509.11543v22025
  41. Self-Improving LLM Agents at Test-Time

    Emre Can Acikgoz, Cheng Qian, Heng Ji +2

    cs.LGcs.AIcs.CLarXiv:2510.07841v12025
  42. InfLLM-V2: Dense-Sparse Switchable Attention for Seamless Short-to-Long Adaptation

    Weilin Zhao, Zihan Zhou, Zhou Su +10

    cs.CLcs.AIcs.LGarXiv:2509.24663v12025
  43. Repurposing Deep Limit Order Book Forecasting for Scenario-Conditioned Market Impact Modeling

    Eljas Linna, Kestutis Baltakys, Derrick Manoharan +2

    cs.LGcs.AIarXiv:2609.16930v12026
  44. Generalization Can Emerge in Tabular Foundation Models From a Single Table

    Junwei Ma, Nour Shaheen, Alex Labach +4

    cs.LGcs.AIarXiv:2511.09665v12025
  45. MUMINS: Metadata-conditioned Uncertainty-aware Medical Image Next-state Synthesis

    Anna Oliveras, Roger Marí, Rafael Redondo +7

    cs.CVcs.AIarXiv:2609.17169v12026
  46. Can AI systems have free will?

    Christian List

    cs.AIcs.CYphysics.soc-pharXiv:2609.15407v12026
  47. Word meaning in minds and machines

    Brenden M. Lake, Gregory L. Murphy

    cs.CLcs.AIcs.LGarXiv:2008.01766v32020
  48. Scalability and Performance Evaluation of Federated Learning Frameworks: A Comparative Analysis

    Bassel Soudan, Sohail Abbas, Ahmed Kubba +2

    cs.DCcs.AIarXiv:2609.15681v12026
  49. Verifiable Social Reasoning for LLM Assistants

    Amir Taubenfeld, Zorik Gekhman, Avigail Grinstein-Dabush +6

    cs.AIcs.CLarXiv:2609.17496v12026
  50. Generalizable machine learning for stress monitoring from wearable devices: A systematic literature review

    Gideon Vos, Kelly Trinh, Zoltan Sarnyai +1

    cs.AIarXiv:2209.15137v32022
  51. DetGPT: Detect What You Need via Reasoning

    Renjie Pi, Jiahui Gao, Shizhe Diao +8

    cs.CVcs.AIarXiv:2305.14167v22023
  52. Towards Optimizing SQL Generation via LLM Routing

    Mohammadhossein Malekpour, Nour Shaheen, Foutse Khomh +1

    cs.DBcs.AIcs.LGarXiv:2411.04319v12024
  53. Representing Numbers in NLP: a Survey and a Vision

    Avijit Thawani, Jay Pujara, Pedro A. Szekely +1

    cs.CLcs.AIcs.LGarXiv:2103.13136v12021
  54. Integrating AI and Learning Analytics for Data-Driven Pedagogical Decisions and Personalized Interventions in Education

    Ramteja Sajja, Yusuf Sermet, David Cwiertny +1

    cs.CYcs.AIcs.HCarXiv:2312.09548v22023
  55. Seeing What Matters: Visual Cue Guided Video Planning for Generalizable Robot Navigation

    Hojin Lee, Sizhe Lester Li, Maximilian Hilger +4

    cs.ROcs.AIcs.CVarXiv:2609.16737v12026
  56. Verified Multi-Agent Orchestration: A Plan-Execute-Verify-Replan Framework for Complex Query Resolution

    Xing Zhang, Yanwei Cui, Guanghui Wang +7

    cs.AIcs.MAarXiv:2603.11445v22026
  57. RewardHackingAgents: Benchmarking Evaluation Integrity for LLM ML-Engineering Agents

    Yonas Atinafu, Robin Cohen

    cs.AIarXiv:2603.11337v12026
  58. Vroom-Vroom at SHROOM-Visions: A Multi-Judge Committee for Detecting Hallucinated Spans in Vision-Language Outputs

    Toqeer Ehsan, Nico Penttilä, Richard Schmidt +2

    cs.CLcs.AIarXiv:2609.17327v12026
  59. Enabling Creative Exploration for Vibe Design Agents

    Yifan Zhang, Nghi D. Q. Bui, Georgios Evangelopoulos +1

    cs.AIarXiv:2609.15078v12026
  60. Intelligence per Watt: Measuring Intelligence Efficiency of Local AI

    Jon Saad-Falcon, Avanika Narayan, Hakki Orhun Akengin +12

    cs.DCcs.AIcs.CLarXiv:2511.07885v62025