Artificial Intelligence

Papers filed under cs.AI on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.

Search paper metadata (including unsummarized papers)

4,621 to 4,680 of 15,246

  1. CacheBridge: Efficient Cross-Model KV Cache Transfer

    Xingyu Qu, Siyuan Lu, Zhiyu Chen +2

    cs.AIarXiv:2609.00891v12026
  2. Transformers are Graph Neural Networks

    Chaitanya K. Joshi

    cs.LGcs.AIarXiv:2506.22084v12025
  3. One Policy, Any Budget: Internalizing Budget-Aware Search via Reinforcement Learning

    Xiaowei Sun, Jin Li, Yili Hong +2

    cs.AIarXiv:2609.00813v12026
  4. Value Over Language Model: Detecting Original Contribution in Writing

    Vibhhu Sharma, Thorsten Joachims, Sarah Dean

    cs.AIcs.CLarXiv:2609.00700v12026
  5. AgentRewardBench: Evaluating Automatic Evaluations of Web Agent Trajectories

    Xing Han Lù, Amirhossein Kazemnejad, Nicholas Meade +7

    cs.LGcs.AIcs.CLarXiv:2504.08942v22025
  6. RL-100: Performant Robotic Manipulation with Real-World Reinforcement Learning

    Kun Lei, Huanyu Li, Dongjie Yu +6

    cs.ROcs.AIcs.LGarXiv:2510.14830v42025
  7. Learning to Predict the Cosmological Structure Formation

    Siyu He, Yin Li, Yu Feng +4

    astro-ph.COcs.AIcs.LGarXiv:1811.06533v22018
  8. Reasoning-SQL: Reinforcement Learning with SQL Tailored Partial Rewards for Reasoning-Enhanced Text-to-SQL

    Mohammadreza Pourreza, Shayan Talaei, Ruoxi Sun +5

    cs.LGcs.AIcs.DBarXiv:2503.23157v22025
  9. GPT Takes the Bar Exam

    Michael Bommarito, Daniel Martin Katz

    cs.CLcs.AIcs.LGarXiv:2212.14402v12022
  10. A Comprehensive Survey of Retrieval-Augmented Generation (RAG): Evolution, Current Landscape and Future Directions

    Shailja Gupta, Rajesh Ranjan, Surya Narayan Singh

    cs.CLcs.AIcs.IRarXiv:2410.12837v12024
  11. Knowledge Graph-Guided Retrieval Augmented Generation

    Xiangrong Zhu, Yuexiang Xie, Yi Liu +2

    cs.CLcs.AIarXiv:2502.06864v12025
  12. A Review on Explainability in Multimodal Deep Neural Nets

    Gargi Joshi, Rahee Walambe, Ketan Kotecha

    cs.AIcs.CVarXiv:2105.07878v22021
  13. VideoVLA: Video Generators Can Be Generalizable Robot Manipulators

    Yichao Shen, Fangyun Wei, Zhiying Du +5

    cs.ROcs.AIcs.CVarXiv:2512.06963v12025
  14. ChartCoder: Advancing Multimodal Large Language Model for Chart-to-Code Generation

    Xuanle Zhao, Xianzhen Luo, Qi Shi +4

    cs.AIarXiv:2501.06598v32025
  15. Authenticated Delegation and Authorized AI Agents

    Tobin South, Samuele Marro, Thomas Hardjono +5

    cs.CYcs.AIcs.NIarXiv:2501.09674v12025
  16. Paper2Poster: Towards Multimodal Poster Automation from Scientific Papers

    Wei Pang, Kevin Qinghong Lin, Xiangru Jian +2

    cs.CVcs.AIcs.CLarXiv:2505.21497v22025
  17. Critique-GRPO: Advancing LLM Reasoning with Natural Language and Numerical Feedback

    Xiaoying Zhang, Yipeng Zhang, Hao Sun +4

    cs.CLcs.AIarXiv:2506.03106v72025
  18. Large Language Model Enhanced Multi-Agent Systems for 6G Communications

    Feibo Jiang, Li Dong, Yubo Peng +5

    cs.AIarXiv:2312.07850v12023
  19. A PSO and Pattern Search based Memetic Algorithm for SVMs Parameters Optimization

    Yukun Bao, Zhongyi Hu, Tao Xiong

    cs.LGcs.AIcs.NEarXiv:1401.1926v12014
  20. Rethinking Weakly-supervised Video Temporal Grounding From a Game Perspective

    Xiang Fang, Zeyu Xiong, Wanlong Fang +7

    cs.CVcs.AIarXiv:2605.26441v12026
  21. Geometry Forcing: Marrying Video Diffusion and 3D Representation for Consistent World Modeling

    Haoyu Wu, Diankun Wu, Tianyu He +4

    cs.CVcs.AIarXiv:2507.07982v22025
  22. Knowledge Graph Embedding: A Survey from the Perspective of Representation Spaces

    Jiahang Cao, Jinyuan Fang, Zaiqiao Meng +1

    cs.LGcs.AIcs.CLarXiv:2211.03536v22022
  23. SciToolAgent: A Knowledge Graph-Driven Scientific Agent for Multi-Tool Integration

    Keyan Ding, Jing Yu, Junjie Huang +3

    cs.AIcs.CLarXiv:2507.20280v12025
  24. On the Adversarial Robustness of Multi-Modal Foundation Models

    Christian Schlarmann, Matthias Hein

    cs.LGcs.AIcs.CRarXiv:2308.10741v12023
  25. Reward Shaping to Mitigate Reward Hacking in RLHF

    Jiayi Fu, Xuandong Zhao, Chengyuan Yao +2

    cs.LGcs.AIcs.CLarXiv:2502.18770v72025
  26. Image-Grounded Conversations: Multimodal Context for Natural Question and Response Generation

    Nasrin Mostafazadeh, Chris Brockett, Bill Dolan +4

    cs.CLcs.AIcs.CVarXiv:1701.08251v22017
  27. CAFE: Catastrophic Data Leakage in Vertical Federated Learning

    Xiao Jin, Pin-Yu Chen, Chia-Yi Hsu +2

    cs.LGcs.AIarXiv:2110.15122v42021
  28. A Survey on Over-the-Air Computation

    Alphan Sahin, Rui Yang

    cs.ITcs.AIeess.SParXiv:2210.11350v52022
  29. An unscented Kalman filter method for real time input-parameter-state estimation

    Marios Impraimakis, Andrew W. Smyth

    eess.SPcs.AIcs.CVarXiv:2511.02717v12025
  30. Hierarchy-of-Groups Policy Optimization for Long-Horizon Agentic Tasks

    Shuo He, Lang Feng, Qi Wei +3

    cs.LGcs.AIarXiv:2602.22817v12026
  31. Factual Error Correction for Abstractive Summarization Models

    Meng Cao, Yue Dong, Jiapeng Wu +1

    cs.CLcs.AIarXiv:2010.08712v22020
  32. MedHELM: Holistic Evaluation of Large Language Models for Medical Tasks

    Suhana Bedi, Hejie Cui, Miguel Fuentes +78

    cs.CLcs.AIarXiv:2505.23802v22025
  33. Gold-medalist Performance in Solving Olympiad Geometry with AlphaGeometry2

    Yuri Chervonyi, Trieu H. Trinh, Miroslav Olšák +8

    cs.AIcs.LGarXiv:2502.03544v32025
  34. Modeling Localness for Self-Attention Networks

    Baosong Yang, Zhaopeng Tu, Derek F. Wong +3

    cs.CLcs.AIarXiv:1810.10182v12018
  35. Adversarial Motion Priors Make Good Substitutes for Complex Reward Functions

    Alejandro Escontrela, Xue Bin Peng, Wenhao Yu +4

    cs.AIcs.ROarXiv:2203.15103v12022
  36. Learning Language-Conditioned Robot Behavior from Offline Data and Crowd-Sourced Annotation

    Suraj Nair, Eric Mitchell, Kevin Chen +3

    cs.ROcs.AIcs.LGarXiv:2109.01115v22021
  37. Mind the Sim2Real Gap in User Simulation for Agentic Tasks

    Xuhui Zhou, Weiwei Sun, Qianou Ma +8

    cs.AIarXiv:2603.11245v22026
  38. LIDA: A Tool for Automatic Generation of Grammar-Agnostic Visualizations and Infographics using Large Language Models

    Victor Dibia

    cs.AIcs.HCcs.PLarXiv:2303.02927v32023
  39. Power Stabilization for AI Training Datacenters

    Esha Choukse, Brijesh Warrier, Scot Heath +54

    cs.ARcs.AIcs.DCarXiv:2508.14318v22025
  40. Recent Advances in Imitation Learning from Observation

    Faraz Torabi, Garrett Warnell, Peter Stone

    cs.ROcs.AIcs.LGarXiv:1905.13566v22019
  41. Probabilistic Similarity Logic

    Matthias Brocheler, Lilyana Mihalkova, Lise Getoor

    cs.AIarXiv:1203.3469v12012
  42. AgentDAM: Privacy Leakage Evaluation for Autonomous Web Agents

    Arman Zharmagambetov, Chuan Guo, Ivan Evtimov +3

    cs.AIarXiv:2503.09780v32025
  43. ACECODER: Acing Coder RL via Automated Test-Case Synthesis

    Huaye Zeng, Dongfu Jiang, Haozhe Wang +3

    cs.SEcs.AIcs.CLarXiv:2502.01718v42025
  44. Universal Actions for Enhanced Embodied Foundation Models

    Jinliang Zheng, Jianxiong Li, Dongxiu Liu +7

    cs.ROcs.AIcs.CVarXiv:2501.10105v22025
  45. OBJECTION! Lawyer Agents Mitigate Guilty Bias in Legal Judgment Prediction

    Jaehoon Jeong, Jay-Yoon Lee

    cs.CLcs.AIarXiv:2609.02158v12026
  46. UTP-Bench: Uncertainty-aware Travel Planning Benchmark

    Etcharla Revanth Rao, Priyanshu Karmakar, Shubhojit Mallick +3

    cs.AIcs.CLarXiv:2609.02421v12026
  47. Memory for Autonomous LLM Agents:Mechanisms, Evaluation, and Emerging Frontiers

    Pengfei Du

    cs.AIarXiv:2603.07670v12026
  48. Airbert: In-domain Pretraining for Vision-and-Language Navigation

    Pierre-Louis Guhur, Makarand Tapaswi, Shizhe Chen +2

    cs.CVcs.AIcs.CLarXiv:2108.09105v12021
  49. InstEditSeg: Instruction-Driven Image Editing for Polyp and Skin Lesion Segmentation

    Ziquan Liu, Zhewei Zhu, Xuyang Shi

    cs.CVcs.AIarXiv:2609.02004v12026
  50. Seaweed-7B: Cost-Effective Training of Video Generation Foundation Model

    Team Seawead, Ceyuan Yang, Zhijie Lin +52

    cs.CVcs.AIarXiv:2504.08685v22025
  51. AceReason-Nemotron 1.1: Advancing Math and Code Reasoning through SFT and RL Synergy

    Zihan Liu, Zhuolin Yang, Yang Chen +4

    cs.CLcs.AIcs.LGarXiv:2506.13284v12025
  52. Residual LSTM: Design of a Deep Recurrent Architecture for Distant Speech Recognition

    Jaeyoung Kim, Mostafa El-Khamy, Jungwon Lee

    cs.LGcs.AIcs.SDarXiv:1701.03360v32017
  53. Do Multilingual LLMs Think In English?

    Lisa Schut, Yarin Gal, Sebastian Farquhar

    cs.CLcs.AIcs.LGarXiv:2502.15603v12025
  54. SMart: A Multi-source Multi-phase Time Series Representation Transfer Framework

    Fang He, Wang-chien Lee

    cs.LGcs.AIarXiv:2609.02203v12026
  55. Schrödinger Bridges on Lie Group Manifolds for Probabilistic Intrinsic Generation

    Shizhe Zhang, Mingyang Zhao, Lei Ma

    stat.MLcs.AIcs.LGarXiv:2609.02196v12026
  56. A Biologically Plausible Supervised Learning Method for Spiking Neural Networks Using the Symmetric STDP Rule

    Yunzhe Hao, Xuhui Huang, Meng Dong +1

    cs.NEcs.AIcs.LGarXiv:1812.06574v32018
  57. Beyond Modality Harmony: Orthogonal Purification and Topology-Guided MoE for Conflict-Aware Multimodal Recommendation

    Jialin Liu, Zhaorui Zhang, Ray C. C. Cheung

    cs.IRcs.AIarXiv:2609.02152v12026
  58. Visual Planning: Let's Think Only with Images

    Yi Xu, Chengzu Li, Han Zhou +4

    cs.LGcs.AIcs.CLarXiv:2505.11409v32025
  59. A Survey of Self-Evolving Agents: What, When, How, and Where to Evolve on the Path to Artificial Super Intelligence

    Huan-ang Gao, Jiayi Geng, Wenyue Hua +24

    cs.AIarXiv:2507.21046v42025
  60. Tell me about yourself: LLMs are aware of their learned behaviors

    Jan Betley, Xuchan Bao, Martín Soto +3

    cs.CLcs.AIcs.CRarXiv:2501.11120v12025