Artificial Intelligence

Papers filed under cs.AI on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.

Search paper metadata (including unsummarized papers)

5,101 to 5,160 of 15,328

  1. A Comprehensive Survey in LLM(-Agent) Full Stack Safety: Data, Training and Deployment

    Kun Wang, Guibin Zhang, Zhenhong Zhou +100

    cs.CRcs.AIcs.CLarXiv:2504.15585v42025
  2. AceReason-Nemotron: Advancing Math and Code Reasoning through Reinforcement Learning

    Yang Chen, Zhuolin Yang, Zihan Liu +5

    cs.LGcs.AIcs.CLarXiv:2505.16400v32025
  3. Towards a Reliable and Practical Eval Pipeline

    Emma Thuong Nguyen, Abhishek Ghose

    cs.AIcs.SEarXiv:2609.00805v12026
  4. MAGIS: LLM-Based Multi-Agent Framework for GitHub Issue Resolution

    Wei Tao, Yucheng Zhou, Yanlin Wang +3

    cs.SEcs.AIarXiv:2403.17927v22024
  5. The Privacy-Hallucination Tradeoff in Differentially Private Language Models

    Krithika Ramesh, Krishna Pillutla, Danish Pruthi +1

    cs.AIcs.CLarXiv:2609.00492v12026
  6. From Prompt Injections to Protocol Exploits: Threats in LLM-Powered AI Agents Workflows

    Mohamed Amine Ferrag, Norbert Tihanyi, Djallel Hamouda +3

    cs.CRcs.AIarXiv:2506.23260v22025
  7. jina-embeddings-v4: Universal Embeddings for Multimodal Multilingual Retrieval

    Michael Günther, Saba Sturua, Mohammad Kalim Akram +8

    cs.AIcs.CLcs.IRarXiv:2506.18902v32025
  8. Kevin: Multi-Turn RL for Generating CUDA Kernels

    Carlo Baronio, Pietro Marsella, Ben Pan +2

    cs.LGcs.AIcs.PFarXiv:2507.11948v12025
  9. MiMo: Unlocking the Reasoning Potential of Language Model -- From Pretraining to Posttraining

    LLM-Core Xiaomi, :, Bingquan Xia +62

    cs.CLcs.AIcs.LGarXiv:2505.07608v22025
  10. Combining Deep Reinforcement Learning and Search for Imperfect-Information Games

    Noam Brown, Anton Bakhtin, Adam Lerer +1

    cs.GTcs.AIcs.LGarXiv:2007.13544v22020
  11. TransZero: Attribute-guided Transformer for Zero-Shot Learning

    Shiming Chen, Ziming Hong, Yang Liu +6

    cs.CVcs.AIarXiv:2112.01683v12021
  12. PPTAgent: Generating and Evaluating Presentations Beyond Text-to-Slides

    Hao Zheng, Xinyan Guan, Hao Kong +7

    cs.AIcs.CLarXiv:2501.03936v32025
  13. Zero-Shot Respiratory Sound Classification through LLM-Augmented Audio-Text Alignment

    Mustafa Talha İlerisoy, Hung Manh Pham, Mathias Funk +2

    cs.CLcs.AIcs.SDarXiv:2609.00055v12026
  14. Diffusion Adversarial Post-Training for One-Step Video Generation

    Shanchuan Lin, Xin Xia, Yuxi Ren +3

    cs.CVcs.AIcs.LGarXiv:2501.08316v32025
  15. GUI-Actor: Coordinate-Free Visual Grounding for GUI Agents

    Qianhui Wu, Kanzhi Cheng, Rui Yang +15

    cs.CLcs.AIcs.CVarXiv:2506.03143v12025
  16. EgoVLA: Learning Vision-Language-Action Models from Egocentric Human Videos

    Ruihan Yang, Qinxi Yu, Yecheng Wu +12

    cs.ROcs.AIcs.CVarXiv:2507.12440v32025
  17. A Survey of AI Agent Protocols

    Yingxuan Yang, Huacan Chai, Yuanyi Song +11

    cs.AIarXiv:2504.16736v32025
  18. TAMI: Temporally Aligned, Missingness-Aware, and Interpretable Multimodal Fusion for Mental Health Assessment in Older Adults with Mild Cognitive Impairment

    Merna Bibars, Bolaji Omofojoye, Allan I. Levey +3

    cs.CVcs.AIarXiv:2608.30857v12026
  19. Process Reward Models That Think

    Muhammad Khalifa, Rishabh Agarwal, Lajanugen Logeswaran +5

    cs.LGcs.AIcs.CLarXiv:2504.16828v52025
  20. MemoryGraft: Persistent Compromise of LLM Agents via Poisoned Experience Retrieval

    Saksham Sahai Srivastava, Haoyu He

    cs.CRcs.AIcs.LGarXiv:2512.16962v12025
  21. Transformer Language Models without Positional Encodings Still Learn Positional Information

    Adi Haviv, Ori Ram, Ofir Press +2

    cs.CLcs.AIcs.LGarXiv:2203.16634v22022
  22. Aether: Geometric-Aware Unified World Modeling

    Aether Team, Haoyi Zhu, Yifan Wang +8

    cs.CVcs.AIcs.LGarXiv:2503.18945v32025
  23. Addressing Complex and Subjective Product-Related Queries with Customer Reviews

    Julian McAuley, Alex Yang

    cs.IRcs.AIcs.SIarXiv:1512.06863v12015
  24. Reward-Guided Speculative Decoding for Efficient LLM Reasoning

    Baohao Liao, Yuhui Xu, Hanze Dong +5

    cs.CLcs.AIarXiv:2501.19324v32025
  25. Progent: Securing AI Agents with Privilege Control

    Tianneng Shi, Jingxuan He, Zhun Wang +4

    cs.CRcs.AIarXiv:2504.11703v32025
  26. DexGraspVLA: A Vision-Language-Action Framework Towards General Dexterous Grasping

    Yifan Zhong, Xuchuan Huang, Ruochong Li +9

    cs.ROcs.AIarXiv:2502.20900v52025
  27. GuardReasoner: Towards Reasoning-based LLM Safeguards

    Yue Liu, Hongcheng Gao, Shengfang Zhai +9

    cs.CRcs.AIcs.LGarXiv:2501.18492v22025
  28. The Road Less Scheduled

    Aaron Defazio, Xingyu Alice Yang, Harsh Mehta +3

    cs.LGcs.AImath.OCarXiv:2405.15682v42024
  29. OstQuant: Refining Large Language Model Quantization with Orthogonal and Scaling Transformations for Better Distribution Fitting

    Xing Hu, Yuan Cheng, Dawei Yang +6

    cs.LGcs.AIarXiv:2501.13987v12025
  30. Scaling up Masked Diffusion Models on Text

    Shen Nie, Fengqi Zhu, Chao Du +5

    cs.AIcs.CLcs.LGarXiv:2410.18514v32024
  31. DARTS-: Robustly Stepping out of Performance Collapse Without Indicators

    Xiangxiang Chu, Xiaoxing Wang, Bo Zhang +3

    cs.LGcs.AIcs.CVarXiv:2009.01027v22020
  32. Small Models Struggle to Learn from Strong Reasoners

    Yuetai Li, Xiang Yue, Zhangchen Xu +5

    cs.AIarXiv:2502.12143v32025
  33. Scaling Spatial Intelligence with Multimodal Foundation Models

    Zhongang Cai, Ruisi Wang, Chenyang Gu +26

    cs.CVcs.AIcs.LGarXiv:2511.13719v42025
  34. Understanding Reasoning in Thinking Language Models via Steering Vectors

    Constantin Venhoff, Iván Arcuschin, Philip Torr +2

    cs.LGcs.AIarXiv:2506.18167v42025
  35. VerlTool: Towards Holistic Agentic Reinforcement Learning with Tool Use

    Dongfu Jiang, Yi Lu, Zhuofeng Li +9

    cs.AIcs.CLcs.CVarXiv:2509.01055v32025
  36. Multi-Agent Collaboration via Evolving Orchestration

    Yufan Dang, Chen Qian, Xueheng Luo +11

    cs.CLcs.AIcs.MAarXiv:2505.19591v22025
  37. What Can We Learn from Collective Human Opinions on Natural Language Inference Data?

    Yixin Nie, Xiang Zhou, Mohit Bansal

    cs.CLcs.AIcs.LGarXiv:2010.03532v22020
  38. A Vision-Language-Action-Critic Model for Robotic Real-World Reinforcement Learning

    Shaopeng Zhai, Qi Zhang, Tianyi Zhang +7

    cs.ROcs.AIarXiv:2509.15937v12025
  39. VLASH: Real-Time VLAs via Future-State-Aware Asynchronous Inference

    Jiaming Tang, Yufei Sun, Yilong Zhao +7

    cs.ROcs.AIcs.LGarXiv:2512.01031v22025
  40. Measuring the environmental impact of delivering AI at Google Scale

    Cooper Elsworth, Keguo Huang, David Patterson +9

    cs.AIarXiv:2508.15734v12025
  41. Latent Diffusion Model without Variational Autoencoder

    Minglei Shi, Haolin Wang, Wenzhao Zheng +6

    cs.CVcs.AIarXiv:2510.15301v42025
  42. Global-Locally Self-Attentive Dialogue State Tracker

    Victor Zhong, Caiming Xiong, Richard Socher

    cs.CLcs.AIarXiv:1805.09655v32018
  43. YuE: Scaling Open Foundation Models for Long-Form Music Generation

    Ruibin Yuan, Hanfeng Lin, Shuyue Guo +55

    eess.AScs.AIcs.MMarXiv:2503.08638v22025
  44. A Survey of Graph Retrieval-Augmented Generation for Customized Large Language Models

    Qinggang Zhang, Shengyuan Chen, Yuanchen Bei +9

    cs.CLcs.AIcs.IRarXiv:2501.13958v32025
  45. Approximate evaluation of marginal association probabilities with belief propagation

    Jason L. Williams, Roslyn A. Lau

    cs.AIcs.CVarXiv:1209.6299v22012
  46. Training Neural Machine Translation To Apply Terminology Constraints

    Georgiana Dinu, Prashant Mathur, Marcello Federico +1

    cs.CLcs.AIcs.LGarXiv:1906.01105v22019
  47. JudgeLRM: Large Reasoning Models as a Judge

    Nuo Chen, Zhiyuan Hu, Qingyun Zou +4

    cs.CLcs.AIarXiv:2504.00050v32025
  48. Spatio-Temporal Wind Speed Forecasting using Graph Networks and Novel Transformer Architectures

    Lars Ødegaard Bentsen, Narada Dilp Warakagoda, Roy Stenbro +1

    cs.LGcs.AIarXiv:2208.13585v22022
  49. Quantum Computing based Hybrid Solution Strategies for Large-scale Discrete-Continuous Optimization Problems

    Akshay Ajagekar, Travis Humble, Fengqi You

    quant-phcs.AImath.OCarXiv:1910.13045v12019
  50. S-GRPO: Early Exit via Reinforcement Learning in Reasoning Models

    Muzhi Dai, Chenxu Yang, Qingyi Si

    cs.AIcs.LGarXiv:2505.07686v22025
  51. How many images do I need? Understanding how sample size per class affects deep learning model performance metrics for balanced designs in autonomous wildlife monitoring

    Saleh Shahinfar, Paul Meek, Greg Falzon

    cs.CVcs.AIcs.LGarXiv:2010.08186v12020
  52. Mechanism Design for Alignment and Control

    Dirk Bergemann, Andrew Koh, Stephen Morris

    econ.THcs.AIcs.GTarXiv:2609.01595v12026
  53. Gemini Embedding: Generalizable Embeddings from Gemini

    Jinhyuk Lee, Feiyang Chen, Sahil Dua +44

    cs.CLcs.AIarXiv:2503.07891v12025
    Summaries:한국어
  54. CaP-X: A Framework for Benchmarking and Improving Coding Agents for Robot Manipulation

    Letian Fu, Justin Yu, Karim El-Refai +13

    cs.ROcs.AIarXiv:2603.22435v22026
  55. RPCBench: A Benchmark for Proactive Premise Critique in LLM-based Recommendation

    Zhongru Chen, Yuan Wu, Yi Chang

    cs.AIcs.CLarXiv:2609.00918v12026
  56. Self-Supervised Hypergraph Transformer for Recommender Systems

    Lianghao Xia, Chao Huang, Chuxu Zhang

    cs.IRcs.AIarXiv:2207.14338v12022
  57. QILP-0: Constructing Observational Declarative Twins of Quantum Circuits

    Marina de la Cruz Echeandía, César Luis Alonso, Tony Ribeiro +1

    cs.AIquant-pharXiv:2609.01049v12026
  58. EdiTikZ: Scientific Figure Editing from Revision Trajectories

    Christian Greisinger, Zhixue Zhao, Steffen Eger

    cs.AIcs.CLcs.CVarXiv:2609.01409v12026
  59. ACON: Optimizing Context Compression for Long-horizon LLM Agents

    Minki Kang, Wei-Ning Chen, Dongge Han +5

    cs.AIcs.CLarXiv:2510.00615v32025
  60. Some Emotions Run Deeper: Layer-wise Probing and Causal Intervention in Large Language Models

    Tian Fang, Gaël Guibon, Davide Buscaldi

    cs.CLcs.AIarXiv:2609.01279v12026