Artificial Intelligence
Papers filed under cs.AI on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.
Search paper metadata (including unsummarized papers)
5,101 to 5,160 of 15,328
A Comprehensive Survey in LLM(-Agent) Full Stack Safety: Data, Training and Deployment
Kun Wang, Guibin Zhang, Zhenhong Zhou +100
cs.CRcs.AIcs.CLarXiv:2504.15585v42025AceReason-Nemotron: Advancing Math and Code Reasoning through Reinforcement Learning
Yang Chen, Zhuolin Yang, Zihan Liu +5
cs.LGcs.AIcs.CLarXiv:2505.16400v32025Towards a Reliable and Practical Eval Pipeline
Emma Thuong Nguyen, Abhishek Ghose
cs.AIcs.SEarXiv:2609.00805v12026MAGIS: LLM-Based Multi-Agent Framework for GitHub Issue Resolution
Wei Tao, Yucheng Zhou, Yanlin Wang +3
cs.SEcs.AIarXiv:2403.17927v22024The Privacy-Hallucination Tradeoff in Differentially Private Language Models
Krithika Ramesh, Krishna Pillutla, Danish Pruthi +1
cs.AIcs.CLarXiv:2609.00492v12026From Prompt Injections to Protocol Exploits: Threats in LLM-Powered AI Agents Workflows
Mohamed Amine Ferrag, Norbert Tihanyi, Djallel Hamouda +3
cs.CRcs.AIarXiv:2506.23260v22025jina-embeddings-v4: Universal Embeddings for Multimodal Multilingual Retrieval
Michael Günther, Saba Sturua, Mohammad Kalim Akram +8
cs.AIcs.CLcs.IRarXiv:2506.18902v32025Kevin: Multi-Turn RL for Generating CUDA Kernels
Carlo Baronio, Pietro Marsella, Ben Pan +2
cs.LGcs.AIcs.PFarXiv:2507.11948v12025MiMo: Unlocking the Reasoning Potential of Language Model -- From Pretraining to Posttraining
LLM-Core Xiaomi, :, Bingquan Xia +62
cs.CLcs.AIcs.LGarXiv:2505.07608v22025Combining Deep Reinforcement Learning and Search for Imperfect-Information Games
Noam Brown, Anton Bakhtin, Adam Lerer +1
cs.GTcs.AIcs.LGarXiv:2007.13544v22020TransZero: Attribute-guided Transformer for Zero-Shot Learning
Shiming Chen, Ziming Hong, Yang Liu +6
cs.CVcs.AIarXiv:2112.01683v12021PPTAgent: Generating and Evaluating Presentations Beyond Text-to-Slides
Hao Zheng, Xinyan Guan, Hao Kong +7
cs.AIcs.CLarXiv:2501.03936v32025Zero-Shot Respiratory Sound Classification through LLM-Augmented Audio-Text Alignment
Mustafa Talha İlerisoy, Hung Manh Pham, Mathias Funk +2
cs.CLcs.AIcs.SDarXiv:2609.00055v12026Diffusion Adversarial Post-Training for One-Step Video Generation
Shanchuan Lin, Xin Xia, Yuxi Ren +3
cs.CVcs.AIcs.LGarXiv:2501.08316v32025GUI-Actor: Coordinate-Free Visual Grounding for GUI Agents
Qianhui Wu, Kanzhi Cheng, Rui Yang +15
cs.CLcs.AIcs.CVarXiv:2506.03143v12025EgoVLA: Learning Vision-Language-Action Models from Egocentric Human Videos
Ruihan Yang, Qinxi Yu, Yecheng Wu +12
cs.ROcs.AIcs.CVarXiv:2507.12440v32025A Survey of AI Agent Protocols
Yingxuan Yang, Huacan Chai, Yuanyi Song +11
cs.AIarXiv:2504.16736v32025TAMI: Temporally Aligned, Missingness-Aware, and Interpretable Multimodal Fusion for Mental Health Assessment in Older Adults with Mild Cognitive Impairment
Merna Bibars, Bolaji Omofojoye, Allan I. Levey +3
cs.CVcs.AIarXiv:2608.30857v12026Process Reward Models That Think
Muhammad Khalifa, Rishabh Agarwal, Lajanugen Logeswaran +5
cs.LGcs.AIcs.CLarXiv:2504.16828v52025MemoryGraft: Persistent Compromise of LLM Agents via Poisoned Experience Retrieval
Saksham Sahai Srivastava, Haoyu He
cs.CRcs.AIcs.LGarXiv:2512.16962v12025Transformer Language Models without Positional Encodings Still Learn Positional Information
Adi Haviv, Ori Ram, Ofir Press +2
cs.CLcs.AIcs.LGarXiv:2203.16634v22022Aether: Geometric-Aware Unified World Modeling
Aether Team, Haoyi Zhu, Yifan Wang +8
cs.CVcs.AIcs.LGarXiv:2503.18945v32025Addressing Complex and Subjective Product-Related Queries with Customer Reviews
Julian McAuley, Alex Yang
cs.IRcs.AIcs.SIarXiv:1512.06863v12015Reward-Guided Speculative Decoding for Efficient LLM Reasoning
Baohao Liao, Yuhui Xu, Hanze Dong +5
cs.CLcs.AIarXiv:2501.19324v32025Progent: Securing AI Agents with Privilege Control
Tianneng Shi, Jingxuan He, Zhun Wang +4
cs.CRcs.AIarXiv:2504.11703v32025DexGraspVLA: A Vision-Language-Action Framework Towards General Dexterous Grasping
Yifan Zhong, Xuchuan Huang, Ruochong Li +9
cs.ROcs.AIarXiv:2502.20900v52025GuardReasoner: Towards Reasoning-based LLM Safeguards
Yue Liu, Hongcheng Gao, Shengfang Zhai +9
cs.CRcs.AIcs.LGarXiv:2501.18492v22025The Road Less Scheduled
Aaron Defazio, Xingyu Alice Yang, Harsh Mehta +3
cs.LGcs.AImath.OCarXiv:2405.15682v42024OstQuant: Refining Large Language Model Quantization with Orthogonal and Scaling Transformations for Better Distribution Fitting
Xing Hu, Yuan Cheng, Dawei Yang +6
cs.LGcs.AIarXiv:2501.13987v12025Scaling up Masked Diffusion Models on Text
Shen Nie, Fengqi Zhu, Chao Du +5
cs.AIcs.CLcs.LGarXiv:2410.18514v32024DARTS-: Robustly Stepping out of Performance Collapse Without Indicators
Xiangxiang Chu, Xiaoxing Wang, Bo Zhang +3
cs.LGcs.AIcs.CVarXiv:2009.01027v22020Small Models Struggle to Learn from Strong Reasoners
Yuetai Li, Xiang Yue, Zhangchen Xu +5
cs.AIarXiv:2502.12143v32025Scaling Spatial Intelligence with Multimodal Foundation Models
Zhongang Cai, Ruisi Wang, Chenyang Gu +26
cs.CVcs.AIcs.LGarXiv:2511.13719v42025Understanding Reasoning in Thinking Language Models via Steering Vectors
Constantin Venhoff, Iván Arcuschin, Philip Torr +2
cs.LGcs.AIarXiv:2506.18167v42025VerlTool: Towards Holistic Agentic Reinforcement Learning with Tool Use
Dongfu Jiang, Yi Lu, Zhuofeng Li +9
cs.AIcs.CLcs.CVarXiv:2509.01055v32025Multi-Agent Collaboration via Evolving Orchestration
Yufan Dang, Chen Qian, Xueheng Luo +11
cs.CLcs.AIcs.MAarXiv:2505.19591v22025What Can We Learn from Collective Human Opinions on Natural Language Inference Data?
Yixin Nie, Xiang Zhou, Mohit Bansal
cs.CLcs.AIcs.LGarXiv:2010.03532v22020A Vision-Language-Action-Critic Model for Robotic Real-World Reinforcement Learning
Shaopeng Zhai, Qi Zhang, Tianyi Zhang +7
cs.ROcs.AIarXiv:2509.15937v12025VLASH: Real-Time VLAs via Future-State-Aware Asynchronous Inference
Jiaming Tang, Yufei Sun, Yilong Zhao +7
cs.ROcs.AIcs.LGarXiv:2512.01031v22025Measuring the environmental impact of delivering AI at Google Scale
Cooper Elsworth, Keguo Huang, David Patterson +9
cs.AIarXiv:2508.15734v12025Latent Diffusion Model without Variational Autoencoder
Minglei Shi, Haolin Wang, Wenzhao Zheng +6
cs.CVcs.AIarXiv:2510.15301v42025Global-Locally Self-Attentive Dialogue State Tracker
Victor Zhong, Caiming Xiong, Richard Socher
cs.CLcs.AIarXiv:1805.09655v32018YuE: Scaling Open Foundation Models for Long-Form Music Generation
Ruibin Yuan, Hanfeng Lin, Shuyue Guo +55
eess.AScs.AIcs.MMarXiv:2503.08638v22025A Survey of Graph Retrieval-Augmented Generation for Customized Large Language Models
Qinggang Zhang, Shengyuan Chen, Yuanchen Bei +9
cs.CLcs.AIcs.IRarXiv:2501.13958v32025Approximate evaluation of marginal association probabilities with belief propagation
Jason L. Williams, Roslyn A. Lau
cs.AIcs.CVarXiv:1209.6299v22012Training Neural Machine Translation To Apply Terminology Constraints
Georgiana Dinu, Prashant Mathur, Marcello Federico +1
cs.CLcs.AIcs.LGarXiv:1906.01105v22019JudgeLRM: Large Reasoning Models as a Judge
Nuo Chen, Zhiyuan Hu, Qingyun Zou +4
cs.CLcs.AIarXiv:2504.00050v32025Spatio-Temporal Wind Speed Forecasting using Graph Networks and Novel Transformer Architectures
Lars Ødegaard Bentsen, Narada Dilp Warakagoda, Roy Stenbro +1
cs.LGcs.AIarXiv:2208.13585v22022Quantum Computing based Hybrid Solution Strategies for Large-scale Discrete-Continuous Optimization Problems
Akshay Ajagekar, Travis Humble, Fengqi You
quant-phcs.AImath.OCarXiv:1910.13045v12019S-GRPO: Early Exit via Reinforcement Learning in Reasoning Models
Muzhi Dai, Chenxu Yang, Qingyi Si
cs.AIcs.LGarXiv:2505.07686v22025How many images do I need? Understanding how sample size per class affects deep learning model performance metrics for balanced designs in autonomous wildlife monitoring
Saleh Shahinfar, Paul Meek, Greg Falzon
cs.CVcs.AIcs.LGarXiv:2010.08186v12020Mechanism Design for Alignment and Control
Dirk Bergemann, Andrew Koh, Stephen Morris
econ.THcs.AIcs.GTarXiv:2609.01595v12026Gemini Embedding: Generalizable Embeddings from Gemini
Jinhyuk Lee, Feiyang Chen, Sahil Dua +44
cs.CLcs.AIarXiv:2503.07891v12025Summaries:한국어CaP-X: A Framework for Benchmarking and Improving Coding Agents for Robot Manipulation
Letian Fu, Justin Yu, Karim El-Refai +13
cs.ROcs.AIarXiv:2603.22435v22026RPCBench: A Benchmark for Proactive Premise Critique in LLM-based Recommendation
Zhongru Chen, Yuan Wu, Yi Chang
cs.AIcs.CLarXiv:2609.00918v12026Self-Supervised Hypergraph Transformer for Recommender Systems
Lianghao Xia, Chao Huang, Chuxu Zhang
cs.IRcs.AIarXiv:2207.14338v12022QILP-0: Constructing Observational Declarative Twins of Quantum Circuits
Marina de la Cruz Echeandía, César Luis Alonso, Tony Ribeiro +1
cs.AIquant-pharXiv:2609.01049v12026EdiTikZ: Scientific Figure Editing from Revision Trajectories
Christian Greisinger, Zhixue Zhao, Steffen Eger
cs.AIcs.CLcs.CVarXiv:2609.01409v12026ACON: Optimizing Context Compression for Long-horizon LLM Agents
Minki Kang, Wei-Ning Chen, Dongge Han +5
cs.AIcs.CLarXiv:2510.00615v32025Some Emotions Run Deeper: Layer-wise Probing and Causal Intervention in Large Language Models
Tian Fang, Gaël Guibon, Davide Buscaldi
cs.CLcs.AIarXiv:2609.01279v12026