Artificial Intelligence
Papers filed under cs.AI on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.
Search paper metadata (including unsummarized papers)
11,941 to 12,000 of 15,247
IQuest-Coder-V1 Technical Report
Jian Yang, Wei Zhang, Shawn Guo +35
cs.AIcs.CLcs.SEarXiv:2603.16733v12026A Survey of Multi-Objective Sequential Decision-Making
Diederik Marijn Roijers, Peter Vamplew, Shimon Whiteson +1
cs.AIarXiv:1402.0590v12014CRITIC: Large Language Models Can Self-Correct with Tool-Interactive Critiquing
Zhibin Gou, Zhihong Shao, Yeyun Gong +4
cs.CLcs.AIarXiv:2305.11738v42023LatentChem: From Textual CoT to Latent Thinking in Chemical Reasoning
Xinwu Ye, Yicheng Mao, Yuxuan Liao +16
physics.chem-phcs.AIcs.CLarXiv:2602.07075v62026FlexGen: High-Throughput Generative Inference of Large Language Models with a Single GPU
Ying Sheng, Lianmin Zheng, Binhang Yuan +11
cs.LGcs.AIcs.PFarXiv:2303.06865v22023Coarse-Guided Visual Generation via Weighted h-Transform Sampling
Yanghao Wang, Ziqi Jiang, Zhen Wang +1
cs.CVcs.AIarXiv:2603.12057v22026HERO: Human-profile Enhanced Retrieval Optimization Framework for Long-term Agent Memory
Yuanhua Lin, Yile Li, Zhiyuan Zhao +2
cs.AIarXiv:2608.22310v12026Clarify User Expertise: Towards Proactive Conversational Agents Tailoring Responses to User Proficiency
Zhihong Cao, Chen Huang
cs.AIcs.CLarXiv:2608.22266v12026Read Less, Solve More: Token-Efficient Sparse Reading for AI Agents
Zedong Liu, Jiaan Wu, Xinyang Ma +5
cs.AIarXiv:2608.22237v12026Beyond What Meets the Eye: Unveiling Situational Illusions for Multimodal Large Language Models
Zhiming Yang, Zhuoxi Xiong, Donglin Zhou +3
cs.AIcs.CLcs.CVarXiv:2608.22232v12026Query-Driven Multimodal Information Extraction from Long Documents
Yikai Gao, Ding Xia, Xi Yang
cs.AIcs.MMarXiv:2608.22214v12026Mitigating Bias in Algorithmic Hiring: Evaluating Claims and Practices
Manish Raghavan, Solon Barocas, Jon Kleinberg +1
cs.CYcs.AIcs.LGarXiv:1906.09208v32019BenchPreS: A Benchmark for Context-Aware Personalized Preference Selectivity of Persistent-Memory LLMs
Sangyeon Yoon, Sunkyoung Kim, Hyesoo Hong +5
cs.AIcs.CLarXiv:2603.16557v12026Unveiling Implicit Advantage Symmetry: Why GRPO Struggles with Exploration and Difficulty Adaptation
Zhiqi Yu, Zhangquan Chen, Mengting Liu +2
cs.LGcs.AIarXiv:2602.05548v32026Disagree to Explore, Agree to Commit: Routing-Guided Test-Time Scaling for Software Agents
Kang Chen, Junjie Nian, Yixin Cao +1
cs.AIcs.SEarXiv:2608.22191v12026Anticipatory Planning for Multimodal AI Agents
Yongyuan Liang, Shijie Zhou, Yu Gu +6
cs.AIarXiv:2603.16777v12026Role-Specialized Mixture-of-Agents with Open-Weight LLMs for Clinical Prediction
Jun Hou, Yi Fang, Xuan Wang
cs.AIcs.LGarXiv:2608.22176v12026Code-Space Response Oracles: Generating Interpretable Multi-Agent Policies with Large Language Models
Daniel Hennes, Zun Li, John Schultz +1
cs.GTcs.AIcs.LGarXiv:2603.10098v12026MCP-Universe RL: A Framework for Training MCP Tool-Use Agents via Reinforcement Learning
Ziyang Luo, Yan Yang, Xiangru Jian +5
cs.AIcs.LGarXiv:2608.22167v12026Small Language Model enabled Autonomous agent for Language-Conditioned Cognitive Radar
Minhaj Uddin Ahmad, Zakia Zaman, Shunqiao Sun +1
eess.SPcs.AIeess.SYarXiv:2608.11596v12026Nanbeige4.1-3B: A Small General Model that Reasons, Aligns, and Acts
Chen Yang, Guangyue Peng, Jiaying Zhu +12
cs.AIcs.CLarXiv:2602.13367v12026Task-Driven 3D Printability Assistance via Geometry- and Knowledge-Grounded LLM Reasoning
Zhaoda Du, Qiaojie Zheng, Xiaoli Zhang
cs.AIarXiv:2608.22128v12026Evaluation of Small Vision-Language Models on Qualitative Mechanical Problems
Henry Fordjour Ansah, Shreya Banerjee, Pranish Ghimire
cs.AIarXiv:2608.22143v12026Aggregation-Aware Synthetic Text Generation Against Authorship Re-Identification
Qian Ma, Anna Squicciarini, Sarah Rajtmajer
cs.AIcs.CLarXiv:2608.22161v12026Flow Equivariant World Models: Memory for Partially Observed Dynamic Environments
Hansen Jin Lillemark, Benhao Huang, Fangneng Zhan +2
cs.LGcs.AIcs.CVarXiv:2601.01075v22026RegNeRF: Regularizing Neural Radiance Fields for View Synthesis from Sparse Inputs
Michael Niemeyer, Jonathan T. Barron, Ben Mildenhall +3
cs.CVcs.AIcs.GRarXiv:2112.00724v12021BrowseComp-$V^3$: A Visual, Vertical, and Verifiable Benchmark for Multimodal Browsing Agents
Huanyao Zhang, Jiepeng Zhou, Bo Li +22
cs.AIarXiv:2602.12876v22026Vega: Learning to Drive with Natural Language Instructions
Sicheng Zuo, Yuxuan Li, Wenzhao Zheng +3
cs.CVcs.AIcs.ROarXiv:2603.25741v22026GISA: A Benchmark for General Information-Seeking Assistant
Yutao Zhu, Xingshuo Zhang, Maosen Zhang +9
cs.CLcs.AIcs.IRarXiv:2602.08543v22026Mixture-of-Experts with Expert Choice Routing
Yanqi Zhou, Tao Lei, Hanxiao Liu +7
cs.LGcs.AIarXiv:2202.09368v22022Self-EvolveRec: Self-Evolving Recommender Systems with LLM-based Directional Feedback
Sein Kim, Sangwu Park, Hongseok Kang +6
cs.IRcs.AIarXiv:2602.12612v22026NExT-GPT: Any-to-Any Multimodal LLM
Shengqiong Wu, Hao Fei, Leigang Qu +2
cs.AIcs.CLcs.LGarXiv:2309.05519v32023AUDITA: certified auditing and causal attribution of adverse outcomes in autonomous multi-agent systems
Zhixu Du, Yiran Chen
cs.AIarXiv:2608.22160v12026The many Shapley values for model explanation
Mukund Sundararajan, Amir Najmi
cs.AIcs.LGecon.THarXiv:1908.08474v22019The Flexibility Trap: Rethinking the Value of Arbitrary Order in Diffusion Language Models
Zanlin Ni, Shenzhi Wang, Yang Yue +8
cs.CLcs.AIcs.LGarXiv:2601.15165v42026Measuring Stability and Failure Behavior in Language Models Under Structured Perturbations
Samira Golsefid
cs.AIcs.CLarXiv:2608.22138v12026The Pensieve Paradigm: Stateful Language Models Mastering Their Own Context
Xiaoyuan Liu, Tian Liang, Dongyang Ma +4
cs.AIarXiv:2602.12108v12026MEMONDEMAND: A Memory Management System for Large-Scale Enterprise Data
Xinyuan Song, Bowen Zhu, Hasibul Haque +1
cs.AIarXiv:2608.22141v12026MegaMem: A Retrieval Solution for Ultra-Large Context Windows
Xinyuan Song, Bowen Zhu, Hasibul Haque +1
cs.AIarXiv:2608.22137v12026Solving math word problems with process- and outcome-based feedback
Jonathan Uesato, Nate Kushman, Ramana Kumar +6
cs.LGcs.AIcs.CLarXiv:2211.14275v12022Zero-Shot Relation Extraction via Reading Comprehension
Omer Levy, Minjoon Seo, Eunsol Choi +1
cs.CLcs.AIcs.LGarXiv:1706.04115v12017Counterfactual Reasoning and Learning Systems
Léon Bottou, Jonas Peters, Joaquin Quiñonero-Candela +6
cs.LGcs.AIcs.IRarXiv:1209.2355v52012A Safety Report on GPT-5.2, Gemini 3 Pro, Qwen3-VL, Grok 4.1 Fast, Nano Banana Pro, and Seedream 4.5
Xingjun Ma, Yixu Wang, Hengyuan Xu +18
cs.AIcs.CLcs.CVarXiv:2601.10527v22026Prompt Injection attack against LLM-integrated Applications
Yi Liu, Gelei Deng, Yuekang Li +9
cs.CRcs.AIcs.CLarXiv:2306.05499v32023TranslateGemma Technical Report
Mara Finkelstein, Isaac Caswell, Tobias Domhan +18
cs.CLcs.AIarXiv:2601.09012v32026DeepSeek-VL: Towards Real-World Vision-Language Understanding
Haoyu Lu, Wen Liu, Bo Zhang +12
cs.AIarXiv:2403.05525v22024MonitorBench: A Comprehensive Benchmark for Chain-of-Thought Monitorability in Large Language Models
Han Wang, Yifan Sun, Brian Ko +8
cs.AIarXiv:2603.28590v32026CUA-Suite: Massive Human-annotated Video Demonstrations for Computer-Use Agents
Xiangru Jian, Shravan Nayak, Kevin Qinghong Lin +5
cs.LGcs.AIcs.CVarXiv:2603.24440v12026Red Teaming Language Models to Reduce Harms: Methods, Scaling Behaviors, and Lessons Learned
Deep Ganguli, Liane Lovitt, Jackson Kernion +33
cs.CLcs.AIcs.CYarXiv:2209.07858v22022Perceiver-Actor: A Multi-Task Transformer for Robotic Manipulation
Mohit Shridhar, Lucas Manuelli, Dieter Fox
cs.ROcs.AIcs.CLarXiv:2209.05451v22022TAG-MoE: Task-Aware Gating for Unified Generative Mixture-of-Experts
Yu Xu, Hongbin Yan, Juan Cao +11
cs.CVcs.AIarXiv:2601.08881v22026DM4CT: Benchmarking Diffusion Models for Computed Tomography Reconstruction
Jiayang Shi, Daniel M. Pelt, K. Joost Batenburg
eess.IVcs.AIcs.CVarXiv:2602.18589v12026MDPBench: A Benchmark for Multilingual Document Parsing in Real-World Scenarios
Zhang Li, Zhibo Lin, Qiang Liu +7
cs.CVcs.AIarXiv:2603.28130v12026MiroEval: Benchmarking Multimodal Deep Research Agents in Process and Outcome
Fangda Ye, Yuxin Hu, Pengxiang Zhu +19
cs.AIcs.CLarXiv:2603.28407v12026How to train your ViT? Data, Augmentation, and Regularization in Vision Transformers
Andreas Steiner, Alexander Kolesnikov, Xiaohua Zhai +3
cs.CVcs.AIcs.LGarXiv:2106.10270v22021MolmoPoint: Better Pointing for VLMs with Grounding Tokens
Christopher Clark, Yue Yang, Jae Sung Park +8
cs.CVcs.AIarXiv:2603.28069v12026Towards a Medical AI Scientist
Hongtao Wu, Boyun Zheng, Dingjie Song +5
cs.AIcs.LGarXiv:2603.28589v12026The Neuro-Symbolic Concept Learner: Interpreting Scenes, Words, and Sentences From Natural Supervision
Jiayuan Mao, Chuang Gan, Pushmeet Kohli +2
cs.CVcs.AIcs.CLarXiv:1904.12584v12019LLVIP: A Visible-infrared Paired Dataset for Low-light Vision
Xinyu Jia, Chuang Zhu, Minzhen Li +3
cs.CVcs.AIarXiv:2108.10831v42021Recommendations as Treatments: Debiasing Learning and Evaluation
Tobias Schnabel, Adith Swaminathan, Ashudeep Singh +2
cs.LGcs.AIcs.IRarXiv:1602.05352v22016