Artificial Intelligence
Papers filed under cs.AI on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.
Search paper metadata (including unsummarized papers)
12,121 to 12,180 of 15,448
DIVE: Scaling Diversity in Agentic Task Synthesis for Generalizable Tool Use
Aili Chen, Chi Zhang, Junteng Liu +11
cs.AIcs.SEarXiv:2603.11076v12026On Distillation of Guided Diffusion Models
Chenlin Meng, Robin Rombach, Ruiqi Gao +4
cs.CVcs.AIcs.LGarXiv:2210.03142v32022Brevity Constraints Reverse Performance Hierarchies in Language Models
MD Azizul Hakim
cs.CLcs.AIarXiv:2604.00025v12026TERMINATOR: Learning Optimal Exit Points for Early Stopping in Chain-of-Thought Reasoning
Alliot Nagle, Jakhongir Saydaliev, Dhia Garbaya +3
cs.LGcs.AIcs.CLarXiv:2603.12529v22026TinyLlama: An Open-Source Small Language Model
Peiyuan Zhang, Guangtao Zeng, Tianduo Wang +1
cs.CLcs.AIarXiv:2401.02385v22024The Web as a Knowledge-base for Answering Complex Questions
Alon Talmor, Jonathan Berant
cs.CLcs.AIcs.LGarXiv:1803.06643v12018PolyChirp: Multi-Species Birdsong Classification Using TinyML on Low-Power Acoustic Sensors
Nathan Duboisset, Zhaolan Huang, Felix Bießmann +3
cs.LGcs.AIarXiv:2608.23101v12026Trust in AI and Its Role in the Acceptance of AI Technologies
Hyesun Choung, Prabu David, Arun Ross
cs.AIcs.HCarXiv:2203.12687v12022Future Querying: Can LLMs Serve as Implicit Medical World Models?
Siri Willems, James Butterworth, Lore Goetschalckx +4
cs.CLcs.AIarXiv:2608.23248v12026Supervised Fine-Tuning versus Reinforcement Learning: A Study of Post-Training Methods for Large Language Models
Haitao Jiang, Wenbo Zhang, Jiarui Yao +3
cs.AIcs.CLarXiv:2603.13985v12026Why Steering Works: Toward a Unified View of Language Model Parameter Dynamics
Ziwen Xu, Chenyan Wu, Hengyu Sun +9
cs.CLcs.AIcs.CVarXiv:2602.02343v32026MedSAM-Agent: Empowering Interactive Medical Image Segmentation with Multi-turn Agentic Reinforcement Learning
Shengyuan Liu, Liuxin Bao, Qi Yang +6
cs.CVcs.AIarXiv:2602.03320v12026Automated Construction of FAIR Digital Object Knowledge Graphs from Flat Cultural Heritage Records
Zeyd Boukhers, Lingxiao Kong, Xenophon Zabulis +1
cs.AIcs.CLcs.DLarXiv:2608.23263v12026Learning Representations for Counterfactual Inference
Fredrik D. Johansson, Uri Shalit, David Sontag
stat.MLcs.AIcs.LGarXiv:1605.03661v32016Rethinking LLM-as-a-Judge: Representation-as-a-Judge with Small Language Models via Semantic Capacity Asymmetry
Zhuochun Li, Yong Zhang, Ming Li +8
cs.CLcs.AIcs.LGarXiv:2601.22588v22026SnapKV: LLM Knows What You are Looking for Before Generation
Yuhong Li, Yingbing Huang, Bowen Yang +6
cs.CLcs.AIarXiv:2404.14469v22024RecGOAT: Graph Optimal Adaptive Transport for LLM-Enhanced Multimodal Recommendation with Dual Semantic Alignment
Yuecheng Li, Hengwei Ju, Zeyu Song +4
cs.IRcs.AIarXiv:2602.00682v22026SCALER:Synthetic Scalable Adaptive Learning Environment for Reasoning
Caijun Xu, Changyi Xiao, Zhongyuan Peng +2
cs.AIarXiv:2601.04809v52026SafePred: A Predictive Guardrail for Computer-Using Agents via World Models
Yurun Chen, Zeyi Liao, Ping Yin +3
cs.CLcs.AIcs.LGarXiv:2602.01725v12026Manipulating and Measuring Model Interpretability
Forough Poursabzi-Sangdeh, Daniel G. Goldstein, Jake M. Hofman +2
cs.AIcs.CYarXiv:1802.07810v52018From Passive Observer to Active Critic: Reinforcement Learning Elicits Process Reasoning for Robotic Manipulation
Yibin Liu, Yaxing Lyu, Daqi Gao +5
cs.ROcs.AIcs.CLarXiv:2603.15600v22026IQuest-Coder-V1 Technical Report
Jian Yang, Wei Zhang, Shawn Guo +35
cs.AIcs.CLcs.SEarXiv:2603.16733v12026A Survey of Multi-Objective Sequential Decision-Making
Diederik Marijn Roijers, Peter Vamplew, Shimon Whiteson +1
cs.AIarXiv:1402.0590v12014CRITIC: Large Language Models Can Self-Correct with Tool-Interactive Critiquing
Zhibin Gou, Zhihong Shao, Yeyun Gong +4
cs.CLcs.AIarXiv:2305.11738v42023LatentChem: From Textual CoT to Latent Thinking in Chemical Reasoning
Xinwu Ye, Yicheng Mao, Yuxuan Liao +16
physics.chem-phcs.AIcs.CLarXiv:2602.07075v62026FlexGen: High-Throughput Generative Inference of Large Language Models with a Single GPU
Ying Sheng, Lianmin Zheng, Binhang Yuan +11
cs.LGcs.AIcs.PFarXiv:2303.06865v22023Coarse-Guided Visual Generation via Weighted h-Transform Sampling
Yanghao Wang, Ziqi Jiang, Zhen Wang +1
cs.CVcs.AIarXiv:2603.12057v22026HERO: Human-profile Enhanced Retrieval Optimization Framework for Long-term Agent Memory
Yuanhua Lin, Yile Li, Zhiyuan Zhao +2
cs.AIarXiv:2608.22310v12026Clarify User Expertise: Towards Proactive Conversational Agents Tailoring Responses to User Proficiency
Zhihong Cao, Chen Huang
cs.AIcs.CLarXiv:2608.22266v12026Read Less, Solve More: Token-Efficient Sparse Reading for AI Agents
Zedong Liu, Jiaan Wu, Xinyang Ma +5
cs.AIarXiv:2608.22237v12026Beyond What Meets the Eye: Unveiling Situational Illusions for Multimodal Large Language Models
Zhiming Yang, Zhuoxi Xiong, Donglin Zhou +3
cs.AIcs.CLcs.CVarXiv:2608.22232v12026Query-Driven Multimodal Information Extraction from Long Documents
Yikai Gao, Ding Xia, Xi Yang
cs.AIcs.MMarXiv:2608.22214v12026Mitigating Bias in Algorithmic Hiring: Evaluating Claims and Practices
Manish Raghavan, Solon Barocas, Jon Kleinberg +1
cs.CYcs.AIcs.LGarXiv:1906.09208v32019BenchPreS: A Benchmark for Context-Aware Personalized Preference Selectivity of Persistent-Memory LLMs
Sangyeon Yoon, Sunkyoung Kim, Hyesoo Hong +5
cs.AIcs.CLarXiv:2603.16557v12026Unveiling Implicit Advantage Symmetry: Why GRPO Struggles with Exploration and Difficulty Adaptation
Zhiqi Yu, Zhangquan Chen, Mengting Liu +2
cs.LGcs.AIarXiv:2602.05548v32026Disagree to Explore, Agree to Commit: Routing-Guided Test-Time Scaling for Software Agents
Kang Chen, Junjie Nian, Yixin Cao +1
cs.AIcs.SEarXiv:2608.22191v12026Anticipatory Planning for Multimodal AI Agents
Yongyuan Liang, Shijie Zhou, Yu Gu +6
cs.AIarXiv:2603.16777v12026Role-Specialized Mixture-of-Agents with Open-Weight LLMs for Clinical Prediction
Jun Hou, Yi Fang, Xuan Wang
cs.AIcs.LGarXiv:2608.22176v12026Code-Space Response Oracles: Generating Interpretable Multi-Agent Policies with Large Language Models
Daniel Hennes, Zun Li, John Schultz +1
cs.GTcs.AIcs.LGarXiv:2603.10098v12026MCP-Universe RL: A Framework for Training MCP Tool-Use Agents via Reinforcement Learning
Ziyang Luo, Yan Yang, Xiangru Jian +5
cs.AIcs.LGarXiv:2608.22167v12026Small Language Model enabled Autonomous agent for Language-Conditioned Cognitive Radar
Minhaj Uddin Ahmad, Zakia Zaman, Shunqiao Sun +1
eess.SPcs.AIeess.SYarXiv:2608.11596v12026Nanbeige4.1-3B: A Small General Model that Reasons, Aligns, and Acts
Chen Yang, Guangyue Peng, Jiaying Zhu +12
cs.AIcs.CLarXiv:2602.13367v12026Task-Driven 3D Printability Assistance via Geometry- and Knowledge-Grounded LLM Reasoning
Zhaoda Du, Qiaojie Zheng, Xiaoli Zhang
cs.AIarXiv:2608.22128v12026Evaluation of Small Vision-Language Models on Qualitative Mechanical Problems
Henry Fordjour Ansah, Shreya Banerjee, Pranish Ghimire
cs.AIarXiv:2608.22143v12026Aggregation-Aware Synthetic Text Generation Against Authorship Re-Identification
Qian Ma, Anna Squicciarini, Sarah Rajtmajer
cs.AIcs.CLarXiv:2608.22161v12026Flow Equivariant World Models: Memory for Partially Observed Dynamic Environments
Hansen Jin Lillemark, Benhao Huang, Fangneng Zhan +2
cs.LGcs.AIcs.CVarXiv:2601.01075v22026RegNeRF: Regularizing Neural Radiance Fields for View Synthesis from Sparse Inputs
Michael Niemeyer, Jonathan T. Barron, Ben Mildenhall +3
cs.CVcs.AIcs.GRarXiv:2112.00724v12021BrowseComp-$V^3$: A Visual, Vertical, and Verifiable Benchmark for Multimodal Browsing Agents
Huanyao Zhang, Jiepeng Zhou, Bo Li +22
cs.AIarXiv:2602.12876v22026Vega: Learning to Drive with Natural Language Instructions
Sicheng Zuo, Yuxuan Li, Wenzhao Zheng +3
cs.CVcs.AIcs.ROarXiv:2603.25741v22026GISA: A Benchmark for General Information-Seeking Assistant
Yutao Zhu, Xingshuo Zhang, Maosen Zhang +9
cs.CLcs.AIcs.IRarXiv:2602.08543v22026Mixture-of-Experts with Expert Choice Routing
Yanqi Zhou, Tao Lei, Hanxiao Liu +7
cs.LGcs.AIarXiv:2202.09368v22022Self-EvolveRec: Self-Evolving Recommender Systems with LLM-based Directional Feedback
Sein Kim, Sangwu Park, Hongseok Kang +6
cs.IRcs.AIarXiv:2602.12612v22026NExT-GPT: Any-to-Any Multimodal LLM
Shengqiong Wu, Hao Fei, Leigang Qu +2
cs.AIcs.CLcs.LGarXiv:2309.05519v32023AUDITA: certified auditing and causal attribution of adverse outcomes in autonomous multi-agent systems
Zhixu Du, Yiran Chen
cs.AIarXiv:2608.22160v12026The many Shapley values for model explanation
Mukund Sundararajan, Amir Najmi
cs.AIcs.LGecon.THarXiv:1908.08474v22019The Flexibility Trap: Rethinking the Value of Arbitrary Order in Diffusion Language Models
Zanlin Ni, Shenzhi Wang, Yang Yue +8
cs.CLcs.AIcs.LGarXiv:2601.15165v42026Measuring Stability and Failure Behavior in Language Models Under Structured Perturbations
Samira Golsefid
cs.AIcs.CLarXiv:2608.22138v12026The Pensieve Paradigm: Stateful Language Models Mastering Their Own Context
Xiaoyuan Liu, Tian Liang, Dongyang Ma +4
cs.AIarXiv:2602.12108v12026MEMONDEMAND: A Memory Management System for Large-Scale Enterprise Data
Xinyuan Song, Bowen Zhu, Hasibul Haque +1
cs.AIarXiv:2608.22141v12026MegaMem: A Retrieval Solution for Ultra-Large Context Windows
Xinyuan Song, Bowen Zhu, Hasibul Haque +1
cs.AIarXiv:2608.22137v12026