Artificial Intelligence
Papers filed under cs.AI on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.
Search paper metadata (including unsummarized papers)
12,541 to 12,600 of 15,247
Revisiting the Platonic Representation Hypothesis: An Aristotelian View
Fabian Gröger, Shuo Wen, Maria Brbić
cs.LGcs.AIcs.CVarXiv:2602.14486v22026A Survey on RAG Meeting LLMs: Towards Retrieval-Augmented Large Language Models
Wenqi Fan, Yujuan Ding, Liangbo Ning +5
cs.CLcs.AIcs.IRarXiv:2405.06211v32024CUA-Skill: Develop Skills for Computer Using Agent
Tianyi Chen, Yinheng Li, Michael Solodko +12
cs.AIarXiv:2601.21123v22026Kernels for Vector-Valued Functions: a Review
Mauricio A. Alvarez, Lorenzo Rosasco, Neil D. Lawrence
stat.MLcs.AImath.STarXiv:1106.6251v22011MiroThinker-1.7 & H1: Towards Heavy-Duty Research Agents via Verification
MiroMind Team, S. Bai, L. Bing +41
cs.CLcs.AIcs.IRarXiv:2603.15726v12026UI-Venus-1.5 Technical Report
Venus Team, Changlong Gao, Zhangxuan Gu +24
cs.CVcs.AIcs.CLarXiv:2602.09082v22026Infinite-World: Scaling Interactive World Models to 1000-Frame Horizons via Pose-Free Hierarchical Memory
Ruiqi Wu, Xuanhua He, Meng Cheng +8
cs.CVcs.AIarXiv:2602.02393v22026TOPReward: Token Probabilities as Hidden Zero-Shot Rewards for Robotics
Shirui Chen, Cole Harrison, Ying-Chun Lee +6
cs.ROcs.AIcs.LGarXiv:2602.19313v22026Sparse but Critical: A Token-Level Analysis of Distributional Shifts in RLVR Fine-Tuning of LLMs
Haoming Meng, Kexin Huang, Shaohang Wei +6
cs.CLcs.AIcs.LGarXiv:2603.22446v12026FeUdal Networks for Hierarchical Reinforcement Learning
Alexander Sasha Vezhnevets, Simon Osindero, Tom Schaul +4
cs.AIarXiv:1703.01161v22017Selective Classification for Deep Neural Networks
Yonatan Geifman, Ran El-Yaniv
cs.LGcs.AIarXiv:1705.08500v22017A General Theoretical Paradigm to Understand Learning from Human Preferences
Mohammad Gheshlaghi Azar, Mark Rowland, Bilal Piot +4
cs.AIcs.LGstat.MLarXiv:2310.12036v22023Pick-a-Pic: An Open Dataset of User Preferences for Text-to-Image Generation
Yuval Kirstain, Adam Polyak, Uriel Singer +3
cs.CVcs.AIarXiv:2305.01569v22023Deep Kernel Learning
Andrew Gordon Wilson, Zhiting Hu, Ruslan Salakhutdinov +1
cs.LGcs.AIstat.MEarXiv:1511.02222v12015Harnessing the Power of LLMs in Practice: A Survey on ChatGPT and Beyond
Jingfeng Yang, Hongye Jin, Ruixiang Tang +5
cs.CLcs.AIcs.LGarXiv:2304.13712v22023Math-Shepherd: Verify and Reinforce LLMs Step-by-step without Human Annotations
Peiyi Wang, Lei Li, Zhihong Shao +6
cs.AIcs.CLcs.LGarXiv:2312.08935v32023Feature Hashing for Large Scale Multitask Learning
Kilian Weinberger, Anirban Dasgupta, Josh Attenberg +2
cs.AIarXiv:0902.2206v52009Neural Network Dynamics for Model-Based Deep Reinforcement Learning with Model-Free Fine-Tuning
Anusha Nagabandi, Gregory Kahn, Ronald S. Fearing +1
cs.LGcs.AIcs.ROarXiv:1708.02596v22017EvoCUA: Evolving Computer Use Agents via Learning from Scalable Synthetic Experience
Taofeng Xue, Chong Peng, Mianqiu Huang +13
cs.AIarXiv:2601.15876v22026Spatial-Temporal Fusion Graph Neural Networks for Traffic Flow Forecasting
Mengzhang Li, Zhanxing Zhu
cs.LGcs.AIarXiv:2012.09641v22020OpenSeeker: Democratizing Frontier Search Agents by Fully Open-Sourcing Training Data
Yuwen Du, Rui Ye, Shuo Tang +4
cs.AIcs.CLarXiv:2603.15594v12026Learning to reinforcement learn
Jane X Wang, Zeb Kurth-Nelson, Dhruva Tirumala +6
cs.LGcs.AIstat.MLarXiv:1611.05763v32016Interactive Attention Networks for Aspect-Level Sentiment Classification
Dehong Ma, Sujian Li, Xiaodong Zhang +1
cs.AIcs.CLarXiv:1709.00893v12017A Survey on Large Language Models for Code Generation
Juyong Jiang, Fan Wang, Jiasi Shen +2
cs.CLcs.AIcs.SEarXiv:2406.00515v22024Kiwi-Edit: Versatile Video Editing via Instruction and Reference Guidance
Yiqi Lin, Guoqiang Liang, Ziyun Zeng +3
cs.CVcs.AIarXiv:2603.02175v42026Men Also Like Shopping: Reducing Gender Bias Amplification using Corpus-level Constraints
Jieyu Zhao, Tianlu Wang, Mark Yatskar +2
cs.AIcs.CLcs.CVarXiv:1707.09457v12017REDSearcher: A Scalable and Cost-Efficient Framework for Long-Horizon Search Agents
Zheng Chu, Xiao Wang, Jack Hong +11
cs.AIcs.CLarXiv:2602.14234v12026Zooming without Zooming: Region-to-Image Distillation for Fine-Grained Multimodal Perception
Lai Wei, Liangbo He, Jun Lan +9
cs.CVcs.AIcs.CLarXiv:2602.11858v22026Hindsight Credit Assignment for Long-Horizon LLM Agents
Hui-Ze Tan, Xiao-Wen Yang, Hao Chen +7
cs.LGcs.AIarXiv:2603.08754v12026Summaries:简体中文Agent World Model: Infinity Synthetic Environments for Agentic Reinforcement Learning
Zhaoyang Wang, Canwen Xu, Boyi Liu +5
cs.AIcs.CLcs.LGarXiv:2602.10090v32026RubricHub: A Comprehensive and Highly Discriminative Rubric Dataset via Automated Coarse-to-Fine Generation
Sunzhu Li, Jiale Zhao, Miteto Wei +6
cs.AIarXiv:2601.08430v22026Large Language Model based Multi-Agents: A Survey of Progress and Challenges
Taicheng Guo, Xiuying Chen, Yaqi Wang +5
cs.CLcs.AIcs.MAarXiv:2402.01680v22024Pre-Trained Models: Past, Present and Future
Xu Han, Zhengyan Zhang, Ning Ding +21
cs.AIcs.CLarXiv:2106.07139v32021Vision-DeepResearch: Incentivizing DeepResearch Capability in Multimodal Large Language Models
Wenxuan Huang, Yu Zeng, Qiuchen Wang +14
cs.CVcs.AIarXiv:2601.22060v32026TS2Vec: Towards Universal Representation of Time Series
Zhihan Yue, Yujing Wang, Juanyong Duan +4
cs.LGcs.AIarXiv:2106.10466v42021RoboMME: Benchmarking and Understanding Memory for Robotic Generalist Policies
Yinpei Dai, Hongze Fu, Jayjun Lee +6
cs.ROcs.AIarXiv:2603.04639v32026Natural-Language Agent Harnesses
Linyue Pan, Lexiao Zou, Shuo Guo +2
cs.CLcs.AIarXiv:2603.25723v22026Kimi K2.5: Visual Agentic Intelligence
Kimi Team, Tongtong Bai, Yifan Bai +334
cs.CLcs.AIcs.LGarXiv:2602.02276v22026Summaries:한국어Revisiting On-Policy Distillation: Empirical Failure Modes and Simple Fixes
Yuqian Fu, Haohuan Huang, Kaiwen Jiang +4
cs.LGcs.AIcs.CLarXiv:2603.25562v22026Meta-Harness: End-to-End Optimization of Model Harnesses
Yoonho Lee, Roshen Nair, Qizheng Zhang +3
cs.AIarXiv:2603.28052v12026Towards a Science of AI Agent Reliability
Stephan Rabanser, Sayash Kapoor, Peter Kirgis +3
cs.AIcs.CYcs.LGarXiv:2602.16666v32026Learning beyond Teacher: Generalized On-Policy Distillation with Reward Extrapolation
Wenkai Yang, Weijie Liu, Ruobing Xie +3
cs.LGcs.AIcs.CLarXiv:2602.12125v22026LLaDA2.1: Speeding Up Text Diffusion via Token Editing
Tiwei Bie, Maosong Cao, Xiang Cao +47
cs.LGcs.AIarXiv:2602.08676v32026SkillsBench: Benchmarking How Well Agent Skills Work Across Diverse Tasks
Xiangyi Li, Yimin Liu, Wenbo Chen +75
cs.AIarXiv:2602.12670v42026Trace2Skill: Distill Trajectory-Local Lessons into Transferable Agent Skills
Jingwei Ni, Yihao Liu, Xinpeng Liu +7
cs.AIarXiv:2603.25158v52026Molmo2: Open Weights and Data for Vision-Language Models with Video Understanding and Grounding
Christopher Clark, Jieyu Zhang, Zixian Ma +18
cs.CVcs.AIarXiv:2601.10611v42026Agent Skills in the Wild: An Empirical Study of Security Vulnerabilities at Scale
Yi Liu, Weizhe Wang, Ruitao Feng +5
cs.CRcs.AIcs.CLarXiv:2601.10338v12026Numina-Lean-Agent: An Open and General Agentic Reasoning System for Formal Mathematics
Junqi Liu, Zihao Zhou, Zekai Zhu +10
cs.AIarXiv:2601.14027v12026CUDA Agent: Large-Scale Agentic RL for High-Performance CUDA Kernel Generation
Weinan Dai, Hanlin Wu, Qiying Yu +13
cs.LGcs.AIarXiv:2602.24286v12026SkillNet: Create, Evaluate, and Connect AI Skills
Yuan Liang, Ruobin Zhong, Haoming Xu +47
cs.AIcs.CLcs.CVarXiv:2603.04448v22026Mobile-Agent-v3.5: Multi-platform Fundamental GUI Agents
Haiyang Xu, Xi Zhang, Haowei Liu +16
cs.AIcs.CLarXiv:2602.16855v12026OpenClaw-RL: Train Any Agent Simply by Talking
Yinjie Wang, Xuyang Chen, Xiaolong Jin +2
cs.CLcs.AIcs.CVarXiv:2603.10165v22026MemSkill: Learning and Evolving Memory Skills for Self-Evolving Agents
Haozhen Zhang, Quanyu Long, Jianzhu Bao +4
cs.CLcs.AIcs.LGarXiv:2602.02474v22026SWE-Skills-Bench: Do Agent Skills Actually Help in Real-World Software Engineering?
Tingxu Han, Yi Zhang, Wei Song +4
cs.SEcs.AIarXiv:2603.15401v12026Cosmos Policy: Fine-Tuning Video Models for Visuomotor Control and Planning
Moo Jin Kim, Yihuai Gao, Tsung-Yi Lin +8
cs.AIcs.ROarXiv:2601.16163v12026DreamDojo: A Generalist Robot World Model from Large-Scale Human Videos
Shenyuan Gao, William Liang, Kaiyuan Zheng +27
cs.ROcs.AIcs.CVarXiv:2602.06949v12026Terminal-Bench: Benchmarking Agents on Hard, Realistic Tasks in Command Line Interfaces
Mike A. Merrill, Alexander G. Shaw, Nicholas Carlini +82
cs.SEcs.AIarXiv:2601.11868v12026EgoSim: Egocentric World Simulator for Embodied Interaction Generation
Jinkun Hao, Mingda Jia, Ruiyan Wang +7
cs.CVcs.AIarXiv:2604.01001v22026PixelPrune: Pixel-Level Adaptive Visual Token Reduction via Predictive Coding
Nan Wang, Zhiwei Jin, Chen Chen +1
cs.CVcs.AIcs.CLarXiv:2604.00886v12026Signals: Trajectory Sampling and Triage for Agentic Interactions
Shuguang Chen, Adil Hafeez, Salman Paracha
cs.AIcs.CLarXiv:2604.00356v12026