Artificial Intelligence
Papers filed under cs.AI on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.
Search paper metadata (including unsummarized papers)
6,661 to 6,720 of 15,403
Pixel Reasoner: Incentivizing Pixel-Space Reasoning with Curiosity-Driven Reinforcement Learning
Haozhe Wang, Alex Su, Weiming Ren +2
cs.CVcs.AIcs.CLarXiv:2505.15966v32025Learning a Recurrent Visual Representation for Image Caption Generation
Xinlei Chen, C. Lawrence Zitnick
cs.CVcs.AIcs.CLarXiv:1411.5654v12014VL-Rethinker: Incentivizing Self-Reflection of Vision-Language Models with Reinforcement Learning
Haozhe Wang, Chao Qu, Zuming Huang +3
cs.LGcs.AIarXiv:2504.08837v32025WebThinker: Empowering Large Reasoning Models with Deep Research Capability
Xiaoxi Li, Jiajie Jin, Guanting Dong +5
cs.CLcs.AIcs.IRarXiv:2504.21776v22025FinBen: A Holistic Financial Benchmark for Large Language Models
Qianqian Xie, Weiguang Han, Zhengyu Chen +31
cs.CLcs.AIcs.CEarXiv:2402.12659v22024Learning to Reason under Off-Policy Guidance
Jianhao Yan, Yafu Li, Zican Hu +5
cs.LGcs.AIcs.CLarXiv:2504.14945v52025LeJEPA: Provable and Scalable Self-Supervised Learning Without the Heuristics
Randall Balestriero, Yann LeCun
cs.LGcs.AIcs.CVarXiv:2511.08544v32025Compute and Energy Consumption Trends in Deep Learning Inference
Radosvet Desislavov, Fernando Martínez-Plumed, José Hernández-Orallo
cs.LGcs.AIarXiv:2109.05472v22021DRACO: Fine-Grained Credit Assignment with Dynamic Rubrics for Long-Horizon Agent Training
Shubham Gandhi, Saurabh Goyal, Kiran Kate +1
cs.AIcs.LGcs.SEarXiv:2609.04094v12026Towards AI-Assisted Clinical Trial Matching: Practical Considerations, Multicenter Evaluation, and Real-World Deployment
Yin Fang, Qiao Jin, Shubo Tian +24
cs.CLcs.AIcs.CYarXiv:2609.01202v12026Accelerating scientific discovery with Co-Scientist
Juraj Gottweis, Wei-Hung Weng, Alexander Daryin +48
cs.AIcs.CLcs.HCarXiv:2502.18864v22025ToolRL: Reward is All Tool Learning Needs
Cheng Qian, Emre Can Acikgoz, Qi He +5
cs.LGcs.AIcs.CLarXiv:2504.13958v12025WorldVLA: Towards Autoregressive Action World Model
Jun Cen, Chaohui Yu, Hangjie Yuan +9
cs.ROcs.AIarXiv:2506.21539v12025Humanity's Last Exam
Long Phan, Alice Gatti, Ziwen Han +1155
cs.LGcs.AIcs.CLarXiv:2501.14249v112025Fast and Eager k-Medoids Clustering: O(k) Runtime Improvement of the PAM, CLARA, and CLARANS Algorithms
Erich Schubert, Peter J. Rousseeuw
cs.LGcs.AIstat.MLarXiv:2008.05171v22020Swin Meets EfficientNet: Lightweight Architectures for GAN-Based Face Forensics
Sejuti Basu, Ashima Sood, Vijay Kumar +1
cs.CVcs.AIarXiv:2609.01749v12026Designing Proactive Thought Partners for Writing
Chao Zhang, Abe Davis, Chih-Wei Chen +1
cs.HCcs.AIcs.CLarXiv:2609.01588v12026The Illusion of Thinking: Understanding the Strengths and Limitations of Reasoning Models via the Lens of Problem Complexity
Parshin Shojaee, Iman Mirzadeh, Keivan Alizadeh +3
cs.AIcs.CLcs.LGarXiv:2506.06941v32025MathArena: Evaluating LLMs on Uncontaminated Math Competitions
Mislav Balunović, Jasper Dekoninck, Ivo Petrov +2
cs.AIcs.CLarXiv:2505.23281v32025"Help Me Help the AI": Understanding How Explainability Can Support Human-AI Interaction
Sunnie S. Y. Kim, Elizabeth Anne Watkins, Olga Russakovsky +2
cs.HCcs.AIcs.CVarXiv:2210.03735v22022Are We There Yet? Assessing Computer-Use Agents for Blind Users' Accessible Interaction with Desktop Applications
Satwik Ram Kodandaram, Monalika Padma Reddy, Xiaojun Bi +3
cs.HCcs.AIarXiv:2609.00524v12026L1: Controlling How Long A Reasoning Model Thinks With Reinforcement Learning
Pranjal Aggarwal, Sean Welleck
cs.CLcs.AIcs.LGarXiv:2503.04697v22025Memory in the Age of AI Agents
Yuyang Hu, Shichun Liu, Yanwei Yue +44
cs.CLcs.AIarXiv:2512.13564v22025DeepSeek-Prover-V2: Advancing Formal Mathematical Reasoning via Reinforcement Learning for Subgoal Decomposition
Z. Z. Ren, Zhihong Shao, Junxiao Song +15
cs.CLcs.AIarXiv:2504.21801v22025Agent Laboratory: Using LLM Agents as Research Assistants
Samuel Schmidgall, Yusheng Su, Ze Wang +7
cs.HCcs.AIcs.CLarXiv:2501.04227v22025X-VLA: Soft-Prompted Transformer as Scalable Cross-Embodiment Vision-Language-Action Model
Jinliang Zheng, Jianxiong Li, Zhihao Wang +12
cs.ROcs.AIcs.CVarXiv:2510.10274v12025LEAP: Likelihood Elicitation and Aggregation for LLM-based Probabilistic Forecasting
Yufei Chen, Yiran Zhao, Xiaogang Xu +3
cs.AIarXiv:2609.01337v12026GLM-4.5V and GLM-4.1V-Thinking: Towards Versatile Multimodal Reasoning with Scalable Reinforcement Learning
GLM-V Team, :, Wenyi Hong +91
cs.CVcs.AIcs.LGarXiv:2507.01006v62025Multi-Agent Collaboration Mechanisms: A Survey of LLMs
Khanh-Tung Tran, Dung Dao, Minh-Duong Nguyen +3
cs.AIarXiv:2501.06322v12025The Regretful Agent: Heuristic-Aided Navigation through Progress Estimation
Chih-Yao Ma, Zuxuan Wu, Ghassan AlRegib +2
cs.AIcs.CVcs.ROarXiv:1903.01602v12019QCell: Recombining and Aligning Cell Queries for Overlapping Instance Segmentation
Yaroslav Prytula, Anton Popov, Dmytro Fishman
cs.CVcs.AIcs.LGarXiv:2608.29253v12026Muon is Scalable for LLM Training
Jingyuan Liu, Jianlin Su, Xingcheng Yao +25
cs.LGcs.AIcs.CLarXiv:2502.16982v12025The Lessons of Developing Process Reward Models in Mathematical Reasoning
Zhenru Zhang, Chujie Zheng, Yangzhen Wu +6
cs.CLcs.AIcs.LGarXiv:2501.07301v22025Search-o1: Agentic Search-Enhanced Large Reasoning Models
Xiaoxi Li, Guanting Dong, Jiajie Jin +5
cs.AIcs.CLcs.IRarXiv:2501.05366v12025Deep Probabilistic Programming
Dustin Tran, Matthew D. Hoffman, Rif A. Saurous +3
stat.MLcs.AIcs.LGarXiv:1701.03757v22017Contrastive Learning for Label-Efficient Semantic Segmentation
Xiangyun Zhao, Raviteja Vemulapalli, Philip Mansfield +4
cs.CVcs.AIcs.LGarXiv:2012.06985v42020Agentic Context Engineering: Evolving Contexts for Self-Improving Language Models
Qizheng Zhang, Changran Hu, Shubhangi Upasani +10
cs.LGcs.AIcs.CLarXiv:2510.04618v32025Block Diffusion: Interpolating Between Autoregressive and Diffusion Language Models
Marianne Arriola, Aaron Gokaslan, Justin T. Chiu +5
cs.LGcs.AIarXiv:2503.09573v32025Who Should I Trust: AI or Myself? Leveraging Human and AI Correctness Likelihood to Promote Appropriate Trust in AI-Assisted Decision-Making
Shuai Ma, Ying Lei, Xinru Wang +4
cs.HCcs.AIcs.LGarXiv:2301.05809v12023Process Reinforcement through Implicit Rewards
Ganqu Cui, Lifan Yuan, Zefan Wang +22
cs.LGcs.AIcs.CLarXiv:2502.01456v22025Modeling Human Motion with Quaternion-based Neural Networks
Dario Pavllo, Christoph Feichtenhofer, Michael Auli +1
cs.CVcs.AIcs.ROarXiv:1901.07677v22019From Reweighting to Rewriting: Unlocking the Intervention Effects of Influential Samples in Training Data Attribution
Yuzhang Luo, Chenpeng Wang, Jianhui Chen +1
cs.CLcs.AIcs.LGarXiv:2609.02771v12026Coverage, Not Targeting: A Structural Regime in Multi-Turn Agent Credit Assignment
Chenyu Zhou, Qiliang Jiang, Shuning Wu +1
cs.LGcs.AIarXiv:2609.02417v12026Political Compass or Spinning Arrow? Towards More Meaningful Evaluations for Values and Opinions in Large Language Models
Paul Röttger, Valentin Hofmann, Valentina Pyatkin +4
cs.CLcs.AIarXiv:2402.16786v22024DeepSeek-Prover-V1.5: Harnessing Proof Assistant Feedback for Reinforcement Learning and Monte-Carlo Tree Search
Huajian Xin, Z. Z. Ren, Junxiao Song +14
cs.CLcs.AIcs.LGarXiv:2408.08152v12024A Network Science Perspective on Evaluating Deep Graph Generative Models
Tianrui Mao, Abele Malan, Megha Khosla +2
cs.SIcs.AIarXiv:2609.01015v12026Dyna-Style Planning with Linear Function Approximation and Prioritized Sweeping
Richard S. Sutton, Csaba Szepesvari, Alborz Geramifard +1
cs.AIcs.LGeess.SYarXiv:1206.3285v12012What Is Worth Representing? Representational Empowerment for Continual Model Construction
Fei Dai, Hanqi Zhou, Alison Gopnik +1
cs.LGcs.AIarXiv:2609.02322v12026Meta-DETR: Image-Level Few-Shot Detection with Inter-Class Correlation Exploitation
Gongjie Zhang, Zhipeng Luo, Kaiwen Cui +2
cs.CVcs.AIcs.LGarXiv:2208.00219v12022VakyArth: Evaluating Pragmatic Competence in LLMs across Indic Languages
Usneek Singh, Poorvaja Veera Balaji Kumar, Parth Nanda +4
cs.CLcs.AIarXiv:2609.01788v12026Semi-Supervised Virtual Staining via Morphology Preservation and Histopathological Realism Constraints
Baoshun Wang, Weiping Lin, Linwu Wang +3
cs.CVcs.AIarXiv:2609.00984v12026Deep Neural Networks for Multiple Speaker Detection and Localization
Weipeng He, Petr Motlicek, Jean-Marc Odobez
cs.SDcs.AIcs.MMarXiv:1711.11565v32017NE-R1: Enhancing Named Entity Recognition Model via Reinforcement Learning
Meixuan Chen, Hehan Li, Ruizhi Zhao +8
cs.CLcs.AIarXiv:2609.02366v12026SAM 3D: 3Dfy Anything in Images
SAM 3D Team, Xingyu Chen, Fu-Jen Chu +20
cs.CVcs.AIarXiv:2511.16624v22025Learn from Whoever Is Right: Answer-Verified Multi-Teacher Distillation for Multi-Domain LLMs
Xixiang He, Xingming Li, Baiqi Wu +4
cs.LGcs.AIarXiv:2609.02548v12026Native Sparse Attention: Hardware-Aligned and Natively Trainable Sparse Attention
Jingyang Yuan, Huazuo Gao, Damai Dai +12
cs.CLcs.AIcs.LGarXiv:2502.11089v22025LHRS-Bot: Empowering Remote Sensing with VGI-Enhanced Large Multimodal Language Model
Dilxat Muhtar, Zhenshi Li, Feng Gu +2
cs.CVcs.AIcs.LGarXiv:2402.02544v42024Multi-class Classification without Multi-class Labels
Yen-Chang Hsu, Zhaoyang Lv, Joel Schlosser +2
cs.LGcs.AIcs.CVarXiv:1901.00544v12019Future progress in artificial intelligence: A survey of expert opinion
Vincent C. Müller, Nick Bostrom
cs.CYcs.AIarXiv:2508.11681v12025Fine-Grained Anomaly Perception in Wild UGC-Enhanced Images: A Comprehensive Dataset and Difference-Fusion Framework
Yan Zhong, Gefei Chen, Qiufang Ma +4
cs.CVcs.AIarXiv:2609.02529v12026