Artificial Intelligence

Papers filed under cs.AI on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.

Search paper metadata (including unsummarized papers)

6,661 to 6,720 of 15,403

  1. Pixel Reasoner: Incentivizing Pixel-Space Reasoning with Curiosity-Driven Reinforcement Learning

    Haozhe Wang, Alex Su, Weiming Ren +2

    cs.CVcs.AIcs.CLarXiv:2505.15966v32025
  2. Learning a Recurrent Visual Representation for Image Caption Generation

    Xinlei Chen, C. Lawrence Zitnick

    cs.CVcs.AIcs.CLarXiv:1411.5654v12014
  3. VL-Rethinker: Incentivizing Self-Reflection of Vision-Language Models with Reinforcement Learning

    Haozhe Wang, Chao Qu, Zuming Huang +3

    cs.LGcs.AIarXiv:2504.08837v32025
  4. WebThinker: Empowering Large Reasoning Models with Deep Research Capability

    Xiaoxi Li, Jiajie Jin, Guanting Dong +5

    cs.CLcs.AIcs.IRarXiv:2504.21776v22025
  5. FinBen: A Holistic Financial Benchmark for Large Language Models

    Qianqian Xie, Weiguang Han, Zhengyu Chen +31

    cs.CLcs.AIcs.CEarXiv:2402.12659v22024
  6. Learning to Reason under Off-Policy Guidance

    Jianhao Yan, Yafu Li, Zican Hu +5

    cs.LGcs.AIcs.CLarXiv:2504.14945v52025
  7. LeJEPA: Provable and Scalable Self-Supervised Learning Without the Heuristics

    Randall Balestriero, Yann LeCun

    cs.LGcs.AIcs.CVarXiv:2511.08544v32025
  8. Compute and Energy Consumption Trends in Deep Learning Inference

    Radosvet Desislavov, Fernando Martínez-Plumed, José Hernández-Orallo

    cs.LGcs.AIarXiv:2109.05472v22021
  9. DRACO: Fine-Grained Credit Assignment with Dynamic Rubrics for Long-Horizon Agent Training

    Shubham Gandhi, Saurabh Goyal, Kiran Kate +1

    cs.AIcs.LGcs.SEarXiv:2609.04094v12026
  10. Towards AI-Assisted Clinical Trial Matching: Practical Considerations, Multicenter Evaluation, and Real-World Deployment

    Yin Fang, Qiao Jin, Shubo Tian +24

    cs.CLcs.AIcs.CYarXiv:2609.01202v12026
  11. Accelerating scientific discovery with Co-Scientist

    Juraj Gottweis, Wei-Hung Weng, Alexander Daryin +48

    cs.AIcs.CLcs.HCarXiv:2502.18864v22025
  12. ToolRL: Reward is All Tool Learning Needs

    Cheng Qian, Emre Can Acikgoz, Qi He +5

    cs.LGcs.AIcs.CLarXiv:2504.13958v12025
  13. WorldVLA: Towards Autoregressive Action World Model

    Jun Cen, Chaohui Yu, Hangjie Yuan +9

    cs.ROcs.AIarXiv:2506.21539v12025
  14. Humanity's Last Exam

    Long Phan, Alice Gatti, Ziwen Han +1155

    cs.LGcs.AIcs.CLarXiv:2501.14249v112025
  15. Fast and Eager k-Medoids Clustering: O(k) Runtime Improvement of the PAM, CLARA, and CLARANS Algorithms

    Erich Schubert, Peter J. Rousseeuw

    cs.LGcs.AIstat.MLarXiv:2008.05171v22020
  16. Swin Meets EfficientNet: Lightweight Architectures for GAN-Based Face Forensics

    Sejuti Basu, Ashima Sood, Vijay Kumar +1

    cs.CVcs.AIarXiv:2609.01749v12026
  17. Designing Proactive Thought Partners for Writing

    Chao Zhang, Abe Davis, Chih-Wei Chen +1

    cs.HCcs.AIcs.CLarXiv:2609.01588v12026
  18. The Illusion of Thinking: Understanding the Strengths and Limitations of Reasoning Models via the Lens of Problem Complexity

    Parshin Shojaee, Iman Mirzadeh, Keivan Alizadeh +3

    cs.AIcs.CLcs.LGarXiv:2506.06941v32025
  19. MathArena: Evaluating LLMs on Uncontaminated Math Competitions

    Mislav Balunović, Jasper Dekoninck, Ivo Petrov +2

    cs.AIcs.CLarXiv:2505.23281v32025
  20. "Help Me Help the AI": Understanding How Explainability Can Support Human-AI Interaction

    Sunnie S. Y. Kim, Elizabeth Anne Watkins, Olga Russakovsky +2

    cs.HCcs.AIcs.CVarXiv:2210.03735v22022
  21. Are We There Yet? Assessing Computer-Use Agents for Blind Users' Accessible Interaction with Desktop Applications

    Satwik Ram Kodandaram, Monalika Padma Reddy, Xiaojun Bi +3

    cs.HCcs.AIarXiv:2609.00524v12026
  22. L1: Controlling How Long A Reasoning Model Thinks With Reinforcement Learning

    Pranjal Aggarwal, Sean Welleck

    cs.CLcs.AIcs.LGarXiv:2503.04697v22025
  23. Memory in the Age of AI Agents

    Yuyang Hu, Shichun Liu, Yanwei Yue +44

    cs.CLcs.AIarXiv:2512.13564v22025
  24. DeepSeek-Prover-V2: Advancing Formal Mathematical Reasoning via Reinforcement Learning for Subgoal Decomposition

    Z. Z. Ren, Zhihong Shao, Junxiao Song +15

    cs.CLcs.AIarXiv:2504.21801v22025
  25. Agent Laboratory: Using LLM Agents as Research Assistants

    Samuel Schmidgall, Yusheng Su, Ze Wang +7

    cs.HCcs.AIcs.CLarXiv:2501.04227v22025
  26. X-VLA: Soft-Prompted Transformer as Scalable Cross-Embodiment Vision-Language-Action Model

    Jinliang Zheng, Jianxiong Li, Zhihao Wang +12

    cs.ROcs.AIcs.CVarXiv:2510.10274v12025
  27. LEAP: Likelihood Elicitation and Aggregation for LLM-based Probabilistic Forecasting

    Yufei Chen, Yiran Zhao, Xiaogang Xu +3

    cs.AIarXiv:2609.01337v12026
  28. GLM-4.5V and GLM-4.1V-Thinking: Towards Versatile Multimodal Reasoning with Scalable Reinforcement Learning

    GLM-V Team, :, Wenyi Hong +91

    cs.CVcs.AIcs.LGarXiv:2507.01006v62025
  29. Multi-Agent Collaboration Mechanisms: A Survey of LLMs

    Khanh-Tung Tran, Dung Dao, Minh-Duong Nguyen +3

    cs.AIarXiv:2501.06322v12025
  30. The Regretful Agent: Heuristic-Aided Navigation through Progress Estimation

    Chih-Yao Ma, Zuxuan Wu, Ghassan AlRegib +2

    cs.AIcs.CVcs.ROarXiv:1903.01602v12019
  31. QCell: Recombining and Aligning Cell Queries for Overlapping Instance Segmentation

    Yaroslav Prytula, Anton Popov, Dmytro Fishman

    cs.CVcs.AIcs.LGarXiv:2608.29253v12026
  32. Muon is Scalable for LLM Training

    Jingyuan Liu, Jianlin Su, Xingcheng Yao +25

    cs.LGcs.AIcs.CLarXiv:2502.16982v12025
  33. The Lessons of Developing Process Reward Models in Mathematical Reasoning

    Zhenru Zhang, Chujie Zheng, Yangzhen Wu +6

    cs.CLcs.AIcs.LGarXiv:2501.07301v22025
  34. Search-o1: Agentic Search-Enhanced Large Reasoning Models

    Xiaoxi Li, Guanting Dong, Jiajie Jin +5

    cs.AIcs.CLcs.IRarXiv:2501.05366v12025
  35. Deep Probabilistic Programming

    Dustin Tran, Matthew D. Hoffman, Rif A. Saurous +3

    stat.MLcs.AIcs.LGarXiv:1701.03757v22017
  36. Contrastive Learning for Label-Efficient Semantic Segmentation

    Xiangyun Zhao, Raviteja Vemulapalli, Philip Mansfield +4

    cs.CVcs.AIcs.LGarXiv:2012.06985v42020
  37. Agentic Context Engineering: Evolving Contexts for Self-Improving Language Models

    Qizheng Zhang, Changran Hu, Shubhangi Upasani +10

    cs.LGcs.AIcs.CLarXiv:2510.04618v32025
  38. Block Diffusion: Interpolating Between Autoregressive and Diffusion Language Models

    Marianne Arriola, Aaron Gokaslan, Justin T. Chiu +5

    cs.LGcs.AIarXiv:2503.09573v32025
  39. Who Should I Trust: AI or Myself? Leveraging Human and AI Correctness Likelihood to Promote Appropriate Trust in AI-Assisted Decision-Making

    Shuai Ma, Ying Lei, Xinru Wang +4

    cs.HCcs.AIcs.LGarXiv:2301.05809v12023
  40. Process Reinforcement through Implicit Rewards

    Ganqu Cui, Lifan Yuan, Zefan Wang +22

    cs.LGcs.AIcs.CLarXiv:2502.01456v22025
  41. Modeling Human Motion with Quaternion-based Neural Networks

    Dario Pavllo, Christoph Feichtenhofer, Michael Auli +1

    cs.CVcs.AIcs.ROarXiv:1901.07677v22019
  42. From Reweighting to Rewriting: Unlocking the Intervention Effects of Influential Samples in Training Data Attribution

    Yuzhang Luo, Chenpeng Wang, Jianhui Chen +1

    cs.CLcs.AIcs.LGarXiv:2609.02771v12026
  43. Coverage, Not Targeting: A Structural Regime in Multi-Turn Agent Credit Assignment

    Chenyu Zhou, Qiliang Jiang, Shuning Wu +1

    cs.LGcs.AIarXiv:2609.02417v12026
  44. Political Compass or Spinning Arrow? Towards More Meaningful Evaluations for Values and Opinions in Large Language Models

    Paul Röttger, Valentin Hofmann, Valentina Pyatkin +4

    cs.CLcs.AIarXiv:2402.16786v22024
  45. DeepSeek-Prover-V1.5: Harnessing Proof Assistant Feedback for Reinforcement Learning and Monte-Carlo Tree Search

    Huajian Xin, Z. Z. Ren, Junxiao Song +14

    cs.CLcs.AIcs.LGarXiv:2408.08152v12024
  46. A Network Science Perspective on Evaluating Deep Graph Generative Models

    Tianrui Mao, Abele Malan, Megha Khosla +2

    cs.SIcs.AIarXiv:2609.01015v12026
  47. Dyna-Style Planning with Linear Function Approximation and Prioritized Sweeping

    Richard S. Sutton, Csaba Szepesvari, Alborz Geramifard +1

    cs.AIcs.LGeess.SYarXiv:1206.3285v12012
  48. What Is Worth Representing? Representational Empowerment for Continual Model Construction

    Fei Dai, Hanqi Zhou, Alison Gopnik +1

    cs.LGcs.AIarXiv:2609.02322v12026
  49. Meta-DETR: Image-Level Few-Shot Detection with Inter-Class Correlation Exploitation

    Gongjie Zhang, Zhipeng Luo, Kaiwen Cui +2

    cs.CVcs.AIcs.LGarXiv:2208.00219v12022
  50. VakyArth: Evaluating Pragmatic Competence in LLMs across Indic Languages

    Usneek Singh, Poorvaja Veera Balaji Kumar, Parth Nanda +4

    cs.CLcs.AIarXiv:2609.01788v12026
  51. Semi-Supervised Virtual Staining via Morphology Preservation and Histopathological Realism Constraints

    Baoshun Wang, Weiping Lin, Linwu Wang +3

    cs.CVcs.AIarXiv:2609.00984v12026
  52. Deep Neural Networks for Multiple Speaker Detection and Localization

    Weipeng He, Petr Motlicek, Jean-Marc Odobez

    cs.SDcs.AIcs.MMarXiv:1711.11565v32017
  53. NE-R1: Enhancing Named Entity Recognition Model via Reinforcement Learning

    Meixuan Chen, Hehan Li, Ruizhi Zhao +8

    cs.CLcs.AIarXiv:2609.02366v12026
  54. SAM 3D: 3Dfy Anything in Images

    SAM 3D Team, Xingyu Chen, Fu-Jen Chu +20

    cs.CVcs.AIarXiv:2511.16624v22025
  55. Learn from Whoever Is Right: Answer-Verified Multi-Teacher Distillation for Multi-Domain LLMs

    Xixiang He, Xingming Li, Baiqi Wu +4

    cs.LGcs.AIarXiv:2609.02548v12026
  56. Native Sparse Attention: Hardware-Aligned and Natively Trainable Sparse Attention

    Jingyang Yuan, Huazuo Gao, Damai Dai +12

    cs.CLcs.AIcs.LGarXiv:2502.11089v22025
  57. LHRS-Bot: Empowering Remote Sensing with VGI-Enhanced Large Multimodal Language Model

    Dilxat Muhtar, Zhenshi Li, Feng Gu +2

    cs.CVcs.AIcs.LGarXiv:2402.02544v42024
  58. Multi-class Classification without Multi-class Labels

    Yen-Chang Hsu, Zhaoyang Lv, Joel Schlosser +2

    cs.LGcs.AIcs.CVarXiv:1901.00544v12019
  59. Future progress in artificial intelligence: A survey of expert opinion

    Vincent C. Müller, Nick Bostrom

    cs.CYcs.AIarXiv:2508.11681v12025
  60. Fine-Grained Anomaly Perception in Wild UGC-Enhanced Images: A Comprehensive Dataset and Difference-Fusion Framework

    Yan Zhong, Gefei Chen, Qiufang Ma +4

    cs.CVcs.AIarXiv:2609.02529v12026