Artificial Intelligence
Papers filed under cs.AI on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.
Search paper metadata (including unsummarized papers)
6,601 to 6,660 of 15,328
Swin Meets EfficientNet: Lightweight Architectures for GAN-Based Face Forensics
Sejuti Basu, Ashima Sood, Vijay Kumar +1
cs.CVcs.AIarXiv:2609.01749v12026Designing Proactive Thought Partners for Writing
Chao Zhang, Abe Davis, Chih-Wei Chen +1
cs.HCcs.AIcs.CLarXiv:2609.01588v12026The Illusion of Thinking: Understanding the Strengths and Limitations of Reasoning Models via the Lens of Problem Complexity
Parshin Shojaee, Iman Mirzadeh, Keivan Alizadeh +3
cs.AIcs.CLcs.LGarXiv:2506.06941v32025MathArena: Evaluating LLMs on Uncontaminated Math Competitions
Mislav Balunović, Jasper Dekoninck, Ivo Petrov +2
cs.AIcs.CLarXiv:2505.23281v32025"Help Me Help the AI": Understanding How Explainability Can Support Human-AI Interaction
Sunnie S. Y. Kim, Elizabeth Anne Watkins, Olga Russakovsky +2
cs.HCcs.AIcs.CVarXiv:2210.03735v22022Are We There Yet? Assessing Computer-Use Agents for Blind Users' Accessible Interaction with Desktop Applications
Satwik Ram Kodandaram, Monalika Padma Reddy, Xiaojun Bi +3
cs.HCcs.AIarXiv:2609.00524v12026L1: Controlling How Long A Reasoning Model Thinks With Reinforcement Learning
Pranjal Aggarwal, Sean Welleck
cs.CLcs.AIcs.LGarXiv:2503.04697v22025Memory in the Age of AI Agents
Yuyang Hu, Shichun Liu, Yanwei Yue +44
cs.CLcs.AIarXiv:2512.13564v22025DeepSeek-Prover-V2: Advancing Formal Mathematical Reasoning via Reinforcement Learning for Subgoal Decomposition
Z. Z. Ren, Zhihong Shao, Junxiao Song +15
cs.CLcs.AIarXiv:2504.21801v22025Agent Laboratory: Using LLM Agents as Research Assistants
Samuel Schmidgall, Yusheng Su, Ze Wang +7
cs.HCcs.AIcs.CLarXiv:2501.04227v22025X-VLA: Soft-Prompted Transformer as Scalable Cross-Embodiment Vision-Language-Action Model
Jinliang Zheng, Jianxiong Li, Zhihao Wang +12
cs.ROcs.AIcs.CVarXiv:2510.10274v12025LEAP: Likelihood Elicitation and Aggregation for LLM-based Probabilistic Forecasting
Yufei Chen, Yiran Zhao, Xiaogang Xu +3
cs.AIarXiv:2609.01337v12026GLM-4.5V and GLM-4.1V-Thinking: Towards Versatile Multimodal Reasoning with Scalable Reinforcement Learning
GLM-V Team, :, Wenyi Hong +91
cs.CVcs.AIcs.LGarXiv:2507.01006v62025Multi-Agent Collaboration Mechanisms: A Survey of LLMs
Khanh-Tung Tran, Dung Dao, Minh-Duong Nguyen +3
cs.AIarXiv:2501.06322v12025The Regretful Agent: Heuristic-Aided Navigation through Progress Estimation
Chih-Yao Ma, Zuxuan Wu, Ghassan AlRegib +2
cs.AIcs.CVcs.ROarXiv:1903.01602v12019QCell: Recombining and Aligning Cell Queries for Overlapping Instance Segmentation
Yaroslav Prytula, Anton Popov, Dmytro Fishman
cs.CVcs.AIcs.LGarXiv:2608.29253v12026Muon is Scalable for LLM Training
Jingyuan Liu, Jianlin Su, Xingcheng Yao +25
cs.LGcs.AIcs.CLarXiv:2502.16982v12025The Lessons of Developing Process Reward Models in Mathematical Reasoning
Zhenru Zhang, Chujie Zheng, Yangzhen Wu +6
cs.CLcs.AIcs.LGarXiv:2501.07301v22025Search-o1: Agentic Search-Enhanced Large Reasoning Models
Xiaoxi Li, Guanting Dong, Jiajie Jin +5
cs.AIcs.CLcs.IRarXiv:2501.05366v12025Deep Probabilistic Programming
Dustin Tran, Matthew D. Hoffman, Rif A. Saurous +3
stat.MLcs.AIcs.LGarXiv:1701.03757v22017Contrastive Learning for Label-Efficient Semantic Segmentation
Xiangyun Zhao, Raviteja Vemulapalli, Philip Mansfield +4
cs.CVcs.AIcs.LGarXiv:2012.06985v42020Agentic Context Engineering: Evolving Contexts for Self-Improving Language Models
Qizheng Zhang, Changran Hu, Shubhangi Upasani +10
cs.LGcs.AIcs.CLarXiv:2510.04618v32025Block Diffusion: Interpolating Between Autoregressive and Diffusion Language Models
Marianne Arriola, Aaron Gokaslan, Justin T. Chiu +5
cs.LGcs.AIarXiv:2503.09573v32025Who Should I Trust: AI or Myself? Leveraging Human and AI Correctness Likelihood to Promote Appropriate Trust in AI-Assisted Decision-Making
Shuai Ma, Ying Lei, Xinru Wang +4
cs.HCcs.AIcs.LGarXiv:2301.05809v12023Process Reinforcement through Implicit Rewards
Ganqu Cui, Lifan Yuan, Zefan Wang +22
cs.LGcs.AIcs.CLarXiv:2502.01456v22025Modeling Human Motion with Quaternion-based Neural Networks
Dario Pavllo, Christoph Feichtenhofer, Michael Auli +1
cs.CVcs.AIcs.ROarXiv:1901.07677v22019From Reweighting to Rewriting: Unlocking the Intervention Effects of Influential Samples in Training Data Attribution
Yuzhang Luo, Chenpeng Wang, Jianhui Chen +1
cs.CLcs.AIcs.LGarXiv:2609.02771v12026Coverage, Not Targeting: A Structural Regime in Multi-Turn Agent Credit Assignment
Chenyu Zhou, Qiliang Jiang, Shuning Wu +1
cs.LGcs.AIarXiv:2609.02417v12026Political Compass or Spinning Arrow? Towards More Meaningful Evaluations for Values and Opinions in Large Language Models
Paul Röttger, Valentin Hofmann, Valentina Pyatkin +4
cs.CLcs.AIarXiv:2402.16786v22024DeepSeek-Prover-V1.5: Harnessing Proof Assistant Feedback for Reinforcement Learning and Monte-Carlo Tree Search
Huajian Xin, Z. Z. Ren, Junxiao Song +14
cs.CLcs.AIcs.LGarXiv:2408.08152v12024A Network Science Perspective on Evaluating Deep Graph Generative Models
Tianrui Mao, Abele Malan, Megha Khosla +2
cs.SIcs.AIarXiv:2609.01015v12026Dyna-Style Planning with Linear Function Approximation and Prioritized Sweeping
Richard S. Sutton, Csaba Szepesvari, Alborz Geramifard +1
cs.AIcs.LGeess.SYarXiv:1206.3285v12012What Is Worth Representing? Representational Empowerment for Continual Model Construction
Fei Dai, Hanqi Zhou, Alison Gopnik +1
cs.LGcs.AIarXiv:2609.02322v12026Meta-DETR: Image-Level Few-Shot Detection with Inter-Class Correlation Exploitation
Gongjie Zhang, Zhipeng Luo, Kaiwen Cui +2
cs.CVcs.AIcs.LGarXiv:2208.00219v12022VakyArth: Evaluating Pragmatic Competence in LLMs across Indic Languages
Usneek Singh, Poorvaja Veera Balaji Kumar, Parth Nanda +4
cs.CLcs.AIarXiv:2609.01788v12026Semi-Supervised Virtual Staining via Morphology Preservation and Histopathological Realism Constraints
Baoshun Wang, Weiping Lin, Linwu Wang +3
cs.CVcs.AIarXiv:2609.00984v12026Deep Neural Networks for Multiple Speaker Detection and Localization
Weipeng He, Petr Motlicek, Jean-Marc Odobez
cs.SDcs.AIcs.MMarXiv:1711.11565v32017NE-R1: Enhancing Named Entity Recognition Model via Reinforcement Learning
Meixuan Chen, Hehan Li, Ruizhi Zhao +8
cs.CLcs.AIarXiv:2609.02366v12026SAM 3D: 3Dfy Anything in Images
SAM 3D Team, Xingyu Chen, Fu-Jen Chu +20
cs.CVcs.AIarXiv:2511.16624v22025Learn from Whoever Is Right: Answer-Verified Multi-Teacher Distillation for Multi-Domain LLMs
Xixiang He, Xingming Li, Baiqi Wu +4
cs.LGcs.AIarXiv:2609.02548v12026Native Sparse Attention: Hardware-Aligned and Natively Trainable Sparse Attention
Jingyang Yuan, Huazuo Gao, Damai Dai +12
cs.CLcs.AIcs.LGarXiv:2502.11089v22025LHRS-Bot: Empowering Remote Sensing with VGI-Enhanced Large Multimodal Language Model
Dilxat Muhtar, Zhenshi Li, Feng Gu +2
cs.CVcs.AIcs.LGarXiv:2402.02544v42024Multi-class Classification without Multi-class Labels
Yen-Chang Hsu, Zhaoyang Lv, Joel Schlosser +2
cs.LGcs.AIcs.CVarXiv:1901.00544v12019Future progress in artificial intelligence: A survey of expert opinion
Vincent C. Müller, Nick Bostrom
cs.CYcs.AIarXiv:2508.11681v12025Fine-Grained Anomaly Perception in Wild UGC-Enhanced Images: A Comprehensive Dataset and Difference-Fusion Framework
Yan Zhong, Gefei Chen, Qiufang Ma +4
cs.CVcs.AIarXiv:2609.02529v12026Spectral Initialization and Scheduled Graph Smoothness for Uncertain Knowledge Graph Completion
Md Abrar Jahin, Taufikur Rahman Fuad, Jay Pujara +1
cs.LGcs.AIarXiv:2609.02519v12026Automated Concatenation of Embeddings for Structured Prediction
Xinyu Wang, Yong Jiang, Nguyen Bach +4
cs.CLcs.AIcs.LGarXiv:2010.05006v42020BLIP3-o: A Family of Fully Open Unified Multimodal Models-Architecture, Training and Dataset
Jiuhai Chen, Zhiyang Xu, Xichen Pan +10
cs.CVcs.AIarXiv:2505.09568v12025ReTool: Reinforcement Learning for Strategic Tool Use in LLMs
Jiazhan Feng, Shijue Huang, Xingwei Qu +6
cs.CLcs.AIarXiv:2504.11536v22025Reasoning Models Don't Always Say What They Think
Yanda Chen, Joe Benton, Ansh Radhakrishnan +12
cs.CLcs.AIcs.LGarXiv:2505.05410v12025Kimi K2: Open Agentic Intelligence
Kimi Team, Yifan Bai, Yiping Bao +197
cs.LGcs.AIcs.CLarXiv:2507.20534v22025Getting aligned on representational alignment
Ilia Sucholutsky, Lukas Muttenthaler, Adrian Weller +30
q-bio.NCcs.AIcs.LGarXiv:2310.13018v32023Scalable Kronecker-Fisher Approximation: Efficient Hessian Analysis for Billion-Parameter Language Models Compression
Viacheslav Yusupov, Daria Cherniuk, Evgeny Frolov
cs.LGcs.AIcs.CLarXiv:2609.02451v12026AI Agents vs. Agentic AI: A Conceptual Taxonomy, Applications and Challenges
Ranjan Sapkota, Konstantinos I. Roumeliotis, Manoj Karkee
cs.AIarXiv:2505.10468v52025CoT-VLA: Visual Chain-of-Thought Reasoning for Vision-Language-Action Models
Qingqing Zhao, Yao Lu, Moo Jin Kim +12
cs.CVcs.AIcs.LGarXiv:2503.22020v12025How AI Prompts Can Teach Us About the Structure of Human Behavior
Matthew O. Jackson, Benjamin S. Manning, Yutong Xie +2
econ.THcs.AIarXiv:2608.18265v12026FLARE MCMC: Fidelity-based Layer-Adaptive REcursive proposals for MCMC
Harini Venkatesan, Christian Shelton, Ming-Feng Ho +2
cs.AIarXiv:2608.13774v12026Emergence of Invariance and Disentanglement in Deep Representations
Alessandro Achille, Stefano Soatto
cs.LGcs.AIstat.MLarXiv:1706.01350v32017Doubly Robust Policy Evaluation and Optimization
Miroslav Dudík, Dumitru Erhan, John Langford +1
stat.MEcs.AIarXiv:1503.02834v12015On the definition of a confounder
Tyler J. VanderWeele, Ilya Shpitser
stat.MEcs.AIarXiv:1304.0564v12013