Artificial Intelligence
Papers filed under cs.AI on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.
Search paper metadata (including unsummarized papers)
5,341 to 5,400 of 15,437
Revisiting Face Recognition for Monozygotic Twins: The Celeb Twins Test Set
Michael Zang, Haiyu Wu, Mrinal Sharma +1
cs.CVcs.AIarXiv:2609.01141v12026Can LLMs Discover Scientific Laws in Real and Parallel Worlds?
Yiming Huang, Ziche Liu, Zhuohang Wu +11
cs.AIcs.LGarXiv:2609.01552v12026A Checklist to assess the energy and carbon impacts of ML/AI applications in Earth System Modeling
Filippo Dainelli, Amirpasha Mozaffari, Marina Castaño +5
physics.ao-phcs.AIcs.LGarXiv:2609.00847v12026Space Generative AI with Solar Energy Harvesting
Jierui Zhang, Jianhao Huang, Zhanwei Wang +1
cs.AIcs.NIeess.SParXiv:2609.01062v12026Cameras as Relative Positional Encoding
Ruilong Li, Brent Yi, Junchen Liu +3
cs.CVcs.AIarXiv:2507.10496v22025Learning What Reinforcement Learning Can't: Interleaved Online Fine-Tuning for Hardest Questions
Lu Ma, Hao Liang, Meiyi Qiang +9
cs.AIcs.LGarXiv:2506.07527v32025OctoThinker: Mid-training Incentivizes Reinforcement Learning Scaling
Zengzhi Wang, Fan Zhou, Xuefeng Li +1
cs.CLcs.AIcs.LGarXiv:2506.20512v12025Expected Attention: KV Cache Compression by Estimating Attention from Future Queries Distribution
Alessio Devoto, Maximilian Jeblick, Simon Jégou
cs.AIcs.CLarXiv:2510.00636v12025Active Learning for Regression Using Greedy Sampling
Dongrui Wu, Chin-Teng Lin, Jian Huang
cs.LGcs.AIstat.MLarXiv:1808.04245v12018VITA: Towards Open-Source Interactive Omni Multimodal LLM
Chaoyou Fu, Haojia Lin, Zuwei Long +16
cs.CVcs.AIcs.CLarXiv:2408.05211v32024Overview of the TREC 2021 deep learning track
Nick Craswell, Bhaskar Mitra, Emine Yilmaz +2
cs.IRcs.AIcs.CLarXiv:2507.08191v12025LlamaFirewall: An open source guardrail system for building secure AI agents
Sahana Chennabasappa, Cyrus Nikolaidis, Daniel Song +16
cs.CRcs.AIarXiv:2505.03574v12025LLM Social Simulations Are a Promising Research Method
Jacy Reese Anthis, Ryan Liu, Sean M. Richardson +5
cs.HCcs.AIcs.CLarXiv:2504.02234v22025MedReason: Eliciting Factual Medical Reasoning Steps in LLMs via Knowledge Graphs
Juncheng Wu, Wenlong Deng, Xingxuan Li +12
cs.CLcs.AIarXiv:2504.00993v22025Multimodal Co-learning: Challenges, Applications with Datasets, Recent Advances and Future Directions
Anil Rahate, Rahee Walambe, Sheela Ramanna +1
cs.LGcs.AIarXiv:2107.13782v32021CoVer: Conflict-Aware Claim Verification
Shuning Zhang, Dai Shi, Bohao Chu +7
cs.AIarXiv:2609.00508v12026PRMBench: A Fine-grained and Challenging Benchmark for Process-Level Reward Models
Mingyang Song, Zhaochen Su, Xiaoye Qu +2
cs.CLcs.AIcs.LGarXiv:2501.03124v52025PHYRE: A New Benchmark for Physical Reasoning
Anton Bakhtin, Laurens van der Maaten, Justin Johnson +2
cs.LGcs.AIstat.MLarXiv:1908.05656v12019Interactive Post-Training for Vision-Language-Action Models
Shuhan Tan, Kairan Dou, Yue Zhao +1
cs.LGcs.AIcs.CVarXiv:2505.17016v12025Denoising Diffusion Bridge Models
Linqi Zhou, Aaron Lou, Samar Khanna +1
cs.CVcs.AIarXiv:2309.16948v32023UniIR: Training and Benchmarking Universal Multimodal Information Retrievers
Cong Wei, Yang Chen, Haonan Chen +5
cs.CVcs.AIcs.CLarXiv:2311.17136v12023Stride-k Subsampling: Train-Free Audio Token Reduction for Whisper
Chanhee Cho, Junhyuk Choi, Bugeun Kim
cs.SDcs.AIarXiv:2608.30927v12026Are You Thinking What I am Thinking? : Examining Conceptual Separation in Neural Architectures
Jaee Ponde, Roshni Agarwal, Subhashis Banerjee
cs.LGcs.AIarXiv:2609.00764v12026Medical Hallucinations in Foundation Models and Their Impact on Healthcare
Yubin Kim, Hyewon Jeong, Shan Chen +24
cs.CLcs.AIcs.CYarXiv:2503.05777v22025MusGU+: Toward a Musician-Centered Evaluation Framework and Discovery Tool for Generative Music AI
Laura Ibáñez-Martínez, Roser Batlle-Roca, Xavier Serra +1
cs.SDcs.AIcs.CYarXiv:2608.30940v12026The Geometry of Ignorance: LLMs Know When to Temper Bayesian Priors
Toni J. B. Liu, Jiajun Bao, Yizhou Liu +4
cs.LGcs.AIcs.CLarXiv:2609.02959v12026AA-CLIP: Enhancing Zero-shot Anomaly Detection via Anomaly-Aware CLIP
Wenxin Ma, Xu Zhang, Qingsong Yao +6
cs.CVcs.AIarXiv:2503.06661v12025Advances and Challenges in Foundation Agents: From Brain-Inspired Intelligence to Evolutionary, Collaborative, and Safe Systems
Bang Liu, Xinfeng Li, Jiayi Zhang +45
cs.AIarXiv:2504.01990v22025Recursive Language Models
Alex L. Zhang, Tim Kraska, Omar Khattab
cs.AIcs.CLarXiv:2512.24601v32025LLMs Can Easily Learn to Reason from Demonstrations Structure, not content, is what matters!
Dacheng Li, Shiyi Cao, Tyler Griggs +9
cs.AIarXiv:2502.07374v22025HAMSTER: Hierarchical Action Models For Open-World Robot Manipulation
Yi Li, Yuquan Deng, Jesse Zhang +9
cs.ROcs.AIcs.CVarXiv:2502.05485v42025Multimodal Recommender Systems: A Survey
Qidong Liu, Jiaxi Hu, Yutian Xiao +5
cs.IRcs.AIarXiv:2302.03883v22023Can LLMs Replace Human Evaluators? An Empirical Study of LLM-as-a-Judge in Software Engineering
Ruiqi Wang, Jiyu Guo, Cuiyun Gao +3
cs.SEcs.AIarXiv:2502.06193v32025LLMs Accelerate Annotation for Medical Information Extraction
Akshay Goel, Almog Gueta, Omry Gilon +10
cs.CLcs.AIcs.LGarXiv:2312.02296v12023A Reinforcement Learning Approach to Weaning of Mechanical Ventilation in Intensive Care Units
Niranjani Prasad, Li-Fang Cheng, Corey Chivers +2
cs.AIarXiv:1704.06300v12017Monet: Reasoning in Latent Visual Space Beyond Images and Language
Qixun Wang, Yang Shi, Yifei Wang +5
cs.CVcs.AIarXiv:2511.21395v22025Llasa: Scaling Train-Time and Inference-Time Compute for Llama-based Speech Synthesis
Zhen Ye, Xinfa Zhu, Chi-Min Chan +17
eess.AScs.AIcs.CLarXiv:2502.04128v22025AMO: Adaptive Motion Optimization for Hyper-Dexterous Humanoid Whole-Body Control
Jialong Li, Xuxin Cheng, Tianshu Huang +3
cs.ROcs.AIcs.LGarXiv:2505.03738v12025Soft Thinking: Unlocking the Reasoning Potential of LLMs in Continuous Concept Space
Zhen Zhang, Xuehai He, Weixiang Yan +5
cs.CLcs.AIarXiv:2505.15778v12025MCP Safety Audit: LLMs with the Model Context Protocol Allow Major Security Exploits
Brandon Radosevich, John Halloran
cs.CRcs.AIcs.LGarXiv:2504.03767v22025HalluLens: LLM Hallucination Benchmark
Yejin Bang, Ziwei Ji, Alan Schelten +5
cs.CLcs.AIarXiv:2504.17550v12025Deep Research Agents: A Systematic Examination And Roadmap
Yuxuan Huang, Yihang Chen, Haozheng Zhang +10
cs.AIarXiv:2506.18096v22025Personalized HeartSteps: A Reinforcement Learning Algorithm for Optimizing Physical Activity
Peng Liao, Kristjan Greenewald, Predrag Klasnja +1
cs.LGcs.AIarXiv:1909.03539v12019HealthGPT: A Medical Large Vision-Language Model for Unifying Comprehension and Generation via Heterogeneous Knowledge Adaptation
Tianwei Lin, Wenqiao Zhang, Sijing Li +12
cs.CVcs.AIarXiv:2502.09838v32025HOMIE: Humanoid Loco-Manipulation with Isomorphic Exoskeleton Cockpit
Qingwei Ben, Feiyu Jia, Jia Zeng +3
cs.ROcs.AIcs.HCarXiv:2502.13013v22025Anomaly Detection of Time Series with Smoothness-Inducing Sequential Variational Auto-Encoder
Longyuan Li, Junchi Yan, Haiyang Wang +1
cs.LGcs.AIarXiv:2102.01331v12021Reasoning with Sampling: Your Base Model is Smarter Than You Think
Aayush Karan, Yilun Du
cs.LGcs.AIcs.CLarXiv:2510.14901v12025frb100-40 After Two Decades: An Optimality Certificate and a Preregistered Search Study
Onur Uğurlu
cs.DMcs.AIarXiv:2609.02804v12026Skill-SD: Skill-Conditioned Self-Distillation for Multi-turn LLM Agents
Hao Wang, Guozhi Wang, Han Xiao +8
cs.LGcs.AIcs.CLarXiv:2604.10674v12026Second-order Non-local Attention Networks for Person Re-identification
Bryan, Xia, Yuan Gong +2
cs.CVcs.AIcs.LGarXiv:1909.00295v12019Modeling What Changes: Sparse, Residual World Models for Object-Centric Manipulation
Param Thakkar, Parsika Paresh Shah, Manisha Sushant Gote
cs.ROcs.AIarXiv:2609.02046v12026Technology Readiness Levels for AI & ML
Alexander Lavin, Gregory Renard
cs.SEcs.AIcs.LGarXiv:2006.12497v32020Measurement-Driven Sub-Network Selection for On-Premise Retrieval-Augmented Factory Agents
Vasileios Rizeakos, Georgios Paisios, Alexandros Machairas +2
cs.AIarXiv:2609.02760v12026More Agents Is All You Need
Junyou Li, Qin Zhang, Yangbin Yu +2
cs.CLcs.AIcs.LGarXiv:2402.05120v22024SEAL: Reinforcing Global Safety in Mixture-of-Experts through Shared Expert ALignment
Qingyu Meng, Yiwei Zha, Jiahuan Pei +3
cs.LGcs.AIcs.CRarXiv:2609.02293v12026Seed1.8 Model Card: Towards Generalized Real-World Agency
Bytedance Seed
cs.AIarXiv:2603.20633v32026Untangling the Mechanisms of Misleading Context in Medical Question Answering
Robin Linzmayer, Noémie Elhadad
cs.CLcs.AIcs.LGarXiv:2609.02754v12026Towards One-for-All Robustness Across a Continuum of Threat Levels
Zhichao Hou, Xiaorui Liu
cs.LGcs.AIarXiv:2609.02440v12026Audio-Reasoner: Improving Reasoning Capability in Large Audio Language Models
Zhifei Xie, Mingbao Lin, Zihang Liu +3
cs.SDcs.AIcs.CLarXiv:2503.02318v22025CALIP: Zero-Shot Enhancement of CLIP with Parameter-free Attention
Ziyu Guo, Renrui Zhang, Longtian Qiu +4
cs.CVcs.AIcs.MMarXiv:2209.14169v22022