Artificial Intelligence

Papers filed under cs.AI on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.

Search paper metadata (including unsummarized papers)

5,341 to 5,400 of 15,437

  1. Revisiting Face Recognition for Monozygotic Twins: The Celeb Twins Test Set

    Michael Zang, Haiyu Wu, Mrinal Sharma +1

    cs.CVcs.AIarXiv:2609.01141v12026
  2. Can LLMs Discover Scientific Laws in Real and Parallel Worlds?

    Yiming Huang, Ziche Liu, Zhuohang Wu +11

    cs.AIcs.LGarXiv:2609.01552v12026
  3. A Checklist to assess the energy and carbon impacts of ML/AI applications in Earth System Modeling

    Filippo Dainelli, Amirpasha Mozaffari, Marina Castaño +5

    physics.ao-phcs.AIcs.LGarXiv:2609.00847v12026
  4. Space Generative AI with Solar Energy Harvesting

    Jierui Zhang, Jianhao Huang, Zhanwei Wang +1

    cs.AIcs.NIeess.SParXiv:2609.01062v12026
  5. Cameras as Relative Positional Encoding

    Ruilong Li, Brent Yi, Junchen Liu +3

    cs.CVcs.AIarXiv:2507.10496v22025
  6. Learning What Reinforcement Learning Can't: Interleaved Online Fine-Tuning for Hardest Questions

    Lu Ma, Hao Liang, Meiyi Qiang +9

    cs.AIcs.LGarXiv:2506.07527v32025
  7. OctoThinker: Mid-training Incentivizes Reinforcement Learning Scaling

    Zengzhi Wang, Fan Zhou, Xuefeng Li +1

    cs.CLcs.AIcs.LGarXiv:2506.20512v12025
  8. Expected Attention: KV Cache Compression by Estimating Attention from Future Queries Distribution

    Alessio Devoto, Maximilian Jeblick, Simon Jégou

    cs.AIcs.CLarXiv:2510.00636v12025
  9. Active Learning for Regression Using Greedy Sampling

    Dongrui Wu, Chin-Teng Lin, Jian Huang

    cs.LGcs.AIstat.MLarXiv:1808.04245v12018
  10. VITA: Towards Open-Source Interactive Omni Multimodal LLM

    Chaoyou Fu, Haojia Lin, Zuwei Long +16

    cs.CVcs.AIcs.CLarXiv:2408.05211v32024
  11. Overview of the TREC 2021 deep learning track

    Nick Craswell, Bhaskar Mitra, Emine Yilmaz +2

    cs.IRcs.AIcs.CLarXiv:2507.08191v12025
  12. LlamaFirewall: An open source guardrail system for building secure AI agents

    Sahana Chennabasappa, Cyrus Nikolaidis, Daniel Song +16

    cs.CRcs.AIarXiv:2505.03574v12025
  13. LLM Social Simulations Are a Promising Research Method

    Jacy Reese Anthis, Ryan Liu, Sean M. Richardson +5

    cs.HCcs.AIcs.CLarXiv:2504.02234v22025
  14. MedReason: Eliciting Factual Medical Reasoning Steps in LLMs via Knowledge Graphs

    Juncheng Wu, Wenlong Deng, Xingxuan Li +12

    cs.CLcs.AIarXiv:2504.00993v22025
  15. Multimodal Co-learning: Challenges, Applications with Datasets, Recent Advances and Future Directions

    Anil Rahate, Rahee Walambe, Sheela Ramanna +1

    cs.LGcs.AIarXiv:2107.13782v32021
  16. CoVer: Conflict-Aware Claim Verification

    Shuning Zhang, Dai Shi, Bohao Chu +7

    cs.AIarXiv:2609.00508v12026
  17. PRMBench: A Fine-grained and Challenging Benchmark for Process-Level Reward Models

    Mingyang Song, Zhaochen Su, Xiaoye Qu +2

    cs.CLcs.AIcs.LGarXiv:2501.03124v52025
  18. PHYRE: A New Benchmark for Physical Reasoning

    Anton Bakhtin, Laurens van der Maaten, Justin Johnson +2

    cs.LGcs.AIstat.MLarXiv:1908.05656v12019
  19. Interactive Post-Training for Vision-Language-Action Models

    Shuhan Tan, Kairan Dou, Yue Zhao +1

    cs.LGcs.AIcs.CVarXiv:2505.17016v12025
  20. Denoising Diffusion Bridge Models

    Linqi Zhou, Aaron Lou, Samar Khanna +1

    cs.CVcs.AIarXiv:2309.16948v32023
  21. UniIR: Training and Benchmarking Universal Multimodal Information Retrievers

    Cong Wei, Yang Chen, Haonan Chen +5

    cs.CVcs.AIcs.CLarXiv:2311.17136v12023
  22. Stride-k Subsampling: Train-Free Audio Token Reduction for Whisper

    Chanhee Cho, Junhyuk Choi, Bugeun Kim

    cs.SDcs.AIarXiv:2608.30927v12026
  23. Are You Thinking What I am Thinking? : Examining Conceptual Separation in Neural Architectures

    Jaee Ponde, Roshni Agarwal, Subhashis Banerjee

    cs.LGcs.AIarXiv:2609.00764v12026
  24. Medical Hallucinations in Foundation Models and Their Impact on Healthcare

    Yubin Kim, Hyewon Jeong, Shan Chen +24

    cs.CLcs.AIcs.CYarXiv:2503.05777v22025
  25. MusGU+: Toward a Musician-Centered Evaluation Framework and Discovery Tool for Generative Music AI

    Laura Ibáñez-Martínez, Roser Batlle-Roca, Xavier Serra +1

    cs.SDcs.AIcs.CYarXiv:2608.30940v12026
  26. The Geometry of Ignorance: LLMs Know When to Temper Bayesian Priors

    Toni J. B. Liu, Jiajun Bao, Yizhou Liu +4

    cs.LGcs.AIcs.CLarXiv:2609.02959v12026
  27. AA-CLIP: Enhancing Zero-shot Anomaly Detection via Anomaly-Aware CLIP

    Wenxin Ma, Xu Zhang, Qingsong Yao +6

    cs.CVcs.AIarXiv:2503.06661v12025
  28. Advances and Challenges in Foundation Agents: From Brain-Inspired Intelligence to Evolutionary, Collaborative, and Safe Systems

    Bang Liu, Xinfeng Li, Jiayi Zhang +45

    cs.AIarXiv:2504.01990v22025
  29. Recursive Language Models

    Alex L. Zhang, Tim Kraska, Omar Khattab

    cs.AIcs.CLarXiv:2512.24601v32025
  30. LLMs Can Easily Learn to Reason from Demonstrations Structure, not content, is what matters!

    Dacheng Li, Shiyi Cao, Tyler Griggs +9

    cs.AIarXiv:2502.07374v22025
  31. HAMSTER: Hierarchical Action Models For Open-World Robot Manipulation

    Yi Li, Yuquan Deng, Jesse Zhang +9

    cs.ROcs.AIcs.CVarXiv:2502.05485v42025
  32. Multimodal Recommender Systems: A Survey

    Qidong Liu, Jiaxi Hu, Yutian Xiao +5

    cs.IRcs.AIarXiv:2302.03883v22023
  33. Can LLMs Replace Human Evaluators? An Empirical Study of LLM-as-a-Judge in Software Engineering

    Ruiqi Wang, Jiyu Guo, Cuiyun Gao +3

    cs.SEcs.AIarXiv:2502.06193v32025
  34. LLMs Accelerate Annotation for Medical Information Extraction

    Akshay Goel, Almog Gueta, Omry Gilon +10

    cs.CLcs.AIcs.LGarXiv:2312.02296v12023
  35. A Reinforcement Learning Approach to Weaning of Mechanical Ventilation in Intensive Care Units

    Niranjani Prasad, Li-Fang Cheng, Corey Chivers +2

    cs.AIarXiv:1704.06300v12017
  36. Monet: Reasoning in Latent Visual Space Beyond Images and Language

    Qixun Wang, Yang Shi, Yifei Wang +5

    cs.CVcs.AIarXiv:2511.21395v22025
  37. Llasa: Scaling Train-Time and Inference-Time Compute for Llama-based Speech Synthesis

    Zhen Ye, Xinfa Zhu, Chi-Min Chan +17

    eess.AScs.AIcs.CLarXiv:2502.04128v22025
  38. AMO: Adaptive Motion Optimization for Hyper-Dexterous Humanoid Whole-Body Control

    Jialong Li, Xuxin Cheng, Tianshu Huang +3

    cs.ROcs.AIcs.LGarXiv:2505.03738v12025
  39. Soft Thinking: Unlocking the Reasoning Potential of LLMs in Continuous Concept Space

    Zhen Zhang, Xuehai He, Weixiang Yan +5

    cs.CLcs.AIarXiv:2505.15778v12025
  40. MCP Safety Audit: LLMs with the Model Context Protocol Allow Major Security Exploits

    Brandon Radosevich, John Halloran

    cs.CRcs.AIcs.LGarXiv:2504.03767v22025
  41. HalluLens: LLM Hallucination Benchmark

    Yejin Bang, Ziwei Ji, Alan Schelten +5

    cs.CLcs.AIarXiv:2504.17550v12025
  42. Deep Research Agents: A Systematic Examination And Roadmap

    Yuxuan Huang, Yihang Chen, Haozheng Zhang +10

    cs.AIarXiv:2506.18096v22025
  43. Personalized HeartSteps: A Reinforcement Learning Algorithm for Optimizing Physical Activity

    Peng Liao, Kristjan Greenewald, Predrag Klasnja +1

    cs.LGcs.AIarXiv:1909.03539v12019
  44. HealthGPT: A Medical Large Vision-Language Model for Unifying Comprehension and Generation via Heterogeneous Knowledge Adaptation

    Tianwei Lin, Wenqiao Zhang, Sijing Li +12

    cs.CVcs.AIarXiv:2502.09838v32025
  45. HOMIE: Humanoid Loco-Manipulation with Isomorphic Exoskeleton Cockpit

    Qingwei Ben, Feiyu Jia, Jia Zeng +3

    cs.ROcs.AIcs.HCarXiv:2502.13013v22025
  46. Anomaly Detection of Time Series with Smoothness-Inducing Sequential Variational Auto-Encoder

    Longyuan Li, Junchi Yan, Haiyang Wang +1

    cs.LGcs.AIarXiv:2102.01331v12021
  47. Reasoning with Sampling: Your Base Model is Smarter Than You Think

    Aayush Karan, Yilun Du

    cs.LGcs.AIcs.CLarXiv:2510.14901v12025
  48. frb100-40 After Two Decades: An Optimality Certificate and a Preregistered Search Study

    Onur Uğurlu

    cs.DMcs.AIarXiv:2609.02804v12026
  49. Skill-SD: Skill-Conditioned Self-Distillation for Multi-turn LLM Agents

    Hao Wang, Guozhi Wang, Han Xiao +8

    cs.LGcs.AIcs.CLarXiv:2604.10674v12026
  50. Second-order Non-local Attention Networks for Person Re-identification

    Bryan, Xia, Yuan Gong +2

    cs.CVcs.AIcs.LGarXiv:1909.00295v12019
  51. Modeling What Changes: Sparse, Residual World Models for Object-Centric Manipulation

    Param Thakkar, Parsika Paresh Shah, Manisha Sushant Gote

    cs.ROcs.AIarXiv:2609.02046v12026
  52. Technology Readiness Levels for AI & ML

    Alexander Lavin, Gregory Renard

    cs.SEcs.AIcs.LGarXiv:2006.12497v32020
  53. Measurement-Driven Sub-Network Selection for On-Premise Retrieval-Augmented Factory Agents

    Vasileios Rizeakos, Georgios Paisios, Alexandros Machairas +2

    cs.AIarXiv:2609.02760v12026
  54. More Agents Is All You Need

    Junyou Li, Qin Zhang, Yangbin Yu +2

    cs.CLcs.AIcs.LGarXiv:2402.05120v22024
  55. SEAL: Reinforcing Global Safety in Mixture-of-Experts through Shared Expert ALignment

    Qingyu Meng, Yiwei Zha, Jiahuan Pei +3

    cs.LGcs.AIcs.CRarXiv:2609.02293v12026
  56. Seed1.8 Model Card: Towards Generalized Real-World Agency

    Bytedance Seed

    cs.AIarXiv:2603.20633v32026
  57. Untangling the Mechanisms of Misleading Context in Medical Question Answering

    Robin Linzmayer, Noémie Elhadad

    cs.CLcs.AIcs.LGarXiv:2609.02754v12026
  58. Towards One-for-All Robustness Across a Continuum of Threat Levels

    Zhichao Hou, Xiaorui Liu

    cs.LGcs.AIarXiv:2609.02440v12026
  59. Audio-Reasoner: Improving Reasoning Capability in Large Audio Language Models

    Zhifei Xie, Mingbao Lin, Zihang Liu +3

    cs.SDcs.AIcs.CLarXiv:2503.02318v22025
  60. CALIP: Zero-Shot Enhancement of CLIP with Parameter-free Attention

    Ziyu Guo, Renrui Zhang, Longtian Qiu +4

    cs.CVcs.AIcs.MMarXiv:2209.14169v22022