Artificial Intelligence

Papers filed under cs.AI on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.

Search paper metadata (including unsummarized papers)

11,701 to 11,760 of 15,487

  1. Loc3R-VLM: Language-based Localization and 3D Reasoning with Vision-Language Models

    Kevin Qu, Haozhe Qi, Mihai Dusmanu +3

    cs.CVcs.AIcs.CLarXiv:2603.18002v12026
  2. s2n-bignum-bench: A practical benchmark for evaluating low-level code reasoning of LLMs

    Balaji Rao, John Harrison, Soonho Kong +2

    cs.PLcs.AIcs.CRarXiv:2603.14628v22026
  3. Socratic Models: Composing Zero-Shot Multimodal Reasoning with Language

    Andy Zeng, Maria Attarian, Brian Ichter +10

    cs.CVcs.AIcs.CLarXiv:2204.00598v22022
  4. AdapterTune: Zero-Initialized Low-Rank Adapters for Frozen Vision Transformers

    Salim Khazem

    cs.CVcs.AIcs.LGarXiv:2603.14706v12026
  5. ReactMotion: Generating Reactive Listener Motions from Speaker Utterance

    Cheng Luo, Bizhu Wu, Bing Li +5

    cs.CVcs.AIcs.HCarXiv:2603.15083v12026
  6. One-Shot Imitation Learning

    Yan Duan, Marcin Andrychowicz, Bradly C. Stadie +5

    cs.AIcs.LGcs.NEarXiv:1703.07326v32017
  7. EG-ARSA: An Expert-Grounded Open Model for Visual Road Safety Auditing in Low-Resource Settings

    Md Thamed Bin Zaman Chowdhury, Moazzem Hossain

    cs.CVcs.AIarXiv:2608.23563v12026
  8. AgentFormer: Agent-Aware Transformers for Socio-Temporal Multi-Agent Forecasting

    Ye Yuan, Xinshuo Weng, Yanglan Ou +1

    cs.AIcs.CVcs.LGarXiv:2103.14023v32021
  9. From Masks to Pixels and Meaning: A New Taxonomy, Benchmark, and Metrics for VLM Image Tampering

    Xinyi Shang, Yi Tang, Jiacheng Cui +9

    cs.CVcs.AIcs.LGarXiv:2603.20193v12026
  10. What's the Catch? Evaluating Temporal Consistency in Vision-Language Models

    Marek Hradil, Danae Sánchez Villegas

    cs.CLcs.AIcs.CVarXiv:2608.23474v22026
  11. Unified Spatio-Temporal Token Scoring for Efficient Video VLMs

    Jianrui Zhang, Yue Yang, Rohun Tripathi +5

    cs.CVcs.AIcs.LGarXiv:2603.18004v12026
  12. Reasoning or Rhetoric? An Empirical Analysis of Moral Reasoning Explanations in Large Language Models

    Aryan Kasat, Smriti Singh, Aman Chadha +1

    cs.AIarXiv:2603.21854v12026
  13. LLM-Based Selection of Incongruent Verbal and Nonverbal Behavior for Virtual Humans

    Parisa Ghanad Torshizi, Stacy Marsella

    cs.AIcs.HCcs.ROarXiv:2608.22731v12026
  14. HiMu: Hierarchical Multimodal Frame Selection for Long Video Question Answering

    Dan Ben-Ami, Gabriele Serussi, Kobi Cohen +1

    cs.CVcs.AIarXiv:2603.18558v22026
  15. Compositional Visual Generation with Composable Diffusion Models

    Nan Liu, Shuang Li, Yilun Du +2

    cs.CVcs.AIcs.LGarXiv:2206.01714v62022
  16. Adapter-Based Few-Shot Continual Learning for Malicious Packet Recognition

    Kyle Stein, Guillermo Francia, III Eman El-Sheikh +1

    cs.CRcs.AIarXiv:2608.23536v12026
  17. Act with Intent: Distilling Behavior Intent for Vision-Language-Action Models

    Sangoh Lee, Sangwoo Mo, Wook-Shin Han

    cs.ROcs.AIcs.CVarXiv:2608.23478v12026
  18. Evaluating AI-based Scientific Knowledge Synthesis with Epidemiological Systematic Reviews

    Shreyansh Padarha, Ryan Othniel Kearns, Tristan Naidoo +13

    cs.IRcs.AIcs.DLarXiv:2603.22327v22026
  19. ViZDoom: A Doom-based AI Research Platform for Visual Reinforcement Learning

    Michał Kempka, Marek Wydmuch, Grzegorz Runc +2

    cs.LGcs.AIcs.CVarXiv:1605.02097v22016
  20. Tree of Attacks: Jailbreaking Black-Box LLMs Automatically

    Anay Mehrotra, Manolis Zampetakis, Paul Kassianik +4

    cs.LGcs.AIcs.CLarXiv:2312.02119v32023
  21. The Emergence of Relevance Through Axiomatic Attention Patterns During LoRA Fine-Tuning

    Matthew Perlman, Atharva Nijasure, James Allan

    cs.CLcs.AIcs.IRarXiv:2608.23338v12026
  22. WorldCache: Content-Aware Caching for Accelerated Video World Models

    Umair Nawaz, Ahmed Heakl, Ufaq Khan +3

    cs.CVcs.AIcs.CLarXiv:2603.22286v12026
  23. SyncDreamer: Generating Multiview-consistent Images from a Single-view Image

    Yuan Liu, Cheng Lin, Zijiao Zeng +4

    cs.CVcs.AIcs.GRarXiv:2309.03453v22023
  24. Robust Reasoning Benchmark

    Pavel Golikov, Evgenii Opryshko, Gennady Pekhimenko +1

    cs.LGcs.AIcs.CLarXiv:2604.08571v32026
  25. STRIDE: When to Speak Meets Sequence Denoising for Streaming Video Understanding

    Junho Kim, Hosu Lee, James M. Rehg +2

    cs.CVcs.AIarXiv:2603.27593v12026
  26. Xpertbench: Expert Level Tasks with Rubrics-Based Evaluation

    Xue Liu, Xin Ma, Yuxin Ma +36

    cs.AIcs.CLarXiv:2604.02368v42026
  27. Going Deeper With Directly-Trained Larger Spiking Neural Networks

    Hanle Zheng, Yujie Wu, Lei Deng +2

    cs.NEcs.AIarXiv:2011.05280v22020
  28. Scaling Teams or Scaling Time? Memory Enabled Lifelong Learning in LLM Multi-Agent Systems

    Shanglin Wu, Yuyang Luo, Yueqing Liang +4

    cs.MAcs.AIarXiv:2604.03295v12026
  29. LAION-BVD: A 10-Million-Hour Open Video Dataset for Multimodal Pre-training

    Andreas Hochlehnert, Marianna Nezhurina, Mehdi Cherti +9

    cs.CVcs.AIcs.LGarXiv:2608.24845v12026
  30. Scalable Training of Artificial Neural Networks with Adaptive Sparse Connectivity inspired by Network Science

    Decebal Constantin Mocanu, Elena Mocanu, Peter Stone +3

    cs.NEcs.AIcs.LGarXiv:1707.04780v22017
  31. Genie: Generative Interactive Environments

    Jake Bruce, Michael Dennis, Ashley Edwards +22

    cs.LGcs.AIcs.CVarXiv:2402.15391v12024
  32. ExecRubrics: Executable Tool-Augmented Rubrics for Verifiable and Efficient Long-Form Evaluation

    Kaustubh D. Dhole, Charles L. A. Clarke, Eugene Y. Agichtein

    cs.AIcs.CLcs.IRarXiv:2608.22559v12026
  33. ViGoR-Bench: How Far Are Visual Generative Models From Zero-Shot Visual Reasoners?

    Haonan Han, Jiancheng Huang, Xiaopeng Sun +7

    cs.CVcs.AIarXiv:2603.25823v12026
  34. The Design and Implementation of XiaoIce, an Empathetic Social Chatbot

    Li Zhou, Jianfeng Gao, Di Li +1

    cs.HCcs.AIcs.CLarXiv:1812.08989v22018
  35. PerceptionComp: A Video Benchmark for Complex Perception-Centric Reasoning

    Shaoxuan Li, Zhixuan Zhao, Hanze Deng +9

    cs.CVcs.AIcs.CLarXiv:2603.26653v12026
  36. Friends and Grandmothers in Silico: Localizing Entity Cells in Language Models

    Itay Yona, Dan Barzilay, Michael Karasik +1

    cs.CLcs.AIarXiv:2604.01404v22026
  37. Spider-Sense: Intrinsic Risk Sensing for Efficient Agent Defense with Hierarchical Adaptive Screening

    Zhenxiong Yu, Zhi Yang, Zhiheng Jin +19

    cs.CRcs.AIarXiv:2602.05386v22026
  38. MemRerank: Preference Memory for Personalized Product Reranking

    Zhiyuan Peng, Xuyang Wu, Huaixiao Tou +2

    cs.CLcs.AIcs.LGarXiv:2603.29247v32026
  39. MAD: Modality-Adaptive Decoding for Mitigating Cross-Modal Hallucinations in Multimodal Large Language Models

    Sangyun Chung, Se Yeon Kim, Youngchae Chee +1

    cs.AIarXiv:2601.21181v12026
  40. WizardMath: Empowering Mathematical Reasoning for Large Language Models via Reinforced Evol-Instruct

    Haipeng Luo, Qingfeng Sun, Can Xu +8

    cs.CLcs.AIcs.LGarXiv:2308.09583v32023
  41. Hybrid Panels: Toward Human-AI Collaboration in Survey Research

    Julia Romberg, Tobias Gummer, Gabriella Lapesa +2

    cs.CLcs.AIcs.CYarXiv:2608.22582v12026
  42. PhotoBench: Beyond Visual Matching Towards Personalized Intent-Driven Photo Retrieval

    Tianyi Xu, Rong Shan, Junjie Wu +11

    cs.IRcs.AIcs.CVarXiv:2603.01493v22026
  43. Global Filter Networks for Image Classification

    Yongming Rao, Wenliang Zhao, Zheng Zhu +2

    cs.CVcs.AIcs.LGarXiv:2107.00645v22021
  44. ProBel: Propaganda Detection with Techniques, Spans, and Explanations

    Mohamed Bayan Kmainasi, Ali Ezzat Shahroor, Elisa Sartori +2

    cs.CLcs.AIcs.LGarXiv:2608.22388v12026
  45. Multiple Instance Learning: A Survey of Problem Characteristics and Applications

    Marc-André Carbonneau, Veronika Cheplygina, Eric Granger +1

    cs.CVcs.AIcs.IRarXiv:1612.03365v12016
  46. AdaptToken: Entropy-based Adaptive Token Selection for MLLM Long Video Understanding

    Haozhe Qi, Kevin Qu, Mahdi Rad +3

    cs.CVcs.AIarXiv:2603.28696v12026
  47. TransHands: Repurposing Human Pose Encoders as Hand Pose Encoders

    Milo Piccioli, Gianluca Amprimo, Claudia Ferraris +1

    cs.CVcs.AIarXiv:2608.22341v12026
  48. Evaluating Very Long-Term Conversational Memory of LLM Agents

    Adyasha Maharana, Dong-Ho Lee, Sergey Tulyakov +3

    cs.CLcs.AIcs.LGarXiv:2402.17753v12024
  49. ROCKET: Rapid Optimization via Calibration-guided Knapsack Enhanced Truncation for Efficient Model Compression

    Ammar Ali, Baher Mohammad, Denis Makhov +3

    cs.LGcs.AIcs.CLarXiv:2602.11008v12026
  50. Rubrics as an Attack Surface: Stealthy Preference Drift in LLM Judges

    Ruomeng Ding, Yifei Pang, He Sun +3

    cs.CRcs.AIcs.CLarXiv:2602.13576v12026
  51. The Internal State of an LLM Knows When It's Lying

    Amos Azaria, Tom Mitchell

    cs.CLcs.AIcs.LGarXiv:2304.13734v22023
  52. Safe RLHF: Safe Reinforcement Learning from Human Feedback

    Josef Dai, Xuehai Pan, Ruiyang Sun +5

    cs.AIcs.LGarXiv:2310.12773v12023
  53. Phi-4 Technical Report

    Marah Abdin, Jyoti Aneja, Harkirat Behl +24

    cs.CLcs.AIarXiv:2412.08905v12024
  54. Sci-Reasoning: A Dataset Decoding AI Innovation Patterns

    Jiachen Liu, Maestro Harmon, Zechen Zhang

    cs.AIcs.LGarXiv:2601.04577v12026
  55. When Personalization Misleads: Understanding and Mitigating Hallucinations in Personalized LLMs

    Zhongxiang Sun, Yi Zhan, Chenglei Shen +4

    cs.CLcs.AIarXiv:2601.11000v12026
  56. Does Inference Scaling Improve Reasoning Faithfulness? A Multi-Model Analysis of Self-Consistency Tradeoffs

    Deep Mehta

    cs.AIarXiv:2601.06423v12026
  57. GutenOCR: A Grounded Vision-Language Front-End for Documents

    Hunter Heidenreich, Ben Elliott, Olivia Dinica +1

    cs.CVcs.AIcs.CLarXiv:2601.14490v22026
  58. Sparking Scientific Creativity via LLM-Driven Interdisciplinary Inspiration

    Priyanka Kargupta, Shuhaib Mehri, Dilek Hakkani-Tur +1

    cs.CLcs.AIarXiv:2603.12226v12026
  59. DARC: Decoupled Asymmetric Reasoning Curriculum for LLM Evolution

    Shengda Fan, Xuyan Ye, Yankai Lin

    cs.AIcs.CLarXiv:2601.13761v22026
  60. Recursive Experiential-Working Memory Evolution for Long-Horizon Agent Harnesses

    Zhaochen Yu, Yingcheng Wu, Zhenfei Yin +5

    cs.AIcs.CLarXiv:2608.24876v12026