Artificial Intelligence
Papers filed under cs.AI on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.
Search paper metadata (including unsummarized papers)
11,701 to 11,760 of 15,487
Loc3R-VLM: Language-based Localization and 3D Reasoning with Vision-Language Models
Kevin Qu, Haozhe Qi, Mihai Dusmanu +3
cs.CVcs.AIcs.CLarXiv:2603.18002v12026s2n-bignum-bench: A practical benchmark for evaluating low-level code reasoning of LLMs
Balaji Rao, John Harrison, Soonho Kong +2
cs.PLcs.AIcs.CRarXiv:2603.14628v22026Socratic Models: Composing Zero-Shot Multimodal Reasoning with Language
Andy Zeng, Maria Attarian, Brian Ichter +10
cs.CVcs.AIcs.CLarXiv:2204.00598v22022AdapterTune: Zero-Initialized Low-Rank Adapters for Frozen Vision Transformers
Salim Khazem
cs.CVcs.AIcs.LGarXiv:2603.14706v12026ReactMotion: Generating Reactive Listener Motions from Speaker Utterance
Cheng Luo, Bizhu Wu, Bing Li +5
cs.CVcs.AIcs.HCarXiv:2603.15083v12026One-Shot Imitation Learning
Yan Duan, Marcin Andrychowicz, Bradly C. Stadie +5
cs.AIcs.LGcs.NEarXiv:1703.07326v32017EG-ARSA: An Expert-Grounded Open Model for Visual Road Safety Auditing in Low-Resource Settings
Md Thamed Bin Zaman Chowdhury, Moazzem Hossain
cs.CVcs.AIarXiv:2608.23563v12026AgentFormer: Agent-Aware Transformers for Socio-Temporal Multi-Agent Forecasting
Ye Yuan, Xinshuo Weng, Yanglan Ou +1
cs.AIcs.CVcs.LGarXiv:2103.14023v32021From Masks to Pixels and Meaning: A New Taxonomy, Benchmark, and Metrics for VLM Image Tampering
Xinyi Shang, Yi Tang, Jiacheng Cui +9
cs.CVcs.AIcs.LGarXiv:2603.20193v12026What's the Catch? Evaluating Temporal Consistency in Vision-Language Models
Marek Hradil, Danae Sánchez Villegas
cs.CLcs.AIcs.CVarXiv:2608.23474v22026Unified Spatio-Temporal Token Scoring for Efficient Video VLMs
Jianrui Zhang, Yue Yang, Rohun Tripathi +5
cs.CVcs.AIcs.LGarXiv:2603.18004v12026Reasoning or Rhetoric? An Empirical Analysis of Moral Reasoning Explanations in Large Language Models
Aryan Kasat, Smriti Singh, Aman Chadha +1
cs.AIarXiv:2603.21854v12026LLM-Based Selection of Incongruent Verbal and Nonverbal Behavior for Virtual Humans
Parisa Ghanad Torshizi, Stacy Marsella
cs.AIcs.HCcs.ROarXiv:2608.22731v12026HiMu: Hierarchical Multimodal Frame Selection for Long Video Question Answering
Dan Ben-Ami, Gabriele Serussi, Kobi Cohen +1
cs.CVcs.AIarXiv:2603.18558v22026Compositional Visual Generation with Composable Diffusion Models
Nan Liu, Shuang Li, Yilun Du +2
cs.CVcs.AIcs.LGarXiv:2206.01714v62022Adapter-Based Few-Shot Continual Learning for Malicious Packet Recognition
Kyle Stein, Guillermo Francia, III Eman El-Sheikh +1
cs.CRcs.AIarXiv:2608.23536v12026Act with Intent: Distilling Behavior Intent for Vision-Language-Action Models
Sangoh Lee, Sangwoo Mo, Wook-Shin Han
cs.ROcs.AIcs.CVarXiv:2608.23478v12026Evaluating AI-based Scientific Knowledge Synthesis with Epidemiological Systematic Reviews
Shreyansh Padarha, Ryan Othniel Kearns, Tristan Naidoo +13
cs.IRcs.AIcs.DLarXiv:2603.22327v22026ViZDoom: A Doom-based AI Research Platform for Visual Reinforcement Learning
Michał Kempka, Marek Wydmuch, Grzegorz Runc +2
cs.LGcs.AIcs.CVarXiv:1605.02097v22016Tree of Attacks: Jailbreaking Black-Box LLMs Automatically
Anay Mehrotra, Manolis Zampetakis, Paul Kassianik +4
cs.LGcs.AIcs.CLarXiv:2312.02119v32023The Emergence of Relevance Through Axiomatic Attention Patterns During LoRA Fine-Tuning
Matthew Perlman, Atharva Nijasure, James Allan
cs.CLcs.AIcs.IRarXiv:2608.23338v12026WorldCache: Content-Aware Caching for Accelerated Video World Models
Umair Nawaz, Ahmed Heakl, Ufaq Khan +3
cs.CVcs.AIcs.CLarXiv:2603.22286v12026SyncDreamer: Generating Multiview-consistent Images from a Single-view Image
Yuan Liu, Cheng Lin, Zijiao Zeng +4
cs.CVcs.AIcs.GRarXiv:2309.03453v22023Robust Reasoning Benchmark
Pavel Golikov, Evgenii Opryshko, Gennady Pekhimenko +1
cs.LGcs.AIcs.CLarXiv:2604.08571v32026STRIDE: When to Speak Meets Sequence Denoising for Streaming Video Understanding
Junho Kim, Hosu Lee, James M. Rehg +2
cs.CVcs.AIarXiv:2603.27593v12026Xpertbench: Expert Level Tasks with Rubrics-Based Evaluation
Xue Liu, Xin Ma, Yuxin Ma +36
cs.AIcs.CLarXiv:2604.02368v42026Going Deeper With Directly-Trained Larger Spiking Neural Networks
Hanle Zheng, Yujie Wu, Lei Deng +2
cs.NEcs.AIarXiv:2011.05280v22020Scaling Teams or Scaling Time? Memory Enabled Lifelong Learning in LLM Multi-Agent Systems
Shanglin Wu, Yuyang Luo, Yueqing Liang +4
cs.MAcs.AIarXiv:2604.03295v12026LAION-BVD: A 10-Million-Hour Open Video Dataset for Multimodal Pre-training
Andreas Hochlehnert, Marianna Nezhurina, Mehdi Cherti +9
cs.CVcs.AIcs.LGarXiv:2608.24845v12026Scalable Training of Artificial Neural Networks with Adaptive Sparse Connectivity inspired by Network Science
Decebal Constantin Mocanu, Elena Mocanu, Peter Stone +3
cs.NEcs.AIcs.LGarXiv:1707.04780v22017Genie: Generative Interactive Environments
Jake Bruce, Michael Dennis, Ashley Edwards +22
cs.LGcs.AIcs.CVarXiv:2402.15391v12024ExecRubrics: Executable Tool-Augmented Rubrics for Verifiable and Efficient Long-Form Evaluation
Kaustubh D. Dhole, Charles L. A. Clarke, Eugene Y. Agichtein
cs.AIcs.CLcs.IRarXiv:2608.22559v12026ViGoR-Bench: How Far Are Visual Generative Models From Zero-Shot Visual Reasoners?
Haonan Han, Jiancheng Huang, Xiaopeng Sun +7
cs.CVcs.AIarXiv:2603.25823v12026The Design and Implementation of XiaoIce, an Empathetic Social Chatbot
Li Zhou, Jianfeng Gao, Di Li +1
cs.HCcs.AIcs.CLarXiv:1812.08989v22018PerceptionComp: A Video Benchmark for Complex Perception-Centric Reasoning
Shaoxuan Li, Zhixuan Zhao, Hanze Deng +9
cs.CVcs.AIcs.CLarXiv:2603.26653v12026Friends and Grandmothers in Silico: Localizing Entity Cells in Language Models
Itay Yona, Dan Barzilay, Michael Karasik +1
cs.CLcs.AIarXiv:2604.01404v22026Spider-Sense: Intrinsic Risk Sensing for Efficient Agent Defense with Hierarchical Adaptive Screening
Zhenxiong Yu, Zhi Yang, Zhiheng Jin +19
cs.CRcs.AIarXiv:2602.05386v22026MemRerank: Preference Memory for Personalized Product Reranking
Zhiyuan Peng, Xuyang Wu, Huaixiao Tou +2
cs.CLcs.AIcs.LGarXiv:2603.29247v32026MAD: Modality-Adaptive Decoding for Mitigating Cross-Modal Hallucinations in Multimodal Large Language Models
Sangyun Chung, Se Yeon Kim, Youngchae Chee +1
cs.AIarXiv:2601.21181v12026WizardMath: Empowering Mathematical Reasoning for Large Language Models via Reinforced Evol-Instruct
Haipeng Luo, Qingfeng Sun, Can Xu +8
cs.CLcs.AIcs.LGarXiv:2308.09583v32023Hybrid Panels: Toward Human-AI Collaboration in Survey Research
Julia Romberg, Tobias Gummer, Gabriella Lapesa +2
cs.CLcs.AIcs.CYarXiv:2608.22582v12026PhotoBench: Beyond Visual Matching Towards Personalized Intent-Driven Photo Retrieval
Tianyi Xu, Rong Shan, Junjie Wu +11
cs.IRcs.AIcs.CVarXiv:2603.01493v22026Global Filter Networks for Image Classification
Yongming Rao, Wenliang Zhao, Zheng Zhu +2
cs.CVcs.AIcs.LGarXiv:2107.00645v22021ProBel: Propaganda Detection with Techniques, Spans, and Explanations
Mohamed Bayan Kmainasi, Ali Ezzat Shahroor, Elisa Sartori +2
cs.CLcs.AIcs.LGarXiv:2608.22388v12026Multiple Instance Learning: A Survey of Problem Characteristics and Applications
Marc-André Carbonneau, Veronika Cheplygina, Eric Granger +1
cs.CVcs.AIcs.IRarXiv:1612.03365v12016AdaptToken: Entropy-based Adaptive Token Selection for MLLM Long Video Understanding
Haozhe Qi, Kevin Qu, Mahdi Rad +3
cs.CVcs.AIarXiv:2603.28696v12026TransHands: Repurposing Human Pose Encoders as Hand Pose Encoders
Milo Piccioli, Gianluca Amprimo, Claudia Ferraris +1
cs.CVcs.AIarXiv:2608.22341v12026Evaluating Very Long-Term Conversational Memory of LLM Agents
Adyasha Maharana, Dong-Ho Lee, Sergey Tulyakov +3
cs.CLcs.AIcs.LGarXiv:2402.17753v12024ROCKET: Rapid Optimization via Calibration-guided Knapsack Enhanced Truncation for Efficient Model Compression
Ammar Ali, Baher Mohammad, Denis Makhov +3
cs.LGcs.AIcs.CLarXiv:2602.11008v12026Rubrics as an Attack Surface: Stealthy Preference Drift in LLM Judges
Ruomeng Ding, Yifei Pang, He Sun +3
cs.CRcs.AIcs.CLarXiv:2602.13576v12026The Internal State of an LLM Knows When It's Lying
Amos Azaria, Tom Mitchell
cs.CLcs.AIcs.LGarXiv:2304.13734v22023Safe RLHF: Safe Reinforcement Learning from Human Feedback
Josef Dai, Xuehai Pan, Ruiyang Sun +5
cs.AIcs.LGarXiv:2310.12773v12023Phi-4 Technical Report
Marah Abdin, Jyoti Aneja, Harkirat Behl +24
cs.CLcs.AIarXiv:2412.08905v12024Sci-Reasoning: A Dataset Decoding AI Innovation Patterns
Jiachen Liu, Maestro Harmon, Zechen Zhang
cs.AIcs.LGarXiv:2601.04577v12026When Personalization Misleads: Understanding and Mitigating Hallucinations in Personalized LLMs
Zhongxiang Sun, Yi Zhan, Chenglei Shen +4
cs.CLcs.AIarXiv:2601.11000v12026Does Inference Scaling Improve Reasoning Faithfulness? A Multi-Model Analysis of Self-Consistency Tradeoffs
Deep Mehta
cs.AIarXiv:2601.06423v12026GutenOCR: A Grounded Vision-Language Front-End for Documents
Hunter Heidenreich, Ben Elliott, Olivia Dinica +1
cs.CVcs.AIcs.CLarXiv:2601.14490v22026Sparking Scientific Creativity via LLM-Driven Interdisciplinary Inspiration
Priyanka Kargupta, Shuhaib Mehri, Dilek Hakkani-Tur +1
cs.CLcs.AIarXiv:2603.12226v12026DARC: Decoupled Asymmetric Reasoning Curriculum for LLM Evolution
Shengda Fan, Xuyan Ye, Yankai Lin
cs.AIcs.CLarXiv:2601.13761v22026Recursive Experiential-Working Memory Evolution for Long-Horizon Agent Harnesses
Zhaochen Yu, Yingcheng Wu, Zhenfei Yin +5
cs.AIcs.CLarXiv:2608.24876v12026