Artificial Intelligence
Papers filed under cs.AI on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.
Search paper metadata (including unsummarized papers)
13,381 to 13,440 of 15,466
The highD Dataset: A Drone Dataset of Naturalistic Vehicle Trajectories on German Highways for Validation of Highly Automated Driving Systems
Robert Krajewski, Julian Bock, Laurent Kloeker +1
cs.CVcs.AIcs.IRarXiv:1810.05642v12018Towards Accurate Generative Models of Video: A New Metric & Challenges
Thomas Unterthiner, Sjoerd van Steenkiste, Karol Kurach +3
cs.CVcs.AIcs.LGarXiv:1812.01717v22018Seeing the Needle in the Haystack: Towards Weakly-Supervised Log Instance Anomaly Localization via Counterfactual Perturbation
Yutszyuk Wong, Wentai Wu, Yuen-Ying Yeung +1
cs.LGcs.AIarXiv:2605.10988v12026Learning Multiagent Communication with Backpropagation
Sainbayar Sukhbaatar, Arthur Szlam, Rob Fergus
cs.LGcs.AIarXiv:1605.07736v22016When Not to Trust Language Models: Investigating Effectiveness of Parametric and Non-Parametric Memories
Alex Mallen, Akari Asai, Victor Zhong +3
cs.CLcs.AIcs.LGarXiv:2212.10511v42022Capabilities of GPT-4 on Medical Challenge Problems
Harsha Nori, Nicholas King, Scott Mayer McKinney +2
cs.CLcs.AIarXiv:2303.13375v22023RigidFormer: Learning Rigid Dynamics using Transformers
Zhiyang Dou, Minghao Guo, Haixu Wu +3
cs.CVcs.AIcs.GRarXiv:2605.09196v12026DiagnosticIQ: A Benchmark for LLM-Based Industrial Maintenance Action Recommendation from Symbolic Rules
Devin Yasith De Silva, Dhaval Patel, Christodoulos Constantinides +7
cs.AIarXiv:2605.08614v12026DeepSeek-V2: A Strong, Economical, and Efficient Mixture-of-Experts Language Model
DeepSeek-AI, Aixin Liu, Bei Feng +154
cs.CLcs.AIarXiv:2405.04434v52024SeePhys Pro: Diagnosing Modality Transfer and Blind-Training Effects in Multimodal RLVR for Physics Reasoning
Kun Xiang, Terry Jingchen Zhang, Zirong Liu +15
cs.AIarXiv:2605.09266v22026VICReg: Variance-Invariance-Covariance Regularization for Self-Supervised Learning
Adrien Bardes, Jean Ponce, Yann LeCun
cs.CVcs.AIcs.LGarXiv:2105.04906v32021LEAD: Length-Efficient Adaptive and Dynamic Reasoning for Large Language Models
Songtao Wei, Yi Li, Zhikai Li +7
cs.LGcs.AIarXiv:2605.09806v12026A Review on Deep Learning Techniques Applied to Semantic Segmentation
Alberto Garcia-Garcia, Sergio Orts-Escolano, Sergiu Oprea +2
cs.CVcs.AIarXiv:1704.06857v12017LIMA: Less Is More for Alignment
Chunting Zhou, Pengfei Liu, Puxin Xu +12
cs.CLcs.AIcs.LGarXiv:2305.11206v12023Taskonomy: Disentangling Task Transfer Learning
Amir Zamir, Alexander Sax, William Shen +3
cs.CVcs.AIcs.LGarXiv:1804.08328v12018How NOT To Evaluate Your Dialogue System: An Empirical Study of Unsupervised Evaluation Metrics for Dialogue Response Generation
Chia-Wei Liu, Ryan Lowe, Iulian V. Serban +3
cs.CLcs.AIcs.LGarXiv:1603.08023v22016MetaFormer Is Actually What You Need for Vision
Weihao Yu, Mi Luo, Pan Zhou +5
cs.CVcs.AIcs.LGarXiv:2111.11418v32021Human-Centered Artificial Intelligence: Reliable, Safe & Trustworthy
Ben Shneiderman
cs.HCcs.AIarXiv:2002.04087v22020Deep Learning in Spiking Neural Networks
Amirhossein Tavanaei, Masoud Ghodrati, Saeed Reza Kheradpisheh +2
cs.NEcs.AIarXiv:1804.08150v42018Stacked Cross Attention for Image-Text Matching
Kuang-Huei Lee, Xi Chen, Gang Hua +2
cs.CVcs.AIcs.LGarXiv:1803.08024v22018Is BERT Really Robust? A Strong Baseline for Natural Language Attack on Text Classification and Entailment
Di Jin, Zhijing Jin, Joey Tianyi Zhou +1
cs.CLcs.AIcs.LGarXiv:1907.11932v62019A Literature Survey of Benchmark Functions For Global Optimization Problems
Momin Jamil, Xin-She Yang
cs.AImath.OCarXiv:1308.4008v12013MemReread: Enhancing Agentic Long-Context Reasoning via Memory-Guided Rereading
Baibei Ji, Xiaoyang Weng, Juntao Li +3
cs.CLcs.AIarXiv:2605.10268v12026DeepRefine: Agent-Compiled Knowledge Refinement via Reinforcement Learning
Haoyu Huang, Jiaxin Bai, Shujie Liu +6
cs.CLcs.AIarXiv:2605.10488v12026Program of Thoughts Prompting: Disentangling Computation from Reasoning for Numerical Reasoning Tasks
Wenhu Chen, Xueguang Ma, Xinyi Wang +1
cs.CLcs.AIarXiv:2211.12588v42022VideoBERT: A Joint Model for Video and Language Representation Learning
Chen Sun, Austin Myers, Carl Vondrick +2
cs.CVcs.AIarXiv:1904.01766v22019Few-Shot Parameter-Efficient Fine-Tuning is Better and Cheaper than In-Context Learning
Haokun Liu, Derek Tam, Mohammed Muqeeth +4
cs.LGcs.AIcs.CLarXiv:2205.05638v22022TMAS: Scaling Test-Time Compute via Multi-Agent Synergy
George Wu, Nan Jing, Qing Yi +7
cs.AIarXiv:2605.10344v22026Scene Representation Networks: Continuous 3D-Structure-Aware Neural Scene Representations
Vincent Sitzmann, Michael Zollhöfer, Gordon Wetzstein
cs.CVcs.AIarXiv:1906.01618v22019Key-Value Means: Transformers with Expandable Block-Recurrent Compressed Memory
Daniel Goldstein, Navneel Singhal, Eugene Cheah
cs.LGcs.AIcs.CLarXiv:2605.09877v52026Active Tabular Augmentation via Policy-Guided Diffusion Inpainting
Zheyu Zhang, Shuo Yang, Bardh Prenkaj +1
cs.LGcs.AIarXiv:2605.10315v12026WildRelight: A Real-World Benchmark and Physics-Guided Adaptation for Single-Image Relighting
Lezhong Wang, Mehmet Onurcan Kaya, Siavash Bigdeli +1
cs.CVcs.AIcs.GRarXiv:2605.11696v12026AlphaGRPO: Unlocking Self-Reflective Multimodal Generation in UMMs via Decompositional Verifiable Reward
Runhui Huang, Jie Wu, Rui Yang +2
cs.CVcs.AIcs.LGarXiv:2605.12495v12026Do Enterprise Systems Need Learned World Models? The Importance of Context to Infer Dynamics
Jishnu Sethumadhavan Nair, Patrice Bechard, Rishabh Maheshwary +14
cs.AIcs.CLcs.LGarXiv:2605.12178v12026CoQA: A Conversational Question Answering Challenge
Siva Reddy, Danqi Chen, Christopher D. Manning
cs.CLcs.AIcs.LGarXiv:1808.07042v22018Learning to Communicate Locally for Large-Scale Multi-Agent Pathfinding
Valeriy Vyaltsev, Alsu Sagirova, Anton Andreychuk +5
cs.AIcs.LGcs.MAarXiv:2605.07637v22026Predicting Decisions of AI Agents from Limited Interaction through Text-Tabular Modeling
Eilam Shapira, Moshe Tennenholtz, Roi Reichart
cs.LGcs.AIcs.CLarXiv:2605.12411v12026Debiased Model-based Representations for Sample-efficient Continuous Control
Jiafei Lyu, Zichuan Lin, Scott Fujimoto +5
cs.LGcs.AIarXiv:2605.11711v12026AST: Audio Spectrogram Transformer
Yuan Gong, Yu-An Chung, James Glass
cs.SDcs.AIarXiv:2104.01778v32021Beyond the Last Layer: Multi-Layer Representation Fusion for Visual Tokenization
Xuanyu Zhu, Yan Bai, Yang Shi +4
cs.CVcs.AIarXiv:2605.10780v22026Improved Knowledge Distillation via Teacher Assistant
Seyed-Iman Mirzadeh, Mehrdad Farajtabar, Ang Li +3
cs.LGcs.AIstat.MLarXiv:1902.03393v22019WriteSAE: Sparse Autoencoders for Recurrent State
Jack Young
cs.LGcs.AIcs.CLarXiv:2605.12770v42026SpectralFormer: Rethinking Hyperspectral Image Classification with Transformers
Danfeng Hong, Zhu Han, Jing Yao +4
cs.CVcs.AIarXiv:2107.02988v22021StereoSet: Measuring stereotypical bias in pretrained language models
Moin Nadeem, Anna Bethke, Siva Reddy
cs.CLcs.AIcs.CYarXiv:2004.09456v12020Orthrus: Memory-Efficient Parallel Token Generation via Dual-View Diffusion
Chien Van Nguyen, Chaitra Hegde, Van Cuong Pham +3
cs.LGcs.AIarXiv:2605.12825v22026BloombergGPT: A Large Language Model for Finance
Shijie Wu, Ozan Irsoy, Steven Lu +6
cs.LGcs.AIcs.CLarXiv:2303.17564v32023MedMNIST v2 -- A large-scale lightweight benchmark for 2D and 3D biomedical image classification
Jiancheng Yang, Rui Shi, Donglai Wei +5
cs.CVcs.AIcs.LGarXiv:2110.14795v22021MODUS: Decoder-Only Any-to-Any Modeling of Diverse Modalities
Mingqiao Ye, Zhaochong An, Zhitong Gao +11
cs.CVcs.AIcs.LGarXiv:2607.25948v12026DeepMind Control Suite
Yuval Tassa, Yotam Doron, Alistair Muldal +9
cs.AIarXiv:1801.00690v12018SPIN: Structural LLM Planning via Iterative Navigation for Industrial Tasks
Yusuke Ozaki, Dhaval Patel
cs.AIarXiv:2605.14051v12026Diffusion-LM Improves Controllable Text Generation
Xiang Lisa Li, John Thickstun, Ishaan Gulrajani +2
cs.CLcs.AIcs.LGarXiv:2205.14217v12022Do Vision Transformers See Like Convolutional Neural Networks?
Maithra Raghu, Thomas Unterthiner, Simon Kornblith +2
cs.CVcs.AIcs.LGarXiv:2108.08810v22021Training Large Language Models to Predict Clinical Events
Benjamin Turtel, Paul Wilczewski, Kris Skotheim
cs.LGcs.AIcs.CLarXiv:2605.12817v12026Vividh-ASR: A Complexity-Tiered Benchmark and Optimization Dynamics for Robust Indic Speech Recognition
Kush Juvekar, Kavya Manohar, Aditya Srinivas Menon +2
cs.CLcs.AIarXiv:2605.13087v22026Context Training with Active Information Seeking
Zeyu Huang, Adhiguna Kuncoro, Qixuan Feng +4
cs.CLcs.AIarXiv:2605.13050v22026The Secret Sharer: Evaluating and Testing Unintended Memorization in Neural Networks
Nicholas Carlini, Chang Liu, Úlfar Erlingsson +2
cs.LGcs.AIcs.CRarXiv:1802.08232v32018RealICU: Do LLM Agents Understand Long-Context ICU Data? A Benchmark Beyond Behavior Imitation
Chengzhi Shen, Weixiang Shen, Tobias Susetzky +8
cs.AIcs.CLcs.LGarXiv:2605.13542v12026MAP: A Map-then-Act Paradigm for Long-Horizon Interactive Agent Reasoning
Yuxin Liu, Ziang Ye, Yueqing Sun +6
cs.AIarXiv:2605.13037v12026Chatbot Arena: An Open Platform for Evaluating LLMs by Human Preference
Wei-Lin Chiang, Lianmin Zheng, Ying Sheng +8
cs.AIcs.CLarXiv:2403.04132v12024Boosting Reinforcement Learning with Verifiable Rewards via Randomly Selected Few-Shot Guidance
Kai Yan, Alexander G. Schwing, Yu-Xiong Wang
cs.LGcs.AIcs.CLarXiv:2605.15012v12026