Artificial Intelligence

Papers filed under cs.AI on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.

Search paper metadata (including unsummarized papers)

13,261 to 13,320 of 15,365

  1. Twins: Revisiting the Design of Spatial Attention in Vision Transformers

    Xiangxiang Chu, Zhi Tian, Yuqing Wang +5

    cs.CVcs.AIcs.LGarXiv:2104.13840v42021
  2. Quantifying Attention Flow in Transformers

    Samira Abnar, Willem Zuidema

    cs.LGcs.AIcs.CLarXiv:2005.00928v22020
  3. The Option-Critic Architecture

    Pierre-Luc Bacon, Jean Harb, Doina Precup

    cs.AIarXiv:1609.05140v22016
  4. Adafactor: Adaptive Learning Rates with Sublinear Memory Cost

    Noam Shazeer, Mitchell Stern

    cs.LGcs.AIstat.MLarXiv:1804.04235v12018
  5. Fast Byte Latent Transformer

    Julie Kallini, Artidoro Pagnoni, Tomasz Limisiewicz +5

    cs.CLcs.AIcs.LGarXiv:2605.08044v12026
  6. Open-vocabulary Object Detection via Vision and Language Knowledge Distillation

    Xiuye Gu, Tsung-Yi Lin, Weicheng Kuo +1

    cs.CVcs.AIcs.LGarXiv:2104.13921v32021
  7. code2vec: Learning Distributed Representations of Code

    Uri Alon, Meital Zilberstein, Omer Levy +1

    cs.LGcs.AIcs.PLarXiv:1803.09473v52018
  8. Do not copy and paste! Rewriting strategies for code retrieval

    Andrea Gurioli, Federico Pennino, Maurizio Gabbrielli

    cs.SEcs.AIarXiv:2605.08299v12026
  9. SafeHarbor: Defining Precise Decision Boundaries via Hierarchical Memory-Augmented Guardrail for LLM Agent Safety

    Zhe Liu, Zonghao Ying, Wenxin Zhang +5

    cs.CRcs.AIarXiv:2605.05704v32026
  10. Trajectory as the Teacher: Few-Step Discrete Flow Matching via Energy-Navigated Distillation

    Amin Karimi Monsefi, Dominic Culver, Nikhil Bhendawade +3

    cs.LGcs.AIcs.CLarXiv:2605.07924v12026
  11. Analyzing Federated Learning through an Adversarial Lens

    Arjun Nitin Bhagoji, Supriyo Chakraborty, Prateek Mittal +1

    cs.LGcs.AIcs.CRarXiv:1811.12470v42018
  12. Who Prices Cognitive Labor in the Age of Agents? Compute-Anchored Wages

    Siqi Zhu

    cs.AIcs.CYarXiv:2605.05558v22026
  13. MM-Vet: Evaluating Large Multimodal Models for Integrated Capabilities

    Weihao Yu, Zhengyuan Yang, Linjie Li +5

    cs.AIcs.CLcs.CVarXiv:2308.02490v42023
  14. Sparse Autoencoders as Plug-and-Play Firewalls for Adversarial Attack Detection in VLMs

    Hao Wang, Yiqun Sun, Pengfei Wei +2

    cs.CVcs.AIcs.CLarXiv:2605.07447v12026
  15. BalCapRL: A Balanced Framework for RL-Based MLLM Image Captioning

    Shaokai Ye, Vasileios Saveris, Yihao Qian +3

    cs.CVcs.AIarXiv:2605.07394v12026
  16. Implicit Preference Alignment for Human Image Animation

    Yuanzhi Wang, Xuhua Ren, Jiaxiang Cheng +5

    cs.CVcs.AIarXiv:2605.07545v12026
  17. MC-RFM: Geometry-Aware Few-Shot Adaptation via Mixed-Curvature Riemannian Flow Matching

    Salim Khazem, Ibrahim Mohamed Serouis, Zakaria Ezzahed

    cs.CVcs.AIcs.LGarXiv:2605.08557v12026
  18. Inference-Time Intervention: Eliciting Truthful Answers from a Language Model

    Kenneth Li, Oam Patel, Fernanda Viégas +2

    cs.LGcs.AIcs.CLarXiv:2306.03341v62023
  19. From Holo Pockets to Electron Density: GPT-style Drug Design with Density

    Jiahao Chen, Letian Gao, Yanhao Zhu +4

    cs.AIarXiv:2605.08767v22026
  20. Prediction Bottlenecks Don't Discover Causal Structure (But Here's What They Actually Do)

    Ankit Hemant Lade, Sai Krishna Jasti, Indar Kumar +1

    cs.LGcs.AIarXiv:2605.09169v22026
  21. The highD Dataset: A Drone Dataset of Naturalistic Vehicle Trajectories on German Highways for Validation of Highly Automated Driving Systems

    Robert Krajewski, Julian Bock, Laurent Kloeker +1

    cs.CVcs.AIcs.IRarXiv:1810.05642v12018
  22. Towards Accurate Generative Models of Video: A New Metric & Challenges

    Thomas Unterthiner, Sjoerd van Steenkiste, Karol Kurach +3

    cs.CVcs.AIcs.LGarXiv:1812.01717v22018
  23. Seeing the Needle in the Haystack: Towards Weakly-Supervised Log Instance Anomaly Localization via Counterfactual Perturbation

    Yutszyuk Wong, Wentai Wu, Yuen-Ying Yeung +1

    cs.LGcs.AIarXiv:2605.10988v12026
  24. Learning Multiagent Communication with Backpropagation

    Sainbayar Sukhbaatar, Arthur Szlam, Rob Fergus

    cs.LGcs.AIarXiv:1605.07736v22016
  25. When Not to Trust Language Models: Investigating Effectiveness of Parametric and Non-Parametric Memories

    Alex Mallen, Akari Asai, Victor Zhong +3

    cs.CLcs.AIcs.LGarXiv:2212.10511v42022
  26. Capabilities of GPT-4 on Medical Challenge Problems

    Harsha Nori, Nicholas King, Scott Mayer McKinney +2

    cs.CLcs.AIarXiv:2303.13375v22023
  27. RigidFormer: Learning Rigid Dynamics using Transformers

    Zhiyang Dou, Minghao Guo, Haixu Wu +3

    cs.CVcs.AIcs.GRarXiv:2605.09196v12026
  28. DiagnosticIQ: A Benchmark for LLM-Based Industrial Maintenance Action Recommendation from Symbolic Rules

    Devin Yasith De Silva, Dhaval Patel, Christodoulos Constantinides +7

    cs.AIarXiv:2605.08614v12026
  29. DeepSeek-V2: A Strong, Economical, and Efficient Mixture-of-Experts Language Model

    DeepSeek-AI, Aixin Liu, Bei Feng +154

    cs.CLcs.AIarXiv:2405.04434v52024
  30. SeePhys Pro: Diagnosing Modality Transfer and Blind-Training Effects in Multimodal RLVR for Physics Reasoning

    Kun Xiang, Terry Jingchen Zhang, Zirong Liu +15

    cs.AIarXiv:2605.09266v22026
  31. VICReg: Variance-Invariance-Covariance Regularization for Self-Supervised Learning

    Adrien Bardes, Jean Ponce, Yann LeCun

    cs.CVcs.AIcs.LGarXiv:2105.04906v32021
  32. LEAD: Length-Efficient Adaptive and Dynamic Reasoning for Large Language Models

    Songtao Wei, Yi Li, Zhikai Li +7

    cs.LGcs.AIarXiv:2605.09806v12026
  33. A Review on Deep Learning Techniques Applied to Semantic Segmentation

    Alberto Garcia-Garcia, Sergio Orts-Escolano, Sergiu Oprea +2

    cs.CVcs.AIarXiv:1704.06857v12017
  34. LIMA: Less Is More for Alignment

    Chunting Zhou, Pengfei Liu, Puxin Xu +12

    cs.CLcs.AIcs.LGarXiv:2305.11206v12023
  35. Taskonomy: Disentangling Task Transfer Learning

    Amir Zamir, Alexander Sax, William Shen +3

    cs.CVcs.AIcs.LGarXiv:1804.08328v12018
  36. How NOT To Evaluate Your Dialogue System: An Empirical Study of Unsupervised Evaluation Metrics for Dialogue Response Generation

    Chia-Wei Liu, Ryan Lowe, Iulian V. Serban +3

    cs.CLcs.AIcs.LGarXiv:1603.08023v22016
  37. MetaFormer Is Actually What You Need for Vision

    Weihao Yu, Mi Luo, Pan Zhou +5

    cs.CVcs.AIcs.LGarXiv:2111.11418v32021
  38. Human-Centered Artificial Intelligence: Reliable, Safe & Trustworthy

    Ben Shneiderman

    cs.HCcs.AIarXiv:2002.04087v22020
  39. Deep Learning in Spiking Neural Networks

    Amirhossein Tavanaei, Masoud Ghodrati, Saeed Reza Kheradpisheh +2

    cs.NEcs.AIarXiv:1804.08150v42018
  40. Stacked Cross Attention for Image-Text Matching

    Kuang-Huei Lee, Xi Chen, Gang Hua +2

    cs.CVcs.AIcs.LGarXiv:1803.08024v22018
  41. Is BERT Really Robust? A Strong Baseline for Natural Language Attack on Text Classification and Entailment

    Di Jin, Zhijing Jin, Joey Tianyi Zhou +1

    cs.CLcs.AIcs.LGarXiv:1907.11932v62019
  42. A Literature Survey of Benchmark Functions For Global Optimization Problems

    Momin Jamil, Xin-She Yang

    cs.AImath.OCarXiv:1308.4008v12013
  43. MemReread: Enhancing Agentic Long-Context Reasoning via Memory-Guided Rereading

    Baibei Ji, Xiaoyang Weng, Juntao Li +3

    cs.CLcs.AIarXiv:2605.10268v12026
  44. DeepRefine: Agent-Compiled Knowledge Refinement via Reinforcement Learning

    Haoyu Huang, Jiaxin Bai, Shujie Liu +6

    cs.CLcs.AIarXiv:2605.10488v12026
  45. Program of Thoughts Prompting: Disentangling Computation from Reasoning for Numerical Reasoning Tasks

    Wenhu Chen, Xueguang Ma, Xinyi Wang +1

    cs.CLcs.AIarXiv:2211.12588v42022
  46. VideoBERT: A Joint Model for Video and Language Representation Learning

    Chen Sun, Austin Myers, Carl Vondrick +2

    cs.CVcs.AIarXiv:1904.01766v22019
  47. Few-Shot Parameter-Efficient Fine-Tuning is Better and Cheaper than In-Context Learning

    Haokun Liu, Derek Tam, Mohammed Muqeeth +4

    cs.LGcs.AIcs.CLarXiv:2205.05638v22022
  48. TMAS: Scaling Test-Time Compute via Multi-Agent Synergy

    George Wu, Nan Jing, Qing Yi +7

    cs.AIarXiv:2605.10344v22026
  49. Scene Representation Networks: Continuous 3D-Structure-Aware Neural Scene Representations

    Vincent Sitzmann, Michael Zollhöfer, Gordon Wetzstein

    cs.CVcs.AIarXiv:1906.01618v22019
  50. Key-Value Means: Transformers with Expandable Block-Recurrent Compressed Memory

    Daniel Goldstein, Navneel Singhal, Eugene Cheah

    cs.LGcs.AIcs.CLarXiv:2605.09877v52026
  51. Active Tabular Augmentation via Policy-Guided Diffusion Inpainting

    Zheyu Zhang, Shuo Yang, Bardh Prenkaj +1

    cs.LGcs.AIarXiv:2605.10315v12026
  52. WildRelight: A Real-World Benchmark and Physics-Guided Adaptation for Single-Image Relighting

    Lezhong Wang, Mehmet Onurcan Kaya, Siavash Bigdeli +1

    cs.CVcs.AIcs.GRarXiv:2605.11696v12026
  53. AlphaGRPO: Unlocking Self-Reflective Multimodal Generation in UMMs via Decompositional Verifiable Reward

    Runhui Huang, Jie Wu, Rui Yang +2

    cs.CVcs.AIcs.LGarXiv:2605.12495v12026
  54. Do Enterprise Systems Need Learned World Models? The Importance of Context to Infer Dynamics

    Jishnu Sethumadhavan Nair, Patrice Bechard, Rishabh Maheshwary +14

    cs.AIcs.CLcs.LGarXiv:2605.12178v12026
  55. CoQA: A Conversational Question Answering Challenge

    Siva Reddy, Danqi Chen, Christopher D. Manning

    cs.CLcs.AIcs.LGarXiv:1808.07042v22018
  56. Learning to Communicate Locally for Large-Scale Multi-Agent Pathfinding

    Valeriy Vyaltsev, Alsu Sagirova, Anton Andreychuk +5

    cs.AIcs.LGcs.MAarXiv:2605.07637v22026
  57. Predicting Decisions of AI Agents from Limited Interaction through Text-Tabular Modeling

    Eilam Shapira, Moshe Tennenholtz, Roi Reichart

    cs.LGcs.AIcs.CLarXiv:2605.12411v12026
  58. Debiased Model-based Representations for Sample-efficient Continuous Control

    Jiafei Lyu, Zichuan Lin, Scott Fujimoto +5

    cs.LGcs.AIarXiv:2605.11711v12026
  59. AST: Audio Spectrogram Transformer

    Yuan Gong, Yu-An Chung, James Glass

    cs.SDcs.AIarXiv:2104.01778v32021
  60. Beyond the Last Layer: Multi-Layer Representation Fusion for Visual Tokenization

    Xuanyu Zhu, Yan Bai, Yang Shi +4

    cs.CVcs.AIarXiv:2605.10780v22026