Artificial Intelligence

Papers filed under cs.AI on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.

Search paper metadata (including unsummarized papers)

13,201 to 13,260 of 15,404

  1. MedSkillAudit: A Domain-Specific Audit Framework for Medical Research Agent Skills

    Yingyong Hou, Xinyuan Lao, Huimei Wang +10

    cs.AIarXiv:2604.20441v12026
  2. Efficient Agent Evaluation via Diversity-Guided User Simulation

    Itay Nakash, George Kour, Ateret Anaby-Tavor

    cs.AIarXiv:2604.21480v12026
  3. ALFRED: A Benchmark for Interpreting Grounded Instructions for Everyday Tasks

    Mohit Shridhar, Jesse Thomason, Daniel Gordon +5

    cs.CVcs.AIcs.CLarXiv:1912.01734v22019
  4. Retrieval Augmentation Reduces Hallucination in Conversation

    Kurt Shuster, Spencer Poff, Moya Chen +2

    cs.CLcs.AIarXiv:2104.07567v12021
  5. Probing Visual Planning in Image Editing Models

    Zhimu Zhou, Yanpeng Zhao, Qiuyu Liao +2

    cs.CVcs.AIarXiv:2604.22868v12026
  6. Discovering Agentic Safety Specifications from 1-Bit Danger Signals

    Víctor Gallego

    cs.AIcs.CLarXiv:2604.23210v12026
  7. IndustryAssetEQA: A Neurosymbolic Operational Intelligence System for Embodied Question Answering in Industrial Asset Maintenance

    Chathurangi Shyalika, Dhaval Patel, Amit Sheth

    cs.AIarXiv:2604.23446v12026
  8. Graph Engineering in the Era of LLM Agents: From Individual Intelligence to System Intelligence

    Yuyuan Feng, Zhishang Xiang, Chaobin Yang +32

    cs.IRcs.AIcs.ETarXiv:2608.21156v12026
  9. AnalogRetriever: Learning Cross-Modal Representations for Analog Circuit Retrieval

    Yihan Wang, Lei Li, Yao Lai +2

    cs.CVcs.AIarXiv:2604.23195v12026
  10. OceanPile: A Large-Scale Multimodal Ocean Corpus for Foundation Models

    Yida Xue, Ningyu Zhang, Tingwei Wu +5

    cs.MMcs.AIcs.CLarXiv:2605.00877v22026
  11. AgentBench: Evaluating LLMs as Agents

    Xiao Liu, Hao Yu, Hanchen Zhang +19

    cs.AIcs.CLcs.LGarXiv:2308.03688v32023
  12. Quantum Kernel Advantage over Classical Collapse in Medical Foundation Model Embeddings

    Sebastian Cajas Ordóñez, Felipe Ocampo Osorio, Dax Enshan Koh +10

    quant-phcs.AIarXiv:2604.24597v12026
  13. StackGAN++: Realistic Image Synthesis with Stacked Generative Adversarial Networks

    Han Zhang, Tao Xu, Hongsheng Li +4

    cs.CVcs.AIstat.MLarXiv:1710.10916v32017
  14. BARRED: Synthetic Training of Custom Policy Guardrails via Asymmetric Debate

    Arnon Mazza, Elad Levi

    cs.CLcs.AIcs.LGarXiv:2604.25203v12026
  15. Gemma: Open Models Based on Gemini Research and Technology

    Gemma Team, Thomas Mesnard, Cassidy Hardin +105

    cs.CLcs.AIarXiv:2403.08295v42024
  16. RL$^2$: Fast Reinforcement Learning via Slow Reinforcement Learning

    Yan Duan, John Schulman, Xi Chen +3

    cs.AIcs.LGcs.NEarXiv:1611.02779v22016
  17. Compliance versus Sensibility: On the Reasoning Controllability in Large Language Models

    Xingwei Tan, Marco Valentino, Mahmud Elahi Akhter +3

    cs.CLcs.AIarXiv:2604.27251v22026
  18. Large Language Models are not Fair Evaluators

    Peiyi Wang, Lei Li, Liang Chen +7

    cs.CLcs.AIcs.IRarXiv:2305.17926v22023
  19. Large Language Models Explore by Latent Distilling

    Yuanhao Zeng, Ao Lu, Lufei Li +3

    cs.CLcs.AIcs.LGarXiv:2604.24927v22026
  20. Deep Learning: A Critical Appraisal

    Gary Marcus

    cs.AIcs.LGstat.MLarXiv:1801.00631v12018
  21. RWKV: Reinventing RNNs for the Transformer Era

    Bo Peng, Eric Alcaide, Quentin Anthony +31

    cs.CLcs.AIarXiv:2305.13048v22023
  22. X2SAM: Any Segmentation in Images and Videos

    Hao Wang, Limeng Qiao, Chi Zhang +4

    cs.CVcs.AIarXiv:2605.00891v12026
  23. LLM.int8(): 8-bit Matrix Multiplication for Transformers at Scale

    Tim Dettmers, Mike Lewis, Younes Belkada +1

    cs.LGcs.AIarXiv:2208.07339v22022
  24. Instruction-Guided Poetry Generation in Arabic and Its Dialects

    Abdelrahman Sadallah, Kareem Elozeiri, Mervat Abassy +5

    cs.CLcs.AIarXiv:2604.27766v12026
  25. Gender Bias in Coreference Resolution: Evaluation and Debiasing Methods

    Jieyu Zhao, Tianlu Wang, Mark Yatskar +2

    cs.CLcs.AIarXiv:1804.06876v12018
  26. Turning the TIDE: Cross-Architecture Distillation for Diffusion Large Language Models

    Gongbo Zhang, Wen Wang, Ye Tian +1

    cs.CLcs.AIcs.LGarXiv:2604.26951v12026
  27. ML-Leaks: Model and Data Independent Membership Inference Attacks and Defenses on Machine Learning Models

    Ahmed Salem, Yang Zhang, Mathias Humbert +3

    cs.CRcs.AIcs.LGarXiv:1806.01246v22018
  28. Captum: A unified and generic model interpretability library for PyTorch

    Narine Kokhlikyan, Vivek Miglani, Miguel Martin +8

    cs.LGcs.AIstat.MLarXiv:2009.07896v12020
  29. End-to-End Attention-based Large Vocabulary Speech Recognition

    Dzmitry Bahdanau, Jan Chorowski, Dmitriy Serdyuk +2

    cs.CLcs.AIcs.LGarXiv:1508.04395v22015
  30. The LAMBADA dataset: Word prediction requiring a broad discourse context

    Denis Paperno, Germán Kruszewski, Angeliki Lazaridou +6

    cs.CLcs.AIcs.LGarXiv:1606.06031v12016
  31. CLEAR: Continuous Latent Adapter Routing for Utility-Preserving LLM Safety Alignment

    Chengxiao Wang, Enyi Jiang, Xiaojing Liao +1

    cs.AIarXiv:2608.21278v12026
  32. PointNeXt: Revisiting PointNet++ with Improved Training and Scaling Strategies

    Guocheng Qian, Yuchen Li, Houwen Peng +4

    cs.CVcs.AIarXiv:2206.04670v22022
  33. Interpretability in the Wild: a Circuit for Indirect Object Identification in GPT-2 small

    Kevin Wang, Alexandre Variengien, Arthur Conmy +2

    cs.LGcs.AIcs.CLarXiv:2211.00593v12022
  34. Audio Adversarial Examples: Targeted Attacks on Speech-to-Text

    Nicholas Carlini, David Wagner

    cs.LGcs.AIcs.CRarXiv:1801.01944v22018
  35. TransBTS: Multimodal Brain Tumor Segmentation Using Transformer

    Wenxuan Wang, Chen Chen, Meng Ding +3

    cs.CVcs.AIarXiv:2103.04430v22021
  36. Bottleneck Transformers for Visual Recognition

    Aravind Srinivas, Tsung-Yi Lin, Niki Parmar +3

    cs.CVcs.AIcs.LGarXiv:2101.11605v22021
  37. Repetition over Diversity: High-Signal Data Filtering for Sample-Efficient German Language Modeling

    Ansar Aynetdinov, Patrick Haller, Alan Akbik

    cs.CLcs.AIarXiv:2604.28075v22026
  38. InteractWeb-Bench: Can Multimodal Agent Escape Blind Execution in Interactive Website Generation?

    Qiyao Wang, Haoran Hu, Longze Chen +4

    cs.AIcs.CLarXiv:2604.27419v12026
  39. A Comprehensive AI Policy Education Framework for University Teaching and Learning

    Cecilia Ka Yuk Chan

    cs.CYcs.AIarXiv:2305.00280v12023
  40. Motion-Aware Caching for Efficient Autoregressive Video Generation

    Jing Xu, Yuexiao Ma, Xuzhe Zheng +7

    cs.CVcs.AIarXiv:2605.01725v22026
  41. BlenderRAG: High-Fidelity 3D Object Generation via Retrieval-Augmented Code Synthesis

    Massimo Rondelli, Francesco Pivi, Maurizio Gabbrielli

    cs.CVcs.AIcs.GRarXiv:2605.00632v12026
  42. Code World Model Preparedness Report

    Daniel Song, Peter Ney, Cristina Menghini +21

    cs.SEcs.AIarXiv:2605.00932v22026
  43. TCDA: Thread-Constrained Discourse-Aware Modeling for Conversational Sentiment Quadruple Analysis

    Xinran Li, Xinze Che, Yifan Lyu +2

    cs.CLcs.AIarXiv:2605.01717v22026
  44. A Minimalist Approach to Offline Reinforcement Learning

    Scott Fujimoto, Shixiang Shane Gu

    cs.LGcs.AIstat.MLarXiv:2106.06860v22021
  45. EDU-CIRCUIT-HW: Evaluating Multimodal Large Language Models on Real-World University-Level STEM Student Handwritten Solutions

    Weiyu Sun, Liangliang Chen, Yongnuo Cai +3

    cs.CVcs.AIcs.CYarXiv:2602.00095v32026
  46. Out-of-Distribution Generalization via Risk Extrapolation (REx)

    David Krueger, Ethan Caballero, Joern-Henrik Jacobsen +5

    cs.LGcs.AIcs.NEarXiv:2003.00688v52020
  47. Tensor field networks: Rotation- and translation-equivariant neural networks for 3D point clouds

    Nathaniel Thomas, Tess Smidt, Steven Kearnes +4

    cs.LGcs.AIcs.CVarXiv:1802.08219v32018
  48. QuoteBench: How Matched Scores Can Hide Command-Path Failures

    Shangao Li, Yao Zhang, Volker Tresp +1

    cs.AIcs.SEarXiv:2608.13547v12026
  49. VBPR: Visual Bayesian Personalized Ranking from Implicit Feedback

    Ruining He, Julian McAuley

    cs.IRcs.AIarXiv:1510.01784v12015
  50. DynamicViT: Efficient Vision Transformers with Dynamic Token Sparsification

    Yongming Rao, Wenliang Zhao, Benlin Liu +3

    cs.CVcs.AIcs.LGarXiv:2106.02034v22021
  51. Spatio-Temporal LSTM with Trust Gates for 3D Human Action Recognition

    Jun Liu, Amir Shahroudy, Dong Xu +1

    cs.CVcs.AIcs.LGarXiv:1607.07043v12016
  52. DN-DETR: Accelerate DETR Training by Introducing Query DeNoising

    Feng Li, Hao Zhang, Shilong Liu +3

    cs.CVcs.AIarXiv:2203.01305v32022
  53. Llama Guard: LLM-based Input-Output Safeguard for Human-AI Conversations

    Hakan Inan, Kartikeya Upasani, Jianfeng Chi +8

    cs.CLcs.AIarXiv:2312.06674v12023
  54. Unmasking Clever Hans Predictors and Assessing What Machines Really Learn

    Sebastian Lapuschkin, Stephan Wäldchen, Alexander Binder +3

    cs.AIcs.CVcs.LGarXiv:1902.10178v12019
  55. Plug and Play Language Models: A Simple Approach to Controlled Text Generation

    Sumanth Dathathri, Andrea Madotto, Janice Lan +5

    cs.CLcs.AIcs.LGarXiv:1912.02164v42019
  56. Holographic Embeddings of Knowledge Graphs

    Maximilian Nickel, Lorenzo Rosasco, Tomaso Poggio

    cs.AIcs.LGstat.MLarXiv:1510.04935v22015
  57. First Order Motion Model for Image Animation

    Aliaksandr Siarohin, Stéphane Lathuilière, Sergey Tulyakov +2

    cs.CVcs.AIarXiv:2003.00196v32020
  58. Representation Engineering: A Top-Down Approach to AI Transparency

    Andy Zou, Long Phan, Sarah Chen +18

    cs.LGcs.AIcs.CLarXiv:2310.01405v42023
  59. Towards AI-Complete Question Answering: A Set of Prerequisite Toy Tasks

    Jason Weston, Antoine Bordes, Sumit Chopra +4

    cs.AIcs.CLstat.MLarXiv:1502.05698v102015
  60. Techniques for Interpretable Machine Learning

    Mengnan Du, Ninghao Liu, Xia Hu

    cs.LGcs.AIstat.MLarXiv:1808.00033v32018