Artificial Intelligence

Papers filed under cs.AI on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.

Search paper metadata (including unsummarized papers)

13,621 to 13,680 of 15,328

  1. Efficient Streaming Language Models with Attention Sinks

    Guangxuan Xiao, Yuandong Tian, Beidi Chen +2

    cs.CLcs.AIarXiv:2309.17453v42023
  2. Understanding the Effective Receptive Field in Deep Convolutional Neural Networks

    Wenjie Luo, Yujia Li, Raquel Urtasun +1

    cs.CVcs.AIcs.LGarXiv:1701.04128v22017
  3. Learning Convolutional Neural Networks for Graphs

    Mathias Niepert, Mohamed Ahmed, Konstantin Kutzkov

    cs.LGcs.AIstat.MLarXiv:1605.05273v42016
  4. Gemma 2: Improving Open Language Models at a Practical Size

    Gemma Team, Morgane Riviere, Shreya Pathak +195

    cs.CLcs.AIarXiv:2408.00118v32024
  5. A Structured Self-attentive Sentence Embedding

    Zhouhan Lin, Minwei Feng, Cicero Nogueira dos Santos +4

    cs.CLcs.AIcs.LGarXiv:1703.03130v12017
  6. SDEdit: Guided Image Synthesis and Editing with Stochastic Differential Equations

    Chenlin Meng, Yutong He, Yang Song +4

    cs.CVcs.AIarXiv:2108.01073v22021
  7. The FF Planning System: Fast Plan Generation Through Heuristic Search

    J. Hoffmann, B. Nebel

    cs.AIarXiv:1106.0675v12011
  8. Deep Reinforcement Learning for Autonomous Driving: A Survey

    B Ravi Kiran, Ibrahim Sobh, Victor Talpaert +4

    cs.LGcs.AIcs.ROarXiv:2002.00444v22020
  9. 4D Spatio-Temporal ConvNets: Minkowski Convolutional Neural Networks

    Christopher Choy, JunYoung Gwak, Silvio Savarese

    cs.CVcs.AIarXiv:1904.08755v42019
  10. Conditional Prompt Learning for Vision-Language Models

    Kaiyang Zhou, Jingkang Yang, Chen Change Loy +1

    cs.CVcs.AIcs.CLarXiv:2203.05557v22022
  11. Ups and Downs: Modeling the Visual Evolution of Fashion Trends with One-Class Collaborative Filtering

    Ruining He, Julian McAuley

    cs.AIcs.IRarXiv:1602.01585v12016
  12. Phi-3 Technical Report: A Highly Capable Language Model Locally on Your Phone

    Marah Abdin, Jyoti Aneja, Hany Awadalla +126

    cs.CLcs.AIarXiv:2404.14219v42024
  13. Qwen2 Technical Report

    An Yang, Baosong Yang, Binyuan Hui +59

    cs.CLcs.AIarXiv:2407.10671v42024
  14. Representation Learning on Graphs with Jumping Knowledge Networks

    Keyulu Xu, Chengtao Li, Yonglong Tian +3

    cs.LGcs.AIcs.CVarXiv:1806.03536v22018
  15. Self-RAG: Learning to Retrieve, Generate, and Critique through Self-Reflection

    Akari Asai, Zeqiu Wu, Yizhong Wang +2

    cs.CLcs.AIcs.LGarXiv:2310.11511v12023
  16. Decision Transformer: Reinforcement Learning via Sequence Modeling

    Lili Chen, Kevin Lu, Aravind Rajeswaran +6

    cs.LGcs.AIarXiv:2106.01345v22021
  17. MMMU: A Massive Multi-discipline Multimodal Understanding and Reasoning Benchmark for Expert AGI

    Xiang Yue, Yuansheng Ni, Kai Zhang +19

    cs.CLcs.AIcs.CVarXiv:2311.16502v42023
  18. MobileViT: Light-weight, General-purpose, and Mobile-friendly Vision Transformer

    Sachin Mehta, Mohammad Rastegari

    cs.CVcs.AIcs.LGarXiv:2110.02178v22021
  19. Deep Spatio-Temporal Residual Networks for Citywide Crowd Flows Prediction

    Junbo Zhang, Yu Zheng, Dekang Qi

    cs.AIcs.LGarXiv:1610.00081v22016
  20. Where Flow Matching Leaks: Characterising Membership Signals Along the Interpolation Path

    Thomas Sesmat, Gabriel Meseguer-Brocal, Geoffroy Peeters

    cs.LGcs.AIcs.SDarXiv:2606.07271v32026
  21. AirSim: High-Fidelity Visual and Physical Simulation for Autonomous Vehicles

    Shital Shah, Debadeepta Dey, Chris Lovett +1

    cs.ROcs.AIcs.CVarXiv:1705.05065v22017
  22. Continual Learning with Deep Generative Replay

    Hanul Shin, Jung Kwon Lee, Jaehong Kim +1

    cs.AIcs.CVcs.LGarXiv:1705.08690v32017
  23. TinyBERT: Distilling BERT for Natural Language Understanding

    Xiaoqi Jiao, Yichun Yin, Lifeng Shang +5

    cs.CLcs.AIcs.LGarXiv:1909.10351v52019
  24. RepVGG: Making VGG-style ConvNets Great Again

    Xiaohan Ding, Xiangyu Zhang, Ningning Ma +3

    cs.CVcs.AIcs.LGarXiv:2101.03697v32021
  25. dots.tts Technical Report

    Shi Lian, Changtao Li, Bohan Li +6

    cs.SDcs.AIeess.ASarXiv:2606.07080v22026
  26. PaperFlow: Profiling, Recommending, and Adapting Across Daily Paper Streams

    Fuqiang Wang, Song Tan, Zheng Guo +8

    cs.IRcs.AIarXiv:2606.07454v12026
  27. Mixed Precision Training

    Paulius Micikevicius, Sharan Narang, Jonah Alben +8

    cs.AIcs.LGstat.MLarXiv:1710.03740v32017
  28. SlimSearcher: Training Efficiency-Aware Web Agents via Adaptive Reward Gating

    Zequn Xie, Junjie Wang, Dan Yang +4

    cs.LGcs.AIarXiv:2606.07074v12026
  29. Set-Based Transformer for Atmospheric Compensation in Standoff LWIR Hyperspectral Imaging

    Fabian Perez, Nicolas Quintero, Jeferson Acevedo +1

    cs.CVcs.AIarXiv:2606.08324v12026
  30. Chiaroscuro Attention: Spending Compute in the Dark

    Prateek Kumar Sikdar

    cs.CLcs.AIcs.LGarXiv:2606.08327v22026
  31. Learn to Explain: Multimodal Reasoning via Thought Chains for Science Question Answering

    Pan Lu, Swaroop Mishra, Tony Xia +6

    cs.CLcs.AIcs.CVarXiv:2209.09513v22022
  32. When Behavioral Safety Evaluation Fails: A Representation-Level Perspective

    Enyi Jiang, Anders Gjølbye, Yibo Jacky Zhang +1

    cs.LGcs.AIcs.CLarXiv:2606.08044v22026
  33. PIPE-Cypher: Automatic Enterprise Benchmark Generation for Text-to-Cypher Systems

    Suraj Ranganath, Anish Raghavendra

    cs.LGcs.AIcs.DBarXiv:2606.08481v12026
  34. Beyond the Imitation Game: Quantifying and extrapolating the capabilities of language models

    Aarohi Srivastava, Abhinav Rastogi, Abhishek Rao +448

    cs.CLcs.AIcs.CYarXiv:2206.04615v32022
  35. SeqGAN: Sequence Generative Adversarial Nets with Policy Gradient

    Lantao Yu, Weinan Zhang, Jun Wang +1

    cs.LGcs.AIarXiv:1609.05473v62016
  36. OmniGameArena: A Unified UE5 Benchmark for VLM Game Agents with Improvement Dynamics

    Mingxian Lin, Shengju Qian, Yuqi Liu +9

    cs.CVcs.AIarXiv:2606.09826v12026
  37. End-to-End Context Compression at Scale

    Ang Li, Sean McLeish, Haozhe Chen +12

    cs.CLcs.AIcs.LGarXiv:2606.09659v12026
  38. Late-Layer Fusion is Enough: Dual-Path Vision Token Routing for Multimodal Large Language Models under Visual Saturation

    Siyuan Liu, Jinyang Wu

    cs.AIcs.CLcs.CVarXiv:2606.09131v12026
  39. CommonsenseQA: A Question Answering Challenge Targeting Commonsense Knowledge

    Alon Talmor, Jonathan Herzig, Nicholas Lourie +1

    cs.CLcs.AIcs.LGarXiv:1811.00937v22018
  40. $τ$-Rec: A Verifiable Benchmark for Agentic Recommender Systems

    Bharath Sivaram Narasimhan, Karthik R Narasimhan

    cs.IRcs.AIcs.CLarXiv:2606.10156v32026
  41. RT-1: Robotics Transformer for Real-World Control at Scale

    Anthony Brohan, Noah Brown, Justice Carbajal +48

    cs.ROcs.AIcs.CLarXiv:2212.06817v22022
  42. A Review of Uncertainty Quantification in Deep Learning: Techniques, Applications and Challenges

    Moloud Abdar, Farhad Pourpanah, Sadiq Hussain +9

    cs.LGcs.AIcs.CVarXiv:2011.06225v42020
  43. 1D Convolutional Neural Networks and Applications: A Survey

    Serkan Kiranyaz, Onur Avci, Osama Abdeljaber +3

    eess.SPcs.AIarXiv:1905.03554v12019
  44. Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems

    Sergey Levine, Aviral Kumar, George Tucker +1

    cs.LGcs.AIstat.MLarXiv:2005.01643v32020
  45. Counterfactual Multi-Agent Policy Gradients

    Jakob Foerster, Gregory Farquhar, Triantafyllos Afouras +2

    cs.AIcs.MAarXiv:1705.08926v32017
  46. Rainbow: Combining Improvements in Deep Reinforcement Learning

    Matteo Hessel, Joseph Modayil, Hado van Hasselt +7

    cs.AIcs.LGarXiv:1710.02298v12017
  47. Unsupervised Data Augmentation for Consistency Training

    Qizhe Xie, Zihang Dai, Eduard Hovy +2

    cs.LGcs.AIcs.CLarXiv:1904.12848v62019
  48. Revisiting Unreasonable Effectiveness of Data in Deep Learning Era

    Chen Sun, Abhinav Shrivastava, Saurabh Singh +1

    cs.CVcs.AIarXiv:1707.02968v22017
  49. Probabilistic Circuits as Reasoning Machines in Artificial Intelligence (Part I)

    Robert Peharz

    cs.AIcs.LGmath.PRarXiv:2608.16565v12026
    Summaries:한국어
  50. Lean Refactor: Multi-Objective Controllable Proof Optimization via Agentic Strategy Search

    Jialin Lu, Soonho Kong, Rodrigo Stehling +4

    cs.LOcs.AIcs.CLarXiv:2605.20244v22026
  51. Reasoning Arena: Trace Tournaments When Verifiable Rewards Fall Short

    Han Zhou, Adam X. Yang, Laurence Aitchison +2

    cs.LGcs.AIcs.CLarXiv:2606.09380v12026
  52. Optical Reasoning: Rethinking Images as an Expressive Reasoning Medium Beyond Text

    Yutong Bian, Dongjie Cheng, Heming Xia +2

    cs.AIarXiv:2606.09585v12026
  53. Interpreting and Steering a Text-to-Speech Language Model with Sparse Autoencoders

    Nikita Koriagin, Georgii Aparin, Nikita Balagansky +1

    cs.LGcs.AIcs.CLarXiv:2606.10029v12026
  54. Your Language Model is Its Own Critic: Reinforcement Learning with Value Estimation from Actor's Internal States

    Yunho Choi, Jongwon Lim, Woojin Ahn +3

    cs.LGcs.AIcs.CLarXiv:2605.07579v22026
  55. Hindsight Experience Replay

    Marcin Andrychowicz, Filip Wolski, Alex Ray +7

    cs.LGcs.AIcs.NEarXiv:1707.01495v32017
  56. TRIAGE: Dialectical Reasoning for Explainable Risk Prediction on Irregularly Sampled Medical Time Series with LLMs

    Hyeongwon Jang, Gyouk Chu, Changhun Kim +3

    cs.LGcs.AIcs.CLarXiv:2606.09030v12026
  57. Event-based Vision: A Survey

    Guillermo Gallego, Tobi Delbruck, Garrick Orchard +8

    cs.CVcs.AIcs.LGarXiv:1904.08405v32019
  58. Video Diffusion Models

    Jonathan Ho, Tim Salimans, Alexey Gritsenko +3

    cs.CVcs.AIcs.LGarXiv:2204.03458v22022
  59. Shallow Prefill, Deep Decoding: Efficient Long-Context Inference via Layer-Asymmetric KV Visibility

    Jungsuk Oh, Hyeseo Jeon, Hyunjune Ji +2

    cs.AIarXiv:2605.06105v12026
  60. Universal adversarial perturbations

    Seyed-Mohsen Moosavi-Dezfooli, Alhussein Fawzi, Omar Fawzi +1

    cs.CVcs.AIcs.LGarXiv:1610.08401v32016