Artificial Intelligence

Papers filed under cs.AI on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.

Search paper metadata (including unsummarized papers)

6,181 to 6,240 of 15,269

  1. Graph Neural Tangent Kernel: Fusing Graph Neural Networks with Graph Kernels

    Simon S. Du, Kangcheng Hou, Barnabás Póczos +3

    cs.LGcs.AIcs.CVarXiv:1905.13192v22019
  2. Text Classification Algorithms: A Survey

    Kamran Kowsari, Kiana Jafari Meimandi, Mojtaba Heidarysafa +3

    cs.LGcs.AIcs.CLarXiv:1904.08067v52019
  3. Deep Reinforcement Learning for Sepsis Treatment

    Aniruddh Raghu, Matthieu Komorowski, Imran Ahmed +3

    cs.AIcs.LGarXiv:1711.09602v12017
  4. Mercury: Ultra-Fast Language Models Based on Diffusion

    Inception Labs, Samar Khanna, Siddhant Kharbanda +10

    cs.CLcs.AIcs.LGarXiv:2506.17298v12025
  5. A Survey on Traffic Signal Control Methods

    Hua Wei, Guanjie Zheng, Vikash Gayah +1

    cs.LGcs.AIstat.MLarXiv:1904.08117v32019
  6. Three scenarios for continual learning

    Gido M. van de Ven, Andreas S. Tolias

    cs.LGcs.AIcs.CVarXiv:1904.07734v12019
  7. Large Batch Optimization for Deep Learning: Training BERT in 76 minutes

    Yang You, Jing Li, Sashank Reddi +7

    cs.LGcs.AIcs.CLarXiv:1904.00962v52019
  8. Soft Actor-Critic Algorithms and Applications

    Tuomas Haarnoja, Aurick Zhou, Kristian Hartikainen +8

    cs.LGcs.AIcs.ROarXiv:1812.05905v22018
  9. The relativistic discriminator: a key element missing from standard GAN

    Alexia Jolicoeur-Martineau

    cs.LGcs.AIcs.CRarXiv:1807.00734v32018
  10. Understanding Batch Normalization

    Johan Bjorck, Carla Gomes, Bart Selman +1

    cs.LGcs.AIstat.MLarXiv:1806.02375v42018
  11. Deep Sequence Learning with Auxiliary Information for Traffic Prediction

    Binbing Liao, Jingqing Zhang, Chao Wu +5

    cs.CVcs.AIarXiv:1806.07380v12018
  12. TADAM: Task dependent adaptive metric for improved few-shot learning

    Boris N. Oreshkin, Pau Rodriguez, Alexandre Lacoste

    cs.LGcs.AIcs.CVarXiv:1805.10123v42018
  13. Data-Efficient Hierarchical Reinforcement Learning

    Ofir Nachum, Shixiang Gu, Honglak Lee +1

    cs.LGcs.AIstat.MLarXiv:1805.08296v42018
  14. Selective Experience Replay for Lifelong Learning

    David Isele, Akansel Cosgun

    cs.AIarXiv:1802.10269v12018
  15. Mean Field Multi-Agent Reinforcement Learning

    Yaodong Yang, Rui Luo, Minne Li +3

    cs.MAcs.AIcs.LGarXiv:1802.05438v52018
  16. Reasoning with Latent Thoughts: On the Power of Looped Transformers

    Nikunj Saunshi, Nishanth Dikkala, Zhiyuan Li +2

    cs.CLcs.AIcs.LGarXiv:2502.17416v12025
  17. Deep Learning for Physical Processes: Incorporating Prior Scientific Knowledge

    Emmanuel de Bezenac, Arthur Pajot, Patrick Gallinari

    cs.AIcs.LGstat.MLarXiv:1711.07970v22017
  18. $Q$- and $A$-Learning Methods for Estimating Optimal Dynamic Treatment Regimes

    Phillip J. Schulte, Anastasios A. Tsiatis, Eric B. Laber +1

    stat.MEcs.AIarXiv:1202.4177v32012
  19. Ensembles of Multiple Models and Architectures for Robust Brain Tumour Segmentation

    Konstantinos Kamnitsas, Wenjia Bai, Enzo Ferrante +8

    cs.CVcs.AIcs.LGarXiv:1711.01468v12017
  20. A systematic study of the class imbalance problem in convolutional neural networks

    Mateusz Buda, Atsuto Maki, Maciej A. Mazurowski

    cs.CVcs.AIcs.LGarXiv:1710.05381v22017
  21. Safe and Nested Subgame Solving for Imperfect-Information Games

    Noam Brown, Tuomas Sandholm

    cs.AIcs.GTarXiv:1705.02955v32017
  22. Minimax Regret Bounds for Reinforcement Learning

    Mohammad Gheshlaghi Azar, Ian Osband, Rémi Munos

    stat.MLcs.AIcs.LGarXiv:1703.05449v22017
  23. Bridging the Gap Between Value and Policy Based Reinforcement Learning

    Ofir Nachum, Mohammad Norouzi, Kelvin Xu +1

    cs.AIcs.LGstat.MLarXiv:1702.08892v32017
  24. MemAgent: Reshaping Long-Context LLM with Multi-Conv RL-based Memory Agent

    Hongli Yu, Tinghong Chen, Jiangtao Feng +8

    cs.CLcs.AIcs.LGarXiv:2507.02259v22025
  25. Generative Adversarial Imitation Learning

    Jonathan Ho, Stefano Ermon

    cs.LGcs.AIarXiv:1606.03476v12016
  26. Unifying Count-Based Exploration and Intrinsic Motivation

    Marc G. Bellemare, Sriram Srinivasan, Georg Ostrovski +3

    cs.AIcs.LGstat.MLarXiv:1606.01868v22016
  27. Adversarial Feature Learning

    Jeff Donahue, Philipp Krähenbühl, Trevor Darrell

    cs.LGcs.AIcs.CVarXiv:1605.09782v72016
  28. Efficient Multi-Scale 3D CNN with Fully Connected CRF for Accurate Brain Lesion Segmentation

    Konstantinos Kamnitsas, Christian Ledig, Virginia F. J. Newcombe +5

    cs.CVcs.AIarXiv:1603.05959v32016
  29. DeepResearcher: Scaling Deep Research via Reinforcement Learning in Real-world Environments

    Yuxiang Zheng, Dayuan Fu, Xiangkun Hu +4

    cs.AIcs.CLcs.LGarXiv:2504.03160v42025
  30. Gated Graph Sequence Neural Networks

    Yujia Li, Daniel Tarlow, Marc Brockschmidt +1

    cs.LGcs.AIcs.NEarXiv:1511.05493v42015
  31. The Ubuntu Dialogue Corpus: A Large Dataset for Research in Unstructured Multi-Turn Dialogue Systems

    Ryan Lowe, Nissan Pow, Iulian Serban +1

    cs.CLcs.AIcs.LGarXiv:1506.08909v32015
  32. Hinge-Loss Markov Random Fields and Probabilistic Soft Logic

    Stephen H. Bach, Matthias Broecheler, Bert Huang +1

    cs.LGcs.AIstat.MLarXiv:1505.04406v32015
  33. Learning Dependency-Based Compositional Semantics

    Percy Liang, Michael I. Jordan, Dan Klein

    cs.AIarXiv:1109.6841v12011
  34. HiLRP: Toward One Trustworthy Explanation for Vision Transformer: Conservation-Valid Attribution via Attention Primitives

    Sathiyamohan Nishankar, Pubudu Sanjeewani, Asanka Perera +1

    cs.CVcs.AIarXiv:2609.01282v12026
  35. StainPresetNet: Stain Preset Network for Fast Multi-to-Multi Stain Normalization

    Hongtao Kang, Die Luo, Li Chen +4

    cs.CVcs.AIarXiv:2609.01146v12026
  36. Adaptive Submodularity: Theory and Applications in Active Learning and Stochastic Optimization

    Daniel Golovin, Andreas Krause

    cs.LGcs.AIcs.DSarXiv:1003.3967v52010
  37. HVI: A New Color Space for Low-light Image Enhancement

    Qingsen Yan, Yixu Feng, Cheng Zhang +6

    cs.CVcs.AIcs.LGarXiv:2502.20272v22025
  38. Monkeypox Skin Lesion Detection Using Deep Learning Models: A Feasibility Study

    Shams Nafisa Ali, Md. Tazuddin Ahmed, Joydip Paul +4

    cs.CVcs.AIeess.IVarXiv:2207.03342v12022
  39. Solaris: Towards Interfaces That Are Generated, Not Coded

    Yuval Alaluf, Omri Avrahami, Guy Bukchin Leshem +18

    cs.CVcs.AIarXiv:2609.00776v12026
  40. AgentFactory: Towards Automated Agentic System Design and Optimization

    Enci Zhang, Haofeng Wang, Yuesheng Zhu +2

    cs.AIarXiv:2609.01045v12026
  41. WISE: A World Knowledge-Informed Semantic Evaluation for Text-to-Image Generation

    Yuwei Niu, Munan Ning, Mengren Zheng +9

    cs.CVcs.AIcs.CLarXiv:2503.07265v42025
  42. Machine Learning for Wireless Connectivity and Security of Cellular-Connected UAVs

    Ursula Challita, Aidin Ferdowsi, Mingzhe Chen +1

    cs.ITcs.AIarXiv:1804.05348v32018
  43. Scalable Best-of-N Selection for Large Language Models via Self-Certainty

    Zhewei Kang, Xuandong Zhao, Dawn Song

    cs.CLcs.AIcs.LGarXiv:2502.18581v32025
  44. SWE-RL: Advancing LLM Reasoning via Reinforcement Learning on Open Software Evolution

    Yuxiang Wei, Olivier Duchenne, Jade Copet +6

    cs.SEcs.AIcs.CLarXiv:2502.18449v22025
  45. Instella-MoE Technical Report

    Jiang Liu, Sudhanshu Ranjan, Prakamya Mishra +10

    cs.CLcs.AIarXiv:2609.00791v12026
  46. Hi Robot: Open-Ended Instruction Following with Hierarchical Vision-Language-Action Models

    Lucy Xiaoyang Shi, Brian Ichter, Michael Equi +12

    cs.ROcs.AIcs.LGarXiv:2502.19417v22025
  47. TripoSG: High-Fidelity 3D Shape Synthesis using Large-Scale Rectified Flow Models

    Yangguang Li, Zi-Xin Zou, Zexiang Liu +8

    cs.CVcs.AIarXiv:2502.06608v32025
  48. ASAP: Aligning Simulation and Real-World Physics for Learning Agile Humanoid Whole-Body Skills

    Tairan He, Jiawei Gao, Wenli Xiao +15

    cs.ROcs.AIcs.LGarXiv:2502.01143v32025
  49. LMM-R1: Empowering 3B LMMs with Strong Reasoning Abilities Through Two-Stage Rule-Based RL

    Yingzhe Peng, Gongrui Zhang, Miaosen Zhang +7

    cs.CLcs.AIarXiv:2503.07536v22025
  50. Learning to Branch

    Maria-Florina Balcan, Travis Dick, Tuomas Sandholm +1

    cs.AIcs.DSarXiv:1803.10150v22018
  51. Goedel-Prover-V2: Scaling Formal Theorem Proving with Scaffolded Data Synthesis and Self-Correction

    Yong Lin, Shange Tang, Bohan Lyu +17

    cs.LGcs.AIarXiv:2508.03613v12025
  52. Meta-ethics and AI: exploring the novel meta-ethical questions in the era of AI

    Shang Lu

    cs.AIcs.CYarXiv:2609.01685v12026
  53. AI as Teammate: Rethinking Task Distribution in Medical Training

    Fendi Tsim, Alina Gutoreva, Anthony Weiss +1

    cs.HCcs.AIarXiv:2608.28373v12026
  54. Can You Say This for Me? Speaking Up by Proxy in Co-Located Discussion

    Yue Shen, Rehema Abulikemu, Ryan P. McMahan +1

    cs.AIcs.ETcs.HCarXiv:2608.26185v12026
  55. Autonomous Mathematical Discovery in an Open-World Multi-Agent Environment

    Stephen Chung, Wenyu Du, William J. Wesley

    cs.AIcs.DMcs.MAarXiv:2608.23691v12026
  56. Safety Hacking in Constrained Best-of-$N$ Inference-time Scaling

    Akifumi Wachi, Takumi Tanabe, Youhei Akimoto

    cs.LGcs.AIcs.CLarXiv:2608.22915v12026
  57. GuardPaint:SpeculativeSafetyDecodingforText-to-ImageGeneration

    Shreyash Dhoot, Paras Dhiman, Arsh Abbas Naqvi +4

    cs.CVcs.AIarXiv:2608.21869v12026
  58. ATHENA: Knowledge-guided agentic neural architecture search for AutoFormer-based electronic health record modeling

    Deyi Li, Qi Xu, Lingyao Li +3

    cs.AIcs.MAarXiv:2608.21712v12026
  59. Testing and Evaluation of Agentic AI Systems In Military Command and Control

    Ulysse Richard, Heather Frase, Sarah Cao +3

    cs.SEcs.AIcs.CYarXiv:2608.20597v12026
  60. Provable Edge-of-Stability for Adam on a One-Dimensional Quadratic

    Yiman Fong, Heng Yang

    cs.LGcs.AImath.OCarXiv:2608.20638v12026