Artificial Intelligence
Papers filed under cs.AI on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.
Search paper metadata (including unsummarized papers)
6,181 to 6,240 of 15,269
Graph Neural Tangent Kernel: Fusing Graph Neural Networks with Graph Kernels
Simon S. Du, Kangcheng Hou, Barnabás Póczos +3
cs.LGcs.AIcs.CVarXiv:1905.13192v22019Text Classification Algorithms: A Survey
Kamran Kowsari, Kiana Jafari Meimandi, Mojtaba Heidarysafa +3
cs.LGcs.AIcs.CLarXiv:1904.08067v52019Deep Reinforcement Learning for Sepsis Treatment
Aniruddh Raghu, Matthieu Komorowski, Imran Ahmed +3
cs.AIcs.LGarXiv:1711.09602v12017Mercury: Ultra-Fast Language Models Based on Diffusion
Inception Labs, Samar Khanna, Siddhant Kharbanda +10
cs.CLcs.AIcs.LGarXiv:2506.17298v12025A Survey on Traffic Signal Control Methods
Hua Wei, Guanjie Zheng, Vikash Gayah +1
cs.LGcs.AIstat.MLarXiv:1904.08117v32019Three scenarios for continual learning
Gido M. van de Ven, Andreas S. Tolias
cs.LGcs.AIcs.CVarXiv:1904.07734v12019Large Batch Optimization for Deep Learning: Training BERT in 76 minutes
Yang You, Jing Li, Sashank Reddi +7
cs.LGcs.AIcs.CLarXiv:1904.00962v52019Soft Actor-Critic Algorithms and Applications
Tuomas Haarnoja, Aurick Zhou, Kristian Hartikainen +8
cs.LGcs.AIcs.ROarXiv:1812.05905v22018The relativistic discriminator: a key element missing from standard GAN
Alexia Jolicoeur-Martineau
cs.LGcs.AIcs.CRarXiv:1807.00734v32018Understanding Batch Normalization
Johan Bjorck, Carla Gomes, Bart Selman +1
cs.LGcs.AIstat.MLarXiv:1806.02375v42018Deep Sequence Learning with Auxiliary Information for Traffic Prediction
Binbing Liao, Jingqing Zhang, Chao Wu +5
cs.CVcs.AIarXiv:1806.07380v12018TADAM: Task dependent adaptive metric for improved few-shot learning
Boris N. Oreshkin, Pau Rodriguez, Alexandre Lacoste
cs.LGcs.AIcs.CVarXiv:1805.10123v42018Data-Efficient Hierarchical Reinforcement Learning
Ofir Nachum, Shixiang Gu, Honglak Lee +1
cs.LGcs.AIstat.MLarXiv:1805.08296v42018Selective Experience Replay for Lifelong Learning
David Isele, Akansel Cosgun
cs.AIarXiv:1802.10269v12018Mean Field Multi-Agent Reinforcement Learning
Yaodong Yang, Rui Luo, Minne Li +3
cs.MAcs.AIcs.LGarXiv:1802.05438v52018Reasoning with Latent Thoughts: On the Power of Looped Transformers
Nikunj Saunshi, Nishanth Dikkala, Zhiyuan Li +2
cs.CLcs.AIcs.LGarXiv:2502.17416v12025Deep Learning for Physical Processes: Incorporating Prior Scientific Knowledge
Emmanuel de Bezenac, Arthur Pajot, Patrick Gallinari
cs.AIcs.LGstat.MLarXiv:1711.07970v22017$Q$- and $A$-Learning Methods for Estimating Optimal Dynamic Treatment Regimes
Phillip J. Schulte, Anastasios A. Tsiatis, Eric B. Laber +1
stat.MEcs.AIarXiv:1202.4177v32012Ensembles of Multiple Models and Architectures for Robust Brain Tumour Segmentation
Konstantinos Kamnitsas, Wenjia Bai, Enzo Ferrante +8
cs.CVcs.AIcs.LGarXiv:1711.01468v12017A systematic study of the class imbalance problem in convolutional neural networks
Mateusz Buda, Atsuto Maki, Maciej A. Mazurowski
cs.CVcs.AIcs.LGarXiv:1710.05381v22017Safe and Nested Subgame Solving for Imperfect-Information Games
Noam Brown, Tuomas Sandholm
cs.AIcs.GTarXiv:1705.02955v32017Minimax Regret Bounds for Reinforcement Learning
Mohammad Gheshlaghi Azar, Ian Osband, Rémi Munos
stat.MLcs.AIcs.LGarXiv:1703.05449v22017Bridging the Gap Between Value and Policy Based Reinforcement Learning
Ofir Nachum, Mohammad Norouzi, Kelvin Xu +1
cs.AIcs.LGstat.MLarXiv:1702.08892v32017MemAgent: Reshaping Long-Context LLM with Multi-Conv RL-based Memory Agent
Hongli Yu, Tinghong Chen, Jiangtao Feng +8
cs.CLcs.AIcs.LGarXiv:2507.02259v22025Generative Adversarial Imitation Learning
Jonathan Ho, Stefano Ermon
cs.LGcs.AIarXiv:1606.03476v12016Unifying Count-Based Exploration and Intrinsic Motivation
Marc G. Bellemare, Sriram Srinivasan, Georg Ostrovski +3
cs.AIcs.LGstat.MLarXiv:1606.01868v22016Adversarial Feature Learning
Jeff Donahue, Philipp Krähenbühl, Trevor Darrell
cs.LGcs.AIcs.CVarXiv:1605.09782v72016Efficient Multi-Scale 3D CNN with Fully Connected CRF for Accurate Brain Lesion Segmentation
Konstantinos Kamnitsas, Christian Ledig, Virginia F. J. Newcombe +5
cs.CVcs.AIarXiv:1603.05959v32016DeepResearcher: Scaling Deep Research via Reinforcement Learning in Real-world Environments
Yuxiang Zheng, Dayuan Fu, Xiangkun Hu +4
cs.AIcs.CLcs.LGarXiv:2504.03160v42025Gated Graph Sequence Neural Networks
Yujia Li, Daniel Tarlow, Marc Brockschmidt +1
cs.LGcs.AIcs.NEarXiv:1511.05493v42015The Ubuntu Dialogue Corpus: A Large Dataset for Research in Unstructured Multi-Turn Dialogue Systems
Ryan Lowe, Nissan Pow, Iulian Serban +1
cs.CLcs.AIcs.LGarXiv:1506.08909v32015Hinge-Loss Markov Random Fields and Probabilistic Soft Logic
Stephen H. Bach, Matthias Broecheler, Bert Huang +1
cs.LGcs.AIstat.MLarXiv:1505.04406v32015Learning Dependency-Based Compositional Semantics
Percy Liang, Michael I. Jordan, Dan Klein
cs.AIarXiv:1109.6841v12011HiLRP: Toward One Trustworthy Explanation for Vision Transformer: Conservation-Valid Attribution via Attention Primitives
Sathiyamohan Nishankar, Pubudu Sanjeewani, Asanka Perera +1
cs.CVcs.AIarXiv:2609.01282v12026StainPresetNet: Stain Preset Network for Fast Multi-to-Multi Stain Normalization
Hongtao Kang, Die Luo, Li Chen +4
cs.CVcs.AIarXiv:2609.01146v12026Adaptive Submodularity: Theory and Applications in Active Learning and Stochastic Optimization
Daniel Golovin, Andreas Krause
cs.LGcs.AIcs.DSarXiv:1003.3967v52010HVI: A New Color Space for Low-light Image Enhancement
Qingsen Yan, Yixu Feng, Cheng Zhang +6
cs.CVcs.AIcs.LGarXiv:2502.20272v22025Monkeypox Skin Lesion Detection Using Deep Learning Models: A Feasibility Study
Shams Nafisa Ali, Md. Tazuddin Ahmed, Joydip Paul +4
cs.CVcs.AIeess.IVarXiv:2207.03342v12022Solaris: Towards Interfaces That Are Generated, Not Coded
Yuval Alaluf, Omri Avrahami, Guy Bukchin Leshem +18
cs.CVcs.AIarXiv:2609.00776v12026AgentFactory: Towards Automated Agentic System Design and Optimization
Enci Zhang, Haofeng Wang, Yuesheng Zhu +2
cs.AIarXiv:2609.01045v12026WISE: A World Knowledge-Informed Semantic Evaluation for Text-to-Image Generation
Yuwei Niu, Munan Ning, Mengren Zheng +9
cs.CVcs.AIcs.CLarXiv:2503.07265v42025Machine Learning for Wireless Connectivity and Security of Cellular-Connected UAVs
Ursula Challita, Aidin Ferdowsi, Mingzhe Chen +1
cs.ITcs.AIarXiv:1804.05348v32018Scalable Best-of-N Selection for Large Language Models via Self-Certainty
Zhewei Kang, Xuandong Zhao, Dawn Song
cs.CLcs.AIcs.LGarXiv:2502.18581v32025SWE-RL: Advancing LLM Reasoning via Reinforcement Learning on Open Software Evolution
Yuxiang Wei, Olivier Duchenne, Jade Copet +6
cs.SEcs.AIcs.CLarXiv:2502.18449v22025Instella-MoE Technical Report
Jiang Liu, Sudhanshu Ranjan, Prakamya Mishra +10
cs.CLcs.AIarXiv:2609.00791v12026Hi Robot: Open-Ended Instruction Following with Hierarchical Vision-Language-Action Models
Lucy Xiaoyang Shi, Brian Ichter, Michael Equi +12
cs.ROcs.AIcs.LGarXiv:2502.19417v22025TripoSG: High-Fidelity 3D Shape Synthesis using Large-Scale Rectified Flow Models
Yangguang Li, Zi-Xin Zou, Zexiang Liu +8
cs.CVcs.AIarXiv:2502.06608v32025ASAP: Aligning Simulation and Real-World Physics for Learning Agile Humanoid Whole-Body Skills
Tairan He, Jiawei Gao, Wenli Xiao +15
cs.ROcs.AIcs.LGarXiv:2502.01143v32025LMM-R1: Empowering 3B LMMs with Strong Reasoning Abilities Through Two-Stage Rule-Based RL
Yingzhe Peng, Gongrui Zhang, Miaosen Zhang +7
cs.CLcs.AIarXiv:2503.07536v22025Learning to Branch
Maria-Florina Balcan, Travis Dick, Tuomas Sandholm +1
cs.AIcs.DSarXiv:1803.10150v22018Goedel-Prover-V2: Scaling Formal Theorem Proving with Scaffolded Data Synthesis and Self-Correction
Yong Lin, Shange Tang, Bohan Lyu +17
cs.LGcs.AIarXiv:2508.03613v12025Meta-ethics and AI: exploring the novel meta-ethical questions in the era of AI
Shang Lu
cs.AIcs.CYarXiv:2609.01685v12026AI as Teammate: Rethinking Task Distribution in Medical Training
Fendi Tsim, Alina Gutoreva, Anthony Weiss +1
cs.HCcs.AIarXiv:2608.28373v12026Can You Say This for Me? Speaking Up by Proxy in Co-Located Discussion
Yue Shen, Rehema Abulikemu, Ryan P. McMahan +1
cs.AIcs.ETcs.HCarXiv:2608.26185v12026Autonomous Mathematical Discovery in an Open-World Multi-Agent Environment
Stephen Chung, Wenyu Du, William J. Wesley
cs.AIcs.DMcs.MAarXiv:2608.23691v12026Safety Hacking in Constrained Best-of-$N$ Inference-time Scaling
Akifumi Wachi, Takumi Tanabe, Youhei Akimoto
cs.LGcs.AIcs.CLarXiv:2608.22915v12026GuardPaint:SpeculativeSafetyDecodingforText-to-ImageGeneration
Shreyash Dhoot, Paras Dhiman, Arsh Abbas Naqvi +4
cs.CVcs.AIarXiv:2608.21869v12026ATHENA: Knowledge-guided agentic neural architecture search for AutoFormer-based electronic health record modeling
Deyi Li, Qi Xu, Lingyao Li +3
cs.AIcs.MAarXiv:2608.21712v12026Testing and Evaluation of Agentic AI Systems In Military Command and Control
Ulysse Richard, Heather Frase, Sarah Cao +3
cs.SEcs.AIcs.CYarXiv:2608.20597v12026Provable Edge-of-Stability for Adam on a One-Dimensional Quadratic
Yiman Fong, Heng Yang
cs.LGcs.AImath.OCarXiv:2608.20638v12026