Artificial Intelligence
Papers filed under cs.AI on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.
Search paper metadata (including unsummarized papers)
4,441 to 4,500 of 15,280
Path Planning for Masked Diffusion Model Sampling
Fred Zhangzhi Peng, Zachary Bezemek, Sawan Patel +5
cs.LGcs.AIarXiv:2502.03540v52025LLMs4OL: Large Language Models for Ontology Learning
Hamed Babaei Giglou, Jennifer D'Souza, Sören Auer
cs.AIcs.CLcs.ITarXiv:2307.16648v22023SkillGLoW: Procedural-Family Skill Consolidation for Self-Improving Agents on Long-Horizon Task Streams
Ao Yan, Xin Zhang, Jiawei Du +1
cs.AIarXiv:2609.02217v12026Recent Advances in Discrete Speech Tokens: A Review
Yiwei Guo, Zhihan Li, Hankun Wang +7
eess.AScs.AIcs.MMarXiv:2502.06490v42025A Framework for Human Evaluation of Large Language Models in Healthcare Derived from Literature Review
Thomas Yu Chow Tam, Sonish Sivarajkumar, Sumit Kapoor +12
cs.CLcs.AIarXiv:2405.02559v22024SIoU Loss: More Powerful Learning for Bounding Box Regression
Zhora Gevorgyan
cs.CVcs.AIarXiv:2205.12740v12022DRIFT: Dynamic Rule-Based Defense with Injection Isolation for Securing LLM Agents
Hao Li, Xiaogeng Liu, Hung-Chun Chiu +3
cs.CRcs.AIarXiv:2506.12104v32025Ethical and social risks of harm from Language Models
Laura Weidinger, John Mellor, Maribeth Rauh +20
cs.CLcs.AIcs.CYarXiv:2112.04359v12021RLVMR: Reinforcement Learning with Verifiable Meta-Reasoning Rewards for Robust Long-Horizon Agents
Zijing Zhang, Ziyang Chen, Mingxiao Li +2
cs.LGcs.AIarXiv:2507.22844v12025Minimizing the Accumulated Trajectory Error to Improve Dataset Distillation
Jiawei Du, Yidi Jiang, Vincent Y. F. Tan +2
cs.LGcs.AIcs.CVarXiv:2211.11004v32022Vision-Language Models are Zero-Shot Reward Models for Reinforcement Learning
Juan Rocamonde, Victoriano Montesinos, Elvis Nava +2
cs.LGcs.AIarXiv:2310.12921v22023The pitfalls of next-token prediction
Gregor Bachmann, Vaishnavh Nagarajan
cs.CLcs.AIcs.LGarXiv:2403.06963v32024PathRAG: Pruning Graph-based Retrieval Augmented Generation with Relational Paths
Boyu Chen, Zirui Guo, Zidan Yang +5
cs.CLcs.AIcs.IRarXiv:2502.14902v22025SAEs Are Good for Steering -- If You Select the Right Features
Dana Arad, Aaron Mueller, Yonatan Belinkov
cs.LGcs.AIcs.CLarXiv:2505.20063v22025CUDA-L1: Improving CUDA Optimization via Contrastive Reinforcement Learning
Xiaoya Li, Albert Wang, Guoyin Wang +2
cs.AIcs.DCcs.LGarXiv:2507.14111v122025PlanGenLLMs: A Modern Survey of LLM Planning Capabilities
Hui Wei, Zihao Zhang, Shenghua He +3
cs.AIcs.CLarXiv:2502.11221v32025StableNormal: Reducing Diffusion Variance for Stable and Sharp Normal
Chongjie Ye, Lingteng Qiu, Xiaodong Gu +6
cs.CVcs.AIcs.GRarXiv:2406.16864v12024Detection of Chagas Disease from the ECG: The George B. Moody PhysioNet Challenge 2025
Matthew A. Reyna, Zuzana Koscova, Jan Pavlus +11
cs.LGcs.AIarXiv:2510.02202v12025Robust Autonomy Emerges from Self-Play
Marco Cusumano-Towner, David Hafner, Alex Hertzberg +9
cs.LGcs.AIcs.ROarXiv:2502.03349v12025SkillForge: Forging Domain-Specific, Self-Evolving Agent Skills in Cloud Technical Support
Xingyan Liu, Xiyue Luo, Linyu Li +3
cs.IRcs.AIcs.SEarXiv:2604.08618v22026CLIP-It! Language-Guided Video Summarization
Medhini Narasimhan, Anna Rohrbach, Trevor Darrell
cs.CVcs.AIcs.MMarXiv:2107.00650v22021Reasoning Models Better Express Their Confidence
Dongkeun Yoon, Seungone Kim, Sohee Yang +6
cs.AIcs.CLarXiv:2505.14489v22025UI-Vision: A Desktop-centric GUI Benchmark for Visual Perception and Interaction
Shravan Nayak, Xiangru Jian, Kevin Qinghong Lin +11
cs.CVcs.AIcs.CLarXiv:2503.15661v22025Vision-R1: Evolving Human-Free Alignment in Large Vision-Language Models via Vision-Guided Reinforcement Learning
Yufei Zhan, Yousong Zhu, Shurong Zheng +4
cs.CVcs.AIarXiv:2503.18013v12025CLAWS: Clustering Assisted Weakly Supervised Learning with Normalcy Suppression for Anomalous Event Detection
Muhammad Zaigham Zaheer, Arif Mahmood, Marcella Astrid +1
cs.CVcs.AIarXiv:2011.12077v42020ML-Master: Towards AI-for-AI via Integration of Exploration and Reasoning
Zexi Liu, Yuzhu Cai, Xinyu Zhu +6
cs.AIcs.LGarXiv:2506.16499v12025Conservative set valued fields, automatic differentiation, stochastic gradient method and deep learning
Jérôme Bolte, Edouard Pauwels
math.OCcs.AIcs.LGarXiv:1909.10300v42019What Makes a Reward Model a Good Teacher? An Optimization Perspective
Noam Razin, Zixuan Wang, Hubert Strauss +3
cs.LGcs.AIcs.CLarXiv:2503.15477v42025Performance assessment and exhaustive listing of 500+ nature inspired metaheuristic algorithms
Zhongqiang Ma, Guohua Wu, Ponnuthurai N. Suganthan +2
cs.NEcs.AIarXiv:2212.09479v12022Agent Learning via Early Experience
Kai Zhang, Xiangchao Chen, Bo Liu +27
cs.AIcs.CLcs.IRarXiv:2510.08558v32025Vending-Bench: A Benchmark for Long-Term Coherence of Autonomous Agents
Axel Backlund, Lukas Petersson
cs.AIarXiv:2502.15840v12025DexGrasp Anything: Towards Universal Robotic Dexterous Grasping with Physics Awareness
Yiming Zhong, Qi Jiang, Jingyi Yu +1
cs.CVcs.AIcs.ROarXiv:2503.08257v22025Comprehensive Comparative Study of Multi-Label Classification Methods
Jasmin Bogatinovski, Ljupčo Todorovski, Sašo Džeroski +1
cs.LGcs.AIcs.CCarXiv:2102.07113v22021BridgeVLA: Input-Output Alignment for Efficient 3D Manipulation Learning with Vision-Language Models
Peiyan Li, Yixiang Chen, Hongtao Wu +6
cs.ROcs.AIarXiv:2506.07961v22025What does it mean to solve the problem of discrimination in hiring? Social, technical and legal perspectives from the UK on automated hiring systems
Javier Sanchez-Monedero, Lina Dencik, Lilian Edwards
cs.CYcs.AIarXiv:1910.06144v22019SAeUron: Interpretable Concept Unlearning in Diffusion Models with Sparse Autoencoders
Bartosz Cywiński, Kamil Deja
cs.LGcs.AIarXiv:2501.18052v32025Can LLMs Design Video Coding Tools? A Case Study on Planar Mode
Yingwen Zhang, Meng Wang, Liqiang He +1
cs.MMcs.AIarXiv:2609.01535v12026Defense-as-Skill: Evolving Runtime Guard Skill for Skill-Augmented Agents
Xiaofang Yang, Ziqi Miao, Dianbo Sui +2
cs.CRcs.AIarXiv:2609.01487v12026Auditing language models for hidden objectives
Samuel Marks, Johannes Treutlein, Trenton Bricken +32
cs.AIcs.CLcs.LGarXiv:2503.10965v22025Semantic-Guided Multimodal Preprocessing for Vision Transformer-Based Clear Cell Renal Cell Carcinoma Grading
Fatemeh Javadian, Zhu Chen, Zahra Aminparast +1
cs.CVcs.AIcs.LGarXiv:2609.01426v12026Measuring consistency via ensemble margin and local prediction variability: Auditing decision systems in the presence of predictive multiplicity
Sinjini Banerjee, Tim Marrinan, Anand D. Sarwate
stat.MLcs.AIcs.LGarXiv:2609.01397v12026One-Prompt-One-Story: Free-Lunch Consistent Text-to-Image Generation Using a Single Prompt
Tao Liu, Kai Wang, Senmao Li +6
cs.CVcs.AIcs.LGarXiv:2501.13554v32025"I'm Not Sure, But...": Examining the Impact of Large Language Models' Uncertainty Expression on User Reliance and Trust
Sunnie S. Y. Kim, Q. Vera Liao, Mihaela Vorvoreanu +2
cs.HCcs.AIarXiv:2405.00623v22024The zbMATH Open Knowledge Graph: Tracing Centuries of Mathematical Research
Yuni Susanti, Moritz Schubotz
cs.DLcs.AIarXiv:2609.00969v12026GUI-G1: Understanding R1-Zero-Like Training for Visual Grounding in GUI Agents
Yuqi Zhou, Sunhao Dai, Shuai Wang +3
cs.CLcs.AIcs.CVarXiv:2505.15810v22025DualStake: Dual-Path Confidence Calibration in Deep Research Agents
Yinuo Xu, Yuwei Liang, Jianjie Cheng +4
cs.CLcs.AIcs.LGarXiv:2609.00935v12026From Assistant to Double Agent: Formalizing and Benchmarking Attacks on OpenClaw for Personalized Local AI Agent
Yuhang Wang, Feiming Xu, Zheng Lin +6
cs.AIarXiv:2602.08412v22026Can We Trust AI Benchmarks? An Interdisciplinary Review of Current Issues in AI Evaluation
Maria Eriksson, Erasmo Purificato, Arman Noroozian +4
cs.AIarXiv:2502.06559v22025Confess What You Know: Forget-Set Misalignment with Model Knowledge in LLM Unlearning
Miso Kim, Georu Lee, Seungwon Jeong +1
cs.LGcs.AIcs.CLarXiv:2609.00605v12026Cardinality Estimation in DBMS: A Comprehensive Benchmark Evaluation
Yuxing Han, Ziniu Wu, Peizhi Wu +11
cs.DBcs.AIarXiv:2109.05877v32021A Study of Hidden-State Optimization Order in Predictive Coding Networks
Xueyuan Li, Danilo Vasconcellos Vargas
cs.LGcs.AIarXiv:2609.00686v12026Skill-Pro: Learning Reusable Skills from Experience via Non-Parametric PPO for LLM Agents
Qirui Mi, Zhijian Ma, Mengyue Yang +4
cs.AIarXiv:2602.01869v32026Agent KB: Leveraging Cross-Domain Experience for Agentic Problem Solving
Xiangru Tang, Tianrui Qin, Tianhao Peng +15
cs.CLcs.AIarXiv:2507.06229v52025PopPert: Population-level Joint-Distribution Modeling for Single-Cell Perturbation Prediction
Handong Wang, Jiaxin Qi, Haochen Feng +1
q-bio.GNcs.AIarXiv:2609.01357v12026ViTAMINS: An Empirical Study of Training Self-Supervised Vision Transformers with Synthetic Hard Negatives
Nikos Giakoumoglou, Andreas Floros, Kleanthis-Marios Papadopoulos +1
cs.CVcs.AIcs.LGarXiv:2609.01041v12026Sparse Autoencoders Do Not Find Canonical Units of Analysis
Patrick Leask, Bart Bussmann, Michael Pearce +5
cs.LGcs.AIarXiv:2502.04878v12025SMART: Self-Aware Agent for Tool Overuse Mitigation
Cheng Qian, Emre Can Acikgoz, Hongru Wang +5
cs.AIcs.CLcs.LGarXiv:2502.11435v22025Learning Generalizable Robotic Reward Functions from "In-The-Wild" Human Videos
Annie S. Chen, Suraj Nair, Chelsea Finn
cs.ROcs.AIcs.CVarXiv:2103.16817v12021Embedded Conditional Independence Tests for Large Language Model Generated Text with an Application to German Parliament Speeches
Marco Simnacher, Georg Keilbar, Benjamin König +2
stat.MLcs.AIcs.LGarXiv:2609.00946v12026Vibe Coding vs. Agentic Coding: Fundamentals and Practical Implications of Agentic AI
Ranjan Sapkota, Konstantinos I. Roumeliotis, Manoj Karkee
cs.SEcs.AIcs.CLarXiv:2505.19443v12025