Artificial Intelligence

Papers filed under cs.AI on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.

Search paper metadata (including unsummarized papers)

4,441 to 4,500 of 15,280

  1. Path Planning for Masked Diffusion Model Sampling

    Fred Zhangzhi Peng, Zachary Bezemek, Sawan Patel +5

    cs.LGcs.AIarXiv:2502.03540v52025
  2. LLMs4OL: Large Language Models for Ontology Learning

    Hamed Babaei Giglou, Jennifer D'Souza, Sören Auer

    cs.AIcs.CLcs.ITarXiv:2307.16648v22023
  3. SkillGLoW: Procedural-Family Skill Consolidation for Self-Improving Agents on Long-Horizon Task Streams

    Ao Yan, Xin Zhang, Jiawei Du +1

    cs.AIarXiv:2609.02217v12026
  4. Recent Advances in Discrete Speech Tokens: A Review

    Yiwei Guo, Zhihan Li, Hankun Wang +7

    eess.AScs.AIcs.MMarXiv:2502.06490v42025
  5. A Framework for Human Evaluation of Large Language Models in Healthcare Derived from Literature Review

    Thomas Yu Chow Tam, Sonish Sivarajkumar, Sumit Kapoor +12

    cs.CLcs.AIarXiv:2405.02559v22024
  6. SIoU Loss: More Powerful Learning for Bounding Box Regression

    Zhora Gevorgyan

    cs.CVcs.AIarXiv:2205.12740v12022
  7. DRIFT: Dynamic Rule-Based Defense with Injection Isolation for Securing LLM Agents

    Hao Li, Xiaogeng Liu, Hung-Chun Chiu +3

    cs.CRcs.AIarXiv:2506.12104v32025
  8. Ethical and social risks of harm from Language Models

    Laura Weidinger, John Mellor, Maribeth Rauh +20

    cs.CLcs.AIcs.CYarXiv:2112.04359v12021
  9. RLVMR: Reinforcement Learning with Verifiable Meta-Reasoning Rewards for Robust Long-Horizon Agents

    Zijing Zhang, Ziyang Chen, Mingxiao Li +2

    cs.LGcs.AIarXiv:2507.22844v12025
  10. Minimizing the Accumulated Trajectory Error to Improve Dataset Distillation

    Jiawei Du, Yidi Jiang, Vincent Y. F. Tan +2

    cs.LGcs.AIcs.CVarXiv:2211.11004v32022
  11. Vision-Language Models are Zero-Shot Reward Models for Reinforcement Learning

    Juan Rocamonde, Victoriano Montesinos, Elvis Nava +2

    cs.LGcs.AIarXiv:2310.12921v22023
  12. The pitfalls of next-token prediction

    Gregor Bachmann, Vaishnavh Nagarajan

    cs.CLcs.AIcs.LGarXiv:2403.06963v32024
  13. PathRAG: Pruning Graph-based Retrieval Augmented Generation with Relational Paths

    Boyu Chen, Zirui Guo, Zidan Yang +5

    cs.CLcs.AIcs.IRarXiv:2502.14902v22025
  14. SAEs Are Good for Steering -- If You Select the Right Features

    Dana Arad, Aaron Mueller, Yonatan Belinkov

    cs.LGcs.AIcs.CLarXiv:2505.20063v22025
  15. CUDA-L1: Improving CUDA Optimization via Contrastive Reinforcement Learning

    Xiaoya Li, Albert Wang, Guoyin Wang +2

    cs.AIcs.DCcs.LGarXiv:2507.14111v122025
  16. PlanGenLLMs: A Modern Survey of LLM Planning Capabilities

    Hui Wei, Zihao Zhang, Shenghua He +3

    cs.AIcs.CLarXiv:2502.11221v32025
  17. StableNormal: Reducing Diffusion Variance for Stable and Sharp Normal

    Chongjie Ye, Lingteng Qiu, Xiaodong Gu +6

    cs.CVcs.AIcs.GRarXiv:2406.16864v12024
  18. Detection of Chagas Disease from the ECG: The George B. Moody PhysioNet Challenge 2025

    Matthew A. Reyna, Zuzana Koscova, Jan Pavlus +11

    cs.LGcs.AIarXiv:2510.02202v12025
  19. Robust Autonomy Emerges from Self-Play

    Marco Cusumano-Towner, David Hafner, Alex Hertzberg +9

    cs.LGcs.AIcs.ROarXiv:2502.03349v12025
  20. SkillForge: Forging Domain-Specific, Self-Evolving Agent Skills in Cloud Technical Support

    Xingyan Liu, Xiyue Luo, Linyu Li +3

    cs.IRcs.AIcs.SEarXiv:2604.08618v22026
  21. CLIP-It! Language-Guided Video Summarization

    Medhini Narasimhan, Anna Rohrbach, Trevor Darrell

    cs.CVcs.AIcs.MMarXiv:2107.00650v22021
  22. Reasoning Models Better Express Their Confidence

    Dongkeun Yoon, Seungone Kim, Sohee Yang +6

    cs.AIcs.CLarXiv:2505.14489v22025
  23. UI-Vision: A Desktop-centric GUI Benchmark for Visual Perception and Interaction

    Shravan Nayak, Xiangru Jian, Kevin Qinghong Lin +11

    cs.CVcs.AIcs.CLarXiv:2503.15661v22025
  24. Vision-R1: Evolving Human-Free Alignment in Large Vision-Language Models via Vision-Guided Reinforcement Learning

    Yufei Zhan, Yousong Zhu, Shurong Zheng +4

    cs.CVcs.AIarXiv:2503.18013v12025
  25. CLAWS: Clustering Assisted Weakly Supervised Learning with Normalcy Suppression for Anomalous Event Detection

    Muhammad Zaigham Zaheer, Arif Mahmood, Marcella Astrid +1

    cs.CVcs.AIarXiv:2011.12077v42020
  26. ML-Master: Towards AI-for-AI via Integration of Exploration and Reasoning

    Zexi Liu, Yuzhu Cai, Xinyu Zhu +6

    cs.AIcs.LGarXiv:2506.16499v12025
  27. Conservative set valued fields, automatic differentiation, stochastic gradient method and deep learning

    Jérôme Bolte, Edouard Pauwels

    math.OCcs.AIcs.LGarXiv:1909.10300v42019
  28. What Makes a Reward Model a Good Teacher? An Optimization Perspective

    Noam Razin, Zixuan Wang, Hubert Strauss +3

    cs.LGcs.AIcs.CLarXiv:2503.15477v42025
  29. Performance assessment and exhaustive listing of 500+ nature inspired metaheuristic algorithms

    Zhongqiang Ma, Guohua Wu, Ponnuthurai N. Suganthan +2

    cs.NEcs.AIarXiv:2212.09479v12022
  30. Agent Learning via Early Experience

    Kai Zhang, Xiangchao Chen, Bo Liu +27

    cs.AIcs.CLcs.IRarXiv:2510.08558v32025
  31. Vending-Bench: A Benchmark for Long-Term Coherence of Autonomous Agents

    Axel Backlund, Lukas Petersson

    cs.AIarXiv:2502.15840v12025
  32. DexGrasp Anything: Towards Universal Robotic Dexterous Grasping with Physics Awareness

    Yiming Zhong, Qi Jiang, Jingyi Yu +1

    cs.CVcs.AIcs.ROarXiv:2503.08257v22025
  33. Comprehensive Comparative Study of Multi-Label Classification Methods

    Jasmin Bogatinovski, Ljupčo Todorovski, Sašo Džeroski +1

    cs.LGcs.AIcs.CCarXiv:2102.07113v22021
  34. BridgeVLA: Input-Output Alignment for Efficient 3D Manipulation Learning with Vision-Language Models

    Peiyan Li, Yixiang Chen, Hongtao Wu +6

    cs.ROcs.AIarXiv:2506.07961v22025
  35. What does it mean to solve the problem of discrimination in hiring? Social, technical and legal perspectives from the UK on automated hiring systems

    Javier Sanchez-Monedero, Lina Dencik, Lilian Edwards

    cs.CYcs.AIarXiv:1910.06144v22019
  36. SAeUron: Interpretable Concept Unlearning in Diffusion Models with Sparse Autoencoders

    Bartosz Cywiński, Kamil Deja

    cs.LGcs.AIarXiv:2501.18052v32025
  37. Can LLMs Design Video Coding Tools? A Case Study on Planar Mode

    Yingwen Zhang, Meng Wang, Liqiang He +1

    cs.MMcs.AIarXiv:2609.01535v12026
  38. Defense-as-Skill: Evolving Runtime Guard Skill for Skill-Augmented Agents

    Xiaofang Yang, Ziqi Miao, Dianbo Sui +2

    cs.CRcs.AIarXiv:2609.01487v12026
  39. Auditing language models for hidden objectives

    Samuel Marks, Johannes Treutlein, Trenton Bricken +32

    cs.AIcs.CLcs.LGarXiv:2503.10965v22025
  40. Semantic-Guided Multimodal Preprocessing for Vision Transformer-Based Clear Cell Renal Cell Carcinoma Grading

    Fatemeh Javadian, Zhu Chen, Zahra Aminparast +1

    cs.CVcs.AIcs.LGarXiv:2609.01426v12026
  41. Measuring consistency via ensemble margin and local prediction variability: Auditing decision systems in the presence of predictive multiplicity

    Sinjini Banerjee, Tim Marrinan, Anand D. Sarwate

    stat.MLcs.AIcs.LGarXiv:2609.01397v12026
  42. One-Prompt-One-Story: Free-Lunch Consistent Text-to-Image Generation Using a Single Prompt

    Tao Liu, Kai Wang, Senmao Li +6

    cs.CVcs.AIcs.LGarXiv:2501.13554v32025
  43. "I'm Not Sure, But...": Examining the Impact of Large Language Models' Uncertainty Expression on User Reliance and Trust

    Sunnie S. Y. Kim, Q. Vera Liao, Mihaela Vorvoreanu +2

    cs.HCcs.AIarXiv:2405.00623v22024
  44. The zbMATH Open Knowledge Graph: Tracing Centuries of Mathematical Research

    Yuni Susanti, Moritz Schubotz

    cs.DLcs.AIarXiv:2609.00969v12026
  45. GUI-G1: Understanding R1-Zero-Like Training for Visual Grounding in GUI Agents

    Yuqi Zhou, Sunhao Dai, Shuai Wang +3

    cs.CLcs.AIcs.CVarXiv:2505.15810v22025
  46. DualStake: Dual-Path Confidence Calibration in Deep Research Agents

    Yinuo Xu, Yuwei Liang, Jianjie Cheng +4

    cs.CLcs.AIcs.LGarXiv:2609.00935v12026
  47. From Assistant to Double Agent: Formalizing and Benchmarking Attacks on OpenClaw for Personalized Local AI Agent

    Yuhang Wang, Feiming Xu, Zheng Lin +6

    cs.AIarXiv:2602.08412v22026
  48. Can We Trust AI Benchmarks? An Interdisciplinary Review of Current Issues in AI Evaluation

    Maria Eriksson, Erasmo Purificato, Arman Noroozian +4

    cs.AIarXiv:2502.06559v22025
  49. Confess What You Know: Forget-Set Misalignment with Model Knowledge in LLM Unlearning

    Miso Kim, Georu Lee, Seungwon Jeong +1

    cs.LGcs.AIcs.CLarXiv:2609.00605v12026
  50. Cardinality Estimation in DBMS: A Comprehensive Benchmark Evaluation

    Yuxing Han, Ziniu Wu, Peizhi Wu +11

    cs.DBcs.AIarXiv:2109.05877v32021
  51. A Study of Hidden-State Optimization Order in Predictive Coding Networks

    Xueyuan Li, Danilo Vasconcellos Vargas

    cs.LGcs.AIarXiv:2609.00686v12026
  52. Skill-Pro: Learning Reusable Skills from Experience via Non-Parametric PPO for LLM Agents

    Qirui Mi, Zhijian Ma, Mengyue Yang +4

    cs.AIarXiv:2602.01869v32026
  53. Agent KB: Leveraging Cross-Domain Experience for Agentic Problem Solving

    Xiangru Tang, Tianrui Qin, Tianhao Peng +15

    cs.CLcs.AIarXiv:2507.06229v52025
  54. PopPert: Population-level Joint-Distribution Modeling for Single-Cell Perturbation Prediction

    Handong Wang, Jiaxin Qi, Haochen Feng +1

    q-bio.GNcs.AIarXiv:2609.01357v12026
  55. ViTAMINS: An Empirical Study of Training Self-Supervised Vision Transformers with Synthetic Hard Negatives

    Nikos Giakoumoglou, Andreas Floros, Kleanthis-Marios Papadopoulos +1

    cs.CVcs.AIcs.LGarXiv:2609.01041v12026
  56. Sparse Autoencoders Do Not Find Canonical Units of Analysis

    Patrick Leask, Bart Bussmann, Michael Pearce +5

    cs.LGcs.AIarXiv:2502.04878v12025
  57. SMART: Self-Aware Agent for Tool Overuse Mitigation

    Cheng Qian, Emre Can Acikgoz, Hongru Wang +5

    cs.AIcs.CLcs.LGarXiv:2502.11435v22025
  58. Learning Generalizable Robotic Reward Functions from "In-The-Wild" Human Videos

    Annie S. Chen, Suraj Nair, Chelsea Finn

    cs.ROcs.AIcs.CVarXiv:2103.16817v12021
  59. Embedded Conditional Independence Tests for Large Language Model Generated Text with an Application to German Parliament Speeches

    Marco Simnacher, Georg Keilbar, Benjamin König +2

    stat.MLcs.AIcs.LGarXiv:2609.00946v12026
  60. Vibe Coding vs. Agentic Coding: Fundamentals and Practical Implications of Agentic AI

    Ranjan Sapkota, Konstantinos I. Roumeliotis, Manoj Karkee

    cs.SEcs.AIcs.CLarXiv:2505.19443v12025