Artificial Intelligence

Papers filed under cs.AI on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.

Search paper metadata (including unsummarized papers)

4,741 to 4,800 of 15,245

  1. CliffRank: A Dual-Branch Framework for Activity-Cliff Ranking Prediction

    Kewei Li, Rongying Zhang, Peiyu Yang +4

    cs.LGcs.AIq-bio.BMarXiv:2609.01673v12026
  2. LLMs for Explainable AI: A Comprehensive Survey

    Ahsan Bilal, David Ebert, Beiyu Lin

    cs.AIcs.CLarXiv:2504.00125v12025
  3. Ranked by the Matcher: A Reproducibility Audit of Knowledge Graph Extraction from Threat Reports

    Safayat Bin Hakim, Houbing Herbert Song

    cs.CRcs.AIcs.CLarXiv:2609.01671v12026
  4. StarVLA-$α$: Reducing Complexity in Vision-Language-Action Systems

    Jinhui Ye, Ning Gao, Senqiao Yang +7

    cs.ROcs.AIcs.CVarXiv:2604.11757v12026
  5. A Survey of WebAgents: Towards Next-Generation AI Agents for Web Automation with Large Foundation Models

    Liangbo Ning, Ziran Liang, Zhuohang Jiang +8

    cs.AIarXiv:2503.23350v42025
  6. StepSearch: Igniting LLMs Search Ability via Step-Wise Proximal Policy Optimization

    Ziliang Wang, Xuhui Zheng, Kang An +4

    cs.CLcs.AIcs.IRarXiv:2505.15107v22025
  7. Understanding Software Engineering Agents: A Study of Thought-Action-Result Trajectories

    Islem Bouzenia, Michael Pradel

    cs.SEcs.AIarXiv:2506.18824v22025
  8. MuQ: Self-Supervised Music Representation Learning with Mel Residual Vector Quantization

    Haina Zhu, Yizhi Zhou, Hangting Chen +6

    cs.SDcs.AIcs.CLarXiv:2501.01108v22025
  9. Retrieval-Augmented Generation with Conflicting Evidence

    Han Wang, Archiki Prasad, Elias Stengel-Eskin +1

    cs.CLcs.AIarXiv:2504.13079v22025
  10. DriveMoE: Mixture-of-Experts for Vision-Language-Action Model in End-to-End Autonomous Driving

    Zhenjie Yang, Yilin Chai, Xiaosong Jia +5

    cs.CVcs.AIcs.ROarXiv:2505.16278v22025
  11. Sparse Meets Dense: Unified Generative Recommendations with Cascaded Sparse-Dense Representations

    Yuhao Yang, Zhi Ji, Zhaopeng Li +8

    cs.IRcs.AIarXiv:2503.02453v12025
  12. Overview of the TREC 2022 deep learning track

    Nick Craswell, Bhaskar Mitra, Emine Yilmaz +4

    cs.IRcs.AIcs.CLarXiv:2507.10865v12025
  13. Reasoning or Memorization? Unreliable Results of Reinforcement Learning Due to Data Contamination

    Mingqi Wu, Zhihao Zhang, Qiaole Dong +11

    cs.LGcs.AIcs.CLarXiv:2507.10532v32025
  14. VideoPainter: Any-length Video Inpainting and Editing with Plug-and-Play Context Control

    Yuxuan Bian, Zhaoyang Zhang, Xuan Ju +4

    cs.CVcs.AIcs.MMarXiv:2503.05639v32025
  15. On Memory Construction and Retrieval for Personalized Conversational Agents

    Zhuoshi Pan, Qianhui Wu, Huiqiang Jiang +8

    cs.CLcs.AIarXiv:2502.05589v32025
  16. ChipGPT: How far are we from natural language hardware design

    Kaiyan Chang, Ying Wang, Haimeng Ren +5

    cs.AIcs.ARcs.PLarXiv:2305.14019v42023
  17. Chebyshev Polynomial-Based Kolmogorov-Arnold Networks: An Efficient Architecture for Nonlinear Function Approximation

    Sidharth SS, Keerthana AR, Gokul R +1

    cs.LGcs.AIarXiv:2405.07200v32024
  18. 3DCodeBench: Benchmarking Agentic Procedural 3D Modeling Via Code

    Yipeng Gao, Lei Shu, Genzhi Ye +5

    cs.CVcs.AIcs.GRarXiv:2606.01057v12026
  19. Deepfake-Eval-2024: A Multi-Modal In-the-Wild Benchmark of Deepfakes Circulated in 2024

    Nuria Alina Chandra, Hannah Lee, Ryan Murtfeldt +10

    cs.CVcs.AIcs.CYarXiv:2503.02857v52025
  20. MARS: Modular Agent with Reflective Search for Automated AI Research

    Jiefeng Chen, Bhavana Dalvi Mishra, Jaehyun Nam +3

    cs.AIarXiv:2602.02660v32026
  21. An Evaluation Framework for National AI Regulation

    Kaushik Sanjay Prabhakar, Tarun Adarsh R S, Amal Dhivyan Gregory +3

    cs.CYcs.AIarXiv:2608.15417v12026
  22. LLM-Grounder: Open-Vocabulary 3D Visual Grounding with Large Language Model as an Agent

    Jianing Yang, Xuweiyi Chen, Shengyi Qian +4

    cs.CVcs.AIcs.CLarXiv:2309.12311v12023
  23. StarCoder: may the source be with you!

    Raymond Li, Loubna Ben Allal, Yangtian Zi +64

    cs.CLcs.AIcs.PLarXiv:2305.06161v22023
  24. A Trajectory-Based Safety Audit of Clawdbot (OpenClaw)

    Tianyu Chen, Dongrui Liu, Xia Hu +2

    cs.CRcs.AIarXiv:2602.14364v12026
  25. Recursive Multi-Agent Systems

    Jiaru Zou, Rui Pan, Ruizhong Qiu +8

    cs.AIcs.CLcs.LGarXiv:2604.25917v22026
    Summaries:한국어
  26. Building Production-Ready Probes For Gemini

    János Kramár, Joshua Engels, Zheng Wang +4

    cs.LGcs.AIcs.CLarXiv:2601.11516v42026
  27. P1-VL: Bridging Visual Perception and Scientific Reasoning in Physics Olympiads

    Yun Luo, Futing Wang, Qianjia Cheng +28

    cs.AIarXiv:2602.09443v12026
  28. EnvHarness: Awakening Static Worlds for Agent Learning

    Chengsong Huang, Zifeng Wang, Rujun Han +14

    cs.AIcs.CLcs.LGarXiv:2608.19880v12026
    Summaries:简体中文
  29. Think Again or Think Longer? Selective Verification for Budget-Aware Reasoning

    Sajib Acharjee Dip, Dawei Zhou, Liqing Zhang

    cs.AIcs.CLarXiv:2606.19808v12026
  30. Ventor-QTest: Threat-Model-Driven Verification of Vendor-Hosted LLM APIs

    Xiangfan Wu, Zonghao Ying, Huiyu Wu +4

    cs.CRcs.AIarXiv:2608.16391v12026
  31. Pruning and Quantization for Deep Neural Network Acceleration: A Survey

    Tailin Liang, John Glossner, Lei Wang +2

    cs.CVcs.AIarXiv:2101.09671v32021
  32. Memory Intelligence Agent

    Jingyang Qiao, Weicheng Meng, Yu Cheng +6

    cs.AIcs.MAarXiv:2604.04503v42026
  33. ScientistOne: Towards Human-Level Autonomous Research via Chain-of-Evidence

    Rui Meng, Bhavana Dalvi Mishra, Jiefeng Chen +10

    cs.AIcs.CLcs.MAarXiv:2605.26340v12026
  34. Co-Director: Agentic Generative Video Storytelling

    Yale Song, Yiwen Song, Nick Losier +13

    cs.AIcs.MAcs.MMarXiv:2604.24842v12026
  35. Transparency of Deep Neural Networks for Medical Image Analysis: A Review of Interpretability Methods

    Zohaib Salahuddin, Henry C Woodruff, Avishek Chatterjee +1

    eess.IVcs.AIcs.CVarXiv:2111.02398v12021
  36. Data-Efficient Autoregressive-to-Diffusion Language Models via On-Policy Distillation

    Xingyu Su, Jacob Helwig, Shubham Parashar +6

    cs.CLcs.AIarXiv:2606.06712v12026
  37. ProEval: Proactive Failure Discovery and Efficient Performance Estimation for Generative AI Evaluation

    Yizheng Huang, Wenjun Zeng, Aditi Kumaresan +1

    cs.LGcs.AIstat.MLarXiv:2604.23099v22026
  38. CLIPO: Contrastive Learning in Policy Optimization Generalizes RLVR

    Sijia Cui, Pengyu Cheng, Jiajun Song +6

    cs.LGcs.AIcs.CLarXiv:2603.10101v12026
  39. Benchmarking Vision-Language Models for Automated Pathology Diagnosis and Report Generation

    Yumi Lee, Harim Oh, Hyoryung Kim +52

    cs.CVcs.AIarXiv:2609.00866v12026
  40. Video models are zero-shot learners and reasoners

    Thaddäus Wiedemer, Yuxuan Li, Paul Vicol +6

    cs.LGcs.AIcs.CVarXiv:2509.20328v22025
  41. World Simulation with Video Foundation Models for Physical AI

    NVIDIA, :, Arslan Ali +87

    cs.CVcs.AIcs.LGarXiv:2511.00062v22025
  42. VBVR-Pro: A Scalable and Verifiable Suite for Native Visual Reasoning

    Junxiang Xu, Ruisi Wang, Fanyi Pu +49

    cs.CVcs.AIcs.LGarXiv:2608.26105v12026
  43. Stitched Value Model for Diffusion Alignment

    Hyojun Go, Hyungjin Chung, Prune Truong +8

    cs.CVcs.AIcs.LGarXiv:2605.19804v12026
  44. Can We Defend Against AI-Generated Video Attacks on Real-World Crisis Events? A Systematic Evaluation of Detectors, Generators and Social Dissemination

    Shuo Liang, Yixing Ma, Pengfei Zhou +32

    cs.CVcs.AIarXiv:2608.14391v12026
  45. Gated DeltaNet-2: Decoupling Erase and Write in Linear Attention

    Ali Hatamizadeh, Yejin Choi, Jan Kautz

    cs.AIarXiv:2605.22791v12026
  46. ABot-World-0: Infinite Interactive World Rollout on a Single Desktop GPU

    Fan Jiang, Zhaoxu Sun, Mengchao Wang +38

    cs.CVcs.AIcs.LGarXiv:2607.19191v12026
  47. SANA-Streaming: Real-time Streaming Video Editing with Hybrid Diffusion Transformer

    Yuyang Zhao, Yicheng Pan, Qiyuan He +6

    cs.CVcs.AIarXiv:2605.30409v12026
  48. Repo-To-Skill: Distilling GitHub Repositories Into AI4AI Skills

    Jianlyu Chen, Yuyang Hu, Hongjin Qian +8

    cs.AIcs.CLarXiv:2609.02749v12026
  49. Pretraining Large Language Models with NVFP4

    NVIDIA, Felix Abecassis, Anjulie Agrusa +87

    cs.CLcs.AIcs.LGarXiv:2509.25149v22025
  50. SkillOS: Learning Skill Curation for Self-Evolving Agents

    Siru Ouyang, Jun Yan, Yanfei Chen +13

    cs.AIcs.CLarXiv:2605.06614v12026
  51. Nemotron 3 Super: Open, Efficient Mixture-of-Experts Hybrid Mamba-Transformer Model for Agentic Reasoning

    NVIDIA, :, Aakshita Chandiramani +544

    cs.LGcs.AIcs.CLarXiv:2604.12374v12026
  52. OR-Space: A Full-Lifecycle Workspace Benchmark for Industrial Optimization Agents

    Chenyu Zhou, Xinyun Lu, Jiangyue Zhao +3

    cs.AIarXiv:2605.28158v12026
  53. Planetary Prediction Engine: Autonomous Geospatial Prediction via Intelligent Data Selection and Foundation Model Embeddings

    Evelyn Ma, Rama Kumar Pasumarthi, Kishwar Shafin +25

    cs.AIcs.LGarXiv:2608.26088v12026
  54. A Very Big Video Reasoning Suite

    Maijunxian Wang, Ruisi Wang, Juyi Lin +53

    cs.CVcs.AIcs.LGarXiv:2602.20159v22026
  55. SWE-Milestone: Evaluating AI Agents on Continuous Software Evolution

    Gangda Deng, Zhaoling Chen, Zhongming Yu +11

    cs.SEcs.AIarXiv:2603.13428v42026
  56. RMBench: Memory-Dependent Robotic Manipulation Benchmark with Insights into Policy Design

    Tianxing Chen, Yuran Wang, Mingleyang Li +16

    cs.ROcs.AIarXiv:2603.01229v32026
  57. PARCEL: Pool-Anchored Resampling with Conditioned Elastic Queries for Efficient Vision-Language Understanding

    Selim Kuzucu, Alessio Tonioni, Vasile Lup +3

    cs.CVcs.AIcs.CLarXiv:2605.30126v12026
  58. Cosmos World Foundation Model Platform for Physical AI

    NVIDIA, :, Niket Agarwal +76

    cs.CVcs.AIcs.LGarXiv:2501.03575v32025
  59. A Survey on Large Language Model (LLM) Security and Privacy: The Good, the Bad, and the Ugly

    Yifan Yao, Jinhao Duan, Kaidi Xu +3

    cs.CRcs.AIarXiv:2312.02003v32023
  60. STAR-1: Safer Alignment of Reasoning LLMs with 1K Data

    Zijun Wang, Haoqin Tu, Yuhan Wang +6

    cs.CLcs.AIarXiv:2504.01903v22025