Artificial Intelligence
Papers filed under cs.AI on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.
Search paper metadata (including unsummarized papers)
4,741 to 4,800 of 15,245
CliffRank: A Dual-Branch Framework for Activity-Cliff Ranking Prediction
Kewei Li, Rongying Zhang, Peiyu Yang +4
cs.LGcs.AIq-bio.BMarXiv:2609.01673v12026LLMs for Explainable AI: A Comprehensive Survey
Ahsan Bilal, David Ebert, Beiyu Lin
cs.AIcs.CLarXiv:2504.00125v12025Ranked by the Matcher: A Reproducibility Audit of Knowledge Graph Extraction from Threat Reports
Safayat Bin Hakim, Houbing Herbert Song
cs.CRcs.AIcs.CLarXiv:2609.01671v12026StarVLA-$α$: Reducing Complexity in Vision-Language-Action Systems
Jinhui Ye, Ning Gao, Senqiao Yang +7
cs.ROcs.AIcs.CVarXiv:2604.11757v12026A Survey of WebAgents: Towards Next-Generation AI Agents for Web Automation with Large Foundation Models
Liangbo Ning, Ziran Liang, Zhuohang Jiang +8
cs.AIarXiv:2503.23350v42025StepSearch: Igniting LLMs Search Ability via Step-Wise Proximal Policy Optimization
Ziliang Wang, Xuhui Zheng, Kang An +4
cs.CLcs.AIcs.IRarXiv:2505.15107v22025Understanding Software Engineering Agents: A Study of Thought-Action-Result Trajectories
Islem Bouzenia, Michael Pradel
cs.SEcs.AIarXiv:2506.18824v22025MuQ: Self-Supervised Music Representation Learning with Mel Residual Vector Quantization
Haina Zhu, Yizhi Zhou, Hangting Chen +6
cs.SDcs.AIcs.CLarXiv:2501.01108v22025Retrieval-Augmented Generation with Conflicting Evidence
Han Wang, Archiki Prasad, Elias Stengel-Eskin +1
cs.CLcs.AIarXiv:2504.13079v22025DriveMoE: Mixture-of-Experts for Vision-Language-Action Model in End-to-End Autonomous Driving
Zhenjie Yang, Yilin Chai, Xiaosong Jia +5
cs.CVcs.AIcs.ROarXiv:2505.16278v22025Sparse Meets Dense: Unified Generative Recommendations with Cascaded Sparse-Dense Representations
Yuhao Yang, Zhi Ji, Zhaopeng Li +8
cs.IRcs.AIarXiv:2503.02453v12025Overview of the TREC 2022 deep learning track
Nick Craswell, Bhaskar Mitra, Emine Yilmaz +4
cs.IRcs.AIcs.CLarXiv:2507.10865v12025Reasoning or Memorization? Unreliable Results of Reinforcement Learning Due to Data Contamination
Mingqi Wu, Zhihao Zhang, Qiaole Dong +11
cs.LGcs.AIcs.CLarXiv:2507.10532v32025VideoPainter: Any-length Video Inpainting and Editing with Plug-and-Play Context Control
Yuxuan Bian, Zhaoyang Zhang, Xuan Ju +4
cs.CVcs.AIcs.MMarXiv:2503.05639v32025On Memory Construction and Retrieval for Personalized Conversational Agents
Zhuoshi Pan, Qianhui Wu, Huiqiang Jiang +8
cs.CLcs.AIarXiv:2502.05589v32025ChipGPT: How far are we from natural language hardware design
Kaiyan Chang, Ying Wang, Haimeng Ren +5
cs.AIcs.ARcs.PLarXiv:2305.14019v42023Chebyshev Polynomial-Based Kolmogorov-Arnold Networks: An Efficient Architecture for Nonlinear Function Approximation
Sidharth SS, Keerthana AR, Gokul R +1
cs.LGcs.AIarXiv:2405.07200v320243DCodeBench: Benchmarking Agentic Procedural 3D Modeling Via Code
Yipeng Gao, Lei Shu, Genzhi Ye +5
cs.CVcs.AIcs.GRarXiv:2606.01057v12026Deepfake-Eval-2024: A Multi-Modal In-the-Wild Benchmark of Deepfakes Circulated in 2024
Nuria Alina Chandra, Hannah Lee, Ryan Murtfeldt +10
cs.CVcs.AIcs.CYarXiv:2503.02857v52025MARS: Modular Agent with Reflective Search for Automated AI Research
Jiefeng Chen, Bhavana Dalvi Mishra, Jaehyun Nam +3
cs.AIarXiv:2602.02660v32026An Evaluation Framework for National AI Regulation
Kaushik Sanjay Prabhakar, Tarun Adarsh R S, Amal Dhivyan Gregory +3
cs.CYcs.AIarXiv:2608.15417v12026LLM-Grounder: Open-Vocabulary 3D Visual Grounding with Large Language Model as an Agent
Jianing Yang, Xuweiyi Chen, Shengyi Qian +4
cs.CVcs.AIcs.CLarXiv:2309.12311v12023StarCoder: may the source be with you!
Raymond Li, Loubna Ben Allal, Yangtian Zi +64
cs.CLcs.AIcs.PLarXiv:2305.06161v22023A Trajectory-Based Safety Audit of Clawdbot (OpenClaw)
Tianyu Chen, Dongrui Liu, Xia Hu +2
cs.CRcs.AIarXiv:2602.14364v12026Recursive Multi-Agent Systems
Jiaru Zou, Rui Pan, Ruizhong Qiu +8
cs.AIcs.CLcs.LGarXiv:2604.25917v22026Summaries:한국어Building Production-Ready Probes For Gemini
János Kramár, Joshua Engels, Zheng Wang +4
cs.LGcs.AIcs.CLarXiv:2601.11516v42026P1-VL: Bridging Visual Perception and Scientific Reasoning in Physics Olympiads
Yun Luo, Futing Wang, Qianjia Cheng +28
cs.AIarXiv:2602.09443v12026EnvHarness: Awakening Static Worlds for Agent Learning
Chengsong Huang, Zifeng Wang, Rujun Han +14
cs.AIcs.CLcs.LGarXiv:2608.19880v12026Summaries:简体中文Think Again or Think Longer? Selective Verification for Budget-Aware Reasoning
Sajib Acharjee Dip, Dawei Zhou, Liqing Zhang
cs.AIcs.CLarXiv:2606.19808v12026Ventor-QTest: Threat-Model-Driven Verification of Vendor-Hosted LLM APIs
Xiangfan Wu, Zonghao Ying, Huiyu Wu +4
cs.CRcs.AIarXiv:2608.16391v12026Pruning and Quantization for Deep Neural Network Acceleration: A Survey
Tailin Liang, John Glossner, Lei Wang +2
cs.CVcs.AIarXiv:2101.09671v32021Memory Intelligence Agent
Jingyang Qiao, Weicheng Meng, Yu Cheng +6
cs.AIcs.MAarXiv:2604.04503v42026ScientistOne: Towards Human-Level Autonomous Research via Chain-of-Evidence
Rui Meng, Bhavana Dalvi Mishra, Jiefeng Chen +10
cs.AIcs.CLcs.MAarXiv:2605.26340v12026Co-Director: Agentic Generative Video Storytelling
Yale Song, Yiwen Song, Nick Losier +13
cs.AIcs.MAcs.MMarXiv:2604.24842v12026Transparency of Deep Neural Networks for Medical Image Analysis: A Review of Interpretability Methods
Zohaib Salahuddin, Henry C Woodruff, Avishek Chatterjee +1
eess.IVcs.AIcs.CVarXiv:2111.02398v12021Data-Efficient Autoregressive-to-Diffusion Language Models via On-Policy Distillation
Xingyu Su, Jacob Helwig, Shubham Parashar +6
cs.CLcs.AIarXiv:2606.06712v12026ProEval: Proactive Failure Discovery and Efficient Performance Estimation for Generative AI Evaluation
Yizheng Huang, Wenjun Zeng, Aditi Kumaresan +1
cs.LGcs.AIstat.MLarXiv:2604.23099v22026CLIPO: Contrastive Learning in Policy Optimization Generalizes RLVR
Sijia Cui, Pengyu Cheng, Jiajun Song +6
cs.LGcs.AIcs.CLarXiv:2603.10101v12026Benchmarking Vision-Language Models for Automated Pathology Diagnosis and Report Generation
Yumi Lee, Harim Oh, Hyoryung Kim +52
cs.CVcs.AIarXiv:2609.00866v12026Video models are zero-shot learners and reasoners
Thaddäus Wiedemer, Yuxuan Li, Paul Vicol +6
cs.LGcs.AIcs.CVarXiv:2509.20328v22025World Simulation with Video Foundation Models for Physical AI
NVIDIA, :, Arslan Ali +87
cs.CVcs.AIcs.LGarXiv:2511.00062v22025VBVR-Pro: A Scalable and Verifiable Suite for Native Visual Reasoning
Junxiang Xu, Ruisi Wang, Fanyi Pu +49
cs.CVcs.AIcs.LGarXiv:2608.26105v12026Stitched Value Model for Diffusion Alignment
Hyojun Go, Hyungjin Chung, Prune Truong +8
cs.CVcs.AIcs.LGarXiv:2605.19804v12026Can We Defend Against AI-Generated Video Attacks on Real-World Crisis Events? A Systematic Evaluation of Detectors, Generators and Social Dissemination
Shuo Liang, Yixing Ma, Pengfei Zhou +32
cs.CVcs.AIarXiv:2608.14391v12026Gated DeltaNet-2: Decoupling Erase and Write in Linear Attention
Ali Hatamizadeh, Yejin Choi, Jan Kautz
cs.AIarXiv:2605.22791v12026ABot-World-0: Infinite Interactive World Rollout on a Single Desktop GPU
Fan Jiang, Zhaoxu Sun, Mengchao Wang +38
cs.CVcs.AIcs.LGarXiv:2607.19191v12026SANA-Streaming: Real-time Streaming Video Editing with Hybrid Diffusion Transformer
Yuyang Zhao, Yicheng Pan, Qiyuan He +6
cs.CVcs.AIarXiv:2605.30409v12026Repo-To-Skill: Distilling GitHub Repositories Into AI4AI Skills
Jianlyu Chen, Yuyang Hu, Hongjin Qian +8
cs.AIcs.CLarXiv:2609.02749v12026Pretraining Large Language Models with NVFP4
NVIDIA, Felix Abecassis, Anjulie Agrusa +87
cs.CLcs.AIcs.LGarXiv:2509.25149v22025SkillOS: Learning Skill Curation for Self-Evolving Agents
Siru Ouyang, Jun Yan, Yanfei Chen +13
cs.AIcs.CLarXiv:2605.06614v12026Nemotron 3 Super: Open, Efficient Mixture-of-Experts Hybrid Mamba-Transformer Model for Agentic Reasoning
NVIDIA, :, Aakshita Chandiramani +544
cs.LGcs.AIcs.CLarXiv:2604.12374v12026OR-Space: A Full-Lifecycle Workspace Benchmark for Industrial Optimization Agents
Chenyu Zhou, Xinyun Lu, Jiangyue Zhao +3
cs.AIarXiv:2605.28158v12026Planetary Prediction Engine: Autonomous Geospatial Prediction via Intelligent Data Selection and Foundation Model Embeddings
Evelyn Ma, Rama Kumar Pasumarthi, Kishwar Shafin +25
cs.AIcs.LGarXiv:2608.26088v12026A Very Big Video Reasoning Suite
Maijunxian Wang, Ruisi Wang, Juyi Lin +53
cs.CVcs.AIcs.LGarXiv:2602.20159v22026SWE-Milestone: Evaluating AI Agents on Continuous Software Evolution
Gangda Deng, Zhaoling Chen, Zhongming Yu +11
cs.SEcs.AIarXiv:2603.13428v42026RMBench: Memory-Dependent Robotic Manipulation Benchmark with Insights into Policy Design
Tianxing Chen, Yuran Wang, Mingleyang Li +16
cs.ROcs.AIarXiv:2603.01229v32026PARCEL: Pool-Anchored Resampling with Conditioned Elastic Queries for Efficient Vision-Language Understanding
Selim Kuzucu, Alessio Tonioni, Vasile Lup +3
cs.CVcs.AIcs.CLarXiv:2605.30126v12026Cosmos World Foundation Model Platform for Physical AI
NVIDIA, :, Niket Agarwal +76
cs.CVcs.AIcs.LGarXiv:2501.03575v32025A Survey on Large Language Model (LLM) Security and Privacy: The Good, the Bad, and the Ugly
Yifan Yao, Jinhao Duan, Kaidi Xu +3
cs.CRcs.AIarXiv:2312.02003v32023STAR-1: Safer Alignment of Reasoning LLMs with 1K Data
Zijun Wang, Haoqin Tu, Yuhan Wang +6
cs.CLcs.AIarXiv:2504.01903v22025