Artificial Intelligence

Papers filed under cs.AI on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.

Search paper metadata (including unsummarized papers)

3,121 to 3,180 of 15,368

  1. OP-Bench: Benchmarking Over-Personalization for Memory-Augmented Personalized Conversational Agents

    Yulin Hu, Zimo Long, Jiahe Guo +5

    cs.CLcs.AIarXiv:2601.13722v12026
  2. Adaptive Multi-Granularity Temporal Modeling for Weakly Supervised Video Anomaly Detection

    Changyi Li, Yu Xiao

    cs.CVcs.AIarXiv:2609.05066v12026
  3. Continual Semantic Segmentation via Repulsion-Attraction of Sparse and Disentangled Latent Representations

    Umberto Michieli, Pietro Zanuttigh

    cs.CVcs.AIcs.LGarXiv:2103.06342v32021
  4. How do LLMs Evaluate Perceived Moral Agency? Investigating Moral Decision-Making in Human-Artificial Agents Interactions

    Fernanda Mansilla, Aloysius Tok, Bahia Guellaï +2

    cs.CLcs.AIarXiv:2609.05037v12026
  5. LIBERO-X: Robustness Litmus for Vision-Language-Action Models

    Guodong Wang, Chenkai Zhang, Qingjie Liu +4

    cs.CVcs.AIcs.ROarXiv:2602.06556v12026
  6. How a Chatbot's Response Style Shapes a Classroom: A Multi-Agent Simulation of Students Consulting AI

    Rin Tamai, Yuya Dan

    cs.HCcs.AIcs.CYarXiv:2609.05018v12026
  7. Goedel-Code-Prover: Hierarchical Proof Search for Open State-of-the-Art Code Verification

    Zenan Li, Ziran Yang, Deyuan He +7

    cs.SEcs.AIarXiv:2603.19329v32026
  8. Adversarial Robustness Against the Union of Multiple Perturbation Models

    Pratyush Maini, Eric Wong, J. Zico Kolter

    cs.LGcs.AIstat.MLarXiv:1909.04068v22019
  9. ARIA - An Agentic Framework for Autonomous Testing of Infotainment Systems

    António Azevedo, Bruno Lima, João Pascoal Faria

    cs.SEcs.AIarXiv:2609.04913v12026
  10. RAM: Recover Any 3D Human Motion in-the-Wild

    Sen Jia, Ning Zhu, Jinqin Zhong +4

    cs.CVcs.AIarXiv:2603.19929v22026
  11. Qlippy: A Retrieval-Augmented GenAI Assistant for Reproducible Quantum Workflows and Experiment Tracking

    Mahee Gamage, Vlad Stirbu

    quant-phcs.AIarXiv:2609.05039v12026
  12. The Paradigm Shift: A Comprehensive Survey on Large Vision Language Models for Multimodal Fake News Detection

    Wei Ai, Yilong Tan, Yuntao Shou +4

    cs.AIcs.CVarXiv:2601.15316v12026
  13. SideQuest: Model-Driven KV Cache Management for Long-Horizon Agentic Reasoning

    Sanjay Kariyappa, G. Edward Suh

    cs.AIcs.LGarXiv:2602.22603v22026
  14. Fine-R1: Make Multi-modal LLMs Excel in Fine-Grained Visual Recognition by Chain-of-Thought Reasoning

    Hulingxiao He, Zijun Geng, Yuxin Peng

    cs.CVcs.AIarXiv:2602.07605v32026
  15. Why Diffusion Language Models Struggle with Truly Parallel (Non-Autoregressive) Decoding?

    Pengxiang Li, Dilxat Muhtar, Tianlong Chen +2

    cs.CLcs.AIarXiv:2602.23225v22026
  16. Amortizing Scaling Law Construction Costs

    Abhash Kumar Jha, Diana Alexandra Onuţu, Neeratyoy Mallik +6

    cs.LGcs.AIarXiv:2609.05016v12026
  17. MCPO: Modality-Contrastive Preference Optimization for Multimodal Chain-of-Thought Compression

    Guangheng Yang, Zhenliang Ni, Zhenkai Wu +4

    cs.CVcs.AIarXiv:2609.04947v12026
  18. Look-Ahead-Bench: a Standardized Benchmark of Look-ahead Bias in Point-in-Time LLMs for Finance

    Mostapha Benhenda

    cs.AIcs.CLcs.LGarXiv:2601.13770v12026
  19. Language-Conditioned World Modeling for Visual Navigation

    Yifei Dong, Fengyi Wu, Yilong Dai +10

    cs.CVcs.AIcs.ROarXiv:2603.26741v12026
  20. Latent Introspection: Models Can Detect Prior Concept Injections

    Theia Pearson-Vogel, Martin Vanek, Raymond Douglas +1

    cs.AIcs.LGarXiv:2602.20031v22026
  21. MUZZLE: Adaptive Agentic Red-Teaming of Web Agents Against Indirect Prompt Injection Attacks

    Georgios Syros, Evan Rose, Brian Grinstead +4

    cs.CRcs.AIarXiv:2602.09222v22026
  22. Latent Adversarial Training Improves Robustness to Persistent Harmful Behaviors in LLMs

    Abhay Sheshadri, Aidan Ewart, Phillip Guo +8

    cs.LGcs.AIcs.CLarXiv:2407.15549v32024
  23. One Diffusion Model, Two Roles: Guided Trajectory Planning and Safety-Critical Scenario Generation in Closed-Loop Simulation

    Arka Pal, Rajesh Kumar, Hannes Eriksson +4

    cs.CVcs.AIcs.LGarXiv:2609.04921v12026
  24. MemPO: Self-Memory Policy Optimization for Long-Horizon Agents

    Ruoran Li, Xinghua Zhang, Haiyang Yu +7

    cs.AIarXiv:2603.00680v42026
  25. Relational Graph Learning for Crowd Navigation

    Changan Chen, Sha Hu, Payam Nikdel +2

    cs.ROcs.AIcs.LGarXiv:1909.13165v32019
  26. FlowBalance: Verifier-Grounded Self-Improvement from On-Policy Reasoning Experience

    Zixun Huang, Kishan Panaganti, Haitao Mi +1

    cs.LGcs.AIarXiv:2609.03241v12026
  27. Better Understanding, Better Fixes? A Study of Hallucination in LLM-based Automated Program Repair

    Xuemeng Cai, Jiakun Liu, Linhan Yang +2

    cs.SEcs.AIarXiv:2609.04909v12026
  28. Sound-based Multi-Person 3D Pose Estimation

    Yusuke Oumi, Yuto Shibata, Go Irie +3

    cs.CVcs.AIcs.LGarXiv:2609.04902v12026
  29. RefactorPlatform: An Open-Source Harness for Controlled Evaluation of Repository-Scale Refactoring Agents

    Aziz Ben Amor, Drish Mali, Mann Acharya +2

    cs.CLcs.AIarXiv:2609.04898v12026
  30. Resource-Efficient Iterative LLM-Based NAS with Feedback Memory

    Xiaojie Gu, Dmitry Ignatov, Radu Timofte

    cs.LGcs.AIarXiv:2603.12091v12026
  31. Institutional AI: Governing LLM Collusion in Multi-Agent Cournot Markets via Public Governance Graphs

    Marcantonio Bracale Syrnikov, Federico Pierucci, Marcello Galisai +6

    cs.GTcs.AIarXiv:2601.11369v22026
  32. Autorubric: A Unifying Framework for Rubric-Based LLM Evaluation on Non-Verifiable Tasks

    Delip Rao, Chris Callison-Burch

    cs.CLcs.AIarXiv:2603.00077v32026
  33. Forgetting Without Restarting: Execution-State Unlearning for Stateful LLM Agents

    Chao Yao, Yangbo Wei, Zhen Huang +5

    cs.CRcs.AIarXiv:2609.04875v12026
  34. ContractSkill: Repairable Contract-Based Skills for Multimodal Web Agents

    Zijian Lu, Yiping Zuo, Yupeng Nie +4

    cs.SEcs.AIarXiv:2603.20340v32026
  35. A Molecular Multimodal Foundation Model Associating Molecule Graphs with Natural Language

    Bing Su, Dazhao Du, Zhao Yang +6

    cs.LGcs.AIarXiv:2209.05481v12022
  36. TreeFI: Value-Aware Statistical Fault Injection for Deep Neural Networks

    Noam Bires, Marcello Traiola, Angeliki Kritikakou +1

    cs.ARcs.AIarXiv:2609.04912v12026
  37. Mousse: Rectifying the Geometry of Muon with Curvature-Aware Preconditioning

    Yechen Zhang, Shuhao Xing, Junhao Huang +5

    cs.LGcs.AIcs.CLarXiv:2603.09697v22026
  38. Learning to Faithfully Rationalize by Construction

    Sarthak Jain, Sarah Wiegreffe, Yuval Pinter +1

    cs.CLcs.AIcs.LGarXiv:2005.00115v12020
  39. Methane Detection On Board Satellites from Unorthorectified Imagery

    Luca Marini, Maggie Chen, Hala Lamdouar +3

    cs.CVcs.AIcs.LGarXiv:2609.04906v12026
  40. Beyond Outcome Verification: Verifiable Process Reward Models for Structured Reasoning

    Massimiliano Pronesti, Anya Belz, Yufang Hou

    cs.CLcs.AIarXiv:2601.17223v12026
  41. Graph is a Substrate Across Data Modalities

    Ziming Li, Xiaoming Wu, Zehong Wang +6

    cs.LGcs.AIarXiv:2601.22384v22026
  42. TFN: An Interpretable Neural Network with Time-Frequency Transform Embedded for Intelligent Fault Diagnosis

    Qian Chen, Xingjian Dong, Guowei Tu +3

    cs.AIcs.LGeess.SParXiv:2209.01992v22022
  43. Olmix: A Framework for Data Mixing Throughout LM Development

    Mayee F. Chen, Tyler Murray, David Heineman +5

    cs.LGcs.AIcs.CLarXiv:2602.12237v12026
  44. Adaptation Interfaces for In-Context Tabular Foundation Models in Time-to-Event Prediction

    Minh-Khoi Pham, Luca Cotugno, Dan Cernei +10

    cs.LGcs.AIarXiv:2609.04901v12026
  45. AgentLeak: A Benchmark for Internal-Channel Privacy Leakage in Multi-Agent LLM Systems

    Faouzi El Yagoubi, Godwin Badu-Marfo, Ranwa Al Mallah

    cs.AIarXiv:2602.11510v32026
  46. Attention-guided super-resolution of 4D flow MRI in carotid arteries

    Ali Mokhtari, Dominik Obrist

    physics.med-phcs.AIphysics.flu-dynarXiv:2609.04891v12026
  47. DRACO: a Cross-Domain Benchmark for Deep Research Accuracy, Completeness, and Objectivity

    Joey Zhong, Hao Zhang, Clare Southern +7

    cs.LGcs.AIarXiv:2602.11685v12026
  48. A very preliminary analysis of DALL-E 2

    Gary Marcus, Ernest Davis, Scott Aaronson

    cs.CVcs.AIarXiv:2204.13807v22022
  49. SimFuse3D: Source-Guided Target Simulation and Confidence-Guided Multi-Stage Localization Reweighting for Cross-Platform 3D Object Detection

    Yongchun Lin, Xinliang Zhang, Yun Zou +7

    cs.CVcs.AIarXiv:2609.04886v12026
  50. ReCAST: Restoration-aware Cascaded Stage-wise Training for Obfuscated SMS Risk Classification

    Jieyun Huang, Yi Shen, Kaikai Zhao +7

    cs.CRcs.AIarXiv:2609.04878v12026
  51. Milestone-Guided Policy Learning for Long-Horizon Language Agents

    Zixuan Wang, Yuchen Yan, Hongxing Li +7

    cs.CLcs.AIarXiv:2605.06078v12026
  52. Automated Optimization Modeling via a Localizable Error-Driven Perspective

    Weiting Liu, Han Wu, Yufei Kuang +4

    cs.LGcs.AIcs.CLarXiv:2602.11164v12026
  53. Mitigating Performance Discrepancy in Cross-Domain 3D Class-Incremental Learning

    Jinge Ma, Gautham Vinod, Bruce Coburn +3

    cs.CVcs.AIarXiv:2609.04860v12026
  54. Cost-Aware Hierarchical Multi-Agent Ransomware Detection and Family Attribution

    Mubashar Iqbal, Asifullah Khan

    cs.CRcs.AIarXiv:2609.04820v12026
  55. MABPD: Multi-Agent Bias Probing & Detection via Structured Argument Debate

    Garvit Joshi, Stavya Dhyani, Jasmine +1

    cs.CLcs.AIarXiv:2609.04841v12026
  56. PRISM-Bench: An Audio-Centric Diagnostic Benchmark for Text-to-Audio-Video Generation

    Yuchen Sun, Qian Yang, Jun Wang +4

    cs.MMcs.AIcs.SDarXiv:2609.04867v12026
  57. A comprehensive survey of research towards AI-enabled unmanned aerial systems in pre-, active-, and post-wildfire management

    Sayed Pedram Haeri Boroujeni, Abolfazl Razi, Sahand Khoshdel +7

    cs.LGcs.AIarXiv:2401.02456v12024
  58. Reinforcement Learning for improving Large Language Models' Catalan text simplification capabilities

    Arnau Ayguadé Domingo, Stefan Bott, Horacio Saggion

    cs.CLcs.AIarXiv:2609.04823v12026
  59. Distributed Linguistic Representations in Decision Making: Taxonomy, Key Elements and Applications, and Challenges in Data Science and Explainable Artificial Intelligence

    Yuzhu Wu, Zhen Zhang, Gang Kou +5

    cs.AIarXiv:2008.01499v22020
  60. CC-Mediation: Evaluating Large Language Models for Cross-Cultural Conflict Mediation

    Suhyun Lee, Wenxuan Zhang, W. Quin Yow +1

    cs.CLcs.AIarXiv:2609.04855v12026