Artificial Intelligence
Papers filed under cs.AI on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.
Search paper metadata (including unsummarized papers)
3,121 to 3,180 of 15,368
OP-Bench: Benchmarking Over-Personalization for Memory-Augmented Personalized Conversational Agents
Yulin Hu, Zimo Long, Jiahe Guo +5
cs.CLcs.AIarXiv:2601.13722v12026Adaptive Multi-Granularity Temporal Modeling for Weakly Supervised Video Anomaly Detection
Changyi Li, Yu Xiao
cs.CVcs.AIarXiv:2609.05066v12026Continual Semantic Segmentation via Repulsion-Attraction of Sparse and Disentangled Latent Representations
Umberto Michieli, Pietro Zanuttigh
cs.CVcs.AIcs.LGarXiv:2103.06342v32021How do LLMs Evaluate Perceived Moral Agency? Investigating Moral Decision-Making in Human-Artificial Agents Interactions
Fernanda Mansilla, Aloysius Tok, Bahia Guellaï +2
cs.CLcs.AIarXiv:2609.05037v12026LIBERO-X: Robustness Litmus for Vision-Language-Action Models
Guodong Wang, Chenkai Zhang, Qingjie Liu +4
cs.CVcs.AIcs.ROarXiv:2602.06556v12026How a Chatbot's Response Style Shapes a Classroom: A Multi-Agent Simulation of Students Consulting AI
Rin Tamai, Yuya Dan
cs.HCcs.AIcs.CYarXiv:2609.05018v12026Goedel-Code-Prover: Hierarchical Proof Search for Open State-of-the-Art Code Verification
Zenan Li, Ziran Yang, Deyuan He +7
cs.SEcs.AIarXiv:2603.19329v32026Adversarial Robustness Against the Union of Multiple Perturbation Models
Pratyush Maini, Eric Wong, J. Zico Kolter
cs.LGcs.AIstat.MLarXiv:1909.04068v22019ARIA - An Agentic Framework for Autonomous Testing of Infotainment Systems
António Azevedo, Bruno Lima, João Pascoal Faria
cs.SEcs.AIarXiv:2609.04913v12026RAM: Recover Any 3D Human Motion in-the-Wild
Sen Jia, Ning Zhu, Jinqin Zhong +4
cs.CVcs.AIarXiv:2603.19929v22026Qlippy: A Retrieval-Augmented GenAI Assistant for Reproducible Quantum Workflows and Experiment Tracking
Mahee Gamage, Vlad Stirbu
quant-phcs.AIarXiv:2609.05039v12026The Paradigm Shift: A Comprehensive Survey on Large Vision Language Models for Multimodal Fake News Detection
Wei Ai, Yilong Tan, Yuntao Shou +4
cs.AIcs.CVarXiv:2601.15316v12026SideQuest: Model-Driven KV Cache Management for Long-Horizon Agentic Reasoning
Sanjay Kariyappa, G. Edward Suh
cs.AIcs.LGarXiv:2602.22603v22026Fine-R1: Make Multi-modal LLMs Excel in Fine-Grained Visual Recognition by Chain-of-Thought Reasoning
Hulingxiao He, Zijun Geng, Yuxin Peng
cs.CVcs.AIarXiv:2602.07605v32026Why Diffusion Language Models Struggle with Truly Parallel (Non-Autoregressive) Decoding?
Pengxiang Li, Dilxat Muhtar, Tianlong Chen +2
cs.CLcs.AIarXiv:2602.23225v22026Amortizing Scaling Law Construction Costs
Abhash Kumar Jha, Diana Alexandra Onuţu, Neeratyoy Mallik +6
cs.LGcs.AIarXiv:2609.05016v12026MCPO: Modality-Contrastive Preference Optimization for Multimodal Chain-of-Thought Compression
Guangheng Yang, Zhenliang Ni, Zhenkai Wu +4
cs.CVcs.AIarXiv:2609.04947v12026Look-Ahead-Bench: a Standardized Benchmark of Look-ahead Bias in Point-in-Time LLMs for Finance
Mostapha Benhenda
cs.AIcs.CLcs.LGarXiv:2601.13770v12026Language-Conditioned World Modeling for Visual Navigation
Yifei Dong, Fengyi Wu, Yilong Dai +10
cs.CVcs.AIcs.ROarXiv:2603.26741v12026Latent Introspection: Models Can Detect Prior Concept Injections
Theia Pearson-Vogel, Martin Vanek, Raymond Douglas +1
cs.AIcs.LGarXiv:2602.20031v22026MUZZLE: Adaptive Agentic Red-Teaming of Web Agents Against Indirect Prompt Injection Attacks
Georgios Syros, Evan Rose, Brian Grinstead +4
cs.CRcs.AIarXiv:2602.09222v22026Latent Adversarial Training Improves Robustness to Persistent Harmful Behaviors in LLMs
Abhay Sheshadri, Aidan Ewart, Phillip Guo +8
cs.LGcs.AIcs.CLarXiv:2407.15549v32024One Diffusion Model, Two Roles: Guided Trajectory Planning and Safety-Critical Scenario Generation in Closed-Loop Simulation
Arka Pal, Rajesh Kumar, Hannes Eriksson +4
cs.CVcs.AIcs.LGarXiv:2609.04921v12026MemPO: Self-Memory Policy Optimization for Long-Horizon Agents
Ruoran Li, Xinghua Zhang, Haiyang Yu +7
cs.AIarXiv:2603.00680v42026Relational Graph Learning for Crowd Navigation
Changan Chen, Sha Hu, Payam Nikdel +2
cs.ROcs.AIcs.LGarXiv:1909.13165v32019FlowBalance: Verifier-Grounded Self-Improvement from On-Policy Reasoning Experience
Zixun Huang, Kishan Panaganti, Haitao Mi +1
cs.LGcs.AIarXiv:2609.03241v12026Better Understanding, Better Fixes? A Study of Hallucination in LLM-based Automated Program Repair
Xuemeng Cai, Jiakun Liu, Linhan Yang +2
cs.SEcs.AIarXiv:2609.04909v12026Sound-based Multi-Person 3D Pose Estimation
Yusuke Oumi, Yuto Shibata, Go Irie +3
cs.CVcs.AIcs.LGarXiv:2609.04902v12026RefactorPlatform: An Open-Source Harness for Controlled Evaluation of Repository-Scale Refactoring Agents
Aziz Ben Amor, Drish Mali, Mann Acharya +2
cs.CLcs.AIarXiv:2609.04898v12026Resource-Efficient Iterative LLM-Based NAS with Feedback Memory
Xiaojie Gu, Dmitry Ignatov, Radu Timofte
cs.LGcs.AIarXiv:2603.12091v12026Institutional AI: Governing LLM Collusion in Multi-Agent Cournot Markets via Public Governance Graphs
Marcantonio Bracale Syrnikov, Federico Pierucci, Marcello Galisai +6
cs.GTcs.AIarXiv:2601.11369v22026Autorubric: A Unifying Framework for Rubric-Based LLM Evaluation on Non-Verifiable Tasks
Delip Rao, Chris Callison-Burch
cs.CLcs.AIarXiv:2603.00077v32026Forgetting Without Restarting: Execution-State Unlearning for Stateful LLM Agents
Chao Yao, Yangbo Wei, Zhen Huang +5
cs.CRcs.AIarXiv:2609.04875v12026ContractSkill: Repairable Contract-Based Skills for Multimodal Web Agents
Zijian Lu, Yiping Zuo, Yupeng Nie +4
cs.SEcs.AIarXiv:2603.20340v32026A Molecular Multimodal Foundation Model Associating Molecule Graphs with Natural Language
Bing Su, Dazhao Du, Zhao Yang +6
cs.LGcs.AIarXiv:2209.05481v12022TreeFI: Value-Aware Statistical Fault Injection for Deep Neural Networks
Noam Bires, Marcello Traiola, Angeliki Kritikakou +1
cs.ARcs.AIarXiv:2609.04912v12026Mousse: Rectifying the Geometry of Muon with Curvature-Aware Preconditioning
Yechen Zhang, Shuhao Xing, Junhao Huang +5
cs.LGcs.AIcs.CLarXiv:2603.09697v22026Learning to Faithfully Rationalize by Construction
Sarthak Jain, Sarah Wiegreffe, Yuval Pinter +1
cs.CLcs.AIcs.LGarXiv:2005.00115v12020Methane Detection On Board Satellites from Unorthorectified Imagery
Luca Marini, Maggie Chen, Hala Lamdouar +3
cs.CVcs.AIcs.LGarXiv:2609.04906v12026Beyond Outcome Verification: Verifiable Process Reward Models for Structured Reasoning
Massimiliano Pronesti, Anya Belz, Yufang Hou
cs.CLcs.AIarXiv:2601.17223v12026Graph is a Substrate Across Data Modalities
Ziming Li, Xiaoming Wu, Zehong Wang +6
cs.LGcs.AIarXiv:2601.22384v22026TFN: An Interpretable Neural Network with Time-Frequency Transform Embedded for Intelligent Fault Diagnosis
Qian Chen, Xingjian Dong, Guowei Tu +3
cs.AIcs.LGeess.SParXiv:2209.01992v22022Olmix: A Framework for Data Mixing Throughout LM Development
Mayee F. Chen, Tyler Murray, David Heineman +5
cs.LGcs.AIcs.CLarXiv:2602.12237v12026Adaptation Interfaces for In-Context Tabular Foundation Models in Time-to-Event Prediction
Minh-Khoi Pham, Luca Cotugno, Dan Cernei +10
cs.LGcs.AIarXiv:2609.04901v12026AgentLeak: A Benchmark for Internal-Channel Privacy Leakage in Multi-Agent LLM Systems
Faouzi El Yagoubi, Godwin Badu-Marfo, Ranwa Al Mallah
cs.AIarXiv:2602.11510v32026Attention-guided super-resolution of 4D flow MRI in carotid arteries
Ali Mokhtari, Dominik Obrist
physics.med-phcs.AIphysics.flu-dynarXiv:2609.04891v12026DRACO: a Cross-Domain Benchmark for Deep Research Accuracy, Completeness, and Objectivity
Joey Zhong, Hao Zhang, Clare Southern +7
cs.LGcs.AIarXiv:2602.11685v12026A very preliminary analysis of DALL-E 2
Gary Marcus, Ernest Davis, Scott Aaronson
cs.CVcs.AIarXiv:2204.13807v22022SimFuse3D: Source-Guided Target Simulation and Confidence-Guided Multi-Stage Localization Reweighting for Cross-Platform 3D Object Detection
Yongchun Lin, Xinliang Zhang, Yun Zou +7
cs.CVcs.AIarXiv:2609.04886v12026ReCAST: Restoration-aware Cascaded Stage-wise Training for Obfuscated SMS Risk Classification
Jieyun Huang, Yi Shen, Kaikai Zhao +7
cs.CRcs.AIarXiv:2609.04878v12026Milestone-Guided Policy Learning for Long-Horizon Language Agents
Zixuan Wang, Yuchen Yan, Hongxing Li +7
cs.CLcs.AIarXiv:2605.06078v12026Automated Optimization Modeling via a Localizable Error-Driven Perspective
Weiting Liu, Han Wu, Yufei Kuang +4
cs.LGcs.AIcs.CLarXiv:2602.11164v12026Mitigating Performance Discrepancy in Cross-Domain 3D Class-Incremental Learning
Jinge Ma, Gautham Vinod, Bruce Coburn +3
cs.CVcs.AIarXiv:2609.04860v12026Cost-Aware Hierarchical Multi-Agent Ransomware Detection and Family Attribution
Mubashar Iqbal, Asifullah Khan
cs.CRcs.AIarXiv:2609.04820v12026MABPD: Multi-Agent Bias Probing & Detection via Structured Argument Debate
Garvit Joshi, Stavya Dhyani, Jasmine +1
cs.CLcs.AIarXiv:2609.04841v12026PRISM-Bench: An Audio-Centric Diagnostic Benchmark for Text-to-Audio-Video Generation
Yuchen Sun, Qian Yang, Jun Wang +4
cs.MMcs.AIcs.SDarXiv:2609.04867v12026A comprehensive survey of research towards AI-enabled unmanned aerial systems in pre-, active-, and post-wildfire management
Sayed Pedram Haeri Boroujeni, Abolfazl Razi, Sahand Khoshdel +7
cs.LGcs.AIarXiv:2401.02456v12024Reinforcement Learning for improving Large Language Models' Catalan text simplification capabilities
Arnau Ayguadé Domingo, Stefan Bott, Horacio Saggion
cs.CLcs.AIarXiv:2609.04823v12026Distributed Linguistic Representations in Decision Making: Taxonomy, Key Elements and Applications, and Challenges in Data Science and Explainable Artificial Intelligence
Yuzhu Wu, Zhen Zhang, Gang Kou +5
cs.AIarXiv:2008.01499v22020CC-Mediation: Evaluating Large Language Models for Cross-Cultural Conflict Mediation
Suhyun Lee, Wenxuan Zhang, W. Quin Yow +1
cs.CLcs.AIarXiv:2609.04855v12026