Artificial Intelligence
Papers filed under cs.AI on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.
Search paper metadata (including unsummarized papers)
2,221 to 2,280 of 15,305
Safe Harness Self-Evolution: A Theoretical Analysis of Feasibility and Limits
Qianshu Cai, Yonggang Zhang, Jun Nie +6
cs.AIarXiv:2609.08175v12026Explainable Deep Learning Methods in Medical Image Classification: A Survey
Cristiano Patrício, João C. Neves, Luís F. Teixeira
eess.IVcs.AIcs.CVarXiv:2205.04766v32022Explainable AI for Bioinformatics: Methods, Tools, and Applications
Md. Rezaul Karim, Tanhim Islam, Oya Beyan +4
q-bio.QMcs.AIcs.LGarXiv:2212.13261v32022SWE-Bench Pro Verified: A Reliable Benchmark for Software Engineering Agents
Pujun Zheng, Zixin Shang, Shufan Jiang +5
cs.AIcs.SEarXiv:2609.08149v12026Constraint-based Causal Discovery from Multiple Interventions over Overlapping Variable Sets
Sofia Triantafillou, Ioannis Tsamardinos
stat.MLcs.AIarXiv:1403.2150v12014CIVI: A Framework for Diagnosing Search Agent Failures in Civic Information
Dingying Liu, Yunshun Zhong, Wentao Zhang +1
cs.AIarXiv:2609.08094v12026OntologyBench: Can Dense Retrieval Satisfy Structured Biomedical Constraints?
Xiao Yu Cindy Zhang, Wyeth Wasserman, Jian Zhu
cs.AIarXiv:2609.08174v12026Key Path Identification for Resolving Knowledge Conflicts via SAE-based Steering
Wenbo Zhang, Zhongxiang Sun, Zhiguang Han +1
cs.AIarXiv:2609.08173v12026WorldAgen: Unified State-Action Prediction with Test-Time World Model Training
Chi Wan, Kangrui Wang, Yuan Si +2
cs.AIarXiv:2609.08162v12026Robot Learning in Homes: Improving Generalization and Reducing Dataset Bias
Abhinav Gupta, Adithyavairavan Murali, Dhiraj Gandhi +1
cs.ROcs.AIcs.CVarXiv:1807.07049v12018CrossCLR: Cross-modal Contrastive Learning For Multi-modal Video Representations
Mohammadreza Zolfaghari, Yi Zhu, Peter Gehler +1
cs.CVcs.AIcs.LGarXiv:2109.14910v12021SchemeArena: Factorized Stress Testing of Scheming in LLM Agents
Jie Ruan, Inderjeet Nair, Amy Liu +3
cs.AIcs.CLarXiv:2609.08126v12026Transformer-Empowered 6G Intelligent Networks: From Massive MIMO Processing to Semantic Communication
Yang Wang, Zhen Gao, Dezhi Zheng +3
cs.ITcs.AIcs.LGarXiv:2205.03770v42022SpectFormer: Frequency and Attention is what you need in a Vision Transformer
Badri N. Patro, Vinay P. Namboodiri, Vijay Srinivas Agneeswaran
cs.CVcs.AIcs.CLarXiv:2304.06446v22023Router Prior Bias: Preserving Base Routing Structure in MoE Post-Training
Jaedeok Lee, Keonwoo Kim, Dongyoon Han +3
cs.AIarXiv:2609.08115v12026Artificial Intelligence-Assisted Digital Inventory of Cultural Heritage & Traditional Knowledge: Case for Indonesian Open Digital Library of Culture
Hokky Situngkir
cs.AIcs.DLcs.HCarXiv:2609.08105v12026Post-Training Ternarization of Qwen3-4B Capability, Effective Bit Budget, Storage Compression, and Deployment
Anirudh Malik, M Sparsh Mehra, Poojith Devan
cs.AIcs.LGarXiv:2609.01962v12026Inference-Time Nash Alignment
Hadi Hosseini, Debmalya Mandal, Duohan Zhang
cs.AIarXiv:2609.08082v12026ResidualAuth: What Authorization State Must Language Agents Preserve under Revocable Delegation?
Moonwon Choi, Seokho Jeong, Seunggeun Lee
cs.AIcs.CRarXiv:2609.08062v12026Eliciting Self-Verification in Multimodal Reasoning Agents with Reinforcement Learning
Vishwas Sathish, Viresh Ranjan, Xinliang Zhu +2
cs.AIcs.CLcs.CVarXiv:2609.08025v12026Online Fair Division: analysing a Food Bank problem
Martin Aleksandrov, Haris Aziz, Serge Gaspers +1
cs.GTcs.AIcs.MAarXiv:1502.07571v22015ManipulaTHOR: A Framework for Visual Object Manipulation
Kiana Ehsani, Winson Han, Alvaro Herrasti +5
cs.CVcs.AIcs.LGarXiv:2104.11213v12021RevalExo: A Functional Daily-Activity Benchmark for Inertial and Visual Locomotion Mode Recognition in Older Adults and Clinical Cohorts
Diwas Lamsal, Juha Carlon, Reinhard Claeys +8
cs.AIcs.CVarXiv:2609.08090v12026Automated Design of Inventory Policy with Large Language Models: An Exploratory Study
Fenghua Yang, Preet Baxi, Yi Zhang +4
cs.AIarXiv:2609.08071v12026Benchmarking datasets for Anomaly-based Network Intrusion Detection: KDD CUP 99 alternatives
Abhishek Divekar, Meet Parekh, Vaibhav Savla +2
cs.LGcs.AIcs.CRarXiv:1811.05372v12018A Layered Analysis of Disagreement And Answer Quality in Multi-Agent LLM Debate
Chen Qian
cs.AIarXiv:2609.08016v12026Diffusion-SDF: Text-to-Shape via Voxelized Diffusion
Muheng Li, Yueqi Duan, Jie Zhou +1
cs.CVcs.AIcs.GRarXiv:2212.03293v22022From Version Conflicts to Decision Conflicts: Selective Revalidation for Long-Running AI Agents
Yongjian Lyu, Yang Ren, Ruofei Lai +1
cs.AIcs.DBarXiv:2609.08015v12026Sparks of In Silico Cognitive Science: Theories from Simulated Data Can Generalize to Humans
Akshay K. Jagadish, Younes Strittmatter, Nori Jacoby +4
cs.AIarXiv:2609.08003v12026Context-Aware Generative Adversarial Privacy
Chong Huang, Peter Kairouz, Xiao Chen +2
cs.LGcs.AIcs.CRarXiv:1710.09549v32017Interpretable and Accurate Fine-grained Recognition via Region Grouping
Zixuan Huang, Yin Li
cs.CVcs.AIcs.LGarXiv:2005.10411v12020BEVBert: Multimodal Map Pre-training for Language-guided Navigation
Dong An, Yuankai Qi, Yangguang Li +4
cs.CVcs.AIcs.CLarXiv:2212.04385v22022OmniACT: A Dataset and Benchmark for Enabling Multimodal Generalist Autonomous Agents for Desktop and Web
Raghav Kapoor, Yash Parag Butala, Melisa Russak +4
cs.AIcs.CLcs.CVarXiv:2402.17553v32024Towards Transferable Adversarial Attacks on Vision Transformers
Zhipeng Wei, Jingjing Chen, Micah Goldblum +3
cs.CVcs.AIarXiv:2109.04176v32021Generalized Preference Optimization: A Unified Approach to Offline Alignment
Yunhao Tang, Zhaohan Daniel Guo, Zeyu Zheng +7
cs.LGcs.AIarXiv:2402.05749v22024Mini-Batch Risk-Averse Deep Q-Learning: A Robot Navigation Case Study
Aayush Patel, Andrzej Ruszczyński
cs.AImath.OCarXiv:2609.07998v12026When Can LLM Digital Twins Reduce Human Measurement? From Behavioral Fidelity to Statistical Substitutability
Steven Wang, Kyle Hunt, Shaojie Tang +1
cs.AIstat.AParXiv:2609.07987v12026From Event Logs to Governed Action: A BlueSky Agenda for Agentic Process Mining
Yiyuan Yang, Zheshun Wu, Yong Chu +3
cs.AIcs.CEarXiv:2609.07984v12026Support Topology and Gradient Mixing in Sinkhorn Layers
Dylan Forde
cs.AIcs.LGarXiv:2609.07954v12026Bayesian Locality Sensitive Hashing for Fast Similarity Search
Venu Satuluri, Srinivasan Parthasarathy
cs.DBcs.AIcs.DSarXiv:1110.1328v32011CausalVerify: An Execution-Grounded Benchmark for LLM Causal Inference Workflows
Yonghong Zhang, Ricardo Correia, Isabel M. Parra +1
cs.AIcs.CLecon.EMarXiv:2609.07944v12026Beliefs and Behavior in Language Models
Alex Smolin, Bryan Wilder
cs.AIcs.LGarXiv:2609.07943v12026FrogNano: Training a 4B Coding Agent via Online Task Synthesis
Minseon Kim, Zhengyan Shi, Emiliano Penaloza +14
cs.AIarXiv:2609.07925v12026PRIMUS: Identity, Governance, and Verification for Multi-Agent Federations
Sasank Annapureddy, Anjaneya Prasad Thamatani
cs.AIarXiv:2609.07910v12026SceneVerse: Scaling 3D Vision-Language Learning for Grounded Scene Understanding
Baoxiong Jia, Yixin Chen, Huangyue Yu +5
cs.CVcs.AIcs.CLarXiv:2401.09340v32024Quantization Amplifies Determinism, Not Bias: Scale-Dependent Behavioral Effects of Serving-Time Weight Compression
Dachi Kurtskhalia
cs.AIarXiv:2609.07901v12026GraphRouter: A Graph-based Router for LLM Selections
Tao Feng, Yanzhen Shen, Jiaxuan You
cs.AIarXiv:2410.03834v22024Do Large Language Models Know What They Don't Know II? A Fully Behavioral, Non-Cognitive Measure of Epistemic Honesty
Ali Şenol, H. Russell Bernard, Huan Liu
cs.AIarXiv:2609.07879v12026Generalizing Dataset Distillation via Deep Generative Prior
George Cazenavette, Tongzhou Wang, Antonio Torralba +2
cs.CVcs.AIcs.LGarXiv:2305.01649v22023RoboCLIP: One Demonstration is Enough to Learn Robot Policies
Sumedh A Sontakke, Jesse Zhang, Sébastien M. R. Arnold +5
cs.AIcs.ROarXiv:2310.07899v12023MM-Embed: Universal Multimodal Retrieval with Multimodal LLMs
Sheng-Chieh Lin, Chankyu Lee, Mohammad Shoeybi +3
cs.CLcs.AIcs.CVarXiv:2411.02571v22024VideoDex: Learning Dexterity from Internet Videos
Kenneth Shaw, Shikhar Bahl, Deepak Pathak
cs.ROcs.AIcs.CVarXiv:2212.04498v12022The Profit Alignment Problem: How Profit Mandates Induce Alignment Failures in LLMs
Eric So
cs.AIecon.GNarXiv:2609.07731v12026A Benchmark for Systematic Generalization in Grounded Language Understanding
Laura Ruis, Jacob Andreas, Marco Baroni +2
cs.CLcs.AIcs.LGarXiv:2003.05161v22020Explainable Temporal Attention-based Defect Detection For Fillet Joints in Real-Time Gas Metal Arc Welding Based on Multi-modal Data
Mobina Mobaraki, Mahyar Asadi, Klaske Van Heusden +1
cs.AIcs.LGarXiv:2609.07893v12026Long-form factuality in large language models
Jerry Wei, Chengrun Yang, Xinying Song +9
cs.CLcs.AIcs.LGarXiv:2403.18802v42024Understanding the Impact of Model Pruning on Long-Tail Forgetting and Explanation Reliability in Medical Imaging
Nazish Khalid, Tausifa Jan Saleem, Amal Saqib +2
cs.AIarXiv:2609.07803v12026R-GAP: Recursive Gradient Attack on Privacy
Junyi Zhu, Matthew Blaschko
cs.LGcs.AIarXiv:2010.07733v32020Text-to-SQL Generation for Question Answering on Electronic Medical Records
Ping Wang, Tian Shi, Chandan K. Reddy
cs.CLcs.AIcs.IRarXiv:1908.01839v22019What Does an LLM-Agent Leaderboard Rank Actually Compare?
Wei-Jung Huang
cs.AIarXiv:2609.07785v12026