Artificial Intelligence

Papers filed under cs.AI on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.

Search paper metadata (including unsummarized papers)

2,221 to 2,280 of 15,305

  1. Safe Harness Self-Evolution: A Theoretical Analysis of Feasibility and Limits

    Qianshu Cai, Yonggang Zhang, Jun Nie +6

    cs.AIarXiv:2609.08175v12026
  2. Explainable Deep Learning Methods in Medical Image Classification: A Survey

    Cristiano Patrício, João C. Neves, Luís F. Teixeira

    eess.IVcs.AIcs.CVarXiv:2205.04766v32022
  3. Explainable AI for Bioinformatics: Methods, Tools, and Applications

    Md. Rezaul Karim, Tanhim Islam, Oya Beyan +4

    q-bio.QMcs.AIcs.LGarXiv:2212.13261v32022
  4. SWE-Bench Pro Verified: A Reliable Benchmark for Software Engineering Agents

    Pujun Zheng, Zixin Shang, Shufan Jiang +5

    cs.AIcs.SEarXiv:2609.08149v12026
  5. Constraint-based Causal Discovery from Multiple Interventions over Overlapping Variable Sets

    Sofia Triantafillou, Ioannis Tsamardinos

    stat.MLcs.AIarXiv:1403.2150v12014
  6. CIVI: A Framework for Diagnosing Search Agent Failures in Civic Information

    Dingying Liu, Yunshun Zhong, Wentao Zhang +1

    cs.AIarXiv:2609.08094v12026
  7. OntologyBench: Can Dense Retrieval Satisfy Structured Biomedical Constraints?

    Xiao Yu Cindy Zhang, Wyeth Wasserman, Jian Zhu

    cs.AIarXiv:2609.08174v12026
  8. Key Path Identification for Resolving Knowledge Conflicts via SAE-based Steering

    Wenbo Zhang, Zhongxiang Sun, Zhiguang Han +1

    cs.AIarXiv:2609.08173v12026
  9. WorldAgen: Unified State-Action Prediction with Test-Time World Model Training

    Chi Wan, Kangrui Wang, Yuan Si +2

    cs.AIarXiv:2609.08162v12026
  10. Robot Learning in Homes: Improving Generalization and Reducing Dataset Bias

    Abhinav Gupta, Adithyavairavan Murali, Dhiraj Gandhi +1

    cs.ROcs.AIcs.CVarXiv:1807.07049v12018
  11. CrossCLR: Cross-modal Contrastive Learning For Multi-modal Video Representations

    Mohammadreza Zolfaghari, Yi Zhu, Peter Gehler +1

    cs.CVcs.AIcs.LGarXiv:2109.14910v12021
  12. SchemeArena: Factorized Stress Testing of Scheming in LLM Agents

    Jie Ruan, Inderjeet Nair, Amy Liu +3

    cs.AIcs.CLarXiv:2609.08126v12026
  13. Transformer-Empowered 6G Intelligent Networks: From Massive MIMO Processing to Semantic Communication

    Yang Wang, Zhen Gao, Dezhi Zheng +3

    cs.ITcs.AIcs.LGarXiv:2205.03770v42022
  14. SpectFormer: Frequency and Attention is what you need in a Vision Transformer

    Badri N. Patro, Vinay P. Namboodiri, Vijay Srinivas Agneeswaran

    cs.CVcs.AIcs.CLarXiv:2304.06446v22023
  15. Router Prior Bias: Preserving Base Routing Structure in MoE Post-Training

    Jaedeok Lee, Keonwoo Kim, Dongyoon Han +3

    cs.AIarXiv:2609.08115v12026
  16. Artificial Intelligence-Assisted Digital Inventory of Cultural Heritage & Traditional Knowledge: Case for Indonesian Open Digital Library of Culture

    Hokky Situngkir

    cs.AIcs.DLcs.HCarXiv:2609.08105v12026
  17. Post-Training Ternarization of Qwen3-4B Capability, Effective Bit Budget, Storage Compression, and Deployment

    Anirudh Malik, M Sparsh Mehra, Poojith Devan

    cs.AIcs.LGarXiv:2609.01962v12026
  18. Inference-Time Nash Alignment

    Hadi Hosseini, Debmalya Mandal, Duohan Zhang

    cs.AIarXiv:2609.08082v12026
  19. ResidualAuth: What Authorization State Must Language Agents Preserve under Revocable Delegation?

    Moonwon Choi, Seokho Jeong, Seunggeun Lee

    cs.AIcs.CRarXiv:2609.08062v12026
  20. Eliciting Self-Verification in Multimodal Reasoning Agents with Reinforcement Learning

    Vishwas Sathish, Viresh Ranjan, Xinliang Zhu +2

    cs.AIcs.CLcs.CVarXiv:2609.08025v12026
  21. Online Fair Division: analysing a Food Bank problem

    Martin Aleksandrov, Haris Aziz, Serge Gaspers +1

    cs.GTcs.AIcs.MAarXiv:1502.07571v22015
  22. ManipulaTHOR: A Framework for Visual Object Manipulation

    Kiana Ehsani, Winson Han, Alvaro Herrasti +5

    cs.CVcs.AIcs.LGarXiv:2104.11213v12021
  23. RevalExo: A Functional Daily-Activity Benchmark for Inertial and Visual Locomotion Mode Recognition in Older Adults and Clinical Cohorts

    Diwas Lamsal, Juha Carlon, Reinhard Claeys +8

    cs.AIcs.CVarXiv:2609.08090v12026
  24. Automated Design of Inventory Policy with Large Language Models: An Exploratory Study

    Fenghua Yang, Preet Baxi, Yi Zhang +4

    cs.AIarXiv:2609.08071v12026
  25. Benchmarking datasets for Anomaly-based Network Intrusion Detection: KDD CUP 99 alternatives

    Abhishek Divekar, Meet Parekh, Vaibhav Savla +2

    cs.LGcs.AIcs.CRarXiv:1811.05372v12018
  26. A Layered Analysis of Disagreement And Answer Quality in Multi-Agent LLM Debate

    Chen Qian

    cs.AIarXiv:2609.08016v12026
  27. Diffusion-SDF: Text-to-Shape via Voxelized Diffusion

    Muheng Li, Yueqi Duan, Jie Zhou +1

    cs.CVcs.AIcs.GRarXiv:2212.03293v22022
  28. From Version Conflicts to Decision Conflicts: Selective Revalidation for Long-Running AI Agents

    Yongjian Lyu, Yang Ren, Ruofei Lai +1

    cs.AIcs.DBarXiv:2609.08015v12026
  29. Sparks of In Silico Cognitive Science: Theories from Simulated Data Can Generalize to Humans

    Akshay K. Jagadish, Younes Strittmatter, Nori Jacoby +4

    cs.AIarXiv:2609.08003v12026
  30. Context-Aware Generative Adversarial Privacy

    Chong Huang, Peter Kairouz, Xiao Chen +2

    cs.LGcs.AIcs.CRarXiv:1710.09549v32017
  31. Interpretable and Accurate Fine-grained Recognition via Region Grouping

    Zixuan Huang, Yin Li

    cs.CVcs.AIcs.LGarXiv:2005.10411v12020
  32. BEVBert: Multimodal Map Pre-training for Language-guided Navigation

    Dong An, Yuankai Qi, Yangguang Li +4

    cs.CVcs.AIcs.CLarXiv:2212.04385v22022
  33. OmniACT: A Dataset and Benchmark for Enabling Multimodal Generalist Autonomous Agents for Desktop and Web

    Raghav Kapoor, Yash Parag Butala, Melisa Russak +4

    cs.AIcs.CLcs.CVarXiv:2402.17553v32024
  34. Towards Transferable Adversarial Attacks on Vision Transformers

    Zhipeng Wei, Jingjing Chen, Micah Goldblum +3

    cs.CVcs.AIarXiv:2109.04176v32021
  35. Generalized Preference Optimization: A Unified Approach to Offline Alignment

    Yunhao Tang, Zhaohan Daniel Guo, Zeyu Zheng +7

    cs.LGcs.AIarXiv:2402.05749v22024
  36. Mini-Batch Risk-Averse Deep Q-Learning: A Robot Navigation Case Study

    Aayush Patel, Andrzej Ruszczyński

    cs.AImath.OCarXiv:2609.07998v12026
  37. When Can LLM Digital Twins Reduce Human Measurement? From Behavioral Fidelity to Statistical Substitutability

    Steven Wang, Kyle Hunt, Shaojie Tang +1

    cs.AIstat.AParXiv:2609.07987v12026
  38. From Event Logs to Governed Action: A BlueSky Agenda for Agentic Process Mining

    Yiyuan Yang, Zheshun Wu, Yong Chu +3

    cs.AIcs.CEarXiv:2609.07984v12026
  39. Support Topology and Gradient Mixing in Sinkhorn Layers

    Dylan Forde

    cs.AIcs.LGarXiv:2609.07954v12026
  40. Bayesian Locality Sensitive Hashing for Fast Similarity Search

    Venu Satuluri, Srinivasan Parthasarathy

    cs.DBcs.AIcs.DSarXiv:1110.1328v32011
  41. CausalVerify: An Execution-Grounded Benchmark for LLM Causal Inference Workflows

    Yonghong Zhang, Ricardo Correia, Isabel M. Parra +1

    cs.AIcs.CLecon.EMarXiv:2609.07944v12026
  42. Beliefs and Behavior in Language Models

    Alex Smolin, Bryan Wilder

    cs.AIcs.LGarXiv:2609.07943v12026
  43. FrogNano: Training a 4B Coding Agent via Online Task Synthesis

    Minseon Kim, Zhengyan Shi, Emiliano Penaloza +14

    cs.AIarXiv:2609.07925v12026
  44. PRIMUS: Identity, Governance, and Verification for Multi-Agent Federations

    Sasank Annapureddy, Anjaneya Prasad Thamatani

    cs.AIarXiv:2609.07910v12026
  45. SceneVerse: Scaling 3D Vision-Language Learning for Grounded Scene Understanding

    Baoxiong Jia, Yixin Chen, Huangyue Yu +5

    cs.CVcs.AIcs.CLarXiv:2401.09340v32024
  46. Quantization Amplifies Determinism, Not Bias: Scale-Dependent Behavioral Effects of Serving-Time Weight Compression

    Dachi Kurtskhalia

    cs.AIarXiv:2609.07901v12026
  47. GraphRouter: A Graph-based Router for LLM Selections

    Tao Feng, Yanzhen Shen, Jiaxuan You

    cs.AIarXiv:2410.03834v22024
  48. Do Large Language Models Know What They Don't Know II? A Fully Behavioral, Non-Cognitive Measure of Epistemic Honesty

    Ali Şenol, H. Russell Bernard, Huan Liu

    cs.AIarXiv:2609.07879v12026
  49. Generalizing Dataset Distillation via Deep Generative Prior

    George Cazenavette, Tongzhou Wang, Antonio Torralba +2

    cs.CVcs.AIcs.LGarXiv:2305.01649v22023
  50. RoboCLIP: One Demonstration is Enough to Learn Robot Policies

    Sumedh A Sontakke, Jesse Zhang, Sébastien M. R. Arnold +5

    cs.AIcs.ROarXiv:2310.07899v12023
  51. MM-Embed: Universal Multimodal Retrieval with Multimodal LLMs

    Sheng-Chieh Lin, Chankyu Lee, Mohammad Shoeybi +3

    cs.CLcs.AIcs.CVarXiv:2411.02571v22024
  52. VideoDex: Learning Dexterity from Internet Videos

    Kenneth Shaw, Shikhar Bahl, Deepak Pathak

    cs.ROcs.AIcs.CVarXiv:2212.04498v12022
  53. The Profit Alignment Problem: How Profit Mandates Induce Alignment Failures in LLMs

    Eric So

    cs.AIecon.GNarXiv:2609.07731v12026
  54. A Benchmark for Systematic Generalization in Grounded Language Understanding

    Laura Ruis, Jacob Andreas, Marco Baroni +2

    cs.CLcs.AIcs.LGarXiv:2003.05161v22020
  55. Explainable Temporal Attention-based Defect Detection For Fillet Joints in Real-Time Gas Metal Arc Welding Based on Multi-modal Data

    Mobina Mobaraki, Mahyar Asadi, Klaske Van Heusden +1

    cs.AIcs.LGarXiv:2609.07893v12026
  56. Long-form factuality in large language models

    Jerry Wei, Chengrun Yang, Xinying Song +9

    cs.CLcs.AIcs.LGarXiv:2403.18802v42024
  57. Understanding the Impact of Model Pruning on Long-Tail Forgetting and Explanation Reliability in Medical Imaging

    Nazish Khalid, Tausifa Jan Saleem, Amal Saqib +2

    cs.AIarXiv:2609.07803v12026
  58. R-GAP: Recursive Gradient Attack on Privacy

    Junyi Zhu, Matthew Blaschko

    cs.LGcs.AIarXiv:2010.07733v32020
  59. Text-to-SQL Generation for Question Answering on Electronic Medical Records

    Ping Wang, Tian Shi, Chandan K. Reddy

    cs.CLcs.AIcs.IRarXiv:1908.01839v22019
  60. What Does an LLM-Agent Leaderboard Rank Actually Compare?

    Wei-Jung Huang

    cs.AIarXiv:2609.07785v12026