Artificial Intelligence

Papers filed under cs.AI on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.

Search paper metadata (including unsummarized papers)

10,621 to 10,680 of 15,485

  1. Reading Is Not Using: Retrieval, Judgment, and the Design of AI Financial Research Workflows

    Miao Liu, Zhizhe Liu

    cs.CLcs.AIarXiv:2608.24842v12026
  2. Knowing When to Ask for Help: Bayesian Self-Escalation in Hierarchical LLM Agents

    Nadeem Shaikh

    cs.LGcs.AIstat.MLarXiv:2608.24087v12026
  3. Constrained Entity Selection under Partial Knowledge for LLM-Based Knowledge Graph QA

    Emanuel Kitzelmann

    cs.AIarXiv:2608.24824v12026
  4. Confident at the moment of action: belief miscalibration in LLM play under hidden information

    Bhushan Kashinath Joshi

    cs.AIcs.CLcs.LGarXiv:2608.24691v12026
  5. Do Recipes Have Personas? Characterizing and Generating Creator Style in Attributed Procedural Graphs

    Lei Jiang

    cs.AIarXiv:2608.24369v12026
  6. Benchmarking LLM Judges for Voice-Agent Evaluation: Reliability, Calibration, and Human Oversight

    Anupam Purwar, Shashank Singh, Kritika Srivastava

    cs.AIcs.ETarXiv:2608.24314v12026
  7. VideoHarness-RSI: Recursive Harness Self-Improvement for Long-Video Understanding with Frozen Vision-Language Models

    Guoyang Xu, Hao Chen

    cs.AIarXiv:2608.24302v12026
  8. Giraffe: A Mapping Architecture from Hidden Text Representations to Visual Embeddings for Efficient Graphic Design

    Nejla Ghaboosi

    cs.AIcs.LGarXiv:2608.23970v12026
  9. MMJailBench: A Factorized Benchmark for Disentangling Multimodal Jailbreak Vulnerabilities

    Tianshi Wang, Jingsong Wang, Yafei Huang +3

    cs.CRcs.AIcs.MMarXiv:2608.25490v12026
  10. RotDroid: Cross-Orientation State Equivalence Testing for Detecting GUI Rotation Bugs in Android Apps

    Mengdi Qin, Bo Jiang

    cs.SEcs.AIarXiv:2608.25425v12026
  11. A Tendon-Driven Five-Fingered Hand with Distributed Tactile Perception for Dexterous Manipulation

    Huayang Chen, Longhui Qin

    cs.ROcs.AIarXiv:2608.25547v12026
  12. PIVOT: A Multi-Trajectory Dataset and Testbed for Pose, Intrinsics, and Novel Viewpoint Evaluation in Real-World 3D Reconstruction

    Mary Raymond

    cs.CVcs.AIarXiv:2608.25401v12026
  13. Neither Precision Nor Architecture Alone: Controlled Tests of Failure Remedies for Physics-Informed Neural Networks

    Jinyuan Zhang, Peng He, He Hu +2

    cs.LGcs.AIarXiv:2608.25327v12026
  14. PonsRAG: A Pons-Inspired RAG Bridging Cognitive Islands for Coordinated Long Narrative Reasoning

    Rongchen Zhao, Yu Chen, Juyuan Wang +4

    cs.AIcs.CLarXiv:2608.25486v12026
  15. Escaping Low-Dimensional Overlap: Multi-Task Model Merging via High-Dimensional Sparse Disentanglement

    Yihang Zhang, Shengke Sun, Junjie Wen +1

    cs.LGcs.AIcs.CLarXiv:2608.25354v12026
  16. When Less Is More: An Empirical Study of Minimal Responses in Counseling Dialogues and the Behavior of LLMs

    Zhiyang Qi

    cs.CLcs.AIarXiv:2608.24080v12026
  17. Repair or Resample? Rethinking Failure Debugging in LLM Multi-Agent Systems

    Zhongwen Luan, Xiaoyu Zhang, Ming Hu +3

    cs.AIcs.SEarXiv:2608.25920v12026
    Summaries:简体中文
  18. A Statistical Audit of Physical AI Benchmark Redundancy

    Zaruhi Navasardyan, Hrant Davtyan

    cs.ROcs.AIarXiv:2608.25940v12026
  19. A Visual Dependence-Aware Framework for Multimodal Unsupervised Continual Post-Training

    Kaichen Li, Zhilin Zhu, Jianhao Huang +7

    cs.CVcs.AIarXiv:2608.26095v12026
  20. TAU-Agent: An Agentic Retrieval-Augmented Framework for Traffic Anomaly Understanding

    Yuqiang Lin, Yan Shi, Sam Lockyer +5

    cs.CVcs.AIarXiv:2608.25935v12026
  21. From General Agents to RCA Experts: A Self-Evolving Harness for Root Cause Analysis

    Haiyu Huang, Jiewei Lyu, Zhihan Jiang +5

    cs.SEcs.AIarXiv:2608.25661v12026
    Summaries:简体中文
  22. Choose Your Game Wisely: Measuring Game-Theoretic Structures in Real-World Vehicle Interactions

    Yueyuan Li, Rongcheng Nie, Weijie Xi +4

    cs.AIcs.ROarXiv:2608.25917v12026
  23. Gating Before Commitment: Anticipating Intent Divergence to Prevent Post-Interaction Decision Failures in Autonomous Driving

    Cong Xu, Ravi Sankar

    cs.ROcs.AIarXiv:2608.26074v12026
  24. SmoothLLM: Defending Large Language Models Against Jailbreaking Attacks

    Alexander Robey, Eric Wong, Hamed Hassani +1

    cs.LGcs.AIstat.MLarXiv:2310.03684v42023
  25. FRAME: separating sampling variation from representational cause in medical imaging fairness

    Mahshad Lotfinia, Daniel Truhn, Andreas Maier +1

    cs.CVcs.AIcs.LGarXiv:2608.25981v12026
  26. Finding and using interpretable latents in a neutrino foundation model with sparse autoencoders

    Raphaël Bonnet-Guerrini, Johann Ioannou-Nikolaides, Inar Timiryasov +1

    astro-ph.HEcs.AIcs.LGarXiv:2608.26090v12026
  27. Imitation Learning for Connection-Tableau Construction

    Fredrik Rømming, Mantas Bakšys, Martin S. Fixman +1

    cs.AIcs.LGcs.LOarXiv:2608.26009v12026
  28. SciMIF: Understanding Multimodal Instruction Following in Scientific Domains

    Ye Shen, Yuting Zheng, Dun Pei +4

    cs.AIcs.LGarXiv:2608.25973v12026
  29. How Robust Are Automated Fact-Checking Systems? A Cross-Benchmark Evaluation

    Aida Usmanova, Zangir Iklassov, Markus Leippold +1

    cs.AIcs.LGarXiv:2608.25934v12026
  30. Mitigating Hallucination in Large Multi-Modal Models via Robust Instruction Tuning

    Fuxiao Liu, Kevin Lin, Linjie Li +3

    cs.CVcs.AIcs.CEarXiv:2306.14565v42023
  31. Communication-Efficient Federated Deep Learning with Asynchronous Model Update and Temporally Weighted Aggregation

    Yang Chen, Xiaoyan Sun, Yaochu Jin

    cs.LGcs.AIcs.DCarXiv:1903.07424v12019
  32. AudioLDM 2: Learning Holistic Audio Generation with Self-supervised Pretraining

    Haohe Liu, Yi Yuan, Xubo Liu +7

    cs.SDcs.AIcs.MMarXiv:2308.05734v32023
  33. MyoMechanix: Biomechanically-Grounded Compositional Skilled Activity Understanding and Coaching

    Hao Yin, Paritosh Parmar, Lijun Gu +6

    cs.CVcs.AIcs.ETarXiv:2608.26094v12026
  34. FuXi: A cascade machine learning forecasting system for 15-day global weather forecast

    Lei Chen, Xiaohui Zhong, Feng Zhang +4

    physics.ao-phcs.AIcs.LGarXiv:2306.12873v32023
  35. Difficulty-Aware Sample Allocation for Adaptive Data Augmentation in Semantic Segmentation

    Olasimbo Ayodeji Arigbabu, Abimbola Ismail Arigbabu

    cs.CVcs.AIarXiv:2608.25710v12026
  36. dm_control: Software and Tasks for Continuous Control

    Yuval Tassa, Saran Tunyasuvunakool, Alistair Muldal +8

    cs.ROcs.AIcs.LGarXiv:2006.12983v22020
  37. Formal Security Analysis of Neural Networks using Symbolic Intervals

    Shiqi Wang, Kexin Pei, Justin Whitehouse +2

    cs.AIcs.LOarXiv:1804.10829v32018
  38. Inductive Relation Prediction by Subgraph Reasoning

    Komal K. Teru, Etienne Denis, William L. Hamilton

    cs.LGcs.AIstat.MLarXiv:1911.06962v22019
  39. MeMark: Membrane-Space Watermarking for Spiking Neural Networks

    Roberto Riaño, Gorka Abad, Stjepan Picek +1

    cs.CRcs.AIcs.LGarXiv:2608.25738v12026
  40. Pointing the Way, Hiding the Destination: Practical Private Dense Retrieval at Scale

    Peichun Hua, Danyang Chen, Junan Zhang +5

    cs.CRcs.AIcs.IRarXiv:2608.25735v12026
  41. Unsupervised Anatomical Feature Learning via Diffusion Models: Enhanced Medical Image Segmentation with Denoising Diffusion Probabilistic Models

    Akshat G, Divyansh Gupta, Shaleen Bhatnagar +2

    cs.CVcs.AIcs.LGarXiv:2608.25693v12026
  42. DualOPSD: Adaptive Privileged Teachers for On-Policy Self-Distillation

    Yutong Chen, Guangfu Guo, Zhichao Xu +1

    cs.LGcs.AIarXiv:2608.26019v12026
  43. Deep Learning with Low Precision by Half-wave Gaussian Quantization

    Zhaowei Cai, Xiaodong He, Jian Sun +1

    cs.CVcs.AIcs.LGarXiv:1702.00953v12017
  44. Towards A Unified Information Bottleneck Framework for Time Series Explanations

    Xu Zheng, Zichuan Liu, Zhuomin Chen +7

    cs.LGcs.AIarXiv:2608.25897v12026
  45. TailSFT: Filtered Fine-Tuning Improves Post-Training Performance

    Sadhika Malladi, Samy Jelassi, Dylan Foster +2

    cs.LGcs.AIarXiv:2608.25756v12026
  46. VT-ADL: A Vision Transformer Network for Image Anomaly Detection and Localization

    Pankaj Mishra, Riccardo Verk, Daniele Fornasier +2

    cs.CVcs.AIcs.LGarXiv:2104.10036v12021
  47. Narcissus: Program Synthesis Using Context-Aware LLM Approximations

    Tilman Hinnerichs, Sebastijan Dumancic, Neil Yorke-Smith

    cs.AIcs.LGcs.PLarXiv:2608.25657v12026
  48. TraceML: An Empirical Analysis of Human-Agent Planning in Machine Learning Development

    Jiarui Yan, Weiwei Sun, Sijie Li +2

    cs.LGcs.AIarXiv:2608.26086v12026
  49. Measuring Faithfulness in Chain-of-Thought Reasoning

    Tamera Lanham, Anna Chen, Ansh Radhakrishnan +27

    cs.AIcs.CLcs.LGarXiv:2307.13702v12023
  50. Massively Parallel Methods for Deep Reinforcement Learning

    Arun Nair, Praveen Srinivasan, Sam Blackwell +11

    cs.LGcs.AIcs.DCarXiv:1507.04296v22015
  51. Towards Optimally Decentralized Multi-Robot Collision Avoidance via Deep Reinforcement Learning

    Pinxin Long, Tingxiang Fan, Xinyi Liao +3

    cs.ROcs.AIcs.LGarXiv:1709.10082v32017
  52. $R^3$: Training Robots to Reason in Natural Language via Reinforcement Learning

    Lehong Wu, Yuxiao Qu, Zheyuan Hu +4

    cs.ROcs.AIcs.CLarXiv:2608.26053v12026
  53. Git Re-Basin: Merging Models modulo Permutation Symmetries

    Samuel K. Ainsworth, Jonathan Hayase, Siddhartha Srinivasa

    cs.LGcs.AIarXiv:2209.04836v62022
  54. Trace Integrity for LLM Data Agents: A Vision for Auditable Structured Reasoning in Real-World Systems

    Srimonti Dutta, Akshata Kishore Moharir

    cs.AIcs.CLarXiv:2608.26036v12026
  55. It's a matter of timescale: non-linear utility in successor features and multi-objective planning and learning

    Liam P. H. Mertens, Lucas N. Alegre, Florent Delgrange +3

    cs.LGcs.AIarXiv:2608.25723v12026
  56. Sionna: An Open-Source Library for Next-Generation Physical Layer Research

    Jakob Hoydis, Sebastian Cammerer, Fayçal Ait Aoudia +4

    cs.ITcs.AIcs.LGarXiv:2203.11854v22022
  57. SwarmWorld: Stigmergic technological evolution in societies of language-model agents

    Subhadeep Pal, Fiona Y. Wang, Markus J. Buehler

    cs.AIcond-mat.mtrl-scics.CLarXiv:2608.26081v12026
  58. Learning Features by Watching Objects Move

    Deepak Pathak, Ross Girshick, Piotr Dollár +2

    cs.CVcs.AIcs.LGarXiv:1612.06370v22016
  59. How Much Rank Does LoRA Need? Rank-Error Bounds for Transformer Attention

    Gerard Conangla Planes

    cs.LGcs.AIcs.CLarXiv:2608.26052v12026
  60. On the Tractability of SHAP Explanations

    Guy Van den Broeck, Anton Lykov, Maximilian Schleich +1

    cs.AIcs.CCcs.LGarXiv:2009.08634v22020