Artificial Intelligence

Papers filed under cs.AI on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.

Search paper metadata (including unsummarized papers)

14,341 to 14,400 of 15,404

  1. Identifying Harm in Personalized, Generative AI Systems Requires User-Centered Auditing at the Interaction Level

    Hannah Cha

    cs.CYcs.AIcs.HCarXiv:2608.14692v12026
  2. Does the Heart Show Your Pain? Tackling the X-ITE Pain Challenge with Self-Supervised ECG Representation Learning

    Dominika Kunc, Przemysław Kazienko, Stanisław Saganowski

    eess.SPcs.AIcs.LGarXiv:2608.14662v12026
  3. Position: AI Agents in Scientific Teams Should Be Studied as Human-Agent Systems

    Patrick Emami, Sameera Horawalavithana, Truc Nguyen +11

    cs.AIcs.HCarXiv:2608.14667v12026
  4. BRA-Audit: Budgeted Runtime Auditing for LLM Multi-Agent Systems via Cumulative-Exposure Audit-Point Placement

    Kaixiang Wang, Yidan Lin, Jiong Lou +1

    cs.MAcs.AIarXiv:2608.14668v12026
  5. Auditing an AI-Generated Mathematical Proof: A Correction to a Greedy Conditioning Lemma in Quantum Parallel Repetition

    Mikołaj Sienicki, Krzysztof Sienicki

    cs.AIquant-pharXiv:2608.14673v12026
  6. Sparse Coverage: Semantic Center Representations for Patent Prior-Art Retrieval

    You Zuo, Kim Gerdes, Éric de la Clergerie +1

    cs.IRcs.AIarXiv:2608.16918v12026
  7. Inference-Time Mitigation of Adversarial Political Bias in Large Language Models

    Tejaswi V. Panchagnula, Bruce Coburn, Bryce J. Dietrich +3

    cs.CLcs.AIarXiv:2608.14629v12026
  8. Ring-based Spatial Transformer: Learning Non-linear Spatial Interactions between Building Distribution and Pedestrian Flow

    Shun Nakayama, Takahiro Kanamori, Wanglin Yan

    cs.LGcs.AIcs.CYarXiv:2608.14660v12026
  9. Cross-Domain Industrial Fault Detection by Causal Mechanism Monitoring

    Dhiraj Neupane, Mohamed Reda Bouadjenek, Richard Dazeley +1

    cs.AIarXiv:2608.14666v12026
  10. Synchronized Logit Steering: Real-world Steganography

    Andrew Rufail, Aadi Dash, Onir Narahari +5

    cs.AIarXiv:2608.14697v12026
  11. Take it Personally: The Limits of General SSL Representations for Real-Life PPG Emotion Detection

    Dominika Kunc, Przemysław Kazienko, Stanisław Saganowski

    cs.LGcs.AIarXiv:2608.14675v12026
  12. Information-Theoretic Causal Modelling of Semiconductor Process Dynamics

    Daniel Sørensen, Giorgio Melchiorre, Sudip Bandyopadhyay +3

    eess.SPcs.AIcs.ITarXiv:2608.14678v12026
  13. Mitigating Rubric Interference in LLM Judges via On-Policy Self-Distillation

    Dingyao Yu, Tong Zhang, Yutao Mou +3

    cs.LGcs.AIarXiv:2608.14684v12026
  14. WIP: LLM Odyssey: A Game-Based Platform for Teaching LLM Engineering Concepts

    Priyamvada Tripathi

    cs.CYcs.AIarXiv:2608.16924v12026
  15. Breaking and Defending LLM-Powered Social Media Bot Detection Systems

    Nof Orenstein, Yoni Birman

    cs.AIarXiv:2608.15893v12026
  16. Pushing the Limits of High-Resolution Weather Forecasting through Data Scaling

    Yang Zhao, Peisong Niu, Tian Zhou +5

    cs.LGcs.AIcs.CVarXiv:2608.14652v12026
  17. When Uncertainty Isn't Enough: An Empirical Study of Self-Correction in Code Generation

    Pranav Rakasi, Maanas Lalwani, Arnav Srivastava +4

    cs.AIcs.LGcs.SEarXiv:2608.14659v12026
  18. pico-type: A 1.5M-Parameter Byte-Level Multi-Head Content Classifier

    Gautam Kishore

    cs.LGcs.AIcs.CLarXiv:2608.14658v12026
  19. Offline Ambient-Controlled Latent Diffusion: Architecture, Telemetry, and On-Device Evaluation

    Lech Kalinowski, Artur Morys-Magiera, Piotr Miłkowski

    eess.SPcs.AIcs.LGarXiv:2608.14677v12026
  20. Characterizing Rhetorical Misalignment in Decision-Making with Language Models

    Zirui Cheng, Joey Chan, Simo Du +3

    cs.CLcs.AIarXiv:2608.14630v12026
  21. ARGUS: Attention-Guided Transformers for Scalable Person Identification Using Wi-Fi Telemetry

    Nayan Sanjay Bhatia, Pranay Kocheta, Yuhan Li +1

    cs.LGcs.AIcs.CVarXiv:2608.14670v12026
  22. When Agentic Executions Fail: Detecting and Localizing Runtime Faults from Telemetry

    Chenkai Zhang, Yiran Li, Yifang Tian +2

    cs.AIcs.SEarXiv:2608.14680v12026
    Summaries:한국어
  23. Valid Per-Field Selective Risk Control for Document Extraction: Three Failure Modes, a Validity Ladder, and When Conditioning Pays

    Bhaskar Gurram

    cs.LGcs.AIcs.CLarXiv:2608.14639v12026
  24. Efficient RLVR Scheduling via Graph-Structured Online Difficulty Estimation

    Zhizhao Liu, Zhiliang Tian, Xi Wang +4

    cs.LGcs.AIcs.CLarXiv:2608.17941v12026
  25. Semantic Uncertainty-Guided Orchestration in Hierarchical Multi-Agent Systems

    John Knowlton, Aritra Guha, Risto Miikkulainen

    cs.AIarXiv:2608.14707v12026
  26. Unraveling the Size Determination Mechanism of Nanocrystal Synthesis via Interpretable Neural Networks

    Kai Gu, Haizheng Zhong

    cs.LGcond-mat.mtrl-scics.AIarXiv:2608.14734v12026
  27. Evaluating Agentic Code Repair Capabilities in Distributed Systems

    Yibo Yan, Huijuan Wang, Junzhou He +4

    cs.SEcs.AIcs.DCarXiv:2608.14863v12026
  28. Adaptive Mixing of Policies from Searching and Policies from Learning

    Gavin B. Rens

    cs.AIarXiv:2608.15700v12026
  29. THESIS-MoE: Trainable Hierarchical Extraction and SteerIng of Sycophancy in Mixture-of-Experts

    Kareem Hassani, Chaymaa Abbas, Lama Mawlawi +1

    cs.AIarXiv:2608.15687v12026
  30. Command-Space Counterfactual Explanations for Pareto-Conditioned Reinforcement Learning

    Joanikij Chulev, Hendrik Baier

    cs.LGcs.AIcs.HCarXiv:2608.14963v12026
  31. SpIn-ViT: Designing a Sparsity-Induced Vision Transformer That Is Mechanistically Interpretable

    Philip H. Lee, Parth Padalkar

    cs.CVcs.AIarXiv:2608.14922v12026
  32. Towards Standardized Evaluation in Automated Domain Modeling: Introducing a Benchmark

    Vasiliy Seibert

    cs.AIarXiv:2608.15255v12026
  33. AudioTQ: A Data-Oblivious 6-Bit CPU Audio Codec via Randomized Hadamard Rotation and Lloyd-Max Quantization

    Sahil Gangurde

    cs.SDcs.AIcs.CRarXiv:2608.15369v12026
  34. MAPLE: MoE Adaptive Plug-and-play Layer-wise Expert allocation

    Lie Li, Wen Li, Junxiao Shen +1

    cs.LGcs.AIarXiv:2608.15299v12026
  35. Handoff-H1: An Orchestrated Vision-Agent System for Material Quantity Takeoff from Construction Blueprints

    Bruno Chicelli, Henrique Alves, Rodrigo Anselmo +3

    cs.CLcs.AIcs.CVarXiv:2608.15032v12026
  36. Agentic-SQL Revisited: Autonomy-Based Taxonomy and Empirical Benchmark Analysis for LLM Text-to-SQL

    Changruo Zhao, Zujun Peng, Yu Tian +5

    cs.AIarXiv:2608.15389v12026
  37. Beyond Pass@k: Measuring Reliability and Security of Agentic Code Generation

    Jiajun Jiang, Sharon Zheng, Natan Vidra +1

    cs.AIarXiv:2608.14711v12026
  38. When Is an Agent Evaluation Over? Outcome Finality and Cross-Unit Separation

    Avyay M. Casheekar

    cs.AIcs.CYarXiv:2608.14940v12026
  39. Frontier AI Forecasting Has a Measurement Problem: An Audit of Progress Evidence

    Fabricio F Costa

    cs.AIarXiv:2608.14903v12026
  40. Class Imbalance and Batch Effects in LLM-Based Screening for Systematic Reviews

    Gilberto Sussumu Hida, Danilo Monteiro Ribeiro, Clayton Suguio Hida

    cs.CLcs.AIarXiv:2608.14737v12026
  41. Identifying Confusion Trends in Concept-based XAI for Multi-Label Classification

    Haadia Amjad, Ronald Tetzlaff

    cs.CVcs.AIarXiv:2608.15731v12026
  42. RRFC: Recursive Refinement via Feedback Conditioning for Iterative Image-to-Image Generation

    Kareem Hassani, Chaymaa Abbas, Hadi Al Mubasher +1

    cs.CVcs.AIarXiv:2608.15694v12026
  43. Not All Attention Is Equal: A Quantitative Survey of the EEI Trade-off

    Aditya Singh

    cs.LGcs.AIarXiv:2608.15459v12026
  44. Solvable Sokoban Without a Solver via Diffusion

    Sina Baghal

    cs.AIcs.GTcs.LGarXiv:2608.15958v12026
  45. Gated Against One Model, Open to the Next: Option-Only Solvability in Legal Multiple-Choice Benchmarks

    Volodymyr Ovcharov

    cs.CLcs.AIcs.CYarXiv:2608.15428v12026
  46. Chameleon: An Adaptive AI-Driven Honeypot Architecture Using Threat-Calibrated Particle Swarm Optimization and Semantic Deception Rapidly-Exploring Random Trees

    Rohit Swami, Tushar Singh, Akash Warde +1

    cs.CRcs.AIcs.NEarXiv:2608.15407v12026
  47. WeSCE: A Benchmark for Measuring Security Drift in LLM-Driven Code Editing

    Zhiyu Zhang, Tingyue Wen, Senke Sun +2

    cs.CRcs.AIcs.SEarXiv:2608.15092v12026
  48. Distinguishing AI-Generated Music from Edited Audio as a Hard-Negative Robustness Task

    Alexandru-Stefan Morosanu, Valerian Cecan, Stefan-Daniel Achirei +1

    cs.SDcs.AIcs.LGarXiv:2608.14916v12026
  49. Trust the Right Teacher: Quality-Aware Self-Distillation for GUI Grounding

    Jingyuan Huang, Zuming Huang, Yucheng Shi +4

    cs.AIarXiv:2606.18101v22026
  50. OGX: An Open-Source, Vendor-Neutral Generative AI Application Server

    Francisco Javier Arceo, Sébastien Han, Matthew Farrellee +8

    cs.AIcs.IRarXiv:2608.14580v12026
  51. Robusto-2: Benchmarking Humans & VLMs for Autonomous Driving in Lima & New York City

    Adrian Cespedes, Marcelo Chincha, Dunant Cusipuma +3

    cs.CVcs.AIcs.ROarXiv:2606.20980v12026
  52. LegalHalluLens: Typed Hallucination Auditing and Calibrated Multi-Agent Debate for Trustworthy Legal AI

    Lalit Yadav, Akshaj Gurugubelli

    cs.AIcs.CLcs.LGarXiv:2606.18021v12026
  53. FAPO: Fully Automated Prompt Optimization of Multi-Step LLM Pipelines

    Paul Kassianik, Baturay Saglam, Huaibo Zhao +4

    cs.SEcs.AIarXiv:2606.19605v22026
  54. Beyond Reward Engineering: A Data Recipe for Long-Context Reinforcement Learning

    Xiaoyue Xu, Sikui Zhang, Xiaorong Wang +2

    cs.CLcs.AIarXiv:2606.18831v12026
  55. Configurable Clinical Information Extraction with Agentic RAG: What Works, What Breaks, and Why

    Osman Alperen Çinar-Koraş, Marie Bauer, Sameh Khattab +7

    cs.AIarXiv:2606.19602v12026
  56. Rethinking Shrinkage Bias in LLM FP4 Pretraining: Geometric Origin, Systemic Impact, and UFP4 Recipe

    Qian Zhao, Kunlong Chen, Changxin Tian +9

    cs.AIarXiv:2606.20381v12026
  57. Reinforcing Dual-Path Reasoning in Spatial Vision Language Models

    Yatai Ji, An-Chieh Cheng, Yang Fu +13

    cs.CVcs.AIarXiv:2606.17539v12026
  58. Morpheus: A Morphology-Aware Neural Tokenizer and Word Embedder for Turkish

    Tolga Şakar

    cs.CLcs.AIarXiv:2606.18717v12026
  59. PerceptionDLM: Parallel Region Perception with Multimodal Diffusion Language Models

    Yueyi Sun, Yuhao Wang, Jason Li +8

    cs.CVcs.AIcs.CLarXiv:2606.19534v12026
  60. DF3DV-1K: A Large-Scale Dataset and Benchmark for Distractor-Free Novel View Synthesis

    Cheng-You Lu, Yi-Shan Hung, Wei-Ling Chi +6

    cs.CVcs.AIarXiv:2604.13416v32026