Artificial Intelligence
Papers filed under cs.AI on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.
Search paper metadata (including unsummarized papers)
14,341 to 14,400 of 15,404
Identifying Harm in Personalized, Generative AI Systems Requires User-Centered Auditing at the Interaction Level
Hannah Cha
cs.CYcs.AIcs.HCarXiv:2608.14692v12026Does the Heart Show Your Pain? Tackling the X-ITE Pain Challenge with Self-Supervised ECG Representation Learning
Dominika Kunc, Przemysław Kazienko, Stanisław Saganowski
eess.SPcs.AIcs.LGarXiv:2608.14662v12026Position: AI Agents in Scientific Teams Should Be Studied as Human-Agent Systems
Patrick Emami, Sameera Horawalavithana, Truc Nguyen +11
cs.AIcs.HCarXiv:2608.14667v12026BRA-Audit: Budgeted Runtime Auditing for LLM Multi-Agent Systems via Cumulative-Exposure Audit-Point Placement
Kaixiang Wang, Yidan Lin, Jiong Lou +1
cs.MAcs.AIarXiv:2608.14668v12026Auditing an AI-Generated Mathematical Proof: A Correction to a Greedy Conditioning Lemma in Quantum Parallel Repetition
Mikołaj Sienicki, Krzysztof Sienicki
cs.AIquant-pharXiv:2608.14673v12026Sparse Coverage: Semantic Center Representations for Patent Prior-Art Retrieval
You Zuo, Kim Gerdes, Éric de la Clergerie +1
cs.IRcs.AIarXiv:2608.16918v12026Inference-Time Mitigation of Adversarial Political Bias in Large Language Models
Tejaswi V. Panchagnula, Bruce Coburn, Bryce J. Dietrich +3
cs.CLcs.AIarXiv:2608.14629v12026Ring-based Spatial Transformer: Learning Non-linear Spatial Interactions between Building Distribution and Pedestrian Flow
Shun Nakayama, Takahiro Kanamori, Wanglin Yan
cs.LGcs.AIcs.CYarXiv:2608.14660v12026Cross-Domain Industrial Fault Detection by Causal Mechanism Monitoring
Dhiraj Neupane, Mohamed Reda Bouadjenek, Richard Dazeley +1
cs.AIarXiv:2608.14666v12026Synchronized Logit Steering: Real-world Steganography
Andrew Rufail, Aadi Dash, Onir Narahari +5
cs.AIarXiv:2608.14697v12026Take it Personally: The Limits of General SSL Representations for Real-Life PPG Emotion Detection
Dominika Kunc, Przemysław Kazienko, Stanisław Saganowski
cs.LGcs.AIarXiv:2608.14675v12026Information-Theoretic Causal Modelling of Semiconductor Process Dynamics
Daniel Sørensen, Giorgio Melchiorre, Sudip Bandyopadhyay +3
eess.SPcs.AIcs.ITarXiv:2608.14678v12026Mitigating Rubric Interference in LLM Judges via On-Policy Self-Distillation
Dingyao Yu, Tong Zhang, Yutao Mou +3
cs.LGcs.AIarXiv:2608.14684v12026WIP: LLM Odyssey: A Game-Based Platform for Teaching LLM Engineering Concepts
Priyamvada Tripathi
cs.CYcs.AIarXiv:2608.16924v12026Breaking and Defending LLM-Powered Social Media Bot Detection Systems
Nof Orenstein, Yoni Birman
cs.AIarXiv:2608.15893v12026Pushing the Limits of High-Resolution Weather Forecasting through Data Scaling
Yang Zhao, Peisong Niu, Tian Zhou +5
cs.LGcs.AIcs.CVarXiv:2608.14652v12026When Uncertainty Isn't Enough: An Empirical Study of Self-Correction in Code Generation
Pranav Rakasi, Maanas Lalwani, Arnav Srivastava +4
cs.AIcs.LGcs.SEarXiv:2608.14659v12026pico-type: A 1.5M-Parameter Byte-Level Multi-Head Content Classifier
Gautam Kishore
cs.LGcs.AIcs.CLarXiv:2608.14658v12026Offline Ambient-Controlled Latent Diffusion: Architecture, Telemetry, and On-Device Evaluation
Lech Kalinowski, Artur Morys-Magiera, Piotr Miłkowski
eess.SPcs.AIcs.LGarXiv:2608.14677v12026Characterizing Rhetorical Misalignment in Decision-Making with Language Models
Zirui Cheng, Joey Chan, Simo Du +3
cs.CLcs.AIarXiv:2608.14630v12026ARGUS: Attention-Guided Transformers for Scalable Person Identification Using Wi-Fi Telemetry
Nayan Sanjay Bhatia, Pranay Kocheta, Yuhan Li +1
cs.LGcs.AIcs.CVarXiv:2608.14670v12026When Agentic Executions Fail: Detecting and Localizing Runtime Faults from Telemetry
Chenkai Zhang, Yiran Li, Yifang Tian +2
cs.AIcs.SEarXiv:2608.14680v12026Summaries:한국어Valid Per-Field Selective Risk Control for Document Extraction: Three Failure Modes, a Validity Ladder, and When Conditioning Pays
Bhaskar Gurram
cs.LGcs.AIcs.CLarXiv:2608.14639v12026Efficient RLVR Scheduling via Graph-Structured Online Difficulty Estimation
Zhizhao Liu, Zhiliang Tian, Xi Wang +4
cs.LGcs.AIcs.CLarXiv:2608.17941v12026Semantic Uncertainty-Guided Orchestration in Hierarchical Multi-Agent Systems
John Knowlton, Aritra Guha, Risto Miikkulainen
cs.AIarXiv:2608.14707v12026Unraveling the Size Determination Mechanism of Nanocrystal Synthesis via Interpretable Neural Networks
Kai Gu, Haizheng Zhong
cs.LGcond-mat.mtrl-scics.AIarXiv:2608.14734v12026Evaluating Agentic Code Repair Capabilities in Distributed Systems
Yibo Yan, Huijuan Wang, Junzhou He +4
cs.SEcs.AIcs.DCarXiv:2608.14863v12026Adaptive Mixing of Policies from Searching and Policies from Learning
Gavin B. Rens
cs.AIarXiv:2608.15700v12026THESIS-MoE: Trainable Hierarchical Extraction and SteerIng of Sycophancy in Mixture-of-Experts
Kareem Hassani, Chaymaa Abbas, Lama Mawlawi +1
cs.AIarXiv:2608.15687v12026Command-Space Counterfactual Explanations for Pareto-Conditioned Reinforcement Learning
Joanikij Chulev, Hendrik Baier
cs.LGcs.AIcs.HCarXiv:2608.14963v12026SpIn-ViT: Designing a Sparsity-Induced Vision Transformer That Is Mechanistically Interpretable
Philip H. Lee, Parth Padalkar
cs.CVcs.AIarXiv:2608.14922v12026Towards Standardized Evaluation in Automated Domain Modeling: Introducing a Benchmark
Vasiliy Seibert
cs.AIarXiv:2608.15255v12026AudioTQ: A Data-Oblivious 6-Bit CPU Audio Codec via Randomized Hadamard Rotation and Lloyd-Max Quantization
Sahil Gangurde
cs.SDcs.AIcs.CRarXiv:2608.15369v12026MAPLE: MoE Adaptive Plug-and-play Layer-wise Expert allocation
Lie Li, Wen Li, Junxiao Shen +1
cs.LGcs.AIarXiv:2608.15299v12026Handoff-H1: An Orchestrated Vision-Agent System for Material Quantity Takeoff from Construction Blueprints
Bruno Chicelli, Henrique Alves, Rodrigo Anselmo +3
cs.CLcs.AIcs.CVarXiv:2608.15032v12026Agentic-SQL Revisited: Autonomy-Based Taxonomy and Empirical Benchmark Analysis for LLM Text-to-SQL
Changruo Zhao, Zujun Peng, Yu Tian +5
cs.AIarXiv:2608.15389v12026Beyond Pass@k: Measuring Reliability and Security of Agentic Code Generation
Jiajun Jiang, Sharon Zheng, Natan Vidra +1
cs.AIarXiv:2608.14711v12026When Is an Agent Evaluation Over? Outcome Finality and Cross-Unit Separation
Avyay M. Casheekar
cs.AIcs.CYarXiv:2608.14940v12026Frontier AI Forecasting Has a Measurement Problem: An Audit of Progress Evidence
Fabricio F Costa
cs.AIarXiv:2608.14903v12026Class Imbalance and Batch Effects in LLM-Based Screening for Systematic Reviews
Gilberto Sussumu Hida, Danilo Monteiro Ribeiro, Clayton Suguio Hida
cs.CLcs.AIarXiv:2608.14737v12026Identifying Confusion Trends in Concept-based XAI for Multi-Label Classification
Haadia Amjad, Ronald Tetzlaff
cs.CVcs.AIarXiv:2608.15731v12026RRFC: Recursive Refinement via Feedback Conditioning for Iterative Image-to-Image Generation
Kareem Hassani, Chaymaa Abbas, Hadi Al Mubasher +1
cs.CVcs.AIarXiv:2608.15694v12026Not All Attention Is Equal: A Quantitative Survey of the EEI Trade-off
Aditya Singh
cs.LGcs.AIarXiv:2608.15459v12026Solvable Sokoban Without a Solver via Diffusion
Sina Baghal
cs.AIcs.GTcs.LGarXiv:2608.15958v12026Gated Against One Model, Open to the Next: Option-Only Solvability in Legal Multiple-Choice Benchmarks
Volodymyr Ovcharov
cs.CLcs.AIcs.CYarXiv:2608.15428v12026Chameleon: An Adaptive AI-Driven Honeypot Architecture Using Threat-Calibrated Particle Swarm Optimization and Semantic Deception Rapidly-Exploring Random Trees
Rohit Swami, Tushar Singh, Akash Warde +1
cs.CRcs.AIcs.NEarXiv:2608.15407v12026WeSCE: A Benchmark for Measuring Security Drift in LLM-Driven Code Editing
Zhiyu Zhang, Tingyue Wen, Senke Sun +2
cs.CRcs.AIcs.SEarXiv:2608.15092v12026Distinguishing AI-Generated Music from Edited Audio as a Hard-Negative Robustness Task
Alexandru-Stefan Morosanu, Valerian Cecan, Stefan-Daniel Achirei +1
cs.SDcs.AIcs.LGarXiv:2608.14916v12026Trust the Right Teacher: Quality-Aware Self-Distillation for GUI Grounding
Jingyuan Huang, Zuming Huang, Yucheng Shi +4
cs.AIarXiv:2606.18101v22026OGX: An Open-Source, Vendor-Neutral Generative AI Application Server
Francisco Javier Arceo, Sébastien Han, Matthew Farrellee +8
cs.AIcs.IRarXiv:2608.14580v12026Robusto-2: Benchmarking Humans & VLMs for Autonomous Driving in Lima & New York City
Adrian Cespedes, Marcelo Chincha, Dunant Cusipuma +3
cs.CVcs.AIcs.ROarXiv:2606.20980v12026LegalHalluLens: Typed Hallucination Auditing and Calibrated Multi-Agent Debate for Trustworthy Legal AI
Lalit Yadav, Akshaj Gurugubelli
cs.AIcs.CLcs.LGarXiv:2606.18021v12026FAPO: Fully Automated Prompt Optimization of Multi-Step LLM Pipelines
Paul Kassianik, Baturay Saglam, Huaibo Zhao +4
cs.SEcs.AIarXiv:2606.19605v22026Beyond Reward Engineering: A Data Recipe for Long-Context Reinforcement Learning
Xiaoyue Xu, Sikui Zhang, Xiaorong Wang +2
cs.CLcs.AIarXiv:2606.18831v12026Configurable Clinical Information Extraction with Agentic RAG: What Works, What Breaks, and Why
Osman Alperen Çinar-Koraş, Marie Bauer, Sameh Khattab +7
cs.AIarXiv:2606.19602v12026Rethinking Shrinkage Bias in LLM FP4 Pretraining: Geometric Origin, Systemic Impact, and UFP4 Recipe
Qian Zhao, Kunlong Chen, Changxin Tian +9
cs.AIarXiv:2606.20381v12026Reinforcing Dual-Path Reasoning in Spatial Vision Language Models
Yatai Ji, An-Chieh Cheng, Yang Fu +13
cs.CVcs.AIarXiv:2606.17539v12026Morpheus: A Morphology-Aware Neural Tokenizer and Word Embedder for Turkish
Tolga Şakar
cs.CLcs.AIarXiv:2606.18717v12026PerceptionDLM: Parallel Region Perception with Multimodal Diffusion Language Models
Yueyi Sun, Yuhao Wang, Jason Li +8
cs.CVcs.AIcs.CLarXiv:2606.19534v12026DF3DV-1K: A Large-Scale Dataset and Benchmark for Distractor-Free Novel View Synthesis
Cheng-You Lu, Yi-Shan Hung, Wei-Ling Chi +6
cs.CVcs.AIarXiv:2604.13416v32026