Artificial Intelligence

Papers filed under cs.AI on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.

Search paper metadata (including unsummarized papers)

11,401 to 11,460 of 15,409

  1. Right for the Right Reasons: Training Differentiable Models by Constraining their Explanations

    Andrew Slavin Ross, Michael C. Hughes, Finale Doshi-Velez

    cs.LGcs.AIstat.MLarXiv:1703.03717v22017
  2. Deep Learning for Time Series Anomaly Detection: A Survey

    Zahra Zamanzadeh Darban, Geoffrey I. Webb, Shirui Pan +2

    cs.LGcs.AIarXiv:2211.05244v32022
  3. RLAIF vs. RLHF: Scaling Reinforcement Learning from Human Feedback with AI Feedback

    Harrison Lee, Samrat Phatale, Hassan Mansoor +8

    cs.CLcs.AIcs.LGarXiv:2309.00267v32023
  4. Can MLLMs Read Students' Minds? Unpacking Multimodal Error Analysis in Handwritten Math

    Dingjie Song, Tianlong Xu, Yi-Fan Zhang +6

    cs.AIcs.CLcs.CVarXiv:2603.24961v12026
  5. Voxtral TTS

    Mistral-AI, :, Alexander H. Liu +186

    cs.AIarXiv:2603.25551v22026
  6. Working Notes on Late Interaction Dynamics: Analyzing Targeted Behaviors of Late Interaction Models

    Antoine Edy, Max Conti, Quentin Macé

    cs.IRcs.AIcs.CLarXiv:2603.26259v22026
  7. HiDiffTIR: Hierarchical Difficulty-Aware Policy Optimization for Multi-Turn Tool-Integrated Reasoning

    Yucan Guo, Xiaohan Wang, Miao Su +8

    cs.CLcs.AIarXiv:2608.21863v12026
  8. DAG-GNN: DAG Structure Learning with Graph Neural Networks

    Yue Yu, Jie Chen, Tian Gao +1

    cs.LGcs.AIstat.MLarXiv:1904.10098v12019
  9. How is ChatGPT's behavior changing over time?

    Lingjiao Chen, Matei Zaharia, James Zou

    cs.CLcs.AIcs.LGarXiv:2307.09009v32023
  10. daVinci-LLM:Towards the Science of Pretraining

    Yiwei Qin, Yixiu Liu, Tiantian Mi +12

    cs.AIcs.CLarXiv:2603.27164v12026
  11. VirtualHome: Simulating Household Activities via Programs

    Xavier Puig, Kevin Ra, Marko Boben +4

    cs.CVcs.AIcs.LGarXiv:1806.07011v12018
  12. Unified Training of Universal Time Series Forecasting Transformers

    Gerald Woo, Chenghao Liu, Akshat Kumar +3

    cs.LGcs.AIarXiv:2402.02592v22024
  13. EpochX: Building the Infrastructure for an Emergent Agent Civilization

    Huacan Wang, Chaofa Yuan, Xialie Zhuang +15

    cs.AIcs.MAarXiv:2603.27304v12026
  14. MatReplace: A Reference-Free, Conditioning-Aligned Benchmark for Material Replacement in Interior Scenes

    Mingzhe Du, Thong Thanh Nguyen, Nguyen Tran Cong Duy +2

    cs.CVcs.AIarXiv:2608.24107v12026
  15. STRIVE: Multi-Agent Structured Temporal Reasoning with Integrated Verification for Longitudinal Radiology Report Generation

    Junyeong Maeng, Eunsong Kang, Heung-Il Suk

    cs.AIarXiv:2608.24237v12026
  16. ESQ-Bench: A Multi-Tier Enterprise Oracle Benchmark for Evaluating NL2SQL Dialect Generalization and Silent Semantic Divergence

    Sanjay Mishra, Divya Chukkapalli, Ganesh R. Naik

    cs.AIarXiv:2608.23569v12026
  17. Generative artificial intelligence enhances creativity but reduces the diversity of novel content

    Anil R. Doshi, Oliver P. Hauser

    cs.HCcs.AIecon.GNarXiv:2312.00506v32023
  18. Decolonial AI: Decolonial Theory as Sociotechnical Foresight in Artificial Intelligence

    Shakir Mohamed, Marie-Therese Png, William Isaac

    cs.CYcs.AIcs.LGarXiv:2007.04068v12020
  19. Syn2RealTrack: Bridging the Gap Between Synthetic and Real-World Datasets for Online Multi-View Multi-Target Tracking

    Duong Nguyen-Ngoc Tran, Ngoc Doan-Minh Huynh, Cu Quoc Le +10

    cs.CVcs.AIarXiv:2608.24130v12026
  20. Rethinking Pre-Training and Augmentation for Zero-Shot Cross-City Object Detection

    Long Hoang Pham, Quoc Pham-Nam Ho, Huy-Hung Nguyen +10

    cs.CVcs.AIarXiv:2608.24154v12026
  21. Infant Care Video Dataset for Classification of Interventions Using Transformers

    Igor Bogdanov, James Green

    cs.CVcs.AIcs.LGarXiv:2608.23838v12026
  22. A Comparative Study in Surgical AI: Potential and Limitations of Data, Compute, and Scaling

    Kirill Skobelev, Eric Fithian, Yegor Baranovski +9

    cs.AIcs.CVcs.LGarXiv:2603.27341v42026
  23. When and why vision-language models behave like bags-of-words, and what to do about it?

    Mert Yuksekgonul, Federico Bianchi, Pratyusha Kalluri +2

    cs.CVcs.AIcs.CLarXiv:2210.01936v32022
  24. Relative Time Intervals Representation for Word-level Timestamping with Masked Training

    Quanwei Tang, Zhiyu Tang, Xu Li +3

    cs.AIarXiv:2608.24041v12026
  25. RENDER: Controlling Reader-Facing Evidence in LLM Memory Evaluation

    Yuan Si, Simeng Han, Daming Li +1

    cs.AIarXiv:2608.23568v12026
  26. Explainable Machine Learning in Deployment

    Umang Bhatt, Alice Xiang, Shubham Sharma +7

    cs.LGcs.AIcs.CYarXiv:1909.06342v42019
  27. ChorusTIC: Training-Free Multivariate Time Series Classification via Chorus In-Context Learning

    Juntao Fang, Shifeng Xie, Ruichu Cai +6

    cs.LGcs.AIstat.MLarXiv:2608.24033v12026
  28. A Human-Factors Guided Cognitive Model of Visuospatial Complexity in Embodied Active Vision

    Vasiliki Kondyli, Jakob Suchan, Mehul Bhatt

    q-bio.NCcs.AIcs.CVarXiv:2608.23572v12026
  29. Understand and Accelerate Memory Processing Pipeline for Large Language Model Inference

    Zifan He, Rui Ma, Yizhou Sun +1

    cs.DCcs.AIarXiv:2603.29002v32026
  30. On Predicting Vulnerability Severity Using In-Context Learning: An Industrial Case Study

    Daniel Rodriguez-Cardenas, David Nader Palacio, Anna Schmedding +8

    cs.CRcs.AIarXiv:2608.22089v12026
  31. Correcting Variable Importance Scored by Random Forests

    Guancheng Zhou, Haiping Xu, Jason Liu +1

    stat.MEcs.AIcs.LGarXiv:2606.10770v12026
  32. CALVIN: A Benchmark for Language-Conditioned Policy Learning for Long-Horizon Robot Manipulation Tasks

    Oier Mees, Lukas Hermann, Erick Rosete-Beas +1

    cs.ROcs.AIcs.CLarXiv:2112.03227v42021
  33. KSE-Web: An Analysis of Hybrid Retrieval and LLM-Assisted Query Expansion for Low-Resource Khmer Semantic Search

    Nimol Thuon

    cs.CLcs.AIcs.IRarXiv:2608.21365v12026
  34. MMFace-DiT: A Dual-Stream Diffusion Transformer for High-Fidelity Multimodal Face Generation

    Bharath Krishnamurthy, Ajita Rattani

    cs.CVcs.AIarXiv:2603.29029v12026
  35. Improving Diffusion Models for Inverse Problems using Manifold Constraints

    Hyungjin Chung, Byeongsu Sim, Dohoon Ryu +1

    cs.LGcs.AIcs.CVarXiv:2206.00941v32022
  36. Joint Extraction of Entities and Relations Based on a Novel Tagging Scheme

    Suncong Zheng, Feng Wang, Hongyun Bao +3

    cs.CLcs.AIcs.LGarXiv:1706.05075v12017
  37. When Will AI Exceed Human Performance? Evidence from AI Experts

    Katja Grace, John Salvatier, Allan Dafoe +2

    cs.AIcs.CYarXiv:1705.08807v32017
  38. Distinguishing Revision and Delayed Elaboration in Incremental Narrative Interpretation

    Yi-Chun Chen

    cs.CLcs.AIcs.MMarXiv:2608.21364v12026
  39. Contrastive Decoding: Open-ended Text Generation as Optimization

    Xiang Lisa Li, Ari Holtzman, Daniel Fried +5

    cs.CLcs.AIcs.LGarXiv:2210.15097v22022
  40. Reinforcement Learning on Benign Facts Amplifies Leakage of Memorized Private Data

    Renfei Zhang, Niloofar Mireshghallah

    cs.LGcs.AIarXiv:2608.21727v12026
  41. Learning by Cheating

    Dian Chen, Brady Zhou, Vladlen Koltun +1

    cs.ROcs.AIcs.CVarXiv:1912.12294v12019
  42. Eureka: Human-Level Reward Design via Coding Large Language Models

    Yecheng Jason Ma, William Liang, Guanzhi Wang +6

    cs.ROcs.AIcs.LGarXiv:2310.12931v22023
  43. Wazobia Eval: A Benchmark for Nigerian Pidgin Emotion Understanding, Sarcasm Detection, and Cultural Reasoning

    Stephanie Okoye

    cs.CLcs.AIarXiv:2608.21369v12026
  44. Why This, Not That? Mining User Profiles for Pair-wise Counterfactuals

    Meysam Varasteh, Veronika Bogina, Noam Koenigstein +1

    cs.IRcs.AIarXiv:2608.21662v12026
  45. Determinants of Starting Salaries for Filipino Graduates: An Explainable Machine Learning Approach

    Alexander Gabriel A. Aranes, John Michael C. Magpantay, Reginald Neil C. Recario +2

    cs.CYcs.AIarXiv:2608.21383v12026
  46. Model of Models: When Does Emitting a Specialist Beat Attending, Adapting, or Tuning?

    John C. Howell

    cs.LGcs.AIarXiv:2608.21386v12026
  47. ADMIL: Attention-Distilled Multiple Instance Learning for Selective Foundation Model Inference in Pathology

    Duncan Stothers, Ren-Chin Wu, William Lotter

    cs.CVcs.AIcs.LGarXiv:2608.22066v12026
  48. The Real-World-Weight Cross-Entropy Loss Function: Modeling the Costs of Mislabeling

    Yaoshiang Ho, Samuel Wookey

    cs.LGcs.AIstat.MLarXiv:2001.00570v12020
  49. GeoRisk-RAG: A Hierarchy-Aware Risk Framework for Improving RAG Reliability through Selective Answering

    Meenu Ravi, Shailik Sarkar, Lulwah AlKulaib +2

    cs.CLcs.AIarXiv:2608.22634v12026
  50. Agentic Scaffolding Amplifies Sycophantic Behavior in Large Language Models

    Thantham Jittham

    cs.CLcs.AIcs.LGarXiv:2608.21377v12026
  51. RoboShape: Information-Theoretic Point Cloud Representations for Privacy-Aware Robot Perception

    Oguzhan Baser, Mirac Sozen, Kaan Kale +2

    cs.ROcs.AIcs.CVarXiv:2608.21380v12026
  52. A Social Media Analysis of Discourse on the Israel--Palestine Conflict on Telegram

    Michail Zafeiropoulos, Despoina Antonakaki, Sotiris Ioannidis

    cs.CLcs.AIcs.CYarXiv:2608.21385v12026
  53. Beyond Two Bytes per Letter: Tokenization Overhead in Cyrillic AI Systems

    Ivan Dobrovolskyi

    cs.CLcs.AIarXiv:2608.21384v12026
  54. Interrupting the Chain: Human Perception of AI-Generated Disinformation Through a Kill Chain Lens

    Alexander Loth, Martin Kappes, Marc-Oliver Pahl

    cs.CYcs.AIcs.CRarXiv:2608.21389v12026
  55. TPU v4: An Optically Reconfigurable Supercomputer for Machine Learning with Hardware Support for Embeddings

    Norman P. Jouppi, George Kurian, Sheng Li +11

    cs.ARcs.AIcs.LGarXiv:2304.01433v32023
  56. A review and comparison of strategies for multi-step ahead time series forecasting based on the NN5 forecasting competition

    Souhaib Ben Taieb, Gianluca Bontempi, Amir Atiya +1

    stat.MLcs.AIcs.LGarXiv:1108.3259v12011
  57. Are LLMs Vulnerable to Preference-Undermining Attacks (PUA)? A Factorial Analysis Methodology for Diagnosing the Trade-off between Preference Alignment and Real-World Validity

    Hongjun An, Yiliang Song, Jiangan Chen +3

    cs.CRcs.AIarXiv:2601.06596v12026
  58. The Agent's First Day: Benchmarking Learning, Exploration, and Scheduling in the Workplace Scenarios

    Daocheng Fu, Jianbiao Mei, Rong Wu +7

    cs.AIarXiv:2601.08173v22026
  59. MAXS: Meta-Adaptive Exploration with LLM Agents

    Jian Zhang, Zhiyuan Wang, Zhangqi Wang +7

    cs.AIarXiv:2601.09259v12026
  60. Architecture as Capability Equalizer for Coding Agents

    Arquimedes Canedo

    cs.SEcs.AIcs.CLarXiv:2608.21747v12026