Artificial Intelligence

Papers filed under cs.AI on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.

Search paper metadata (including unsummarized papers)

11,461 to 11,520 of 15,447

  1. When and why vision-language models behave like bags-of-words, and what to do about it?

    Mert Yuksekgonul, Federico Bianchi, Pratyusha Kalluri +2

    cs.CVcs.AIcs.CLarXiv:2210.01936v32022
  2. Relative Time Intervals Representation for Word-level Timestamping with Masked Training

    Quanwei Tang, Zhiyu Tang, Xu Li +3

    cs.AIarXiv:2608.24041v12026
  3. RENDER: Controlling Reader-Facing Evidence in LLM Memory Evaluation

    Yuan Si, Simeng Han, Daming Li +1

    cs.AIarXiv:2608.23568v12026
  4. Explainable Machine Learning in Deployment

    Umang Bhatt, Alice Xiang, Shubham Sharma +7

    cs.LGcs.AIcs.CYarXiv:1909.06342v42019
  5. ChorusTIC: Training-Free Multivariate Time Series Classification via Chorus In-Context Learning

    Juntao Fang, Shifeng Xie, Ruichu Cai +6

    cs.LGcs.AIstat.MLarXiv:2608.24033v12026
  6. A Human-Factors Guided Cognitive Model of Visuospatial Complexity in Embodied Active Vision

    Vasiliki Kondyli, Jakob Suchan, Mehul Bhatt

    q-bio.NCcs.AIcs.CVarXiv:2608.23572v12026
  7. Understand and Accelerate Memory Processing Pipeline for Large Language Model Inference

    Zifan He, Rui Ma, Yizhou Sun +1

    cs.DCcs.AIarXiv:2603.29002v32026
  8. On Predicting Vulnerability Severity Using In-Context Learning: An Industrial Case Study

    Daniel Rodriguez-Cardenas, David Nader Palacio, Anna Schmedding +8

    cs.CRcs.AIarXiv:2608.22089v12026
  9. Correcting Variable Importance Scored by Random Forests

    Guancheng Zhou, Haiping Xu, Jason Liu +1

    stat.MEcs.AIcs.LGarXiv:2606.10770v12026
  10. CALVIN: A Benchmark for Language-Conditioned Policy Learning for Long-Horizon Robot Manipulation Tasks

    Oier Mees, Lukas Hermann, Erick Rosete-Beas +1

    cs.ROcs.AIcs.CLarXiv:2112.03227v42021
  11. KSE-Web: An Analysis of Hybrid Retrieval and LLM-Assisted Query Expansion for Low-Resource Khmer Semantic Search

    Nimol Thuon

    cs.CLcs.AIcs.IRarXiv:2608.21365v12026
  12. MMFace-DiT: A Dual-Stream Diffusion Transformer for High-Fidelity Multimodal Face Generation

    Bharath Krishnamurthy, Ajita Rattani

    cs.CVcs.AIarXiv:2603.29029v12026
  13. Improving Diffusion Models for Inverse Problems using Manifold Constraints

    Hyungjin Chung, Byeongsu Sim, Dohoon Ryu +1

    cs.LGcs.AIcs.CVarXiv:2206.00941v32022
  14. Joint Extraction of Entities and Relations Based on a Novel Tagging Scheme

    Suncong Zheng, Feng Wang, Hongyun Bao +3

    cs.CLcs.AIcs.LGarXiv:1706.05075v12017
  15. When Will AI Exceed Human Performance? Evidence from AI Experts

    Katja Grace, John Salvatier, Allan Dafoe +2

    cs.AIcs.CYarXiv:1705.08807v32017
  16. Distinguishing Revision and Delayed Elaboration in Incremental Narrative Interpretation

    Yi-Chun Chen

    cs.CLcs.AIcs.MMarXiv:2608.21364v12026
  17. Contrastive Decoding: Open-ended Text Generation as Optimization

    Xiang Lisa Li, Ari Holtzman, Daniel Fried +5

    cs.CLcs.AIcs.LGarXiv:2210.15097v22022
  18. Reinforcement Learning on Benign Facts Amplifies Leakage of Memorized Private Data

    Renfei Zhang, Niloofar Mireshghallah

    cs.LGcs.AIarXiv:2608.21727v12026
  19. Learning by Cheating

    Dian Chen, Brady Zhou, Vladlen Koltun +1

    cs.ROcs.AIcs.CVarXiv:1912.12294v12019
  20. Eureka: Human-Level Reward Design via Coding Large Language Models

    Yecheng Jason Ma, William Liang, Guanzhi Wang +6

    cs.ROcs.AIcs.LGarXiv:2310.12931v22023
  21. Wazobia Eval: A Benchmark for Nigerian Pidgin Emotion Understanding, Sarcasm Detection, and Cultural Reasoning

    Stephanie Okoye

    cs.CLcs.AIarXiv:2608.21369v12026
  22. Why This, Not That? Mining User Profiles for Pair-wise Counterfactuals

    Meysam Varasteh, Veronika Bogina, Noam Koenigstein +1

    cs.IRcs.AIarXiv:2608.21662v12026
  23. Determinants of Starting Salaries for Filipino Graduates: An Explainable Machine Learning Approach

    Alexander Gabriel A. Aranes, John Michael C. Magpantay, Reginald Neil C. Recario +2

    cs.CYcs.AIarXiv:2608.21383v12026
  24. Model of Models: When Does Emitting a Specialist Beat Attending, Adapting, or Tuning?

    John C. Howell

    cs.LGcs.AIarXiv:2608.21386v12026
  25. ADMIL: Attention-Distilled Multiple Instance Learning for Selective Foundation Model Inference in Pathology

    Duncan Stothers, Ren-Chin Wu, William Lotter

    cs.CVcs.AIcs.LGarXiv:2608.22066v12026
  26. The Real-World-Weight Cross-Entropy Loss Function: Modeling the Costs of Mislabeling

    Yaoshiang Ho, Samuel Wookey

    cs.LGcs.AIstat.MLarXiv:2001.00570v12020
  27. GeoRisk-RAG: A Hierarchy-Aware Risk Framework for Improving RAG Reliability through Selective Answering

    Meenu Ravi, Shailik Sarkar, Lulwah AlKulaib +2

    cs.CLcs.AIarXiv:2608.22634v12026
  28. Agentic Scaffolding Amplifies Sycophantic Behavior in Large Language Models

    Thantham Jittham

    cs.CLcs.AIcs.LGarXiv:2608.21377v12026
  29. RoboShape: Information-Theoretic Point Cloud Representations for Privacy-Aware Robot Perception

    Oguzhan Baser, Mirac Sozen, Kaan Kale +2

    cs.ROcs.AIcs.CVarXiv:2608.21380v12026
  30. A Social Media Analysis of Discourse on the Israel--Palestine Conflict on Telegram

    Michail Zafeiropoulos, Despoina Antonakaki, Sotiris Ioannidis

    cs.CLcs.AIcs.CYarXiv:2608.21385v12026
  31. Beyond Two Bytes per Letter: Tokenization Overhead in Cyrillic AI Systems

    Ivan Dobrovolskyi

    cs.CLcs.AIarXiv:2608.21384v12026
  32. Interrupting the Chain: Human Perception of AI-Generated Disinformation Through a Kill Chain Lens

    Alexander Loth, Martin Kappes, Marc-Oliver Pahl

    cs.CYcs.AIcs.CRarXiv:2608.21389v12026
  33. TPU v4: An Optically Reconfigurable Supercomputer for Machine Learning with Hardware Support for Embeddings

    Norman P. Jouppi, George Kurian, Sheng Li +11

    cs.ARcs.AIcs.LGarXiv:2304.01433v32023
  34. A review and comparison of strategies for multi-step ahead time series forecasting based on the NN5 forecasting competition

    Souhaib Ben Taieb, Gianluca Bontempi, Amir Atiya +1

    stat.MLcs.AIcs.LGarXiv:1108.3259v12011
  35. Are LLMs Vulnerable to Preference-Undermining Attacks (PUA)? A Factorial Analysis Methodology for Diagnosing the Trade-off between Preference Alignment and Real-World Validity

    Hongjun An, Yiliang Song, Jiangan Chen +3

    cs.CRcs.AIarXiv:2601.06596v12026
  36. The Agent's First Day: Benchmarking Learning, Exploration, and Scheduling in the Workplace Scenarios

    Daocheng Fu, Jianbiao Mei, Rong Wu +7

    cs.AIarXiv:2601.08173v22026
  37. MAXS: Meta-Adaptive Exploration with LLM Agents

    Jian Zhang, Zhiyuan Wang, Zhangqi Wang +7

    cs.AIarXiv:2601.09259v12026
  38. Architecture as Capability Equalizer for Coding Agents

    Arquimedes Canedo

    cs.SEcs.AIcs.CLarXiv:2608.21747v12026
  39. Self-Supervised Graph Representation Learning for In-The-Wild Wearable and Smartphone based Emotion Recognition

    Ioannis N. Ziogas, Leontios J. Hadjileontiadis, Ahsan H. Khandoker +1

    cs.LGcs.AIeess.SParXiv:2608.22387v12026
  40. Artificial Entanglement in the Fine-Tuning of Large Language Models

    Min Chen, Zihan Wang, Canyu Chen +3

    cs.LGcs.AIhep-tharXiv:2601.06788v12026
  41. What makes ImageNet good for transfer learning?

    Minyoung Huh, Pulkit Agrawal, Alexei A. Efros

    cs.CVcs.AIcs.LGarXiv:1608.08614v22016
  42. YaPO: Learnable Sparse Activation Steering Vectors for Domain Adaptation

    Abdelaziz Bounhar, Rania Hossam Elmohamady Elbadry, Hadi Abdine +3

    cs.AIarXiv:2601.08441v12026
  43. sui-1: Grounded and Verifiable Long-Form Summarization

    Benedikt Droste, Jan Philipp Harries, Maximilian Idahl +1

    cs.CLcs.AIarXiv:2601.08472v12026
  44. SkinFlow: Efficient Information Transmission for Open Dermatological Diagnosis via Dynamic Visual Encoding and Staged RL

    Lijun Liu, Linwei Chen, Zhishou Zhang +7

    cs.CVcs.AIarXiv:2601.09136v12026
  45. Bulbul: A Dataset for Dialectal Arabic Speech Recognition

    Ahmed Ashraf, Aisha Alansari, Fadel Al Abbas +30

    cs.CLcs.AIarXiv:2608.21950v12026
  46. TabFact: A Large-scale Dataset for Table-based Fact Verification

    Wenhu Chen, Hongmin Wang, Jianshu Chen +5

    cs.CLcs.AIarXiv:1909.02164v52019
  47. EvoFSM: Controllable Self-Evolution for Deep Research with Finite State Machines

    Shuo Zhang, Chaofa Yuan, Ryan Guo +11

    cs.AIarXiv:2601.09465v22026
  48. Scalable quantum simulation of continuous-time generative models via tensor networks

    Nathan X. Kodama, L. Andrew Wray, Sam Cochran +3

    quant-phcs.AIcs.LGarXiv:2608.21700v12026
  49. robosuite: A Modular Simulation Framework and Benchmark for Robot Learning

    Yuke Zhu, Josiah Wong, Ajay Mandlekar +6

    cs.ROcs.AIcs.LGarXiv:2009.12293v32020
  50. Sequential LLM Release Facilitates Manipulation in Regulated Markets

    Eilam Shapira, Moshe Tennenholtz, Roi Reichart

    cs.GTcs.AIcs.CLarXiv:2601.11496v32026
  51. Exploring the Limitations of Behavior Cloning for Autonomous Driving

    Felipe Codevilla, Eder Santana, Antonio M. López +1

    cs.CVcs.AIarXiv:1904.08980v12019
  52. Memory-V2V: Memory-Augmented Video-to-Video Diffusion for Consistent Multi-Turn Editing

    Dohun Lee, Chun-Hao Paul Huang, Xuelin Chen +3

    cs.CVcs.AIcs.LGarXiv:2601.16296v22026
  53. $A^3$-Bench: Benchmarking Memory-Driven Scientific Reasoning via Anchor and Attractor Activation

    Jian Zhang, Yu He, Zhiyuan Wang +5

    cs.AIarXiv:2601.09274v12026
  54. Explainable AI (XAI): A Systematic Meta-Survey of Current Challenges and Future Opportunities

    Waddah Saeed, Christian Omlin

    cs.LGcs.AIarXiv:2111.06420v12021
  55. A Survey on Metric Learning for Feature Vectors and Structured Data

    Aurélien Bellet, Amaury Habrard, Marc Sebban

    cs.LGcs.AIstat.MLarXiv:1306.6709v42013
  56. DanQing: An Up-to-Date Large-Scale Chinese Vision-Language Pre-training Dataset

    Hengyu Shen, Tiancheng Gu, Bin Qin +10

    cs.CVcs.AIarXiv:2601.10305v32026
  57. LIBERTy: A Causal Framework for Benchmarking Concept-Based Explanations of LLMs with Structural Counterfactuals

    Gilat Toker, Nitay Calderon, Ohad Amosy +1

    cs.CLcs.AIarXiv:2601.10700v22026
  58. Register Shifts Break LLM Safety: A Bengali Benchmark with Culturally Grounded Harms

    Naymul Islam, Nusrat Jahan Lia, Shubhashis Roy Dipta +2

    cs.CLcs.AIarXiv:2608.22335v12026
  59. AstroReason-Bench: Evaluating Unified Agentic Planning across Heterogeneous Space Planning Problems

    Weiyi Wang, Xinchi Chen, Jingjing Gong +2

    cs.AIcs.CLarXiv:2601.11354v12026
  60. Less Is More -- Until It Breaks: Security Pitfalls of Vision Token Compression in Large Vision-Language Models

    Xiaomei Zhang, Zhaoxi Zhang, Leo Yu Zhang +3

    cs.CRcs.AIarXiv:2601.12042v12026