Artificial Intelligence
Papers filed under cs.AI on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.
Search paper metadata (including unsummarized papers)
11,461 to 11,520 of 15,447
When and why vision-language models behave like bags-of-words, and what to do about it?
Mert Yuksekgonul, Federico Bianchi, Pratyusha Kalluri +2
cs.CVcs.AIcs.CLarXiv:2210.01936v32022Relative Time Intervals Representation for Word-level Timestamping with Masked Training
Quanwei Tang, Zhiyu Tang, Xu Li +3
cs.AIarXiv:2608.24041v12026RENDER: Controlling Reader-Facing Evidence in LLM Memory Evaluation
Yuan Si, Simeng Han, Daming Li +1
cs.AIarXiv:2608.23568v12026Explainable Machine Learning in Deployment
Umang Bhatt, Alice Xiang, Shubham Sharma +7
cs.LGcs.AIcs.CYarXiv:1909.06342v42019ChorusTIC: Training-Free Multivariate Time Series Classification via Chorus In-Context Learning
Juntao Fang, Shifeng Xie, Ruichu Cai +6
cs.LGcs.AIstat.MLarXiv:2608.24033v12026A Human-Factors Guided Cognitive Model of Visuospatial Complexity in Embodied Active Vision
Vasiliki Kondyli, Jakob Suchan, Mehul Bhatt
q-bio.NCcs.AIcs.CVarXiv:2608.23572v12026Understand and Accelerate Memory Processing Pipeline for Large Language Model Inference
Zifan He, Rui Ma, Yizhou Sun +1
cs.DCcs.AIarXiv:2603.29002v32026On Predicting Vulnerability Severity Using In-Context Learning: An Industrial Case Study
Daniel Rodriguez-Cardenas, David Nader Palacio, Anna Schmedding +8
cs.CRcs.AIarXiv:2608.22089v12026Correcting Variable Importance Scored by Random Forests
Guancheng Zhou, Haiping Xu, Jason Liu +1
stat.MEcs.AIcs.LGarXiv:2606.10770v12026CALVIN: A Benchmark for Language-Conditioned Policy Learning for Long-Horizon Robot Manipulation Tasks
Oier Mees, Lukas Hermann, Erick Rosete-Beas +1
cs.ROcs.AIcs.CLarXiv:2112.03227v42021KSE-Web: An Analysis of Hybrid Retrieval and LLM-Assisted Query Expansion for Low-Resource Khmer Semantic Search
Nimol Thuon
cs.CLcs.AIcs.IRarXiv:2608.21365v12026MMFace-DiT: A Dual-Stream Diffusion Transformer for High-Fidelity Multimodal Face Generation
Bharath Krishnamurthy, Ajita Rattani
cs.CVcs.AIarXiv:2603.29029v12026Improving Diffusion Models for Inverse Problems using Manifold Constraints
Hyungjin Chung, Byeongsu Sim, Dohoon Ryu +1
cs.LGcs.AIcs.CVarXiv:2206.00941v32022Joint Extraction of Entities and Relations Based on a Novel Tagging Scheme
Suncong Zheng, Feng Wang, Hongyun Bao +3
cs.CLcs.AIcs.LGarXiv:1706.05075v12017When Will AI Exceed Human Performance? Evidence from AI Experts
Katja Grace, John Salvatier, Allan Dafoe +2
cs.AIcs.CYarXiv:1705.08807v32017Distinguishing Revision and Delayed Elaboration in Incremental Narrative Interpretation
Yi-Chun Chen
cs.CLcs.AIcs.MMarXiv:2608.21364v12026Contrastive Decoding: Open-ended Text Generation as Optimization
Xiang Lisa Li, Ari Holtzman, Daniel Fried +5
cs.CLcs.AIcs.LGarXiv:2210.15097v22022Reinforcement Learning on Benign Facts Amplifies Leakage of Memorized Private Data
Renfei Zhang, Niloofar Mireshghallah
cs.LGcs.AIarXiv:2608.21727v12026Learning by Cheating
Dian Chen, Brady Zhou, Vladlen Koltun +1
cs.ROcs.AIcs.CVarXiv:1912.12294v12019Eureka: Human-Level Reward Design via Coding Large Language Models
Yecheng Jason Ma, William Liang, Guanzhi Wang +6
cs.ROcs.AIcs.LGarXiv:2310.12931v22023Wazobia Eval: A Benchmark for Nigerian Pidgin Emotion Understanding, Sarcasm Detection, and Cultural Reasoning
Stephanie Okoye
cs.CLcs.AIarXiv:2608.21369v12026Why This, Not That? Mining User Profiles for Pair-wise Counterfactuals
Meysam Varasteh, Veronika Bogina, Noam Koenigstein +1
cs.IRcs.AIarXiv:2608.21662v12026Determinants of Starting Salaries for Filipino Graduates: An Explainable Machine Learning Approach
Alexander Gabriel A. Aranes, John Michael C. Magpantay, Reginald Neil C. Recario +2
cs.CYcs.AIarXiv:2608.21383v12026Model of Models: When Does Emitting a Specialist Beat Attending, Adapting, or Tuning?
John C. Howell
cs.LGcs.AIarXiv:2608.21386v12026ADMIL: Attention-Distilled Multiple Instance Learning for Selective Foundation Model Inference in Pathology
Duncan Stothers, Ren-Chin Wu, William Lotter
cs.CVcs.AIcs.LGarXiv:2608.22066v12026The Real-World-Weight Cross-Entropy Loss Function: Modeling the Costs of Mislabeling
Yaoshiang Ho, Samuel Wookey
cs.LGcs.AIstat.MLarXiv:2001.00570v12020GeoRisk-RAG: A Hierarchy-Aware Risk Framework for Improving RAG Reliability through Selective Answering
Meenu Ravi, Shailik Sarkar, Lulwah AlKulaib +2
cs.CLcs.AIarXiv:2608.22634v12026Agentic Scaffolding Amplifies Sycophantic Behavior in Large Language Models
Thantham Jittham
cs.CLcs.AIcs.LGarXiv:2608.21377v12026RoboShape: Information-Theoretic Point Cloud Representations for Privacy-Aware Robot Perception
Oguzhan Baser, Mirac Sozen, Kaan Kale +2
cs.ROcs.AIcs.CVarXiv:2608.21380v12026A Social Media Analysis of Discourse on the Israel--Palestine Conflict on Telegram
Michail Zafeiropoulos, Despoina Antonakaki, Sotiris Ioannidis
cs.CLcs.AIcs.CYarXiv:2608.21385v12026Beyond Two Bytes per Letter: Tokenization Overhead in Cyrillic AI Systems
Ivan Dobrovolskyi
cs.CLcs.AIarXiv:2608.21384v12026Interrupting the Chain: Human Perception of AI-Generated Disinformation Through a Kill Chain Lens
Alexander Loth, Martin Kappes, Marc-Oliver Pahl
cs.CYcs.AIcs.CRarXiv:2608.21389v12026TPU v4: An Optically Reconfigurable Supercomputer for Machine Learning with Hardware Support for Embeddings
Norman P. Jouppi, George Kurian, Sheng Li +11
cs.ARcs.AIcs.LGarXiv:2304.01433v32023A review and comparison of strategies for multi-step ahead time series forecasting based on the NN5 forecasting competition
Souhaib Ben Taieb, Gianluca Bontempi, Amir Atiya +1
stat.MLcs.AIcs.LGarXiv:1108.3259v12011Are LLMs Vulnerable to Preference-Undermining Attacks (PUA)? A Factorial Analysis Methodology for Diagnosing the Trade-off between Preference Alignment and Real-World Validity
Hongjun An, Yiliang Song, Jiangan Chen +3
cs.CRcs.AIarXiv:2601.06596v12026The Agent's First Day: Benchmarking Learning, Exploration, and Scheduling in the Workplace Scenarios
Daocheng Fu, Jianbiao Mei, Rong Wu +7
cs.AIarXiv:2601.08173v22026MAXS: Meta-Adaptive Exploration with LLM Agents
Jian Zhang, Zhiyuan Wang, Zhangqi Wang +7
cs.AIarXiv:2601.09259v12026Architecture as Capability Equalizer for Coding Agents
Arquimedes Canedo
cs.SEcs.AIcs.CLarXiv:2608.21747v12026Self-Supervised Graph Representation Learning for In-The-Wild Wearable and Smartphone based Emotion Recognition
Ioannis N. Ziogas, Leontios J. Hadjileontiadis, Ahsan H. Khandoker +1
cs.LGcs.AIeess.SParXiv:2608.22387v12026Artificial Entanglement in the Fine-Tuning of Large Language Models
Min Chen, Zihan Wang, Canyu Chen +3
cs.LGcs.AIhep-tharXiv:2601.06788v12026What makes ImageNet good for transfer learning?
Minyoung Huh, Pulkit Agrawal, Alexei A. Efros
cs.CVcs.AIcs.LGarXiv:1608.08614v22016YaPO: Learnable Sparse Activation Steering Vectors for Domain Adaptation
Abdelaziz Bounhar, Rania Hossam Elmohamady Elbadry, Hadi Abdine +3
cs.AIarXiv:2601.08441v12026sui-1: Grounded and Verifiable Long-Form Summarization
Benedikt Droste, Jan Philipp Harries, Maximilian Idahl +1
cs.CLcs.AIarXiv:2601.08472v12026SkinFlow: Efficient Information Transmission for Open Dermatological Diagnosis via Dynamic Visual Encoding and Staged RL
Lijun Liu, Linwei Chen, Zhishou Zhang +7
cs.CVcs.AIarXiv:2601.09136v12026Bulbul: A Dataset for Dialectal Arabic Speech Recognition
Ahmed Ashraf, Aisha Alansari, Fadel Al Abbas +30
cs.CLcs.AIarXiv:2608.21950v12026TabFact: A Large-scale Dataset for Table-based Fact Verification
Wenhu Chen, Hongmin Wang, Jianshu Chen +5
cs.CLcs.AIarXiv:1909.02164v52019EvoFSM: Controllable Self-Evolution for Deep Research with Finite State Machines
Shuo Zhang, Chaofa Yuan, Ryan Guo +11
cs.AIarXiv:2601.09465v22026Scalable quantum simulation of continuous-time generative models via tensor networks
Nathan X. Kodama, L. Andrew Wray, Sam Cochran +3
quant-phcs.AIcs.LGarXiv:2608.21700v12026robosuite: A Modular Simulation Framework and Benchmark for Robot Learning
Yuke Zhu, Josiah Wong, Ajay Mandlekar +6
cs.ROcs.AIcs.LGarXiv:2009.12293v32020Sequential LLM Release Facilitates Manipulation in Regulated Markets
Eilam Shapira, Moshe Tennenholtz, Roi Reichart
cs.GTcs.AIcs.CLarXiv:2601.11496v32026Exploring the Limitations of Behavior Cloning for Autonomous Driving
Felipe Codevilla, Eder Santana, Antonio M. López +1
cs.CVcs.AIarXiv:1904.08980v12019Memory-V2V: Memory-Augmented Video-to-Video Diffusion for Consistent Multi-Turn Editing
Dohun Lee, Chun-Hao Paul Huang, Xuelin Chen +3
cs.CVcs.AIcs.LGarXiv:2601.16296v22026$A^3$-Bench: Benchmarking Memory-Driven Scientific Reasoning via Anchor and Attractor Activation
Jian Zhang, Yu He, Zhiyuan Wang +5
cs.AIarXiv:2601.09274v12026Explainable AI (XAI): A Systematic Meta-Survey of Current Challenges and Future Opportunities
Waddah Saeed, Christian Omlin
cs.LGcs.AIarXiv:2111.06420v12021A Survey on Metric Learning for Feature Vectors and Structured Data
Aurélien Bellet, Amaury Habrard, Marc Sebban
cs.LGcs.AIstat.MLarXiv:1306.6709v42013DanQing: An Up-to-Date Large-Scale Chinese Vision-Language Pre-training Dataset
Hengyu Shen, Tiancheng Gu, Bin Qin +10
cs.CVcs.AIarXiv:2601.10305v32026LIBERTy: A Causal Framework for Benchmarking Concept-Based Explanations of LLMs with Structural Counterfactuals
Gilat Toker, Nitay Calderon, Ohad Amosy +1
cs.CLcs.AIarXiv:2601.10700v22026Register Shifts Break LLM Safety: A Bengali Benchmark with Culturally Grounded Harms
Naymul Islam, Nusrat Jahan Lia, Shubhashis Roy Dipta +2
cs.CLcs.AIarXiv:2608.22335v12026AstroReason-Bench: Evaluating Unified Agentic Planning across Heterogeneous Space Planning Problems
Weiyi Wang, Xinchi Chen, Jingjing Gong +2
cs.AIcs.CLarXiv:2601.11354v12026Less Is More -- Until It Breaks: Security Pitfalls of Vision Token Compression in Large Vision-Language Models
Xiaomei Zhang, Zhaoxi Zhang, Leo Yu Zhang +3
cs.CRcs.AIarXiv:2601.12042v12026