Artificial Intelligence

Papers filed under cs.AI on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.

Search paper metadata (including unsummarized papers)

1,381 to 1,440 of 15,379

  1. Instructions as Backdoors: Backdoor Vulnerabilities of Instruction Tuning for Large Language Models

    Jiashu Xu, Mingyu Derek Ma, Fei Wang +2

    cs.CLcs.AIcs.CRarXiv:2305.14710v22023
  2. Reasoning-Driven Multimodal LLM for Domain Generalization

    Zhipeng Xu, Zilong Wang, Xinyang Jiang +3

    cs.AIarXiv:2602.23777v12026
  3. CoDAR: Continuous Diffusion Language Models are More Powerful Than You Think

    Junzhe Shen, Jieru Zhao, Ziwei He +1

    cs.CLcs.AIcs.LGarXiv:2603.02547v12026
  4. Does Linguistic Structure Enrichment Enhance Coherence Assessment? Not With Current Architectures

    Victor Mazzotti, Luiz Pereira, Marina Bitencourt dos Santos +4

    cs.CLcs.AIarXiv:2609.10893v12026
  5. MMedAgent: Learning to Use Medical Tools with Multi-modal Agent

    Binxu Li, Tiankai Yan, Yuanting Pan +8

    cs.CLcs.AIarXiv:2407.02483v22024
  6. Generating Benchmarks for Factuality Evaluation of Language Models

    Dor Muhlgay, Ori Ram, Inbal Magar +7

    cs.CLcs.AIarXiv:2307.06908v22023
  7. RiskOracle: A Minute-level Citywide Traffic Accident Forecasting Framework

    Zhengyang Zhou, Yang Wang, Xike Xie +2

    cs.AIeess.SParXiv:2003.00819v12020
  8. Multilingual in Name Only? Cultural and Linguistic Weaknesses of LLMs in Urdu

    Farah Adeeba, Abdul Rafae Khan, Rajesh Bhatt +1

    cs.CLcs.AIcs.LGarXiv:2609.10758v12026
  9. GCGNet: Graph-Consistent Generative Network for Time Series Forecasting with Exogenous Variables

    Zhengyu Li, Xiangfei Qiu, Yuhan Zhu +4

    cs.LGcs.AIarXiv:2603.08032v22026
  10. ReviewerGPT? An Exploratory Study on Using Large Language Models for Paper Reviewing

    Ryan Liu, Nihar B. Shah

    cs.CLcs.AIcs.DLarXiv:2306.00622v12023
  11. Radiomics in Medical Imaging: Methods, Applications, and Challenges

    Fnu Neha, Deepak kumar Shukla

    eess.IVcs.AIcs.LGarXiv:2602.00102v12026
  12. A Comprehensive Survey on Data Augmentation

    Zaitian Wang, Pengfei Wang, Kunpeng Liu +6

    cs.LGcs.AIarXiv:2405.09591v42024
  13. On the Creativity of Large Language Models

    Giorgio Franceschelli, Mirco Musolesi

    cs.AIcs.CLcs.CYarXiv:2304.00008v52023
  14. Using Text-to-Image Generation for Architectural Design Ideation

    Ville Paananen, Jonas Oppenlaender, Aku Visuri

    cs.HCcs.AIcs.CVarXiv:2304.10182v12023
  15. Better to Ask in English: Cross-Lingual Evaluation of Large Language Models for Healthcare Queries

    Yiqiao Jin, Mohit Chandra, Gaurav Verma +3

    cs.CLcs.AIarXiv:2310.13132v22023
  16. Colosseum: Auditing Collusion in Cooperative Multi-Agent Systems

    Mason Nakamura, Abhinav Kumar, Saswat Das +5

    cs.MAcs.AIcs.CLarXiv:2602.15198v22026
  17. Data-Efficient Language Modeling: From Frontier Advancement to Principle-Guided Model Improvement

    Shuxing Yang, Kaihao Zhu, Junjie Yang +13

    cs.CLcs.AIarXiv:2609.10702v12026
  18. HIQL: Offline Goal-Conditioned RL with Latent States as Actions

    Seohong Park, Dibya Ghosh, Benjamin Eysenbach +1

    cs.LGcs.AIcs.ROarXiv:2307.11949v42023
  19. Robotic World Model: A Neural Network Simulator for Robust Policy Optimization in Robotics

    Chenhao Li, Andreas Krause, Marco Hutter

    cs.ROcs.AIcs.LGarXiv:2501.10100v52025
  20. Taking the Next Step with Generative Artificial Intelligence: The Transformative Role of Multimodal Large Language Models in Science Education

    Arne Bewersdorff, Christian Hartmann, Marie Hornberger +6

    cs.AIcs.CYarXiv:2401.00832v32024
  21. Unmasking Bias in AI: A Systematic Review of Bias Detection and Mitigation Strategies in Electronic Health Record-based Models

    Feng Chen, Liqin Wang, Julie Hong +2

    cs.AIcs.CYcs.LGarXiv:2310.19917v32023
  22. Lag-Llama: Towards Foundation Models for Probabilistic Time Series Forecasting

    Kashif Rasul, Arjun Ashok, Andrew Robert Williams +15

    cs.LGcs.AIarXiv:2310.08278v32023
  23. M2F: Automated Formalization of Mathematical Literature at Scale

    Zichen Wang, Wanli Ma, Zhenyu Ming +3

    cs.AIarXiv:2602.17016v12026
  24. Not All Experts are Equal: Efficient Expert Pruning and Skipping for Mixture-of-Experts Large Language Models

    Xudong Lu, Qi Liu, Yuhui Xu +5

    cs.CLcs.AIcs.LGarXiv:2402.14800v22024
  25. DRG-MAPPO: Hierarchical Dynamic Role-Graph Multi-Agent Reinforcement Learning for Cooperative Air Combat

    Junlin Liu, Chengwei Li, Yang Gao +4

    cs.AIarXiv:2609.11155v12026
  26. Generative agent-based modeling with actions grounded in physical, social, or digital space using Concordia

    Alexander Sasha Vezhnevets, John P. Agapiou, Avia Aharon +7

    cs.AIcs.CLarXiv:2312.03664v22023
  27. Flow Policy Gradients for Robot Control

    Brent Yi, Hongsuk Choi, Himanshu Gaurav Singh +9

    cs.ROcs.AIarXiv:2602.02481v12026
  28. AnyV2V: A Tuning-Free Framework For Any Video-to-Video Editing Tasks

    Max Ku, Cong Wei, Weiming Ren +2

    cs.CVcs.AIcs.MMarXiv:2403.14468v42024
  29. ChipBench: A Next-Step Benchmark for Evaluating LLM Performance in AI-Aided Chip Design

    Zhongkai Yu, Chenyang Zhou, Yichen Lin +6

    cs.AIcs.ARarXiv:2601.21448v22026
  30. An Open Recipe for IMO Gold: Training Nemotron for Olympiad Mathematics

    Ivan Moshkov, Stephen Ge, George Armstrong +3

    cs.AIarXiv:2609.10712v12026
  31. Decoupling KL and Trajectories: A Unified Perspective for SFT, DAgger, Offline RL, and OPD in LLM Distillation

    Anhao Zhao, Haoran Xin, Yingqi Fan +3

    cs.LGcs.AIcs.CLarXiv:2605.16826v12026
  32. Large Language Models and Knowledge Graphs: Opportunities and Challenges

    Jeff Z. Pan, Simon Razniewski, Jan-Christoph Kalo +13

    cs.AIcs.CLarXiv:2308.06374v12023
  33. What Counts as AI Sycophancy? A Taxonomy and Expert Survey of a Fragmented Construct

    Meryl Ye, Lujain Ibrahim, Jessica Y. Bo +5

    cs.AIarXiv:2605.21778v12026
  34. DynamicStereo: Consistent Dynamic Depth from Stereo Videos

    Nikita Karaev, Ignacio Rocco, Benjamin Graham +3

    cs.CVcs.AIarXiv:2305.02296v12023
  35. CycleResearcher: Improving Automated Research via Automated Review

    Yixuan Weng, Minjun Zhu, Guangsheng Bao +4

    cs.CLcs.AIcs.CYarXiv:2411.00816v32024
  36. We're Different, We're the Same: Creative Homogeneity Across LLMs

    Emily Wenger, Yoed Kenett

    cs.CYcs.AIcs.CLarXiv:2501.19361v12025
  37. A Review of Deep Transfer Learning and Recent Advancements

    Mohammadreza Iman, Khaled Rasheed, Hamid R. Arabnia

    cs.LGcs.AIcs.CVarXiv:2201.09679v22022
  38. Force Prompting: Video Generation Models Can Learn and Generalize Physics-based Control Signals

    Nate Gillman, Charles Herrmann, Michael Freeman +4

    cs.CVcs.AIarXiv:2505.19386v22025
  39. X-AuT: Progressive Audio-Encoder Compression for Speech LLMs with Cross-Scale Distillation

    Haojun Zhang, Yi Zou, Min Chen +7

    cs.SDcs.AIarXiv:2609.11412v12026
  40. Self-Improving AI Coding Agents Through Accumulated Behavioral Rules: A Closed-Loop Framework

    Aditya Aggarwal, Nahid Farhady Ghalaty

    cs.SEcs.AIarXiv:2607.13091v12026
  41. The AI Scientist-v2: Workshop-Level Automated Scientific Discovery via Agentic Tree Search

    Yutaro Yamada, Robert Tjarko Lange, Cong Lu +5

    cs.AIcs.CLcs.LGarXiv:2504.08066v12025
    Summaries:한국어
  42. BioT5: Enriching Cross-modal Integration in Biology with Chemical Knowledge and Natural Language Associations

    Qizhi Pei, Wei Zhang, Jinhua Zhu +5

    cs.CLcs.AIcs.LGarXiv:2310.07276v32023
  43. SpatialBlock: Enhancing Spatial Intelligence in LVLMs via Synthetic Block-Stacking Problem

    Soohyun Ryu, Sohee Kim, Eunho Yang

    cs.CVcs.AIarXiv:2609.07064v12026
  44. A Comparative study Between Fuzzy Clustering Algorithm and Hard Clustering Algorithm

    Dibya Jyoti Bora, Dr. Anil Kumar Gupta

    cs.AIarXiv:1404.6059v12014
  45. Nested conformal prediction and quantile out-of-bag ensemble methods

    Chirag Gupta, Arun K. Kuchibhotla, Aaditya K. Ramdas

    stat.MEcs.AImath.STarXiv:1910.10562v42019
  46. The Causal-Neural Connection: Expressiveness, Learnability, and Inference

    Kevin Xia, Kai-Zhan Lee, Yoshua Bengio +1

    cs.LGcs.AIarXiv:2107.00793v32021
  47. X-OPD: Cross-Modal On-Policy Distillation for Capability Alignment in Speech LLMs

    Di Cao, Dongjie Fu, Hai Yu +3

    eess.AScs.AIcs.CLarXiv:2603.24596v32026
  48. Few-shot In-context Learning for Knowledge Base Question Answering

    Tianle Li, Xueguang Ma, Alex Zhuang +3

    cs.CLcs.AIarXiv:2305.01750v22023
  49. MAD: A Scalable Dataset for Language Grounding in Videos from Movie Audio Descriptions

    Mattia Soldan, Alejandro Pardo, Juan León Alcázar +4

    cs.CVcs.AIarXiv:2112.00431v22021
  50. DeCap: Decoding CLIP Latents for Zero-Shot Captioning via Text-Only Training

    Wei Li, Linchao Zhu, Longyin Wen +1

    cs.CVcs.AIcs.CLarXiv:2303.03032v12023
  51. B-Pref: Benchmarking Preference-Based Reinforcement Learning

    Kimin Lee, Laura Smith, Anca Dragan +1

    cs.LGcs.AIcs.HCarXiv:2111.03026v12021
  52. Large Language Models and the Reverse Turing Test

    Terrence Sejnowski

    cs.CLcs.AIcs.LGarXiv:2207.14382v92022
  53. Memory-Efficient Fine-Tuning of Compressed Large Language Models via sub-4-bit Integer Quantization

    Jeonghoon Kim, Jung Hyun Lee, Sungdong Kim +4

    cs.LGcs.AIarXiv:2305.14152v22023
  54. Formal Verification of Autonomous Vehicle Platooning

    Maryam Kamali, Louise A. Dennis, Owen McAree +2

    cs.AIcs.SEarXiv:1602.01718v12016
  55. A Multi-Objective Deep Reinforcement Learning Framework

    Thanh Thi Nguyen, Ngoc Duy Nguyen, Peter Vamplew +3

    cs.LGcs.AIstat.MLarXiv:1803.02965v32018
  56. IPIGuard: A Novel Tool Dependency Graph-Based Defense Against Indirect Prompt Injection in LLM Agents

    Hengyu An, Jinghuai Zhang, Tianyu Du +4

    cs.CRcs.AIcs.CLarXiv:2508.15310v12025
  57. UNR-Explainer: Counterfactual Explanations for Unsupervised Node Representation Learning Models

    Hyunju Kang, Geonhee Han, Hogun Park

    cs.LGcs.AIarXiv:2605.17285v12026
  58. A spelling correction model for end-to-end speech recognition

    Jinxi Guo, Tara N. Sainath, Ron J. Weiss

    eess.AScs.AIcs.CLarXiv:1902.07178v12019
  59. LLM-based Agents Suffer from Hallucinations: A Survey of Taxonomy, Methods, and Directions

    Xixun Lin, Yucheng Ning, Jingwen Zhang +21

    cs.AIarXiv:2509.18970v22025
  60. ORGANA: A Robotic Assistant for Automated Chemistry Experimentation and Characterization

    Kourosh Darvish, Marta Skreta, Yuchi Zhao +9

    cs.ROcs.AIarXiv:2401.06949v22024