Artificial Intelligence
Papers filed under cs.AI on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.
Search paper metadata (including unsummarized papers)
4,921 to 4,980 of 15,269
LightEMMA: A Longitudinal Evaluation of Vision-Language Models for Autonomous Driving
Zhijie Qiao, Haowei Li, Zhong Cao +1
cs.ROcs.AIarXiv:2505.00284v32025The Transformative Potential of Artificial Intelligence
Ross Gruetzemacher, Jess Whittlestone
cs.CYcs.AIarXiv:1912.00747v32019Darknet and Deepnet Mining for Proactive Cybersecurity Threat Intelligence
Eric Nunes, Ahmad Diab, Andrew Gunn +7
cs.CRcs.AIcs.CYarXiv:1607.08583v12016Persona Features Control Emergent Misalignment
Miles Wang, Tom Dupré la Tour, Olivia Watkins +8
cs.LGcs.AIarXiv:2506.19823v22025EditReward: A Human-Aligned Reward Model for Instruction-Guided Image Editing
Keming Wu, Sicong Jiang, Max Ku +3
cs.CVcs.AIcs.CLarXiv:2509.26346v22025RailSyn: Diagnosis-Guided Image Generation for Traceable Data Completion in Railway Foreign Object Detection
Quan Hao, Chenxi Zhang, Ziyang Tao +6
cs.CVcs.AIarXiv:2608.30709v12026General Video Game AI: a Multi-Track Framework for Evaluating Agents, Games and Content Generation Algorithms
Diego Perez-Liebana, Jialin Liu, Ahmed Khalifa +3
cs.AIarXiv:1802.10363v42018Artificial Intelligence in the Battle against Coronavirus (COVID-19): A Survey and Future Research Directions
Thanh Thi Nguyen, Quoc Viet Hung Nguyen, Dung Tien Nguyen +7
cs.CYcs.AIcs.LGarXiv:2008.07343v42020AgentAuditor: Human-Level Safety and Security Evaluation for LLM Agents
Hanjun Luo, Shenyu Dai, Chiming Ni +5
cs.AIarXiv:2506.00641v32025Probabilistic Model Checking of Autoregressive Neural Sequence Models
Helge Spieker, Dennis Gross, Arnaud Gotlieb
cs.SEcs.AIarXiv:2609.00838v12026jina-embeddings-v3: Multilingual Embeddings With Task LoRA
Saba Sturua, Isabelle Mohr, Mohammad Kalim Akram +8
cs.CLcs.AIcs.IRarXiv:2409.10173v32024Barbarians at the Gate: How AI is Upending Systems Research
Audrey Cheng, Shu Liu, Melissa Pan +14
cs.AIarXiv:2510.06189v32025WorldModelBench: Judging Video Generation Models As World Models
Dacheng Li, Yunhao Fang, Yukang Chen +10
cs.CVcs.AIarXiv:2502.20694v12025mimeo: Compiling Public Expert Corpora into Agent Skills and Testing What Transfers
Timothy Kassis
cs.AIarXiv:2609.00453v12026Distilled One-Shot Federated Learning
Yanlin Zhou, George Pu, Xiyao Ma +2
cs.LGcs.AIstat.MLarXiv:2009.07999v32020Reinforcement Learning for Self-Improving Agent with Skill Library
Jiongxiao Wang, Qiaojing Yan, Yawei Wang +6
cs.AIarXiv:2512.17102v22025Causal Abstraction: A Theoretical Foundation for Mechanistic Interpretability
Atticus Geiger, Duligur Ibeling, Amir Zur +8
cs.AIarXiv:2301.04709v42023Stress Testing Deliberative Alignment for Anti-Scheming Training
Bronson Schoen, Evgenia Nitishinskaya, Mikita Balesni +16
cs.AIarXiv:2509.15541v12025Is Best-of-N the Best of Them? Coverage, Scaling, and Optimality in Inference-Time Alignment
Audrey Huang, Adam Block, Qinghua Liu +3
cs.AIcs.LGstat.MLarXiv:2503.21878v22025hLLM: Single Pass Decoding for Generative Reranking
Emil Laftchiev, Prachi Agrawal, Moe Kayali +7
cs.LGcs.AIcs.IRarXiv:2609.01807v12026Supply-Chain Poisoning Attacks Against LLM Coding Agent Skill Ecosystems
Yubin Qu, Yi Liu, Tongcheng Geng +5
cs.CRcs.AIcs.CLarXiv:2604.03081v12026Agent Lightning: Train ANY AI Agents with Reinforcement Learning
Xufang Luo, Yuge Zhang, Zhiyuan He +5
cs.AIcs.LGarXiv:2508.03680v12025Evaluation of Pose Tracking Accuracy in the First and Second Generations of Microsoft Kinect
Qifei Wang, Gregorij Kurillo, Ferda Ofli +1
cs.CVcs.AIarXiv:1512.04134v12015Scene-LLM: Extending Language Model for 3D Visual Understanding and Reasoning
Rao Fu, Jingyu Liu, Xilun Chen +2
cs.CVcs.AIarXiv:2403.11401v22024CURIOUS: Intrinsically Motivated Modular Multi-Goal Reinforcement Learning
Cédric Colas, Pierre Fournier, Olivier Sigaud +2
cs.AIarXiv:1810.06284v52018The Hidden Risks of Large Reasoning Models: A Safety Assessment of R1
Kaiwen Zhou, Chengzhi Liu, Xuandong Zhao +5
cs.CYcs.AIarXiv:2502.12659v42025Composite Task-Completion Dialogue Policy Learning via Hierarchical Deep Reinforcement Learning
Baolin Peng, Xiujun Li, Lihong Li +4
cs.CLcs.AIcs.LGarXiv:1704.03084v32017Social Information Processing in Social News Aggregation
Kristina Lerman
cs.CYcs.AIcs.HCarXiv:cs/0703087v22007CausaLM: Causal Model Explanation Through Counterfactual Language Models
Amir Feder, Nadav Oved, Uri Shalit +1
cs.CLcs.AIcs.LGarXiv:2005.13407v52020ShareGPT-4o-Image: Aligning Multimodal Models with GPT-4o-Level Image Generation
Junying Chen, Zhenyang Cai, Pengcheng Chen +5
cs.CVcs.AIcs.LGarXiv:2506.18095v12025The Tool Decathlon: Benchmarking Language Agents for Diverse, Realistic, and Long-Horizon Task Execution
Junlong Li, Wenshuo Zhao, Jian Zhao +18
cs.CLcs.AIarXiv:2510.25726v22025Big-Math: A Large-Scale, High-Quality Math Dataset for Reinforcement Learning in Language Models
Alon Albalak, Duy Phung, Nathan Lile +8
cs.LGcs.AIcs.CLarXiv:2502.17387v12025Decision Theory with Prospect Interference and Entanglement
V. I. Yukalov, D. Sornette
math-phcs.AIphysics.soc-pharXiv:1102.2738v12011SpatialLadder: Progressive Training for Spatial Reasoning in Vision-Language Models
Hongxing Li, Dingming Li, Zixuan Wang +7
cs.CVcs.AIcs.CLarXiv:2510.08531v12025Masked Autoencoders Are Effective Tokenizers for Diffusion Models
Hao Chen, Yujin Han, Fangyi Chen +7
cs.CVcs.AIcs.LGarXiv:2502.03444v22025DeepSeekMath-V2: Towards Self-Verifiable Mathematical Reasoning
Zhihong Shao, Yuxiang Luo, Chengda Lu +6
cs.AIcs.CLarXiv:2511.22570v12025Text-guided flow matching enables sample-efficient crystal structure generation
Wentao Li
cond-mat.mtrl-scics.AIarXiv:2609.01076v12026Chain-of-Visual-Thought: Teaching VLMs to See and Think Better with Continuous Visual Tokens
Yiming Qin, Bomin Wei, Jiaxin Ge +4
cs.CVcs.AIcs.LGarXiv:2511.19418v32025Ad Headline Generation using Self-Critical Masked Language Model
Yashal Shakti Kanungo, Sumit Negi, Aruna Rajan
cs.CLcs.AIcs.LGarXiv:2607.06818v12026Apple Intelligence Foundation Language Models: Tech Report 2025
Ethan Li, Anders Boesen Lindbo Larsen, Chen Zhang +395
cs.LGcs.AIarXiv:2507.13575v32025TempCloze: Can Video-LLMs Identify the Missing Middle?
Wenqi Pei, Henry Hengyuan Zhao, Yilai Liu +4
cs.CVcs.AIarXiv:2609.01515v12026Domino: Discovering Systematic Errors with Cross-Modal Embeddings
Sabri Eyuboglu, Maya Varma, Khaled Saab +5
cs.LGcs.AIarXiv:2203.14960v32022Solving 3D Inverse Problems using Pre-trained 2D Diffusion Models
Hyungjin Chung, Dohoon Ryu, Michael T. McCann +2
cs.CVcs.AIcs.LGarXiv:2211.10655v12022MATH-Perturb: Benchmarking LLMs' Math Reasoning Abilities against Hard Perturbations
Kaixuan Huang, Jiacheng Guo, Zihao Li +15
cs.LGcs.AIcs.CLarXiv:2502.06453v22025Personalized Transformer for Explainable Recommendation
Lei Li, Yongfeng Zhang, Li Chen
cs.IRcs.AIcs.CLarXiv:2105.11601v22021ResearchClawBench: A Benchmark for End-to-End Autonomous Scientific Research
Wanghan Xu, Shuo Li, Tianlin Ye +48
cs.LGcs.AIcs.CLarXiv:2606.07591v52026Towards End-to-End Automation of AI Research
Yutaro Yamada, Robert Tjarko Lange, Cong Lu +5
cs.AIarXiv:2606.15497v12026ViSAR: Training-Free Adaptive-$k$ Retrieval for Visual Document Question Answering
Adrien Mialland, Marc Plantevit, Julien Gallois +1
cs.IRcs.AIcs.CLarXiv:2609.02486v12026Narrative-Driven Paper-to-Slide Generation via ArcDeck
Tarik Can Ozden, Sachidanand VS, Furkan Horoz +3
cs.AIarXiv:2604.11969v12026Llama-Nemotron: Efficient Reasoning Models
Akhiad Bercovich, Itay Levy, Izik Golan +133
cs.CLcs.AIcs.LGarXiv:2505.00949v52025Beware of Metacognitive Laziness: Effects of Generative Artificial Intelligence on Learning Motivation, Processes, and Performance
Yizhou Fan, Luzhen Tang, Huixiao Le +6
cs.AIcs.HCarXiv:2412.09315v12024Principles and Guidelines for Evaluating Social Robot Navigation Algorithms
Anthony Francis, Claudia Pérez-D'Arpino, Chengshu Li +28
cs.ROcs.AIcs.HCarXiv:2306.16740v42023NeuroLKH: Combining Deep Learning Model with Lin-Kernighan-Helsgaun Heuristic for Solving the Traveling Salesman Problem
Liang Xin, Wen Song, Zhiguang Cao +1
cs.AIcs.LGarXiv:2110.07983v12021Diffuse and Disperse: Image Generation with Representation Regularization
Runqian Wang, Kaiming He
cs.CVcs.AIcs.LGarXiv:2506.09027v22025Reconciling Reality through Simulation: A Real-to-Sim-to-Real Approach for Robust Manipulation
Marcel Torne, Anthony Simeonov, Zechu Li +4
cs.ROcs.AIcs.LGarXiv:2403.03949v32024End-to-End Training of Multi-Document Reader and Retriever for Open-Domain Question Answering
Devendra Singh Sachan, Siva Reddy, William Hamilton +2
cs.CLcs.AIcs.IRarXiv:2106.05346v22021When Chain of Thought is Necessary, Language Models Struggle to Evade Monitors
Scott Emmons, Erik Jenner, David K. Elson +5
cs.AIcs.CLarXiv:2507.05246v12025Capability-Gated Language Models: Security Composes, Utility Does Not
Patrikas Vanagas, Augustas Mačijauskas, Laurynas Lopata
cs.CRcs.AIcs.LGarXiv:2609.00445v12026Rethinking Rubric Generation for Improving LLM Judge and Reward Modeling for Open-ended Tasks
William F. Shen, Xinchi Qiu, Chenxi Whitehouse +6
cs.LGcs.AIarXiv:2602.05125v12026An Agentic System for Rare Disease Diagnosis with Traceable Reasoning
Weike Zhao, Chaoyi Wu, Yanjie Fan +10
cs.CLcs.AIcs.CVarXiv:2506.20430v32025