Artificial Intelligence
Papers filed under cs.AI on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.
Search paper metadata (including unsummarized papers)
6,721 to 6,780 of 15,326
Under the Hood: Using Diagnostic Classifiers to Investigate and Improve how Language Models Track Agreement Information
Mario Giulianelli, Jacqueline Harding, Florian Mohnert +2
cs.CLcs.AIarXiv:1808.08079v32018Give Us the Facts: Enhancing Large Language Models with Knowledge Graphs for Fact-aware Language Modeling
Linyao Yang, Hongyang Chen, Zhao Li +2
cs.CLcs.AIarXiv:2306.11489v22023Automated Speed and Lane Change Decision Making using Deep Reinforcement Learning
Carl-Johan Hoel, Krister Wolff, Leo Laine
cs.ROcs.AIcs.LGarXiv:1803.10056v22018On the Interaction Between Model Compression and Test-Time Adaptation
Francesco Corti, Dong Wang, Young D. Kwon +2
cs.LGcs.AIarXiv:2609.03604v12026Beyond "Made with AI": Visualizing Provenance Density to Mitigate the Transparency Penalty
Qing Zhang, Yifei Huang, Juyoung Lee +2
cs.AIarXiv:2609.03460v12026Predictive Coding: a Theoretical and Experimental Review
Beren Millidge, Anil Seth, Christopher L Buckley
cs.AIcs.NEq-bio.NCarXiv:2107.12979v42021Towards Human-Level Bimanual Dexterous Manipulation with Reinforcement Learning
Yuanpei Chen, Tianhao Wu, Shengjie Wang +8
cs.ROcs.AIcs.LGarXiv:2206.08686v22022User Representation via Cross Multi-source Behavior Pre-training for Mobile Games
Chengqi Yang, Yiran Qiao, Feng Liu +7
cs.AIarXiv:2609.01057v12026GazeRefine: Expert Gaze as a Test-Time Prompt for Training-Free Medical Image Segmentation
Mohammed Oussama Benyahia, Marouane Tliba, Mohamed Amine Kerkouri +10
eess.IVcs.AIcs.CVarXiv:2609.01310v12026Enabling Conversational Interaction with Mobile UI using Large Language Models
Bryan Wang, Gang Li, Yang Li
cs.HCcs.AIarXiv:2209.08655v22022Grounded Semantic Composition for Visual Scenes
P. Gorniak, D. Roy
cs.AIarXiv:1107.0031v12011Safe Exploration in Finite Markov Decision Processes with Gaussian Processes
Matteo Turchetta, Felix Berkenkamp, Andreas Krause
cs.LGcs.AIcs.ROarXiv:1606.04753v22016TUTA: Tree-based Transformers for Generally Structured Table Pre-training
Zhiruo Wang, Haoyu Dong, Ran Jia +4
cs.IRcs.AIcs.DBarXiv:2010.12537v42020Kimi k1.5: Scaling Reinforcement Learning with LLMs
Kimi Team, Angang Du, Bofei Gao +93
cs.AIcs.LGarXiv:2501.12599v42025SAM 3: Segment Anything with Concepts
Nicolas Carion, Laura Gustafson, Yuan-Ting Hu +35
cs.CVcs.AIarXiv:2511.16719v22025Summarizing Opinions: Aspect Extraction Meets Sentiment Prediction and They Are Both Weakly Supervised
Stefanos Angelidis, Mirella Lapata
cs.CLcs.AIcs.LGarXiv:1808.08858v12018Universal Self-Consistency for Large Language Model Generation
Xinyun Chen, Renat Aksitov, Uri Alon +7
cs.CLcs.AIarXiv:2311.17311v12023Self Forcing: Bridging the Train-Test Gap in Autoregressive Video Diffusion
Xun Huang, Zhengqi Li, Guande He +2
cs.CVcs.AIcs.LGarXiv:2506.08009v22025Janus-Pro: Unified Multimodal Understanding and Generation with Data and Model Scaling
Xiaokang Chen, Zhiyu Wu, Xingchao Liu +5
cs.AIcs.CLcs.CVarXiv:2501.17811v12025Fine-Tuning Vision-Language-Action Models: Optimizing Speed and Success
Moo Jin Kim, Chelsea Finn, Percy Liang
cs.ROcs.AIcs.CVarXiv:2502.19645v22025Neighborhood Contrastive Learning for Novel Class Discovery
Zhun Zhong, Enrico Fini, Subhankar Roy +3
cs.CVcs.AIcs.LGarXiv:2106.10731v12021Does Reinforcement Learning Really Incentivize Reasoning Capacity in LLMs Beyond the Base Model?
Yang Yue, Zhiqi Chen, Rui Lu +5
cs.AIcs.CLcs.CVarXiv:2504.13837v52025SigLIP 2: Multilingual Vision-Language Encoders with Improved Semantic Understanding, Localization, and Dense Features
Michael Tschannen, Alexey Gritsenko, Xiao Wang +11
cs.CVcs.AIarXiv:2502.14786v12025GR00T N1: An Open Foundation Model for Generalist Humanoid Robots
NVIDIA, :, Johan Bjorck +40
cs.ROcs.AIcs.LGarXiv:2503.14734v22025Improving the Accuracy and Efficiency of MAP Inference for Markov Logic
Sebastian Riedel
cs.AIarXiv:1206.3282v12012Benchmarking Robustness of 3D Object Detection to Common Corruptions in Autonomous Driving
Yinpeng Dong, Caixin Kang, Jinlai Zhang +6
cs.CVcs.AIcs.CRarXiv:2303.11040v12023Understanding R1-Zero-Like Training: A Critical Perspective
Zichen Liu, Changyu Chen, Wenjun Li +5
cs.LGcs.AIcs.CLarXiv:2503.20783v22025Solving Imperfect-Information Games via Discounted Regret Minimization
Noam Brown, Tuomas Sandholm
cs.GTcs.AIarXiv:1809.04040v32018Parsing the Stream: A Live Trace Model for Long-Horizon Agents and Their Observers
Egor Pakhomov, Erik Nijkamp
cs.AIarXiv:2609.01466v12026RL-GAN-Net: A Reinforcement Learning Agent Controlled GAN Network for Real-Time Point Cloud Shape Completion
Muhammad Sarmad, Hyunjoo Jenny Lee, Young Min Kim
cs.CVcs.AIarXiv:1904.12304v12019OpenMathInstruct-2: Accelerating AI for Math with Massive Open-Source Instruction Data
Shubham Toshniwal, Wei Du, Ivan Moshkov +3
cs.CLcs.AIcs.LGarXiv:2410.01560v22024FurnitureBench: Reproducible Real-World Benchmark for Long-Horizon Complex Manipulation
Minho Heo, Youngwoon Lee, Doohyun Lee +1
cs.ROcs.AIcs.LGarXiv:2305.12821v12023The Rise of Verbal Reinforcement Learning
Kshitij Tayal, Arun Sharma, Genta Indra Winata +2
cs.CLcs.AIarXiv:2609.01597v12026One Prompt Is Enough: Watermark Laundering Through Foundation Image Models
Jidong Yang, Qi Li, Wei Zong +5
cs.CVcs.AIcs.CRarXiv:2609.01249v12026Illuminating Generalization in Deep Reinforcement Learning through Procedural Level Generation
Niels Justesen, Ruben Rodriguez Torrado, Philip Bontrager +3
cs.LGcs.AIstat.MLarXiv:1806.10729v52018A Variable Neighborhood Search for Flying Sidekick Traveling Salesman Problem
Julia C. Freitas, Puca Huachi V. Penna
math.OCcs.AIarXiv:1804.03954v22018TimeSteer: Inference-Time Speech Scheduling in Joint Audio-Visual Diffusion Models
Chao Zhou, Yiling Chen, Qi Chu +3
cs.CVcs.AIcs.MMarXiv:2609.01277v12026DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning
DeepSeek-AI, Daya Guo, Dejian Yang +197
cs.CLcs.AIcs.LGarXiv:2501.12948v22025Search-R1: Training LLMs to Reason and Leverage Search Engines with Reinforcement Learning
Bowen Jin, Hansi Zeng, Zhenrui Yue +5
cs.CLcs.AIcs.IRarXiv:2503.09516v52025Gemma 3 Technical Report
Gemma Team, Aishwarya Kamath, Johan Ferret +213
cs.CLcs.AIarXiv:2503.19786v12025Qwen3-VL Technical Report
Shuai Bai, Yuxuan Cai, Ruizhe Chen +61
cs.CVcs.AIarXiv:2511.21631v22025LoRA Fine-tuning Efficiently Undoes Safety Training in Llama 2-Chat 70B
Simon Lermen, Charlie Rogers-Smith, Jeffrey Ladish
cs.LGcs.AIarXiv:2310.20624v22023Knapsack based Optimal Policies for Budget-Limited Multi-Armed Bandits
Long Tran-Thanh, Archie Chapman, Alex Rogers +1
cs.AIcs.LGarXiv:1204.1909v12012OpenAI GPT-5 System Card
Aaditya Singh, Adam Fry, Adam Perelman +483
cs.CLcs.AIarXiv:2601.03267v22025DeepSeek-V4: Towards Highly Efficient Million-Token Context Intelligence
DeepSeek-AI, Anyi Xu, Bangcai Lin +316
cs.CLcs.AIarXiv:2606.19348v12026gpt-oss-120b & gpt-oss-20b Model Card
OpenAI, :, Sandhini Agarwal +124
cs.CLcs.AIarXiv:2508.10925v12025QA Dataset Explosion: A Taxonomy of NLP Resources for Question Answering and Reading Comprehension
Anna Rogers, Matt Gardner, Isabelle Augenstein
cs.CLcs.AIarXiv:2107.12708v22021Efficient SWE Agent Benchmarking via Trajectory-Aware Evaluation
Kefeng Duan, Dewu Zheng, Yanlin Wang +7
cs.SEcs.AIcs.CLarXiv:2609.01603v12026Time Travel in LLMs: Tracing Data Contamination in Large Language Models
Shahriar Golchin, Mihai Surdeanu
cs.CLcs.AIcs.CRarXiv:2308.08493v32023Adaptive Critical Token-Aware Retrieval for Repository-Level Code Generation
Kefeng Duan, Dewu Zheng, Yanlin Wang +8
cs.SEcs.AIcs.CLarXiv:2609.01601v12026YOLOv12: Attention-Centric Real-Time Object Detectors
Yunjie Tian, Qixiang Ye, David Doermann
cs.CVcs.AIarXiv:2502.12524v12025MIDR: Enrichment-Augmented Indexing for Multimodal Document Retrieval
Debanjan Mahata, Atharva Tendle, Daniel Preotiuc-Pietro +2
cs.IRcs.AIcs.CLarXiv:2609.01316v12026Rubrics as Rewards: Reinforcement Learning Beyond Verifiable Domains
Anisha Gunjal, Anthony Wang, Elaine Lau +4
cs.LGcs.AIcs.CLarXiv:2507.17746v22025Scaling Monosemanticity: Extracting Interpretable Features from Claude 3 Sonnet
Adly Templeton, Tom Conerly, Jonathan Marcus +23
cs.AIarXiv:2605.29358v12026Designing Fair AI for Managing Employees in Organizations: A Review, Critique, and Design Agenda
Lionel P. Robert, Casey Pierce, Liz Morris +2
cs.HCcs.AIcs.CYarXiv:2002.09054v12020NUWA-XL: Diffusion over Diffusion for eXtremely Long Video Generation
Shengming Yin, Chenfei Wu, Huan Yang +13
cs.CVcs.AIarXiv:2303.12346v12023StoryBuddy: A Human-AI Collaborative Chatbot for Parent-Child Interactive Storytelling with Flexible Parental Involvement
Zheng Zhang, Ying Xu, Yanhao Wang +6
cs.HCcs.AIcs.CLarXiv:2202.06205v22022StudyBench: Can Self-Evolution Squeeze Textbooks for Olympiad Capability?
Yinghao Chen, Zixi Chen, Bingxiang He +7
cs.AIarXiv:2609.00787v12026Evaluating Explainability for Graph Neural Networks
Chirag Agarwal, Owen Queen, Himabindu Lakkaraju +1
cs.LGcs.AIarXiv:2208.09339v22022HarmoFL: Harmonizing Local and Global Drifts in Federated Learning on Heterogeneous Medical Images
Meirui Jiang, Zirui Wang, Qi Dou
eess.IVcs.AIcs.CVarXiv:2112.10775v32021