Artificial Intelligence
Papers filed under cs.AI on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.
Search paper metadata (including unsummarized papers)
5,461 to 5,520 of 15,365
Conditional Hypothesis Generation for LLM-Based Text Analysis with Researcher-Specified Covariates
Paiheng Xu, Jing Liu, Wei Ai
cs.CLcs.AIarXiv:2606.03029v12026CroCo: Cross-Lingual Contrastive Preference Tuning on Self-Generations
Mike Zhang, Ali Basirat, Desmond Elliott
cs.CLcs.AIarXiv:2605.26293v12026MobileGym: A Verifiable and Highly Parallel Simulation Platform for Mobile GUI Agent Research
Dingbang Wu, Rui Hao, Haiyang Wang +8
cs.AIcs.CLarXiv:2605.26114v22026Foundation Protocol: A Coordination Layer for Agentic Society
Bang Liu, Yongfeng Gu, Jiayi Zhang +26
cs.AIarXiv:2605.23218v12026CHI-Bench: Can AI Agents Automate End-to-End, Long-Horizon, Policy-Rich Healthcare Workflows?
Haolin Chen, Deon Metelski, Leon Qi +30
cs.CLcs.AIarXiv:2605.16679v22026TriPlay-RL: Tri-Role Self-Play Reinforcement Learning for LLM Safety Alignment
Zhewen Tan, Wenhan Yu, Jianfeng Si +9
cs.LGcs.AIarXiv:2601.18292v22026KV Packet: Recomputation-Free Context-Independent KV Caching for LLMs
Chuangtao Chen, Grace Li Zhang, Xunzhao Yin +3
cs.LGcs.AIarXiv:2604.13226v22026LightMem: Lightweight and Efficient Memory-Augmented Generation
Jizhan Fang, Xinle Deng, Haoming Xu +9
cs.CLcs.AIcs.CVarXiv:2510.18866v42025Type-Checked Compliance: Deterministic Guardrails for Agentic Financial Systems Using Lean 4 Theorem Proving
Devakh Rashie, Veda Rashi
cs.LOcs.AIcs.CRarXiv:2604.01483v12026REVERE: Reflective Evolving Research Engineer
Balaji Dinesh Gangireddi, Aniketh Garikaparthi, Manasi Patwardhan +1
cs.SEcs.AIarXiv:2603.20667v22026Reaching Beyond the Mode: RL for Distributional Reasoning in Language Models
Isha Puri, Mehul Damani, Idan Shenfeld +3
cs.LGcs.AIcs.CLarXiv:2603.24844v12026Memento-Skills: Let Agents Design Agents
Huichi Zhou, Siyuan Guo, Anjie Liu +14
cs.AIcs.CLcs.LGarXiv:2603.18743v12026Vision-Language-Action Models for Robotics: A Review Towards Real-World Applications
Kento Kawaharazuka, Jihoon Oh, Jun Yamada +2
cs.ROcs.AIcs.CVarXiv:2510.07077v12025Designing ECG Monitoring Healthcare System with Federated Transfer Learning and Explainable AI
Ali Raza, Kim Phuc Tran, Ludovic Koehl +1
cs.LGcs.AIeess.SParXiv:2105.12497v22021Unified Vision-Language Modeling via Concept Space Alignment
Yifu Qiu, Paul-Ambroise Duquenne, Holger Schwenk
cs.CVcs.AIcs.CLarXiv:2603.01096v12026A Mixed Diet Makes DINO An Omnivorous Vision Encoder
Rishabh Kabra, Maks Ovsjanikov, Drew A. Hudson +5
cs.CVcs.AIarXiv:2602.24181v22026DREAM: Deep Research Evaluation with Agentic Metrics
Elad Ben Avraham, Changhao Li, Ron Dorfman +8
cs.AIarXiv:2602.18940v12026Implicit Intelligence -- Evaluating Agents on What Users Don't Say
Ved Sirdeshmukh, Marc Wetter
cs.AIarXiv:2602.20424v12026References Improve LLM Alignment in Non-Verifiable Domains
Kejian Shi, Yixin Liu, Peifeng Wang +3
cs.CLcs.AIcs.LGarXiv:2602.16802v12026scPilot: Large Language Model Reasoning Toward Automated Single-Cell Analysis and Discovery
Yiming Gao, Zhen Wang, Jefferson Chen +8
cs.AIq-bio.GNarXiv:2602.11609v12026s1: Simple test-time scaling
Niklas Muennighoff, Zitong Yang, Weijia Shi +7
cs.CLcs.AIcs.LGarXiv:2501.19393v32025AudioSAE: Towards Understanding of Audio-Processing Models with Sparse AutoEncoders
Georgii Aparin, Tasnima Sadekova, Alexey Rukhovich +5
cs.SDcs.AIarXiv:2602.05027v22026Industrial Internet of Things Intelligence Empowering Smart Manufacturing: A Literature Review
Yujiao Hu, Qingmin Jia, Yuao Yao +6
cs.AIcs.CYarXiv:2312.16174v22023TTCS: Test-Time Curriculum Synthesis for Self-Evolving
Chengyi Yang, Zhishang Xiang, Yunbo Tang +5
cs.LGcs.AIcs.CLarXiv:2601.22628v12026REASONING GYM: Reasoning Environments for Reinforcement Learning with Verifiable Rewards
Zafir Stojanovski, Oliver Stanley, Joe Sharratt +4
cs.LGcs.AIcs.CLarXiv:2505.24760v22025A Mechanistic View on Video Generation as World Models: State and Dynamics
Luozhou Wang, Zhifei Chen, Yihua Du +11
cs.CVcs.AIarXiv:2601.17067v12026Multiplex Thinking: Reasoning via Token-wise Branch-and-Merge
Yao Tang, Li Dong, Yaru Hao +3
cs.CLcs.AIcs.LGarXiv:2601.08808v12026Reinforcement Learning with Action Chunking
Qiyang Li, Zhiyuan Zhou, Sergey Levine
cs.LGcs.AIcs.ROarXiv:2507.07969v42025VIBE: Visual Instruction Based Editor
Grigorii Alekseenko, Aleksandr Gordeev, Irina Tolstykh +7
cs.CVcs.AIcs.LGarXiv:2601.02242v12026GAIA-2: A Controllable Multi-View Generative World Model for Autonomous Driving
Lloyd Russell, Anthony Hu, Lorenzo Bertoni +4
cs.CVcs.AIcs.ROarXiv:2503.20523v12025OmniRetarget: Interaction-Preserving Data Generation for Humanoid Whole-Body Loco-Manipulation and Scene Interaction
Lujie Yang, Xiaoyu Huang, Zhen Wu +6
cs.ROcs.AIcs.LGarXiv:2509.26633v32025MultiAgentBench: Evaluating the Collaboration and Competition of LLM agents
Kunlun Zhu, Hongyi Du, Zhaochen Hong +8
cs.MAcs.AIcs.CLarXiv:2503.01935v12025Distributed Implicit Harm: A Compositional Safety Blind Spot in MLLM-Based Video Moderation
Ruotong Wang, Zihao Zhu, Siwei Lyu +2
cs.CVcs.AIarXiv:2609.00206v12026Learning a visuomotor controller for real world robotic grasping using simulated depth images
Ulrich Viereck, Andreas ten Pas, Kate Saenko +1
cs.ROcs.AIarXiv:1706.04652v32017PaliGemma: A versatile 3B VLM for transfer
Lucas Beyer, Andreas Steiner, André Susano Pinto +32
cs.CVcs.AIcs.CLarXiv:2407.07726v22024ReFT: Representation Finetuning for Language Models
Zhengxuan Wu, Aryaman Arora, Zheng Wang +4
cs.CLcs.AIcs.LGarXiv:2404.03592v32024V-STaR: Training Verifiers for Self-Taught Reasoners
Arian Hosseini, Xingdi Yuan, Nikolay Malkin +3
cs.LGcs.AIcs.CLarXiv:2402.06457v22024Advancements in Generative AI: A Comprehensive Review of GANs, GPT, Autoencoders, Diffusion Model, and Transformers
Staphord Bengesi, Hoda El-Sayed, Md Kamruzzaman Sarker +3
cs.LGcs.AIarXiv:2311.10242v22023Convolutional Recurrent Neural Networks for Small-Footprint Keyword Spotting
Sercan O. Arik, Markus Kliegl, Rewon Child +5
cs.CLcs.AIcs.LGarXiv:1703.05390v32017Large Language Models as Optimizers
Chengrun Yang, Xuezhi Wang, Yifeng Lu +4
cs.LGcs.AIcs.CLarXiv:2309.03409v32023LIMR: Less is More for RL Scaling
Xuefeng Li, Haoyang Zou, Pengfei Liu
cs.LGcs.AIcs.CLarXiv:2502.11886v12025A Comprehensive Survey of Deep Transfer Learning for Anomaly Detection in Industrial Time Series: Methods, Applications, and Directions
Peng Yan, Ahmed Abdulkadir, Paul-Philipp Luley +4
cs.LGcs.AIarXiv:2307.05638v22023Language is All a Graph Needs
Ruosong Ye, Caiqi Zhang, Runhui Wang +2
cs.CLcs.AIcs.IRarXiv:2308.07134v52023Is Self-Repair a Silver Bullet for Code Generation?
Theo X. Olausson, Jeevana Priya Inala, Chenglong Wang +2
cs.CLcs.AIcs.PLarXiv:2306.09896v52023Can Language Models Solve Graph Problems in Natural Language?
Heng Wang, Shangbin Feng, Tianxing He +3
cs.CLcs.AIarXiv:2305.10037v32023Open-vocabulary Queryable Scene Representations for Real World Planning
Boyuan Chen, Fei Xia, Brian Ichter +5
cs.ROcs.AIcs.CVarXiv:2209.09874v22022Data Augmentation techniques in time series domain: A survey and taxonomy
Guillermo Iglesias, Edgar Talavera, Ángel González-Prieto +2
cs.LGcs.AIarXiv:2206.13508v42022Towards Unified Conversational Recommender Systems via Knowledge-Enhanced Prompt Learning
Xiaolei Wang, Kun Zhou, Ji-Rong Wen +1
cs.CLcs.AIcs.IRarXiv:2206.09363v12022Deep ROC Analysis and AUC as Balanced Average Accuracy to Improve Model Selection, Understanding and Interpretation
André M. Carrington, Douglas G. Manuel, Paul W. Fieguth +9
stat.MEcs.AIcs.LGarXiv:2103.11357v12021Distributional Soft Actor-Critic: Off-Policy Reinforcement Learning for Addressing Value Estimation Errors
Jingliang Duan, Yang Guan, Shengbo Eben Li +2
cs.LGcs.AIeess.SYarXiv:2001.02811v32020Attributed Graph Clustering via Adaptive Graph Convolution
Xiaotong Zhang, Han Liu, Qimai Li +1
cs.LGcs.AIstat.MLarXiv:1906.01210v12019Adversarial Attack and Defense on Graph Data: A Survey
Lichao Sun, Yingtong Dou, Carl Yang +5
cs.CRcs.AIcs.SIarXiv:1812.10528v42018Beyond Binary Rewards: Training LMs to Reason About Their Uncertainty
Mehul Damani, Isha Puri, Stewart Slocum +4
cs.LGcs.AIcs.CLarXiv:2507.16806v22025Flow: A Modular Learning Framework for Mixed Autonomy Traffic
Cathy Wu, Aboudy Kreidieh, Kanaad Parvate +2
cs.AIcs.ROeess.SYarXiv:1710.05465v42017MedSAM2: Segment Anything in 3D Medical Images and Videos
Jun Ma, Zongxin Yang, Sumin Kim +6
eess.IVcs.AIcs.CVarXiv:2504.03600v12025Echo-4o: Harnessing the Power of GPT-4o Synthetic Images for Improved Image Generation
Junyan Ye, Dongzhi Jiang, Zihao Wang +9
cs.CVcs.AIcs.CLarXiv:2508.09987v12025RoboCasa365: A Large-Scale Simulation Framework for Training and Benchmarking Generalist Robots
Soroush Nasiriany, Sepehr Nasiriany, Abhiram Maddukuri +1
cs.ROcs.AIcs.LGarXiv:2603.04356v12026OVO-Bench: How Far is Your Video-LLMs from Real-World Online Video Understanding?
Yifei Li, Junbo Niu, Ziyang Miao +12
cs.CVcs.AIarXiv:2501.05510v22025Subliminal Learning: Language models transmit behavioral traits via hidden signals in data
Alex Cloud, Minh Le, James Chua +5
cs.LGcs.AIarXiv:2507.14805v12025Do generative video models understand physical principles?
Saman Motamed, Laura Culp, Kevin Swersky +2
cs.CVcs.AIcs.GRarXiv:2501.09038v32025