Artificial Intelligence

Papers filed under cs.AI on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.

Search paper metadata (including unsummarized papers)

5,461 to 5,520 of 15,365

  1. Conditional Hypothesis Generation for LLM-Based Text Analysis with Researcher-Specified Covariates

    Paiheng Xu, Jing Liu, Wei Ai

    cs.CLcs.AIarXiv:2606.03029v12026
  2. CroCo: Cross-Lingual Contrastive Preference Tuning on Self-Generations

    Mike Zhang, Ali Basirat, Desmond Elliott

    cs.CLcs.AIarXiv:2605.26293v12026
  3. MobileGym: A Verifiable and Highly Parallel Simulation Platform for Mobile GUI Agent Research

    Dingbang Wu, Rui Hao, Haiyang Wang +8

    cs.AIcs.CLarXiv:2605.26114v22026
  4. Foundation Protocol: A Coordination Layer for Agentic Society

    Bang Liu, Yongfeng Gu, Jiayi Zhang +26

    cs.AIarXiv:2605.23218v12026
  5. CHI-Bench: Can AI Agents Automate End-to-End, Long-Horizon, Policy-Rich Healthcare Workflows?

    Haolin Chen, Deon Metelski, Leon Qi +30

    cs.CLcs.AIarXiv:2605.16679v22026
  6. TriPlay-RL: Tri-Role Self-Play Reinforcement Learning for LLM Safety Alignment

    Zhewen Tan, Wenhan Yu, Jianfeng Si +9

    cs.LGcs.AIarXiv:2601.18292v22026
  7. KV Packet: Recomputation-Free Context-Independent KV Caching for LLMs

    Chuangtao Chen, Grace Li Zhang, Xunzhao Yin +3

    cs.LGcs.AIarXiv:2604.13226v22026
  8. LightMem: Lightweight and Efficient Memory-Augmented Generation

    Jizhan Fang, Xinle Deng, Haoming Xu +9

    cs.CLcs.AIcs.CVarXiv:2510.18866v42025
  9. Type-Checked Compliance: Deterministic Guardrails for Agentic Financial Systems Using Lean 4 Theorem Proving

    Devakh Rashie, Veda Rashi

    cs.LOcs.AIcs.CRarXiv:2604.01483v12026
  10. REVERE: Reflective Evolving Research Engineer

    Balaji Dinesh Gangireddi, Aniketh Garikaparthi, Manasi Patwardhan +1

    cs.SEcs.AIarXiv:2603.20667v22026
  11. Reaching Beyond the Mode: RL for Distributional Reasoning in Language Models

    Isha Puri, Mehul Damani, Idan Shenfeld +3

    cs.LGcs.AIcs.CLarXiv:2603.24844v12026
  12. Memento-Skills: Let Agents Design Agents

    Huichi Zhou, Siyuan Guo, Anjie Liu +14

    cs.AIcs.CLcs.LGarXiv:2603.18743v12026
  13. Vision-Language-Action Models for Robotics: A Review Towards Real-World Applications

    Kento Kawaharazuka, Jihoon Oh, Jun Yamada +2

    cs.ROcs.AIcs.CVarXiv:2510.07077v12025
  14. Designing ECG Monitoring Healthcare System with Federated Transfer Learning and Explainable AI

    Ali Raza, Kim Phuc Tran, Ludovic Koehl +1

    cs.LGcs.AIeess.SParXiv:2105.12497v22021
  15. Unified Vision-Language Modeling via Concept Space Alignment

    Yifu Qiu, Paul-Ambroise Duquenne, Holger Schwenk

    cs.CVcs.AIcs.CLarXiv:2603.01096v12026
  16. A Mixed Diet Makes DINO An Omnivorous Vision Encoder

    Rishabh Kabra, Maks Ovsjanikov, Drew A. Hudson +5

    cs.CVcs.AIarXiv:2602.24181v22026
  17. DREAM: Deep Research Evaluation with Agentic Metrics

    Elad Ben Avraham, Changhao Li, Ron Dorfman +8

    cs.AIarXiv:2602.18940v12026
  18. Implicit Intelligence -- Evaluating Agents on What Users Don't Say

    Ved Sirdeshmukh, Marc Wetter

    cs.AIarXiv:2602.20424v12026
  19. References Improve LLM Alignment in Non-Verifiable Domains

    Kejian Shi, Yixin Liu, Peifeng Wang +3

    cs.CLcs.AIcs.LGarXiv:2602.16802v12026
  20. scPilot: Large Language Model Reasoning Toward Automated Single-Cell Analysis and Discovery

    Yiming Gao, Zhen Wang, Jefferson Chen +8

    cs.AIq-bio.GNarXiv:2602.11609v12026
  21. s1: Simple test-time scaling

    Niklas Muennighoff, Zitong Yang, Weijia Shi +7

    cs.CLcs.AIcs.LGarXiv:2501.19393v32025
  22. AudioSAE: Towards Understanding of Audio-Processing Models with Sparse AutoEncoders

    Georgii Aparin, Tasnima Sadekova, Alexey Rukhovich +5

    cs.SDcs.AIarXiv:2602.05027v22026
  23. Industrial Internet of Things Intelligence Empowering Smart Manufacturing: A Literature Review

    Yujiao Hu, Qingmin Jia, Yuao Yao +6

    cs.AIcs.CYarXiv:2312.16174v22023
  24. TTCS: Test-Time Curriculum Synthesis for Self-Evolving

    Chengyi Yang, Zhishang Xiang, Yunbo Tang +5

    cs.LGcs.AIcs.CLarXiv:2601.22628v12026
  25. REASONING GYM: Reasoning Environments for Reinforcement Learning with Verifiable Rewards

    Zafir Stojanovski, Oliver Stanley, Joe Sharratt +4

    cs.LGcs.AIcs.CLarXiv:2505.24760v22025
  26. A Mechanistic View on Video Generation as World Models: State and Dynamics

    Luozhou Wang, Zhifei Chen, Yihua Du +11

    cs.CVcs.AIarXiv:2601.17067v12026
  27. Multiplex Thinking: Reasoning via Token-wise Branch-and-Merge

    Yao Tang, Li Dong, Yaru Hao +3

    cs.CLcs.AIcs.LGarXiv:2601.08808v12026
  28. Reinforcement Learning with Action Chunking

    Qiyang Li, Zhiyuan Zhou, Sergey Levine

    cs.LGcs.AIcs.ROarXiv:2507.07969v42025
  29. VIBE: Visual Instruction Based Editor

    Grigorii Alekseenko, Aleksandr Gordeev, Irina Tolstykh +7

    cs.CVcs.AIcs.LGarXiv:2601.02242v12026
  30. GAIA-2: A Controllable Multi-View Generative World Model for Autonomous Driving

    Lloyd Russell, Anthony Hu, Lorenzo Bertoni +4

    cs.CVcs.AIcs.ROarXiv:2503.20523v12025
  31. OmniRetarget: Interaction-Preserving Data Generation for Humanoid Whole-Body Loco-Manipulation and Scene Interaction

    Lujie Yang, Xiaoyu Huang, Zhen Wu +6

    cs.ROcs.AIcs.LGarXiv:2509.26633v32025
  32. MultiAgentBench: Evaluating the Collaboration and Competition of LLM agents

    Kunlun Zhu, Hongyi Du, Zhaochen Hong +8

    cs.MAcs.AIcs.CLarXiv:2503.01935v12025
  33. Distributed Implicit Harm: A Compositional Safety Blind Spot in MLLM-Based Video Moderation

    Ruotong Wang, Zihao Zhu, Siwei Lyu +2

    cs.CVcs.AIarXiv:2609.00206v12026
  34. Learning a visuomotor controller for real world robotic grasping using simulated depth images

    Ulrich Viereck, Andreas ten Pas, Kate Saenko +1

    cs.ROcs.AIarXiv:1706.04652v32017
  35. PaliGemma: A versatile 3B VLM for transfer

    Lucas Beyer, Andreas Steiner, André Susano Pinto +32

    cs.CVcs.AIcs.CLarXiv:2407.07726v22024
  36. ReFT: Representation Finetuning for Language Models

    Zhengxuan Wu, Aryaman Arora, Zheng Wang +4

    cs.CLcs.AIcs.LGarXiv:2404.03592v32024
  37. V-STaR: Training Verifiers for Self-Taught Reasoners

    Arian Hosseini, Xingdi Yuan, Nikolay Malkin +3

    cs.LGcs.AIcs.CLarXiv:2402.06457v22024
  38. Advancements in Generative AI: A Comprehensive Review of GANs, GPT, Autoencoders, Diffusion Model, and Transformers

    Staphord Bengesi, Hoda El-Sayed, Md Kamruzzaman Sarker +3

    cs.LGcs.AIarXiv:2311.10242v22023
  39. Convolutional Recurrent Neural Networks for Small-Footprint Keyword Spotting

    Sercan O. Arik, Markus Kliegl, Rewon Child +5

    cs.CLcs.AIcs.LGarXiv:1703.05390v32017
  40. Large Language Models as Optimizers

    Chengrun Yang, Xuezhi Wang, Yifeng Lu +4

    cs.LGcs.AIcs.CLarXiv:2309.03409v32023
  41. LIMR: Less is More for RL Scaling

    Xuefeng Li, Haoyang Zou, Pengfei Liu

    cs.LGcs.AIcs.CLarXiv:2502.11886v12025
  42. A Comprehensive Survey of Deep Transfer Learning for Anomaly Detection in Industrial Time Series: Methods, Applications, and Directions

    Peng Yan, Ahmed Abdulkadir, Paul-Philipp Luley +4

    cs.LGcs.AIarXiv:2307.05638v22023
  43. Language is All a Graph Needs

    Ruosong Ye, Caiqi Zhang, Runhui Wang +2

    cs.CLcs.AIcs.IRarXiv:2308.07134v52023
  44. Is Self-Repair a Silver Bullet for Code Generation?

    Theo X. Olausson, Jeevana Priya Inala, Chenglong Wang +2

    cs.CLcs.AIcs.PLarXiv:2306.09896v52023
  45. Can Language Models Solve Graph Problems in Natural Language?

    Heng Wang, Shangbin Feng, Tianxing He +3

    cs.CLcs.AIarXiv:2305.10037v32023
  46. Open-vocabulary Queryable Scene Representations for Real World Planning

    Boyuan Chen, Fei Xia, Brian Ichter +5

    cs.ROcs.AIcs.CVarXiv:2209.09874v22022
  47. Data Augmentation techniques in time series domain: A survey and taxonomy

    Guillermo Iglesias, Edgar Talavera, Ángel González-Prieto +2

    cs.LGcs.AIarXiv:2206.13508v42022
  48. Towards Unified Conversational Recommender Systems via Knowledge-Enhanced Prompt Learning

    Xiaolei Wang, Kun Zhou, Ji-Rong Wen +1

    cs.CLcs.AIcs.IRarXiv:2206.09363v12022
  49. Deep ROC Analysis and AUC as Balanced Average Accuracy to Improve Model Selection, Understanding and Interpretation

    André M. Carrington, Douglas G. Manuel, Paul W. Fieguth +9

    stat.MEcs.AIcs.LGarXiv:2103.11357v12021
  50. Distributional Soft Actor-Critic: Off-Policy Reinforcement Learning for Addressing Value Estimation Errors

    Jingliang Duan, Yang Guan, Shengbo Eben Li +2

    cs.LGcs.AIeess.SYarXiv:2001.02811v32020
  51. Attributed Graph Clustering via Adaptive Graph Convolution

    Xiaotong Zhang, Han Liu, Qimai Li +1

    cs.LGcs.AIstat.MLarXiv:1906.01210v12019
  52. Adversarial Attack and Defense on Graph Data: A Survey

    Lichao Sun, Yingtong Dou, Carl Yang +5

    cs.CRcs.AIcs.SIarXiv:1812.10528v42018
  53. Beyond Binary Rewards: Training LMs to Reason About Their Uncertainty

    Mehul Damani, Isha Puri, Stewart Slocum +4

    cs.LGcs.AIcs.CLarXiv:2507.16806v22025
  54. Flow: A Modular Learning Framework for Mixed Autonomy Traffic

    Cathy Wu, Aboudy Kreidieh, Kanaad Parvate +2

    cs.AIcs.ROeess.SYarXiv:1710.05465v42017
  55. MedSAM2: Segment Anything in 3D Medical Images and Videos

    Jun Ma, Zongxin Yang, Sumin Kim +6

    eess.IVcs.AIcs.CVarXiv:2504.03600v12025
  56. Echo-4o: Harnessing the Power of GPT-4o Synthetic Images for Improved Image Generation

    Junyan Ye, Dongzhi Jiang, Zihao Wang +9

    cs.CVcs.AIcs.CLarXiv:2508.09987v12025
  57. RoboCasa365: A Large-Scale Simulation Framework for Training and Benchmarking Generalist Robots

    Soroush Nasiriany, Sepehr Nasiriany, Abhiram Maddukuri +1

    cs.ROcs.AIcs.LGarXiv:2603.04356v12026
  58. OVO-Bench: How Far is Your Video-LLMs from Real-World Online Video Understanding?

    Yifei Li, Junbo Niu, Ziyang Miao +12

    cs.CVcs.AIarXiv:2501.05510v22025
  59. Subliminal Learning: Language models transmit behavioral traits via hidden signals in data

    Alex Cloud, Minh Le, James Chua +5

    cs.LGcs.AIarXiv:2507.14805v12025
  60. Do generative video models understand physical principles?

    Saman Motamed, Laura Culp, Kevin Swersky +2

    cs.CVcs.AIcs.GRarXiv:2501.09038v32025