Artificial Intelligence

Papers filed under cs.AI on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.

Search paper metadata (including unsummarized papers)

5,221 to 5,280 of 15,440

  1. MiMo: Unlocking the Reasoning Potential of Language Model -- From Pretraining to Posttraining

    LLM-Core Xiaomi, :, Bingquan Xia +62

    cs.CLcs.AIcs.LGarXiv:2505.07608v22025
  2. Combining Deep Reinforcement Learning and Search for Imperfect-Information Games

    Noam Brown, Anton Bakhtin, Adam Lerer +1

    cs.GTcs.AIcs.LGarXiv:2007.13544v22020
  3. TransZero: Attribute-guided Transformer for Zero-Shot Learning

    Shiming Chen, Ziming Hong, Yang Liu +6

    cs.CVcs.AIarXiv:2112.01683v12021
  4. PPTAgent: Generating and Evaluating Presentations Beyond Text-to-Slides

    Hao Zheng, Xinyan Guan, Hao Kong +7

    cs.AIcs.CLarXiv:2501.03936v32025
  5. Zero-Shot Respiratory Sound Classification through LLM-Augmented Audio-Text Alignment

    Mustafa Talha İlerisoy, Hung Manh Pham, Mathias Funk +2

    cs.CLcs.AIcs.SDarXiv:2609.00055v12026
  6. Diffusion Adversarial Post-Training for One-Step Video Generation

    Shanchuan Lin, Xin Xia, Yuxi Ren +3

    cs.CVcs.AIcs.LGarXiv:2501.08316v32025
  7. GUI-Actor: Coordinate-Free Visual Grounding for GUI Agents

    Qianhui Wu, Kanzhi Cheng, Rui Yang +15

    cs.CLcs.AIcs.CVarXiv:2506.03143v12025
  8. EgoVLA: Learning Vision-Language-Action Models from Egocentric Human Videos

    Ruihan Yang, Qinxi Yu, Yecheng Wu +12

    cs.ROcs.AIcs.CVarXiv:2507.12440v32025
  9. A Survey of AI Agent Protocols

    Yingxuan Yang, Huacan Chai, Yuanyi Song +11

    cs.AIarXiv:2504.16736v32025
  10. TAMI: Temporally Aligned, Missingness-Aware, and Interpretable Multimodal Fusion for Mental Health Assessment in Older Adults with Mild Cognitive Impairment

    Merna Bibars, Bolaji Omofojoye, Allan I. Levey +3

    cs.CVcs.AIarXiv:2608.30857v12026
  11. Process Reward Models That Think

    Muhammad Khalifa, Rishabh Agarwal, Lajanugen Logeswaran +5

    cs.LGcs.AIcs.CLarXiv:2504.16828v52025
  12. MemoryGraft: Persistent Compromise of LLM Agents via Poisoned Experience Retrieval

    Saksham Sahai Srivastava, Haoyu He

    cs.CRcs.AIcs.LGarXiv:2512.16962v12025
  13. Transformer Language Models without Positional Encodings Still Learn Positional Information

    Adi Haviv, Ori Ram, Ofir Press +2

    cs.CLcs.AIcs.LGarXiv:2203.16634v22022
  14. Aether: Geometric-Aware Unified World Modeling

    Aether Team, Haoyi Zhu, Yifan Wang +8

    cs.CVcs.AIcs.LGarXiv:2503.18945v32025
  15. Addressing Complex and Subjective Product-Related Queries with Customer Reviews

    Julian McAuley, Alex Yang

    cs.IRcs.AIcs.SIarXiv:1512.06863v12015
  16. Reward-Guided Speculative Decoding for Efficient LLM Reasoning

    Baohao Liao, Yuhui Xu, Hanze Dong +5

    cs.CLcs.AIarXiv:2501.19324v32025
  17. Progent: Securing AI Agents with Privilege Control

    Tianneng Shi, Jingxuan He, Zhun Wang +4

    cs.CRcs.AIarXiv:2504.11703v32025
  18. DexGraspVLA: A Vision-Language-Action Framework Towards General Dexterous Grasping

    Yifan Zhong, Xuchuan Huang, Ruochong Li +9

    cs.ROcs.AIarXiv:2502.20900v52025
  19. GuardReasoner: Towards Reasoning-based LLM Safeguards

    Yue Liu, Hongcheng Gao, Shengfang Zhai +9

    cs.CRcs.AIcs.LGarXiv:2501.18492v22025
  20. The Road Less Scheduled

    Aaron Defazio, Xingyu Alice Yang, Harsh Mehta +3

    cs.LGcs.AImath.OCarXiv:2405.15682v42024
  21. OstQuant: Refining Large Language Model Quantization with Orthogonal and Scaling Transformations for Better Distribution Fitting

    Xing Hu, Yuan Cheng, Dawei Yang +6

    cs.LGcs.AIarXiv:2501.13987v12025
  22. Scaling up Masked Diffusion Models on Text

    Shen Nie, Fengqi Zhu, Chao Du +5

    cs.AIcs.CLcs.LGarXiv:2410.18514v32024
  23. DARTS-: Robustly Stepping out of Performance Collapse Without Indicators

    Xiangxiang Chu, Xiaoxing Wang, Bo Zhang +3

    cs.LGcs.AIcs.CVarXiv:2009.01027v22020
  24. Small Models Struggle to Learn from Strong Reasoners

    Yuetai Li, Xiang Yue, Zhangchen Xu +5

    cs.AIarXiv:2502.12143v32025
  25. Scaling Spatial Intelligence with Multimodal Foundation Models

    Zhongang Cai, Ruisi Wang, Chenyang Gu +26

    cs.CVcs.AIcs.LGarXiv:2511.13719v42025
  26. Understanding Reasoning in Thinking Language Models via Steering Vectors

    Constantin Venhoff, Iván Arcuschin, Philip Torr +2

    cs.LGcs.AIarXiv:2506.18167v42025
  27. VerlTool: Towards Holistic Agentic Reinforcement Learning with Tool Use

    Dongfu Jiang, Yi Lu, Zhuofeng Li +9

    cs.AIcs.CLcs.CVarXiv:2509.01055v32025
  28. Multi-Agent Collaboration via Evolving Orchestration

    Yufan Dang, Chen Qian, Xueheng Luo +11

    cs.CLcs.AIcs.MAarXiv:2505.19591v22025
  29. What Can We Learn from Collective Human Opinions on Natural Language Inference Data?

    Yixin Nie, Xiang Zhou, Mohit Bansal

    cs.CLcs.AIcs.LGarXiv:2010.03532v22020
  30. A Vision-Language-Action-Critic Model for Robotic Real-World Reinforcement Learning

    Shaopeng Zhai, Qi Zhang, Tianyi Zhang +7

    cs.ROcs.AIarXiv:2509.15937v12025
  31. VLASH: Real-Time VLAs via Future-State-Aware Asynchronous Inference

    Jiaming Tang, Yufei Sun, Yilong Zhao +7

    cs.ROcs.AIcs.LGarXiv:2512.01031v22025
  32. Measuring the environmental impact of delivering AI at Google Scale

    Cooper Elsworth, Keguo Huang, David Patterson +9

    cs.AIarXiv:2508.15734v12025
  33. Latent Diffusion Model without Variational Autoencoder

    Minglei Shi, Haolin Wang, Wenzhao Zheng +6

    cs.CVcs.AIarXiv:2510.15301v42025
  34. Global-Locally Self-Attentive Dialogue State Tracker

    Victor Zhong, Caiming Xiong, Richard Socher

    cs.CLcs.AIarXiv:1805.09655v32018
  35. YuE: Scaling Open Foundation Models for Long-Form Music Generation

    Ruibin Yuan, Hanfeng Lin, Shuyue Guo +55

    eess.AScs.AIcs.MMarXiv:2503.08638v22025
  36. A Survey of Graph Retrieval-Augmented Generation for Customized Large Language Models

    Qinggang Zhang, Shengyuan Chen, Yuanchen Bei +9

    cs.CLcs.AIcs.IRarXiv:2501.13958v32025
  37. Approximate evaluation of marginal association probabilities with belief propagation

    Jason L. Williams, Roslyn A. Lau

    cs.AIcs.CVarXiv:1209.6299v22012
  38. Training Neural Machine Translation To Apply Terminology Constraints

    Georgiana Dinu, Prashant Mathur, Marcello Federico +1

    cs.CLcs.AIcs.LGarXiv:1906.01105v22019
  39. JudgeLRM: Large Reasoning Models as a Judge

    Nuo Chen, Zhiyuan Hu, Qingyun Zou +4

    cs.CLcs.AIarXiv:2504.00050v32025
  40. Spatio-Temporal Wind Speed Forecasting using Graph Networks and Novel Transformer Architectures

    Lars Ødegaard Bentsen, Narada Dilp Warakagoda, Roy Stenbro +1

    cs.LGcs.AIarXiv:2208.13585v22022
  41. Quantum Computing based Hybrid Solution Strategies for Large-scale Discrete-Continuous Optimization Problems

    Akshay Ajagekar, Travis Humble, Fengqi You

    quant-phcs.AImath.OCarXiv:1910.13045v12019
  42. S-GRPO: Early Exit via Reinforcement Learning in Reasoning Models

    Muzhi Dai, Chenxu Yang, Qingyi Si

    cs.AIcs.LGarXiv:2505.07686v22025
  43. How many images do I need? Understanding how sample size per class affects deep learning model performance metrics for balanced designs in autonomous wildlife monitoring

    Saleh Shahinfar, Paul Meek, Greg Falzon

    cs.CVcs.AIcs.LGarXiv:2010.08186v12020
  44. Mechanism Design for Alignment and Control

    Dirk Bergemann, Andrew Koh, Stephen Morris

    econ.THcs.AIcs.GTarXiv:2609.01595v12026
  45. Gemini Embedding: Generalizable Embeddings from Gemini

    Jinhyuk Lee, Feiyang Chen, Sahil Dua +44

    cs.CLcs.AIarXiv:2503.07891v12025
    Summaries:한국어
  46. CaP-X: A Framework for Benchmarking and Improving Coding Agents for Robot Manipulation

    Letian Fu, Justin Yu, Karim El-Refai +13

    cs.ROcs.AIarXiv:2603.22435v22026
  47. RPCBench: A Benchmark for Proactive Premise Critique in LLM-based Recommendation

    Zhongru Chen, Yuan Wu, Yi Chang

    cs.AIcs.CLarXiv:2609.00918v12026
  48. Self-Supervised Hypergraph Transformer for Recommender Systems

    Lianghao Xia, Chao Huang, Chuxu Zhang

    cs.IRcs.AIarXiv:2207.14338v12022
  49. QILP-0: Constructing Observational Declarative Twins of Quantum Circuits

    Marina de la Cruz Echeandía, César Luis Alonso, Tony Ribeiro +1

    cs.AIquant-pharXiv:2609.01049v12026
  50. EdiTikZ: Scientific Figure Editing from Revision Trajectories

    Christian Greisinger, Zhixue Zhao, Steffen Eger

    cs.AIcs.CLcs.CVarXiv:2609.01409v12026
  51. ACON: Optimizing Context Compression for Long-horizon LLM Agents

    Minki Kang, Wei-Ning Chen, Dongge Han +5

    cs.AIcs.CLarXiv:2510.00615v32025
  52. Some Emotions Run Deeper: Layer-wise Probing and Causal Intervention in Large Language Models

    Tian Fang, Gaël Guibon, Davide Buscaldi

    cs.CLcs.AIarXiv:2609.01279v12026
  53. Mixture of Contexts for Long Video Generation

    Shengqu Cai, Ceyuan Yang, Lvmin Zhang +10

    cs.GRcs.AIcs.CVarXiv:2508.21058v32025
  54. Chain-of-Thought Reasoning In The Wild Is Not Always Faithful

    Iván Arcuschin, Jett Janiak, Robert Krzyzanowski +3

    cs.AIcs.CLcs.LGarXiv:2503.08679v62025
  55. All-atom Diffusion Transformers: Unified generative modelling of molecules and materials

    Chaitanya K. Joshi, Xiang Fu, Yi-Lun Liao +4

    cs.LGcs.AIarXiv:2503.03965v22025
  56. A Composable Evaluation System for Reproducible Omni-Modal Foundation Model Evaluation

    Hodong Lee, Sanghee Park, Dohoon Ryu +4

    cs.AIarXiv:2609.01315v12026
  57. FractalNet-Based Heterogeneous Federated Learning for Orbital Edge Intelligence in Satellite Mega-Constellations: A Wildfire Case Study

    Sai Puppala, Koushik Sinha

    cs.AIcs.DCcs.ETarXiv:2609.00875v12026
  58. A Survey on Embedding Dynamic Graphs

    Claudio D. T. Barros, Matheus R. F. Mendonça, Alex B. Vieira +1

    cs.LGcs.AIarXiv:2101.01229v22021
  59. Invalidation Contracts for Cross-Episode Agent Memory

    Michael Wu, Arquimedes Canedo

    cs.AIarXiv:2609.00243v12026
  60. MIRAGE: The Illusion of Visual Understanding

    Mohammad Asadi, Jack W. O'Sullivan, Fang Cao +5

    cs.AIarXiv:2603.21687v32026