Artificial Intelligence

Papers filed under cs.AI on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.

Search paper metadata (including unsummarized papers)

5,821 to 5,880 of 15,320

  1. ReViV: Reconstructing the Viewer and the View in 4D from Monocular Egocentric Video

    Xiaozhong Lyu, Gen Li, Zhiyin Qian +3

    cs.CVcs.AIarXiv:2607.17790v12026
  2. Harness Handbook: Making Evolving Agent Harnesses Readable,Navigable, and Editable

    Ruhan Wang, Yucheng Shi, Zongxia Li +7

    cs.AIcs.SEarXiv:2607.13285v12026
  3. ShortOPD: Recovering Pruned LLMs with Short-to-Long On-Policy Distillation

    Qingyu Zhang, Qianhao Yuan, Hongyu Lin +5

    cs.LGcs.AIcs.CLarXiv:2607.13124v22026
    Summaries:한국어
  4. Function-Aware Fill-in-the-Middle as Mid-Training for Coding Agent Foundation Models

    Yubo Wang, Jiarong Liang, Yuxuan Zhang +5

    cs.AIcs.CLarXiv:2607.12463v32026
  5. The AI Index 2021 Annual Report

    Daniel Zhang, Saurabh Mishra, Erik Brynjolfsson +10

    cs.AIcs.GLarXiv:2103.06312v12021
  6. Accurate, Interdisciplinary and Transparent Structure-property Understanding with Deep Native Structural Reasoning

    Chen Tang, Yizhou Wang, Jianyu Wu +26

    cs.CLcs.AIcs.CEarXiv:2607.07708v12026
    Summaries:한국어
  7. AgentLens: Production-Assessed Trajectory Reviews for Coding Agent Evaluation

    Andrey Podivilov, Vadim Lomshakov, Sergey Savin +4

    cs.AIcs.LGcs.SEarXiv:2607.06624v22026
  8. What LLM Forecasters Know but Don't Say: Probing Internal Representations for Calibration and Faithfulness

    Raphaël Sarfati, Pratyush Ranjan Tiwari, Siddharth Boppana +3

    cs.CLcs.AIarXiv:2607.08046v12026
  9. Multiplayer Interactive World Models with Representation Autoencoders

    Anthony Hu, Václav Volhejn, Adrien Ramanana Rahary +24

    cs.CVcs.AIcs.LGarXiv:2607.05352v22026
  10. Attending to Multimodal Generation One Token at a Time

    Varun Gupta, Vineet Gandhi, Makarand Tapaswi

    cs.CVcs.AIarXiv:2607.03738v12026
  11. Learning to Move Before Learning to Do: Task-Agnostic pretraining for VLAs

    Junhao Shi, Siyin Wang, Xiaopeng Yu +3

    cs.ROcs.AIarXiv:2607.02466v12026
  12. OrbitQuant: Data-Agnostic Quantization for Image and Video Diffusion Transformers

    Donghyun Lee, Jitesh Chavan, Duy Nguyen +5

    cs.CVcs.AIcs.LGarXiv:2607.02461v12026
  13. PACE: A Proxy for Agentic Capability Evaluation

    Yueqi Song, Lintang Sutawika, Jiarui Liu +8

    cs.AIcs.CLarXiv:2607.02032v22026
  14. Discrete Diffusion Language Models for Interactive Radiology Report Drafting

    Max Van Puyvelde, Halil Ibrahim Gulluk, Wim Van Criekinge +1

    cs.AIcs.LGarXiv:2607.01436v12026
  15. Are Performance-Optimization Benchmarks Reliably Measuring Coding Agents?

    Zhi Chen, Zhensu Sun, Yuling Shi +2

    cs.SEcs.AIarXiv:2607.01211v12026
  16. MemSyco-Bench: Benchmarking Sycophancy in Agent Memory

    Zhishang Xiang, Zerui Chen, Yunbo Tang +5

    cs.IRcs.AIarXiv:2607.01071v22026
    Summaries:한국어
  17. Logit-Contribution Scoring Identifies Non-Literal Retrieval Heads

    Aryo Pradipta Gema, Beatrice Alex, Pasquale Minervini

    cs.CLcs.AIcs.LGarXiv:2607.01002v12026
  18. Personalization as Inverse Planning: Learning Latent Design Intents for Agentic Slide Generation via Structural Denoising

    Tianci Liu, Zihan Dong, Linjun Zhang +5

    cs.AIarXiv:2607.00407v12026
  19. Seed2.0 Model Card: Towards Intelligence Frontier for Real-World Complexity

    Bytedance Seed

    cs.AIarXiv:2607.00248v12026
    Summaries:한국어
  20. ASPIRE: Agentic /Skills Discovery for Robotics

    Runyu Lu, Yubo Wu, Ethan Kou +11

    cs.ROcs.AIcs.MAarXiv:2607.00272v12026
  21. QVal: Cheaply Evaluating Dense Supervision Signals for Long-Horizon LLM Agents

    Sergio Hernández-Gutiérrez, Matteo Merler, Ilze Amanda Auzina +3

    cs.LGcs.AIcs.CLarXiv:2606.32034v12026
  22. Reinforcement Learning with Metacognitive Feedback Elicits Faithful Uncertainty Expression in LLMs

    Gabrielle Kaili-May Liu, Avi Caciularu, Gal Yona +2

    cs.CLcs.AIarXiv:2606.32032v12026
  23. Breaking Failure Cascades: Step-Aware Reinforcement Learning for Medical Multimodal Reasoning

    Junha Jung, Minbyul Jeong, Suhyeon Lim +5

    cs.CVcs.AIarXiv:2606.31825v12026
    Summaries:한국어
  24. TRIAGE: Role-Typed Credit Assignment for Agentic Reinforcement Learning

    Yuanda Xu, Zhengze Zhou, Hejian Sang +6

    cs.LGcs.AIarXiv:2606.32017v22026
  25. Rank-Aware Hyperbolic Alignment for Vision-Language Dataset Distillation

    Jongoh Jeong, Sun-Kyung Lee, Kuk-Jin Yoon

    cs.CVcs.AIarXiv:2606.29464v12026
  26. Agentic Abstention: Do Agents Know When to Stop Instead of Act?

    Han Luo, Bingbing Wen, Lucy Lu Wang

    cs.AIarXiv:2606.28733v12026
  27. PhysisForcing: Physics Reinforced World Simulator for Robotic Manipulation

    Peiwen Zhang, Yufan Deng, Shangkun Sun +11

    cs.CVcs.AIcs.ROarXiv:2606.28128v12026
    Summaries:한국어
  28. Plans Don't Persist: Why Context Management Is Load Bearing for LLM Agents

    Aman Mehta, Anupam Datta

    cs.AIcs.CLarXiv:2606.22953v12026
  29. FLUX3D: High-Fidelity 3D Gaussian Generation with Diffusion-Aligned Sparse Representation

    Haorui Ji, Weizhe Liu, Hongdong Li +1

    cs.CVcs.AIarXiv:2606.24874v12026
  30. Hy-Embodied-0.5-VLA: From Vision-Language-Action Models to a Real-World Robot Learning Stack

    He Zhang, Lingzhu Xiang, Haitao Lin +23

    cs.ROcs.AIarXiv:2606.14409v22026
  31. Regret Minimization with Adaptive Opponents in Repeated Games

    Mingyang Liu, Asuman Ozdaglar, Tiancheng Yu +1

    cs.LGcs.AIcs.GTarXiv:2606.06486v12026
  32. Reinforcement Learning from Rich Feedback with Distributional DAgger

    Rishabh Agrawal, Jacob Fein-Ashley, Paria Rashidinejad

    cs.LGcs.AIcs.CLarXiv:2606.05152v22026
  33. Unlocking Feature Learning in Gated Delta Networks at Scale

    Yifeng Liu, Quanquan Gu

    cs.LGcs.AIarXiv:2606.04048v12026
  34. SDR: Set-Distance Rewards for Radiology Report Generation

    Halil Ibrahim Gulluk, Max Van Puyvelde, Wim Van Criekinge +1

    cs.AIarXiv:2606.00440v12026
  35. Unified Neural Scaling Laws

    Ethan Caballero, Priyank Jaini, David Krueger +1

    cs.LGcs.AIcs.NEarXiv:2605.26248v12026
  36. Efficient Agentic Reasoning Through Self-Regulated Simulative Planning

    Mingkai Deng, Jinyu Hou, Lara Sá Neves +4

    cs.AIcs.CLcs.LGarXiv:2605.22138v12026
  37. From Reasoning Chains to Verifiable Subproblems: Curriculum Reinforcement Learning Enables Credit Assignment for LLM Reasoning

    Xitai Jiang, Zihan Tang, Wenze Lin +3

    cs.LGcs.AIcs.CLarXiv:2605.22074v12026
  38. Agent-ValueBench: A Comprehensive Benchmark for Evaluating Agent Values

    Haonan Dong, Qiguan Feng, Kehan Jiang +3

    cs.AIarXiv:2605.10365v12026
    Summaries:한국어
  39. DynMuon: A Dynamic Spectral Shaping View of Muon

    Fangzhou Wu, Rikhav Shah, Sandeep Silwal +1

    cs.LGcs.AIarXiv:2605.17109v32026
  40. FutureSim: Replaying World Events to Evaluate Adaptive Agents

    Shashwat Goel, Nikhil Chandak, Arvindh Arun +5

    cs.LGcs.AIcs.CLarXiv:2605.15188v12026
  41. MinT: Managed Infrastructure for Training and Serving Millions of LLMs

    Mind Lab, :, Song Cao +60

    cs.LGcs.AIcs.DCarXiv:2605.13779v22026
  42. Retrieval is Cheap, Show Me the Code: Executable Multi-Hop Reasoning for Retrieval-Augmented Generation

    Jiashuo Sun, Jimeng Shi, Yixuan Xie +10

    cs.AIarXiv:2605.12975v12026
  43. What if AI systems weren't chatbots?

    Sourojit Ghosh, Pranav Narayanan Venkit, Sanjana Gautam +1

    cs.CYcs.AIarXiv:2605.07896v12026
  44. Automating Database-Native Function Code Synthesis with LLMs

    Wei Zhou, Xuanhe Zhou, Qikang He +4

    cs.DBcs.AIcs.CLarXiv:2604.06231v12026
  45. GBQA: A Game Benchmark for Evaluating LLMs as Quality Assurance Engineers

    Shufan Jiang, Chios Chen, Zhiyang Chen

    cs.SEcs.AIarXiv:2604.02648v12026
    Summaries:한국어
  46. The Model Says Walk: How Surface Heuristics Override Implicit Constraints in LLM Reasoning

    Yubo Li, Lu Zhang, Tianchong Jiang +2

    cs.CLcs.AIarXiv:2603.29025v32026
  47. From Truncation to Commitment: Persistent Context in Uniform Discrete Diffusion

    Satoshi Hayakawa

    cs.LGcs.AImath.PRarXiv:2609.01043v12026
  48. Not All Layers Are Created Equal: Adaptive LoRA Ranks for Personalized Image Generation

    Donald Shenaj, Federico Errica, Antonio Carta

    cs.CVcs.AIcs.LGarXiv:2603.21884v12026
  49. I Know What I Don't Know: Latent Posterior Factor Models for Multi-Evidence Probabilistic Reasoning

    Aliyu Agboola Alege

    cs.AIcs.LGarXiv:2603.15670v22026
  50. SurvHTE-Bench: A Benchmark for Heterogeneous Treatment Effect Estimation in Survival Analysis

    Shahriar Noroozizadeh, Xiaobin Shen, Jeremy C. Weiss +1

    cs.LGcs.AIstat.MLarXiv:2603.05483v12026
  51. MMTEB: Massive Multilingual Text Embedding Benchmark

    Kenneth Enevoldsen, Isaac Chung, Imene Kerboua +83

    cs.CLcs.AIcs.IRarXiv:2502.13595v42025
  52. Context Length Alone Hurts LLM Performance Despite Perfect Retrieval

    Yufeng Du, Minyang Tian, Srikanth Ronanki +7

    cs.CLcs.AIarXiv:2510.05381v12025
  53. AgentSociety: Large-Scale Simulation of LLM-Driven Generative Agents Advances Understanding of Human Behaviors and Society

    Jinghua Piao, Yuwei Yan, Jun Zhang +13

    cs.SIcs.AIarXiv:2502.08691v22025
  54. Surprised by Attention: Predictable Query Dynamics for Time Series Anomaly Detection

    Kadir-Kaan Özer, René Ebeling, Markus Enzweiler

    cs.LGcs.AIarXiv:2603.12916v32026
  55. ELEPHANT: Measuring and understanding social sycophancy in LLMs

    Myra Cheng, Sunny Yu, Cinoo Lee +3

    cs.CLcs.AIcs.CYarXiv:2505.13995v22025
  56. Agentic programs: an emerging form of scientific software in computational materials science

    Yunsung Lim, Haekwan Jeon, Jaesun Kim +2

    cond-mat.mtrl-scics.AIarXiv:2609.00795v12026
  57. Scale Space Diffusion

    Soumik Mukhopadhyay, Prateksha Udhayanan, Abhinav Shrivastava

    cs.CVcs.AIarXiv:2603.08709v12026
  58. Denoising Diffusion Generative Models Secretly Calculate Attentions

    Farzan Haddadi, Leila Monfared, Ebrahim Rezaii +3

    cs.AIcs.CVcs.LGarXiv:2609.00885v12026
  59. See and Fix the Flaws: Enabling VLMs and Diffusion Models to Comprehend Visual Artifacts via Agentic Data Synthesis

    Jaehyun Park, Minyoung Ahn, Minkyu Kim +3

    cs.CVcs.AIarXiv:2602.20951v22026
  60. No Language Left Behind: Scaling Human-Centered Machine Translation

    NLLB Team, Marta R. Costa-jussà, James Cross +36

    cs.CLcs.AIarXiv:2207.04672v32022