Artificial Intelligence

Papers filed under cs.AI on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.

Search paper metadata (including unsummarized papers)

5,881 to 5,940 of 15,364

  1. Logit-Contribution Scoring Identifies Non-Literal Retrieval Heads

    Aryo Pradipta Gema, Beatrice Alex, Pasquale Minervini

    cs.CLcs.AIcs.LGarXiv:2607.01002v12026
  2. Personalization as Inverse Planning: Learning Latent Design Intents for Agentic Slide Generation via Structural Denoising

    Tianci Liu, Zihan Dong, Linjun Zhang +5

    cs.AIarXiv:2607.00407v12026
  3. Seed2.0 Model Card: Towards Intelligence Frontier for Real-World Complexity

    Bytedance Seed

    cs.AIarXiv:2607.00248v12026
    Summaries:한국어
  4. ASPIRE: Agentic /Skills Discovery for Robotics

    Runyu Lu, Yubo Wu, Ethan Kou +11

    cs.ROcs.AIcs.MAarXiv:2607.00272v12026
  5. QVal: Cheaply Evaluating Dense Supervision Signals for Long-Horizon LLM Agents

    Sergio Hernández-Gutiérrez, Matteo Merler, Ilze Amanda Auzina +3

    cs.LGcs.AIcs.CLarXiv:2606.32034v12026
  6. Reinforcement Learning with Metacognitive Feedback Elicits Faithful Uncertainty Expression in LLMs

    Gabrielle Kaili-May Liu, Avi Caciularu, Gal Yona +2

    cs.CLcs.AIarXiv:2606.32032v12026
  7. Breaking Failure Cascades: Step-Aware Reinforcement Learning for Medical Multimodal Reasoning

    Junha Jung, Minbyul Jeong, Suhyeon Lim +5

    cs.CVcs.AIarXiv:2606.31825v12026
    Summaries:한국어
  8. TRIAGE: Role-Typed Credit Assignment for Agentic Reinforcement Learning

    Yuanda Xu, Zhengze Zhou, Hejian Sang +6

    cs.LGcs.AIarXiv:2606.32017v22026
  9. Rank-Aware Hyperbolic Alignment for Vision-Language Dataset Distillation

    Jongoh Jeong, Sun-Kyung Lee, Kuk-Jin Yoon

    cs.CVcs.AIarXiv:2606.29464v12026
  10. Agentic Abstention: Do Agents Know When to Stop Instead of Act?

    Han Luo, Bingbing Wen, Lucy Lu Wang

    cs.AIarXiv:2606.28733v12026
  11. PhysisForcing: Physics Reinforced World Simulator for Robotic Manipulation

    Peiwen Zhang, Yufan Deng, Shangkun Sun +11

    cs.CVcs.AIcs.ROarXiv:2606.28128v12026
    Summaries:한국어
  12. Plans Don't Persist: Why Context Management Is Load Bearing for LLM Agents

    Aman Mehta, Anupam Datta

    cs.AIcs.CLarXiv:2606.22953v12026
  13. FLUX3D: High-Fidelity 3D Gaussian Generation with Diffusion-Aligned Sparse Representation

    Haorui Ji, Weizhe Liu, Hongdong Li +1

    cs.CVcs.AIarXiv:2606.24874v12026
  14. Hy-Embodied-0.5-VLA: From Vision-Language-Action Models to a Real-World Robot Learning Stack

    He Zhang, Lingzhu Xiang, Haitao Lin +23

    cs.ROcs.AIarXiv:2606.14409v22026
  15. Regret Minimization with Adaptive Opponents in Repeated Games

    Mingyang Liu, Asuman Ozdaglar, Tiancheng Yu +1

    cs.LGcs.AIcs.GTarXiv:2606.06486v12026
  16. Reinforcement Learning from Rich Feedback with Distributional DAgger

    Rishabh Agrawal, Jacob Fein-Ashley, Paria Rashidinejad

    cs.LGcs.AIcs.CLarXiv:2606.05152v22026
  17. Unlocking Feature Learning in Gated Delta Networks at Scale

    Yifeng Liu, Quanquan Gu

    cs.LGcs.AIarXiv:2606.04048v12026
  18. SDR: Set-Distance Rewards for Radiology Report Generation

    Halil Ibrahim Gulluk, Max Van Puyvelde, Wim Van Criekinge +1

    cs.AIarXiv:2606.00440v12026
  19. Unified Neural Scaling Laws

    Ethan Caballero, Priyank Jaini, David Krueger +1

    cs.LGcs.AIcs.NEarXiv:2605.26248v12026
  20. Efficient Agentic Reasoning Through Self-Regulated Simulative Planning

    Mingkai Deng, Jinyu Hou, Lara Sá Neves +4

    cs.AIcs.CLcs.LGarXiv:2605.22138v12026
  21. From Reasoning Chains to Verifiable Subproblems: Curriculum Reinforcement Learning Enables Credit Assignment for LLM Reasoning

    Xitai Jiang, Zihan Tang, Wenze Lin +3

    cs.LGcs.AIcs.CLarXiv:2605.22074v12026
  22. Agent-ValueBench: A Comprehensive Benchmark for Evaluating Agent Values

    Haonan Dong, Qiguan Feng, Kehan Jiang +3

    cs.AIarXiv:2605.10365v12026
    Summaries:한국어
  23. DynMuon: A Dynamic Spectral Shaping View of Muon

    Fangzhou Wu, Rikhav Shah, Sandeep Silwal +1

    cs.LGcs.AIarXiv:2605.17109v32026
  24. FutureSim: Replaying World Events to Evaluate Adaptive Agents

    Shashwat Goel, Nikhil Chandak, Arvindh Arun +5

    cs.LGcs.AIcs.CLarXiv:2605.15188v12026
  25. MinT: Managed Infrastructure for Training and Serving Millions of LLMs

    Mind Lab, :, Song Cao +60

    cs.LGcs.AIcs.DCarXiv:2605.13779v22026
  26. Retrieval is Cheap, Show Me the Code: Executable Multi-Hop Reasoning for Retrieval-Augmented Generation

    Jiashuo Sun, Jimeng Shi, Yixuan Xie +10

    cs.AIarXiv:2605.12975v12026
  27. What if AI systems weren't chatbots?

    Sourojit Ghosh, Pranav Narayanan Venkit, Sanjana Gautam +1

    cs.CYcs.AIarXiv:2605.07896v12026
  28. Automating Database-Native Function Code Synthesis with LLMs

    Wei Zhou, Xuanhe Zhou, Qikang He +4

    cs.DBcs.AIcs.CLarXiv:2604.06231v12026
  29. GBQA: A Game Benchmark for Evaluating LLMs as Quality Assurance Engineers

    Shufan Jiang, Chios Chen, Zhiyang Chen

    cs.SEcs.AIarXiv:2604.02648v12026
    Summaries:한국어
  30. The Model Says Walk: How Surface Heuristics Override Implicit Constraints in LLM Reasoning

    Yubo Li, Lu Zhang, Tianchong Jiang +2

    cs.CLcs.AIarXiv:2603.29025v32026
  31. From Truncation to Commitment: Persistent Context in Uniform Discrete Diffusion

    Satoshi Hayakawa

    cs.LGcs.AImath.PRarXiv:2609.01043v12026
  32. Not All Layers Are Created Equal: Adaptive LoRA Ranks for Personalized Image Generation

    Donald Shenaj, Federico Errica, Antonio Carta

    cs.CVcs.AIcs.LGarXiv:2603.21884v12026
  33. I Know What I Don't Know: Latent Posterior Factor Models for Multi-Evidence Probabilistic Reasoning

    Aliyu Agboola Alege

    cs.AIcs.LGarXiv:2603.15670v22026
  34. SurvHTE-Bench: A Benchmark for Heterogeneous Treatment Effect Estimation in Survival Analysis

    Shahriar Noroozizadeh, Xiaobin Shen, Jeremy C. Weiss +1

    cs.LGcs.AIstat.MLarXiv:2603.05483v12026
  35. MMTEB: Massive Multilingual Text Embedding Benchmark

    Kenneth Enevoldsen, Isaac Chung, Imene Kerboua +83

    cs.CLcs.AIcs.IRarXiv:2502.13595v42025
  36. Context Length Alone Hurts LLM Performance Despite Perfect Retrieval

    Yufeng Du, Minyang Tian, Srikanth Ronanki +7

    cs.CLcs.AIarXiv:2510.05381v12025
  37. AgentSociety: Large-Scale Simulation of LLM-Driven Generative Agents Advances Understanding of Human Behaviors and Society

    Jinghua Piao, Yuwei Yan, Jun Zhang +13

    cs.SIcs.AIarXiv:2502.08691v22025
  38. Surprised by Attention: Predictable Query Dynamics for Time Series Anomaly Detection

    Kadir-Kaan Özer, René Ebeling, Markus Enzweiler

    cs.LGcs.AIarXiv:2603.12916v32026
  39. ELEPHANT: Measuring and understanding social sycophancy in LLMs

    Myra Cheng, Sunny Yu, Cinoo Lee +3

    cs.CLcs.AIcs.CYarXiv:2505.13995v22025
  40. Agentic programs: an emerging form of scientific software in computational materials science

    Yunsung Lim, Haekwan Jeon, Jaesun Kim +2

    cond-mat.mtrl-scics.AIarXiv:2609.00795v12026
  41. Scale Space Diffusion

    Soumik Mukhopadhyay, Prateksha Udhayanan, Abhinav Shrivastava

    cs.CVcs.AIarXiv:2603.08709v12026
  42. Denoising Diffusion Generative Models Secretly Calculate Attentions

    Farzan Haddadi, Leila Monfared, Ebrahim Rezaii +3

    cs.AIcs.CVcs.LGarXiv:2609.00885v12026
  43. See and Fix the Flaws: Enabling VLMs and Diffusion Models to Comprehend Visual Artifacts via Agentic Data Synthesis

    Jaehyun Park, Minyoung Ahn, Minkyu Kim +3

    cs.CVcs.AIarXiv:2602.20951v22026
  44. No Language Left Behind: Scaling Human-Centered Machine Translation

    NLLB Team, Marta R. Costa-jussà, James Cross +36

    cs.CLcs.AIarXiv:2207.04672v32022
  45. Simple Contrastive Graph Clustering

    Yue Liu, Xihong Yang, Sihang Zhou +1

    cs.LGcs.AIarXiv:2205.07865v32022
  46. Flexible Entropy Control in RLVR with a Gradient-Preserving Perspective

    Kun Chen, Peng Shi, Fanfan Liu +4

    cs.LGcs.AIcs.CLarXiv:2602.09782v22026
  47. Trust The Typical

    Debargha Ganguly, Sreehari Sankar, Biyao Zhang +8

    cs.CLcs.AIcs.DCarXiv:2602.04581v12026
  48. Agentic AI: A Comprehensive Survey of Architectures, Applications, and Future Directions

    Mohamad Abou Ali, Fadi Dornaika

    cs.AIcs.LGarXiv:2510.25445v12025
  49. Context Learning for Multi-Agent Discussion

    Xingyuan Hua, Sheng Yue, Xinyi Li +3

    cs.AIcs.LGcs.MAarXiv:2602.02350v32026
  50. Steering LLMs via Scalable Interactive Oversight

    Enyu Zhou, Zhiheng Xi, Long Ma +9

    cs.AIcs.LGarXiv:2602.04210v22026
  51. Ethics-Based Auditing to Develop Trustworthy AI

    Jakob Mokander, Luciano Floridi

    cs.CYcs.AIarXiv:2105.00002v12021
  52. Your Brain on ChatGPT: Accumulation of Cognitive Debt when Using an AI Assistant for Essay Writing Task

    Nataliya Kosmyna, Eugene Hauptmann, Ye Tong Yuan +5

    cs.AIarXiv:2506.08872v22025
  53. InT: Self-Proposed Interventions Enable Credit Assignment in LLM Reasoning

    Matthew Y. R. Yang, Hao Bai, Ian Wu +3

    cs.LGcs.AIcs.CLarXiv:2601.14209v12026
  54. A Survey on LLM-as-a-Judge

    Jiawei Gu, Xuhui Jiang, Zhichao Shi +13

    cs.CLcs.AIarXiv:2411.15594v62024
  55. An All-in-One Network for Dehazing and Beyond

    Boyi Li, Xiulian Peng, Zhangyang Wang +2

    cs.CVcs.AIarXiv:1707.06543v12017
  56. LLM Agents Grounded in Self-Reports Enable General-Purpose Simulation of Individuals

    Joon Sung Park, Carolyn Q. Zou, Jonne Kamphorst +8

    cs.AIcs.HCcs.LGarXiv:2411.10109v32024
  57. Taming Visually Guided Sound Generation

    Vladimir Iashin, Esa Rahtu

    cs.CVcs.AIcs.LGarXiv:2110.08791v12021
  58. From PINNs to PIKANs: Recent Advances in Physics-Informed Machine Learning

    Juan Diego Toscano, Vivek Oommen, Alan John Varghese +4

    cs.LGcs.AIphysics.comp-pharXiv:2410.13228v22024
  59. A Survey on Diffusion Models for Inverse Problems

    Giannis Daras, Hyungjin Chung, Chieh-Hsin Lai +5

    cs.LGcs.AIcs.CVarXiv:2410.00083v12024
  60. Magpie: Alignment Data Synthesis from Scratch by Prompting Aligned LLMs with Nothing

    Zhangchen Xu, Fengqing Jiang, Luyao Niu +4

    cs.CLcs.AIarXiv:2406.08464v22024
    Summaries:한국어