Artificial Intelligence

Papers filed under cs.AI on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.

Search paper metadata (including unsummarized papers)

4,801 to 4,860 of 15,246

  1. STAR-1: Safer Alignment of Reasoning LLMs with 1K Data

    Zijun Wang, Haoqin Tu, Yuhan Wang +6

    cs.CLcs.AIarXiv:2504.01903v22025
  2. RoboDojo: A Unified Sim-and-Real Benchmark for Comprehensive Evaluation of Generalist Robot Manipulation Policies

    Tianxing Chen, Yue Chen, Zixuan Li +41

    cs.ROcs.AIcs.CVarXiv:2607.04434v32026
  3. Learning Humanoid Standing-up Control across Diverse Postures

    Tao Huang, Junli Ren, Huayi Wang +6

    cs.ROcs.AIcs.LGarXiv:2502.08378v22025
  4. Context-Grounding Gains Are Mediated by Pre-existing Machinery: Auditing GRPO, SFT, and DPO

    Prakhar Gupta, Vaibhav Gupta

    cs.CLcs.AIcs.LGarXiv:2609.00925v12026
  5. Unsupervised Anomaly Detection for Image Dataset Quality Assurance in Multi-Center Breast MRI

    Chiara Tappermann, Steffen Renisch, Lars Ole Schwen +3

    cs.CVcs.AIarXiv:2608.16725v12026
  6. Effectiveness of IoT and Deep Learning for Detection and Severity Assessment of Postelectrotermes militaris in Tea Plantations

    D. K. C. Senevirathna, A. A. E. Nanayakkara, H. M. C. K. Kulathunga +7

    cs.AIcs.LGcs.SDarXiv:2608.27480v12026
  7. Generative to Agentic AI: Survey, Conceptualization, and Challenges

    Johannes Schneider

    cs.AIarXiv:2504.18875v12025
  8. Exploring the Role of LLMs in HPC Programming: A Survey

    Strahinja Ljaljevic, Josep Jorba, Sergio Iserte

    cs.DCcs.AIarXiv:2608.26110v12026
  9. Tiny but Trusted: Efficient Vision-Language Reasoning for Time-Series Anomaly Detection

    Xiaona Zhou, Muntasir Wahed, Tianjiao Yu +2

    cs.AIarXiv:2605.30344v12026
  10. Layered LLM Defenses as an Ensemble: Access Tiers, Inference Cost, and the Measured Failure Correlation Between Defense Layers

    Abrar Alotaibi, Muhammad Shahid Jabbar, Sadam Al-Azani +1

    cs.CRcs.AIcs.CLarXiv:2608.28327v12026
  11. wd1: Weighted Policy Optimization for Reasoning in Diffusion Language Models

    Xiaohang Tang, Rares Dolga, Sangwoong Yoon +1

    cs.LGcs.AIstat.MLarXiv:2507.08838v22025
  12. NeuroCogMap Reveals Cognitive Organization of Large Language Models

    Zhongxiang Sun, Haolang Lu, Qiang Ma +11

    q-bio.NCcs.AIcs.CLarXiv:2607.00397v12026
    Summaries:한국어
  13. A Comprehensive Survey of Mixture-of-Experts: Algorithms, Theory, and Applications

    Siyuan Mu, Sen Lin

    cs.LGcs.AIarXiv:2503.07137v42025
  14. Machine learning meets network science: dimensionality reduction for fast and efficient embedding of networks in the hyperbolic space

    Josephine Maria Thomas, Alessandro Muscoloni, Sara Ciucci +2

    cond-mat.dis-nncs.AIcs.LGarXiv:1602.06522v12016
  15. Intelligent AI Delegation

    Nenad Tomašev, Matija Franklin, Simon Osindero

    cs.AIarXiv:2602.11865v12026
  16. Provably Safe Sim-to-Real Transfer

    Tingting Ni, Maryam Kamgarpour

    cs.LGcs.AIarXiv:2609.01418v12026
  17. Bandits in Prod: Hyperparameter Optimization at Inference Time

    Louis Abraham, Tuan-Anh Nguyen, Nicolas Devatine

    cs.LGcs.AIarXiv:2609.01335v22026
  18. Superposed Latent Autoencoder

    Quanling Zhao, Jiaying Yang, Tianqi Zhang +4

    cs.LGcs.AIarXiv:2609.01158v12026
  19. Reinforcement Learning Outperforms Supervised Fine-Tuning: A Case Study on Audio Question Answering

    Gang Li, Jizhong Liu, Heinrich Dinkel +3

    cs.SDcs.AIcs.CLarXiv:2503.11197v42025
  20. Not All Rollouts are Useful: Down-Sampling Rollouts in LLM Reinforcement Learning

    Yixuan Even Xu, Yash Savani, Fei Fang +1

    cs.LGcs.AIcs.CLarXiv:2504.13818v52025
  21. Beyond the Image Plane: World-Grounded Queries for Multi-Object Tracking

    Orcun Cetintas, Guillem Brasó, Tim Meinhardt +1

    cs.CVcs.AIarXiv:2609.00924v12026
  22. Restrict, Don't Retrain: Inference-Time VLM Guidance for Zero-Shot Aerial Segmentation

    Teresa DiMeola, Charles Walter, Hong Xiao

    cs.CVcs.AIarXiv:2609.00628v12026
  23. VDocRAG: Retrieval-Augmented Generation over Visually-Rich Documents

    Ryota Tanaka, Taichi Iki, Taku Hasegawa +3

    cs.CLcs.AIcs.CVarXiv:2504.09795v12025
  24. H2Table: Hierarchical Hypergraph-Enhanced Large Language Models for Complex Table Reasoning

    Jia Ling, Yangfan Wang, Chen Tang +4

    cs.AIarXiv:2609.01216v12026
  25. More Thinking, Less Seeing? Assessing Amplified Hallucination in Multimodal Reasoning Models

    Chengzhi Liu, Zhongxing Xu, Qingyue Wei +5

    cs.CLcs.AIcs.CVarXiv:2505.21523v32025
  26. Data-Driven Persona-Conditioned Agents for A/B Test Simulation

    Ziyad Benomar, Weronika Łajewska, Leonardo Perelli +1

    cs.AIarXiv:2609.01038v12026
  27. Neural Rough Differential Equations for Long Time Series

    James Morrill, Cristopher Salvi, Patrick Kidger +2

    cs.LGcs.AImath.DSarXiv:2009.08295v42020
  28. FLaG: Frequency-Domain Latent-attention Gated Pooling for Token Aggregation

    Kewei Li, Rongying Zhang, Xueli Wang +6

    cs.AIq-bio.BMarXiv:2609.00831v12026
  29. WritingBench: A Comprehensive Benchmark for Generative Writing

    Yuning Wu, Jiahao Mei, Ming Yan +8

    cs.AIcs.CLarXiv:2503.05244v42025
  30. Feedback-Assisted Trust Propagation over Document Relation Graphs for Retrieval-Augmented Generation

    Zhuoheng Li, Ying Chen

    cs.AIarXiv:2609.00543v12026
  31. Agentic Software Engineering: Foundational Pillars and a Research Roadmap

    Ahmed E. Hassan, Hao Li, Dayi Lin +4

    cs.SEcs.AIarXiv:2509.06216v32025
  32. On the Adversarial Robustness of Vision Transformers

    Rulin Shao, Zhouxing Shi, Jinfeng Yi +2

    cs.CVcs.AIcs.LGarXiv:2103.15670v32021
  33. Sparse Autoencoders Learn Monosemantic Features in Vision-Language Models

    Mateusz Pach, Shyamgopal Karthik, Quentin Bouniot +2

    cs.CVcs.AIcs.LGarXiv:2504.02821v32025
  34. Assessing Alignment and Stability of Feature Importance Explanations via Weight of Evidence

    Eddie Conti, Claudio Daka, Álvaro Parafita +3

    cs.LGcs.AIarXiv:2609.00090v12026
  35. On the Reliable Detection of Concept Drift from Streaming Unlabeled Data

    Tegjyot Singh Sethi, Mehmed Kantardzic

    stat.MLcs.AIcs.LGarXiv:1704.00023v12017
  36. Investigating Affective Use and Emotional Well-being on ChatGPT

    Jason Phang, Michael Lampe, Lama Ahmad +8

    cs.HCcs.AIarXiv:2504.03888v12025
  37. Unsupervised Control Through Non-Parametric Discriminative Rewards

    David Warde-Farley, Tom Van de Wiele, Tejas Kulkarni +3

    cs.LGcs.AIstat.MLarXiv:1811.11359v12018
  38. Vevo: Controllable Zero-Shot Voice Imitation with Self-Supervised Disentanglement

    Xueyao Zhang, Xiaohui Zhang, Kainan Peng +10

    cs.SDcs.AIeess.ASarXiv:2502.07243v12025
  39. Interpretable to Whom? A Role-based Model for Analyzing Interpretable Machine Learning Systems

    Richard Tomsett, Dave Braines, Dan Harborne +2

    cs.AIarXiv:1806.07552v12018
  40. Lingua Franca or Probing Artifact? Rethinking Latent Language in Multilingual LLMs

    Deniz Bayazit, Badr AlKhamissi, Antoine Bosselut

    cs.CLcs.AIcs.LGarXiv:2609.00155v12026
  41. How Well do LLMs Compress Their Own Chain-of-Thought? A Token Complexity Approach

    Ayeong Lee, Ethan Che, Tianyi Peng

    cs.CLcs.AIarXiv:2503.01141v22025
  42. PyKEEN 1.0: A Python Library for Training and Evaluating Knowledge Graph Embeddings

    Mehdi Ali, Max Berrendorf, Charles Tapley Hoyt +4

    cs.LGcs.AIstat.MLarXiv:2007.14175v22020
  43. Are Sparse Autoencoders Useful? A Case Study in Sparse Probing

    Subhash Kantamneni, Joshua Engels, Senthooran Rajamanoharan +2

    cs.LGcs.AIarXiv:2502.16681v12025
  44. Winning Gold at IMO 2025 with a Model-Agnostic Verification-and-Refinement Pipeline

    Yichen Huang, Lin F. Yang

    cs.AIarXiv:2507.15855v42025
  45. Dish-TS: A General Paradigm for Alleviating Distribution Shift in Time Series Forecasting

    Wei Fan, Pengyang Wang, Dongkun Wang +3

    cs.LGcs.AIarXiv:2302.14829v32023
  46. KItCAT: Knowledge Injection via Input Corruption for Auto-regressive Training

    Meghanadh Pulivarthi, Kushagra Bhushan, Vineet Kumar +5

    cs.CLcs.AIarXiv:2609.00082v12026
  47. Aristotle: IMO-level Automated Theorem Proving

    Tudor Achim, Alex Best, Alberto Bietti +20

    cs.AIcs.CLarXiv:2510.01346v22025
  48. Kosmos: An AI Scientist for Autonomous Discovery

    Ludovico Mitchener, Angela Yiu, Benjamin Chang +34

    cs.AIarXiv:2511.02824v22025
  49. Vision Is Not Overhead: One-Pass Block Drafting for Lossless Speculative Decoding in Vision-Language Models

    Jungseob Lee, Seongtae Hong, Dongyub Jude Lee +4

    cs.AIcs.CLcs.CVarXiv:2609.00355v12026
  50. Deformable ProtoPNet: An Interpretable Image Classifier Using Deformable Prototypes

    Jon Donnelly, Alina Jade Barnett, Chaofan Chen

    cs.CVcs.AIcs.LGarXiv:2111.15000v32021
  51. DiTAR: Diffusion Transformer Autoregressive Modeling for Speech Generation

    Dongya Jia, Zhuo Chen, Jiawei Chen +8

    eess.AScs.AIcs.CLarXiv:2502.03930v42025
  52. AdaCoT: Pareto-Optimal Adaptive Chain-of-Thought Triggering via Reinforcement Learning

    Chenwei Lou, Zewei Sun, Xinnian Liang +6

    cs.LGcs.AIarXiv:2505.11896v22025
  53. Chain-of-Agents: End-to-End Agent Foundation Models via Multi-Agent Distillation and Agentic RL

    Weizhen Li, Jianbo Lin, Zhuosong Jiang +27

    cs.AIcs.CLarXiv:2508.13167v12025
  54. SafeArena: Evaluating the Safety of Autonomous Web Agents

    Ada Defne Tur, Nicholas Meade, Xing Han Lù +6

    cs.LGcs.AIcs.CLarXiv:2503.04957v12025
  55. Token Assorted: Mixing Latent and Text Tokens for Improved Language Model Reasoning

    DiJia Su, Hanlin Zhu, Yingchen Xu +3

    cs.CLcs.AIcs.LGarXiv:2502.03275v22025
  56. AdaEvolve: Adaptive LLM Driven Zeroth-Order Optimization

    Mert Cemri, Shubham Agrawal, Akshat Gupta +9

    cs.NEcs.AIcs.CLarXiv:2602.20133v12026
  57. MMC: Advancing Multimodal Chart Understanding with Large-scale Instruction Tuning

    Fuxiao Liu, Xiaoyang Wang, Wenlin Yao +5

    cs.CLcs.AIarXiv:2311.10774v22023
  58. SEED-GRPO: Semantic Entropy Enhanced GRPO for Uncertainty-Aware Policy Optimization

    Minghan Chen, Guikun Chen, Wenguan Wang +1

    cs.AIarXiv:2505.12346v12025
  59. S*: Test Time Scaling for Code Generation

    Dacheng Li, Shiyi Cao, Chengkun Cao +6

    cs.LGcs.AIarXiv:2502.14382v12025
  60. BeamDojo: Learning Agile Humanoid Locomotion on Sparse Footholds

    Huayi Wang, Zirui Wang, Junli Ren +4

    cs.ROcs.AIcs.LGarXiv:2502.10363v32025