Artificial Intelligence
Papers filed under cs.AI on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.
Search paper metadata (including unsummarized papers)
4,801 to 4,860 of 15,246
STAR-1: Safer Alignment of Reasoning LLMs with 1K Data
Zijun Wang, Haoqin Tu, Yuhan Wang +6
cs.CLcs.AIarXiv:2504.01903v22025RoboDojo: A Unified Sim-and-Real Benchmark for Comprehensive Evaluation of Generalist Robot Manipulation Policies
Tianxing Chen, Yue Chen, Zixuan Li +41
cs.ROcs.AIcs.CVarXiv:2607.04434v32026Learning Humanoid Standing-up Control across Diverse Postures
Tao Huang, Junli Ren, Huayi Wang +6
cs.ROcs.AIcs.LGarXiv:2502.08378v22025Context-Grounding Gains Are Mediated by Pre-existing Machinery: Auditing GRPO, SFT, and DPO
Prakhar Gupta, Vaibhav Gupta
cs.CLcs.AIcs.LGarXiv:2609.00925v12026Unsupervised Anomaly Detection for Image Dataset Quality Assurance in Multi-Center Breast MRI
Chiara Tappermann, Steffen Renisch, Lars Ole Schwen +3
cs.CVcs.AIarXiv:2608.16725v12026Effectiveness of IoT and Deep Learning for Detection and Severity Assessment of Postelectrotermes militaris in Tea Plantations
D. K. C. Senevirathna, A. A. E. Nanayakkara, H. M. C. K. Kulathunga +7
cs.AIcs.LGcs.SDarXiv:2608.27480v12026Generative to Agentic AI: Survey, Conceptualization, and Challenges
Johannes Schneider
cs.AIarXiv:2504.18875v12025Exploring the Role of LLMs in HPC Programming: A Survey
Strahinja Ljaljevic, Josep Jorba, Sergio Iserte
cs.DCcs.AIarXiv:2608.26110v12026Tiny but Trusted: Efficient Vision-Language Reasoning for Time-Series Anomaly Detection
Xiaona Zhou, Muntasir Wahed, Tianjiao Yu +2
cs.AIarXiv:2605.30344v12026Layered LLM Defenses as an Ensemble: Access Tiers, Inference Cost, and the Measured Failure Correlation Between Defense Layers
Abrar Alotaibi, Muhammad Shahid Jabbar, Sadam Al-Azani +1
cs.CRcs.AIcs.CLarXiv:2608.28327v12026wd1: Weighted Policy Optimization for Reasoning in Diffusion Language Models
Xiaohang Tang, Rares Dolga, Sangwoong Yoon +1
cs.LGcs.AIstat.MLarXiv:2507.08838v22025NeuroCogMap Reveals Cognitive Organization of Large Language Models
Zhongxiang Sun, Haolang Lu, Qiang Ma +11
q-bio.NCcs.AIcs.CLarXiv:2607.00397v12026Summaries:한국어A Comprehensive Survey of Mixture-of-Experts: Algorithms, Theory, and Applications
Siyuan Mu, Sen Lin
cs.LGcs.AIarXiv:2503.07137v42025Machine learning meets network science: dimensionality reduction for fast and efficient embedding of networks in the hyperbolic space
Josephine Maria Thomas, Alessandro Muscoloni, Sara Ciucci +2
cond-mat.dis-nncs.AIcs.LGarXiv:1602.06522v12016Intelligent AI Delegation
Nenad Tomašev, Matija Franklin, Simon Osindero
cs.AIarXiv:2602.11865v12026Provably Safe Sim-to-Real Transfer
Tingting Ni, Maryam Kamgarpour
cs.LGcs.AIarXiv:2609.01418v12026Bandits in Prod: Hyperparameter Optimization at Inference Time
Louis Abraham, Tuan-Anh Nguyen, Nicolas Devatine
cs.LGcs.AIarXiv:2609.01335v22026Superposed Latent Autoencoder
Quanling Zhao, Jiaying Yang, Tianqi Zhang +4
cs.LGcs.AIarXiv:2609.01158v12026Reinforcement Learning Outperforms Supervised Fine-Tuning: A Case Study on Audio Question Answering
Gang Li, Jizhong Liu, Heinrich Dinkel +3
cs.SDcs.AIcs.CLarXiv:2503.11197v42025Not All Rollouts are Useful: Down-Sampling Rollouts in LLM Reinforcement Learning
Yixuan Even Xu, Yash Savani, Fei Fang +1
cs.LGcs.AIcs.CLarXiv:2504.13818v52025Beyond the Image Plane: World-Grounded Queries for Multi-Object Tracking
Orcun Cetintas, Guillem Brasó, Tim Meinhardt +1
cs.CVcs.AIarXiv:2609.00924v12026Restrict, Don't Retrain: Inference-Time VLM Guidance for Zero-Shot Aerial Segmentation
Teresa DiMeola, Charles Walter, Hong Xiao
cs.CVcs.AIarXiv:2609.00628v12026VDocRAG: Retrieval-Augmented Generation over Visually-Rich Documents
Ryota Tanaka, Taichi Iki, Taku Hasegawa +3
cs.CLcs.AIcs.CVarXiv:2504.09795v12025H2Table: Hierarchical Hypergraph-Enhanced Large Language Models for Complex Table Reasoning
Jia Ling, Yangfan Wang, Chen Tang +4
cs.AIarXiv:2609.01216v12026More Thinking, Less Seeing? Assessing Amplified Hallucination in Multimodal Reasoning Models
Chengzhi Liu, Zhongxing Xu, Qingyue Wei +5
cs.CLcs.AIcs.CVarXiv:2505.21523v32025Data-Driven Persona-Conditioned Agents for A/B Test Simulation
Ziyad Benomar, Weronika Łajewska, Leonardo Perelli +1
cs.AIarXiv:2609.01038v12026Neural Rough Differential Equations for Long Time Series
James Morrill, Cristopher Salvi, Patrick Kidger +2
cs.LGcs.AImath.DSarXiv:2009.08295v42020FLaG: Frequency-Domain Latent-attention Gated Pooling for Token Aggregation
Kewei Li, Rongying Zhang, Xueli Wang +6
cs.AIq-bio.BMarXiv:2609.00831v12026WritingBench: A Comprehensive Benchmark for Generative Writing
Yuning Wu, Jiahao Mei, Ming Yan +8
cs.AIcs.CLarXiv:2503.05244v42025Feedback-Assisted Trust Propagation over Document Relation Graphs for Retrieval-Augmented Generation
Zhuoheng Li, Ying Chen
cs.AIarXiv:2609.00543v12026Agentic Software Engineering: Foundational Pillars and a Research Roadmap
Ahmed E. Hassan, Hao Li, Dayi Lin +4
cs.SEcs.AIarXiv:2509.06216v32025On the Adversarial Robustness of Vision Transformers
Rulin Shao, Zhouxing Shi, Jinfeng Yi +2
cs.CVcs.AIcs.LGarXiv:2103.15670v32021Sparse Autoencoders Learn Monosemantic Features in Vision-Language Models
Mateusz Pach, Shyamgopal Karthik, Quentin Bouniot +2
cs.CVcs.AIcs.LGarXiv:2504.02821v32025Assessing Alignment and Stability of Feature Importance Explanations via Weight of Evidence
Eddie Conti, Claudio Daka, Álvaro Parafita +3
cs.LGcs.AIarXiv:2609.00090v12026On the Reliable Detection of Concept Drift from Streaming Unlabeled Data
Tegjyot Singh Sethi, Mehmed Kantardzic
stat.MLcs.AIcs.LGarXiv:1704.00023v12017Investigating Affective Use and Emotional Well-being on ChatGPT
Jason Phang, Michael Lampe, Lama Ahmad +8
cs.HCcs.AIarXiv:2504.03888v12025Unsupervised Control Through Non-Parametric Discriminative Rewards
David Warde-Farley, Tom Van de Wiele, Tejas Kulkarni +3
cs.LGcs.AIstat.MLarXiv:1811.11359v12018Vevo: Controllable Zero-Shot Voice Imitation with Self-Supervised Disentanglement
Xueyao Zhang, Xiaohui Zhang, Kainan Peng +10
cs.SDcs.AIeess.ASarXiv:2502.07243v12025Interpretable to Whom? A Role-based Model for Analyzing Interpretable Machine Learning Systems
Richard Tomsett, Dave Braines, Dan Harborne +2
cs.AIarXiv:1806.07552v12018Lingua Franca or Probing Artifact? Rethinking Latent Language in Multilingual LLMs
Deniz Bayazit, Badr AlKhamissi, Antoine Bosselut
cs.CLcs.AIcs.LGarXiv:2609.00155v12026How Well do LLMs Compress Their Own Chain-of-Thought? A Token Complexity Approach
Ayeong Lee, Ethan Che, Tianyi Peng
cs.CLcs.AIarXiv:2503.01141v22025PyKEEN 1.0: A Python Library for Training and Evaluating Knowledge Graph Embeddings
Mehdi Ali, Max Berrendorf, Charles Tapley Hoyt +4
cs.LGcs.AIstat.MLarXiv:2007.14175v22020Are Sparse Autoencoders Useful? A Case Study in Sparse Probing
Subhash Kantamneni, Joshua Engels, Senthooran Rajamanoharan +2
cs.LGcs.AIarXiv:2502.16681v12025Winning Gold at IMO 2025 with a Model-Agnostic Verification-and-Refinement Pipeline
Yichen Huang, Lin F. Yang
cs.AIarXiv:2507.15855v42025Dish-TS: A General Paradigm for Alleviating Distribution Shift in Time Series Forecasting
Wei Fan, Pengyang Wang, Dongkun Wang +3
cs.LGcs.AIarXiv:2302.14829v32023KItCAT: Knowledge Injection via Input Corruption for Auto-regressive Training
Meghanadh Pulivarthi, Kushagra Bhushan, Vineet Kumar +5
cs.CLcs.AIarXiv:2609.00082v12026Aristotle: IMO-level Automated Theorem Proving
Tudor Achim, Alex Best, Alberto Bietti +20
cs.AIcs.CLarXiv:2510.01346v22025Kosmos: An AI Scientist for Autonomous Discovery
Ludovico Mitchener, Angela Yiu, Benjamin Chang +34
cs.AIarXiv:2511.02824v22025Vision Is Not Overhead: One-Pass Block Drafting for Lossless Speculative Decoding in Vision-Language Models
Jungseob Lee, Seongtae Hong, Dongyub Jude Lee +4
cs.AIcs.CLcs.CVarXiv:2609.00355v12026Deformable ProtoPNet: An Interpretable Image Classifier Using Deformable Prototypes
Jon Donnelly, Alina Jade Barnett, Chaofan Chen
cs.CVcs.AIcs.LGarXiv:2111.15000v32021DiTAR: Diffusion Transformer Autoregressive Modeling for Speech Generation
Dongya Jia, Zhuo Chen, Jiawei Chen +8
eess.AScs.AIcs.CLarXiv:2502.03930v42025AdaCoT: Pareto-Optimal Adaptive Chain-of-Thought Triggering via Reinforcement Learning
Chenwei Lou, Zewei Sun, Xinnian Liang +6
cs.LGcs.AIarXiv:2505.11896v22025Chain-of-Agents: End-to-End Agent Foundation Models via Multi-Agent Distillation and Agentic RL
Weizhen Li, Jianbo Lin, Zhuosong Jiang +27
cs.AIcs.CLarXiv:2508.13167v12025SafeArena: Evaluating the Safety of Autonomous Web Agents
Ada Defne Tur, Nicholas Meade, Xing Han Lù +6
cs.LGcs.AIcs.CLarXiv:2503.04957v12025Token Assorted: Mixing Latent and Text Tokens for Improved Language Model Reasoning
DiJia Su, Hanlin Zhu, Yingchen Xu +3
cs.CLcs.AIcs.LGarXiv:2502.03275v22025AdaEvolve: Adaptive LLM Driven Zeroth-Order Optimization
Mert Cemri, Shubham Agrawal, Akshat Gupta +9
cs.NEcs.AIcs.CLarXiv:2602.20133v12026MMC: Advancing Multimodal Chart Understanding with Large-scale Instruction Tuning
Fuxiao Liu, Xiaoyang Wang, Wenlin Yao +5
cs.CLcs.AIarXiv:2311.10774v22023SEED-GRPO: Semantic Entropy Enhanced GRPO for Uncertainty-Aware Policy Optimization
Minghan Chen, Guikun Chen, Wenguan Wang +1
cs.AIarXiv:2505.12346v12025S*: Test Time Scaling for Code Generation
Dacheng Li, Shiyi Cao, Chengkun Cao +6
cs.LGcs.AIarXiv:2502.14382v12025BeamDojo: Learning Agile Humanoid Locomotion on Sparse Footholds
Huayi Wang, Zirui Wang, Junli Ren +4
cs.ROcs.AIcs.LGarXiv:2502.10363v32025