Artificial Intelligence
Papers filed under cs.AI on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.
Search paper metadata (including unsummarized papers)
1,021 to 1,080 of 15,283
Explainable AI for clinical and remote health applications: a survey on tabular and time series data
Flavio Di Martino, Franca Delmastro
cs.LGcs.AIarXiv:2209.06528v12022VitaBench: Benchmarking LLM Agents with Versatile Interactive Tasks in Real-world Applications
Wei He, Yueqing Sun, Hongyan Hao +13
cs.CLcs.AIarXiv:2509.26490v22025Stabilizing Reinforcement Learning with LLMs: Formulation and Practices
Chujie Zheng, Kai Dang, Bowen Yu +10
cs.LGcs.AIcs.CLarXiv:2512.01374v32025Beyond Visual Quality: Evaluating Physical Consistency under Ego-Motion with EgoGenEval
Yilin Long, Chenming Zhu, Zitang Gou +2
cs.CVcs.AIarXiv:2609.11172v12026Your Agent May Misevolve: Emergent Risks in Self-evolving LLM Agents
Shuai Shao, Qihan Ren, Chen Qian +8
cs.AIcs.CLcs.LGarXiv:2509.26354v22025Beyond Benchmarks: Using VLMs to Reveal Systematic Classification Failures Under Real World Conditions
Dieuwertje Alblas, Alma M. Liezenga, Jan Erik van Woerden +3
cs.CVcs.AIcs.ETarXiv:2609.11126v12026Scaling LLM Multi-turn RL with End-to-end Summarization-based Context Management
Miao Lu, Weiwei Sun, Weihua Du +4
cs.CLcs.AIcs.LGarXiv:2510.06727v12025Toward Interpretable Multimodal Fusion: Heat Conduction Modeling for Hyperspectral and LiDAR Joint Classification
Kan Wei, Jiahui Cui, Jing Yao +3
cs.CVcs.AIarXiv:2609.11040v12026Premise Selection for Theorem Proving by Deep Graph Embedding
Mingzhe Wang, Yihe Tang, Jian Wang +1
cs.AIcs.LGcs.LOarXiv:1709.09994v12017AI and 6G into the Metaverse: Fundamentals, Challenges and Future Research Trends
Muhammad Zawish, Fayaz Ali Dharejo, Sunder Ali Khowaja +4
cs.AIcs.HCcs.NIarXiv:2208.10921v22022KVCOMM: Online Cross-context KV-cache Communication for Efficient LLM-based Multi-agent Systems
Hancheng Ye, Zhengqi Gao, Mingyuan Ma +8
cs.MAcs.AIstat.MLarXiv:2510.12872v22025Black-Box On-Policy Distillation of Large Language Models
Tianzhu Ye, Li Dong, Zewen Chi +3
cs.CLcs.AIarXiv:2511.10643v32025Prosperity before Collapse: How Far Can Off-Policy RL Reach with Stale Data on LLMs?
Haizhong Zheng, Jiawei Zhao, Beidi Chen
cs.LGcs.AIarXiv:2510.01161v22025Conditional LSTM-GAN for Melody Generation from Lyrics
Yi Yu, Abhishek Srivastava, Simon Canales
cs.AIcs.SDeess.ASarXiv:1908.05551v22019No Prompt Left Behind: Exploiting Zero-Variance Prompts in LLM Reinforcement Learning via Entropy-Guided Advantage Shaping
Thanh-Long V. Le, Myeongho Jeon, Kim Vu +2
cs.CLcs.AIcs.LGarXiv:2509.21880v32025SWE-EVO: Benchmarking Coding Agents in Long-Horizon Software Evolution Scenarios
Tue Le, Minh V. T. Thai, Dung Nguyen Manh +2
cs.SEcs.AIcs.MAarXiv:2512.18470v62025From $f(x)$ and $g(x)$ to $f(g(x))$: LLMs Learn New Skills in RL by Composing Old Ones
Lifan Yuan, Weize Chen, Yuchen Zhang +7
cs.AIcs.CLarXiv:2509.25123v32025Explainability Assistant: A Conversational XAI Interface for Interpreting Energy Consumption Models
Rodion Krjutškov, Eduard Barbu, Nikos Sakkas +1
cs.AIcs.LGarXiv:2609.11860v12026Step-Audio-R1 Technical Report
Fei Tian, Xiangyu Tony Zhang, Yuxin Zhang +14
cs.AIcs.CLcs.SDarXiv:2511.15848v22025Generative Marketing Mix Modeling: A Causal Inference Framework Linking GEO and GEM to Business Impact
Masahiro Kato, Daiki Honma, Taka Kato
stat.MLcs.AIcs.LGarXiv:2609.11915v12026Language Models are Injective and Hence Invertible
Giorgos Nikolaou, Tommaso Mencattini, Donato Crisostomi +3
cs.LGcs.AIarXiv:2510.15511v42025TiDAR: Think in Diffusion, Talk in Autoregression
Jingyu Liu, Xin Dong, Zhifan Ye +6
cs.CLcs.AIarXiv:2511.08923v12025Large Language Model Hacking: Quantifying the Hidden Risks of Using LLMs for Text Annotation
Joachim Baumann, Paul Röttger, Aleksandra Urman +4
cs.CLcs.AIcs.LGarXiv:2509.08825v22025ORCH: Organizational Principles Enable Collective Intelligence in Embodied AI
Zhengran Ji, Jonathan Hyun, Boyuan Chen
cs.MAcs.AIcs.LGarXiv:2609.11737v12026Logit Refiner: Improving Visual Autoregressive Models via Intra-Scale Dependency Modeling
Meimingwei Li, Stefan Andreas Baumann, Felix Krause +1
cs.CVcs.AIcs.LGarXiv:2609.11804v12026AMO-Bench: Large Language Models Still Struggle in High School Math Competitions
Shengnan An, Xunliang Cai, Xuezhi Cao +8
cs.CLcs.AIarXiv:2510.26768v12025UniME-V2: MLLM-as-a-Judge for Universal Multimodal Embedding Learning
Tiancheng Gu, Kaicheng Yang, Kaichen Zhang +6
cs.CVcs.AIarXiv:2510.13515v32025Evaluating Gemini Robotics Policies in a Veo World Simulator
Gemini Robotics Team, Krzysztof Choromanski, Coline Devin +20
cs.ROcs.AIcs.CVarXiv:2512.10675v22025Dream2Flow: Bridging Video Generation and Open-World Manipulation with 3D Object Flow
Karthik Dharmarajan, Wenlong Huang, Jiajun Wu +2
cs.ROcs.AIcs.CVarXiv:2512.24766v12025VeRO: A Harness for Agents to Optimize Agents
Varun Ursekar, Apaar Shanker, Veronica Chatrath +2
cs.AIcs.CLcs.LGarXiv:2602.22480v42026ZipCodec: Ultra-Low-Frame-Rate Streaming Speech Coding
Luca Della Libera, Cem Subakan, Mirco Ravanelli
cs.SDcs.AIcs.LGarXiv:2609.11642v12026Great Models Think Alike and this Undermines AI Oversight
Shashwat Goel, Joschka Struber, Ilze Amanda Auzina +6
cs.LGcs.AIcs.CLarXiv:2502.04313v22025WebInject: Prompt Injection Attack to Web Agents
Xilong Wang, John Bloch, Zedian Shao +3
cs.LGcs.AIcs.CLarXiv:2505.11717v42025Distributed Optimization of Modular Production Systems using Model-based Reinforcement Learning with Inverse Models
Andreas Schwung, Steve Yuwono, Sofiene Lassoued +1
cs.AIcs.LGeess.SYarXiv:2609.11615v12026Conformity Assessments and Post-market Monitoring: A Guide to the Role of Auditing in the Proposed European AI Regulation
Jakob Mokander, Maria Axente, Federico Casolari +1
cs.CYcs.AIarXiv:2111.05071v12021ExGRPO: Learning to Reason from Experience
Runzhe Zhan, Yafu Li, Zhi Wang +5
cs.LGcs.AIcs.CLarXiv:2510.02245v22025UME-R1: Exploring Reasoning-Driven Generative Multimodal Embeddings
Zhibin Lan, Liqiang Niu, Fandong Meng +2
cs.LGcs.AIarXiv:2511.00405v22025FineVision: Open Data Is All You Need
Luis Wiedmann, Orr Zohar, Amir Mahla +6
cs.CVcs.AIarXiv:2510.17269v22025Published Unlearning Numbers Move Per Checkpoint, and Not Because the Removed Data Survives: An Audit of 263 Released Batch-Normalized Checkpoints
Junlong Shen Xingyu Li
cs.AIcs.LGarXiv:2609.11490v12026Enabling Knowledge Graph Understanding at Scale with the EXplore Your Graphs ENgine (EXYGEN)
Harshdeep Singh, Yurui Zhu, Giovanni Colavizza +1
cs.AIcs.LGarXiv:2609.11569v12026The Path Not Taken: RLVR Provably Learns Off the Principals
Hanqing Zhu, Zhenyu Zhang, Hanxian Huang +11
cs.LGcs.AIarXiv:2511.08567v12025Think with 3D: Geometric Imagination Grounded Spatial Reasoning from Limited Views
Zhangquan Chen, Manyuan Zhang, Xinlei Yu +8
cs.CVcs.AIarXiv:2510.18632v42025Your Model Already Knows Don't Teach It, Learn to Ask It: Soft Prompting for Few-Shot Adaptation of Vision-Language Models
Gautam Rajendrakumar Gare, Siyi Li, Hewei Wang +5
cs.CVcs.AIcs.LGarXiv:2609.11310v12026General Agentic Memory Via Deep Research
B. Y. Yan, Chaofan Li, Hongjin Qian +2
cs.CLcs.AIcs.IRarXiv:2511.18423v12025Open-o3-Video: Grounded Video Reasoning with Explicit Spatio-Temporal Evidence
Jiahao Meng, Xiangtai Li, Haochen Wang +8
cs.CVcs.AIcs.MMarXiv:2510.20579v22025OmniVideoBench: Towards Audio-Visual Understanding Evaluation for Omni MLLMs
Caorui Li, Yu Chen, Yiyan Ji +40
cs.AIarXiv:2510.10689v32025Predicting Train Delays in Finland Using Machine Learning and Weather Data
Vinicius Pozzobon Borin, Jean Michel de Souza Sant'Ana, Nurul Huda Mahmood
cs.AIcs.LGarXiv:2609.11277v12026Improving Faint Object Detection for Space Situational Awareness with Variational Autoencoders
Angela Cratere, Luca Ghilardi, Vishnu Reddy +3
cs.CVcs.AIcs.LGarXiv:2609.11269v12026Bio-inspired Learning and Decision-Making with Probabilistic In-Memory Computing Hardware: Part 1
Thomas Dalgaty, Eiji Kawasaki, Miguel de Prado +2
cs.AIcs.LGarXiv:2609.11281v12026Kangaroo: A Powerful Video-Language Model Supporting Long-context Video Input
Jiajun Liu, Yibing Wang, Hanghang Ma +6
cs.CVcs.AIcs.MMarXiv:2408.15542v12024SIM-CoT: Supervised Implicit Chain-of-Thought
Xilin Wei, Xiaoran Liu, Yuhang Zang +5
cs.CLcs.AIarXiv:2509.20317v22025Generative Replay Mitigates Sample Starvation in Quantum Architecture Search
Akash Kundu, Amit Kumar Jaiswal, Sebastian Feld +1
quant-phcs.AIcs.ETarXiv:2609.11248v12026CryptoL: Towards Scale Dominance and Physics Constraints Mitigation in Financial Multivariate Time Series Forecasting
Yalda Taheri, Mohammad Hassan Heydari, Armon Rasooli +3
cs.AIcs.CEcs.LGarXiv:2609.11206v12026WMPO: World Model-based Policy Optimization for Vision-Language-Action Models
Fangqi Zhu, Zhengyang Yan, Zicong Hong +3
cs.ROcs.AIarXiv:2511.09515v12025Training Deeper Neural Machine Translation Models with Transparent Attention
Ankur Bapna, Mia Xu Chen, Orhan Firat +2
cs.CLcs.AIcs.LGarXiv:1808.07561v22018VLA-0: Building State-of-the-Art VLAs with Zero Modification
Ankit Goyal, Hugo Hadfield, Xuning Yang +2
cs.ROcs.AIarXiv:2510.13054v12025Are We Really Doing Few-Shot Learning? A Critical Examination of Pre-Training Assumptions
Alejandro Galan-Cuenca, Marcelo Saval-Calvo, Antonio Javier Gallego
cs.CVcs.AIcs.LGarXiv:2609.10851v12026DriftNet: A Dual-Head Trajectory Transformer for Detecting and Localizing Prompt Injection in LLM Agents
Asif Pinjari, Mithun Paul Saint-Germain
cs.CRcs.AIcs.LGarXiv:2609.10892v12026TOUCAN: Synthesizing 1.5M Tool-Agentic Data from Real-World MCP Environments
Zhangchen Xu, Adriana Meza Soria, Shawn Tan +4
cs.LGcs.AIcs.CLarXiv:2510.01179v12025SLA: Beyond Sparsity in Diffusion Transformers via Fine-Tunable Sparse-Linear Attention
Jintao Zhang, Haoxu Wang, Kai Jiang +10
cs.LGcs.AIcs.CVarXiv:2509.24006v22025