Artificial Intelligence

Papers filed under cs.AI on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.

Search paper metadata (including unsummarized papers)

1,021 to 1,080 of 15,283

  1. Explainable AI for clinical and remote health applications: a survey on tabular and time series data

    Flavio Di Martino, Franca Delmastro

    cs.LGcs.AIarXiv:2209.06528v12022
  2. VitaBench: Benchmarking LLM Agents with Versatile Interactive Tasks in Real-world Applications

    Wei He, Yueqing Sun, Hongyan Hao +13

    cs.CLcs.AIarXiv:2509.26490v22025
  3. Stabilizing Reinforcement Learning with LLMs: Formulation and Practices

    Chujie Zheng, Kai Dang, Bowen Yu +10

    cs.LGcs.AIcs.CLarXiv:2512.01374v32025
  4. Beyond Visual Quality: Evaluating Physical Consistency under Ego-Motion with EgoGenEval

    Yilin Long, Chenming Zhu, Zitang Gou +2

    cs.CVcs.AIarXiv:2609.11172v12026
  5. Your Agent May Misevolve: Emergent Risks in Self-evolving LLM Agents

    Shuai Shao, Qihan Ren, Chen Qian +8

    cs.AIcs.CLcs.LGarXiv:2509.26354v22025
  6. Beyond Benchmarks: Using VLMs to Reveal Systematic Classification Failures Under Real World Conditions

    Dieuwertje Alblas, Alma M. Liezenga, Jan Erik van Woerden +3

    cs.CVcs.AIcs.ETarXiv:2609.11126v12026
  7. Scaling LLM Multi-turn RL with End-to-end Summarization-based Context Management

    Miao Lu, Weiwei Sun, Weihua Du +4

    cs.CLcs.AIcs.LGarXiv:2510.06727v12025
  8. Toward Interpretable Multimodal Fusion: Heat Conduction Modeling for Hyperspectral and LiDAR Joint Classification

    Kan Wei, Jiahui Cui, Jing Yao +3

    cs.CVcs.AIarXiv:2609.11040v12026
  9. Premise Selection for Theorem Proving by Deep Graph Embedding

    Mingzhe Wang, Yihe Tang, Jian Wang +1

    cs.AIcs.LGcs.LOarXiv:1709.09994v12017
  10. AI and 6G into the Metaverse: Fundamentals, Challenges and Future Research Trends

    Muhammad Zawish, Fayaz Ali Dharejo, Sunder Ali Khowaja +4

    cs.AIcs.HCcs.NIarXiv:2208.10921v22022
  11. KVCOMM: Online Cross-context KV-cache Communication for Efficient LLM-based Multi-agent Systems

    Hancheng Ye, Zhengqi Gao, Mingyuan Ma +8

    cs.MAcs.AIstat.MLarXiv:2510.12872v22025
  12. Black-Box On-Policy Distillation of Large Language Models

    Tianzhu Ye, Li Dong, Zewen Chi +3

    cs.CLcs.AIarXiv:2511.10643v32025
  13. Prosperity before Collapse: How Far Can Off-Policy RL Reach with Stale Data on LLMs?

    Haizhong Zheng, Jiawei Zhao, Beidi Chen

    cs.LGcs.AIarXiv:2510.01161v22025
  14. Conditional LSTM-GAN for Melody Generation from Lyrics

    Yi Yu, Abhishek Srivastava, Simon Canales

    cs.AIcs.SDeess.ASarXiv:1908.05551v22019
  15. No Prompt Left Behind: Exploiting Zero-Variance Prompts in LLM Reinforcement Learning via Entropy-Guided Advantage Shaping

    Thanh-Long V. Le, Myeongho Jeon, Kim Vu +2

    cs.CLcs.AIcs.LGarXiv:2509.21880v32025
  16. SWE-EVO: Benchmarking Coding Agents in Long-Horizon Software Evolution Scenarios

    Tue Le, Minh V. T. Thai, Dung Nguyen Manh +2

    cs.SEcs.AIcs.MAarXiv:2512.18470v62025
  17. From $f(x)$ and $g(x)$ to $f(g(x))$: LLMs Learn New Skills in RL by Composing Old Ones

    Lifan Yuan, Weize Chen, Yuchen Zhang +7

    cs.AIcs.CLarXiv:2509.25123v32025
  18. Explainability Assistant: A Conversational XAI Interface for Interpreting Energy Consumption Models

    Rodion Krjutškov, Eduard Barbu, Nikos Sakkas +1

    cs.AIcs.LGarXiv:2609.11860v12026
  19. Step-Audio-R1 Technical Report

    Fei Tian, Xiangyu Tony Zhang, Yuxin Zhang +14

    cs.AIcs.CLcs.SDarXiv:2511.15848v22025
  20. Generative Marketing Mix Modeling: A Causal Inference Framework Linking GEO and GEM to Business Impact

    Masahiro Kato, Daiki Honma, Taka Kato

    stat.MLcs.AIcs.LGarXiv:2609.11915v12026
  21. Language Models are Injective and Hence Invertible

    Giorgos Nikolaou, Tommaso Mencattini, Donato Crisostomi +3

    cs.LGcs.AIarXiv:2510.15511v42025
  22. TiDAR: Think in Diffusion, Talk in Autoregression

    Jingyu Liu, Xin Dong, Zhifan Ye +6

    cs.CLcs.AIarXiv:2511.08923v12025
  23. Large Language Model Hacking: Quantifying the Hidden Risks of Using LLMs for Text Annotation

    Joachim Baumann, Paul Röttger, Aleksandra Urman +4

    cs.CLcs.AIcs.LGarXiv:2509.08825v22025
  24. ORCH: Organizational Principles Enable Collective Intelligence in Embodied AI

    Zhengran Ji, Jonathan Hyun, Boyuan Chen

    cs.MAcs.AIcs.LGarXiv:2609.11737v12026
  25. Logit Refiner: Improving Visual Autoregressive Models via Intra-Scale Dependency Modeling

    Meimingwei Li, Stefan Andreas Baumann, Felix Krause +1

    cs.CVcs.AIcs.LGarXiv:2609.11804v12026
  26. AMO-Bench: Large Language Models Still Struggle in High School Math Competitions

    Shengnan An, Xunliang Cai, Xuezhi Cao +8

    cs.CLcs.AIarXiv:2510.26768v12025
  27. UniME-V2: MLLM-as-a-Judge for Universal Multimodal Embedding Learning

    Tiancheng Gu, Kaicheng Yang, Kaichen Zhang +6

    cs.CVcs.AIarXiv:2510.13515v32025
  28. Evaluating Gemini Robotics Policies in a Veo World Simulator

    Gemini Robotics Team, Krzysztof Choromanski, Coline Devin +20

    cs.ROcs.AIcs.CVarXiv:2512.10675v22025
  29. Dream2Flow: Bridging Video Generation and Open-World Manipulation with 3D Object Flow

    Karthik Dharmarajan, Wenlong Huang, Jiajun Wu +2

    cs.ROcs.AIcs.CVarXiv:2512.24766v12025
  30. VeRO: A Harness for Agents to Optimize Agents

    Varun Ursekar, Apaar Shanker, Veronica Chatrath +2

    cs.AIcs.CLcs.LGarXiv:2602.22480v42026
  31. ZipCodec: Ultra-Low-Frame-Rate Streaming Speech Coding

    Luca Della Libera, Cem Subakan, Mirco Ravanelli

    cs.SDcs.AIcs.LGarXiv:2609.11642v12026
  32. Great Models Think Alike and this Undermines AI Oversight

    Shashwat Goel, Joschka Struber, Ilze Amanda Auzina +6

    cs.LGcs.AIcs.CLarXiv:2502.04313v22025
  33. WebInject: Prompt Injection Attack to Web Agents

    Xilong Wang, John Bloch, Zedian Shao +3

    cs.LGcs.AIcs.CLarXiv:2505.11717v42025
  34. Distributed Optimization of Modular Production Systems using Model-based Reinforcement Learning with Inverse Models

    Andreas Schwung, Steve Yuwono, Sofiene Lassoued +1

    cs.AIcs.LGeess.SYarXiv:2609.11615v12026
  35. Conformity Assessments and Post-market Monitoring: A Guide to the Role of Auditing in the Proposed European AI Regulation

    Jakob Mokander, Maria Axente, Federico Casolari +1

    cs.CYcs.AIarXiv:2111.05071v12021
  36. ExGRPO: Learning to Reason from Experience

    Runzhe Zhan, Yafu Li, Zhi Wang +5

    cs.LGcs.AIcs.CLarXiv:2510.02245v22025
  37. UME-R1: Exploring Reasoning-Driven Generative Multimodal Embeddings

    Zhibin Lan, Liqiang Niu, Fandong Meng +2

    cs.LGcs.AIarXiv:2511.00405v22025
  38. FineVision: Open Data Is All You Need

    Luis Wiedmann, Orr Zohar, Amir Mahla +6

    cs.CVcs.AIarXiv:2510.17269v22025
  39. Published Unlearning Numbers Move Per Checkpoint, and Not Because the Removed Data Survives: An Audit of 263 Released Batch-Normalized Checkpoints

    Junlong Shen Xingyu Li

    cs.AIcs.LGarXiv:2609.11490v12026
  40. Enabling Knowledge Graph Understanding at Scale with the EXplore Your Graphs ENgine (EXYGEN)

    Harshdeep Singh, Yurui Zhu, Giovanni Colavizza +1

    cs.AIcs.LGarXiv:2609.11569v12026
  41. The Path Not Taken: RLVR Provably Learns Off the Principals

    Hanqing Zhu, Zhenyu Zhang, Hanxian Huang +11

    cs.LGcs.AIarXiv:2511.08567v12025
  42. Think with 3D: Geometric Imagination Grounded Spatial Reasoning from Limited Views

    Zhangquan Chen, Manyuan Zhang, Xinlei Yu +8

    cs.CVcs.AIarXiv:2510.18632v42025
  43. Your Model Already Knows Don't Teach It, Learn to Ask It: Soft Prompting for Few-Shot Adaptation of Vision-Language Models

    Gautam Rajendrakumar Gare, Siyi Li, Hewei Wang +5

    cs.CVcs.AIcs.LGarXiv:2609.11310v12026
  44. General Agentic Memory Via Deep Research

    B. Y. Yan, Chaofan Li, Hongjin Qian +2

    cs.CLcs.AIcs.IRarXiv:2511.18423v12025
  45. Open-o3-Video: Grounded Video Reasoning with Explicit Spatio-Temporal Evidence

    Jiahao Meng, Xiangtai Li, Haochen Wang +8

    cs.CVcs.AIcs.MMarXiv:2510.20579v22025
  46. OmniVideoBench: Towards Audio-Visual Understanding Evaluation for Omni MLLMs

    Caorui Li, Yu Chen, Yiyan Ji +40

    cs.AIarXiv:2510.10689v32025
  47. Predicting Train Delays in Finland Using Machine Learning and Weather Data

    Vinicius Pozzobon Borin, Jean Michel de Souza Sant'Ana, Nurul Huda Mahmood

    cs.AIcs.LGarXiv:2609.11277v12026
  48. Improving Faint Object Detection for Space Situational Awareness with Variational Autoencoders

    Angela Cratere, Luca Ghilardi, Vishnu Reddy +3

    cs.CVcs.AIcs.LGarXiv:2609.11269v12026
  49. Bio-inspired Learning and Decision-Making with Probabilistic In-Memory Computing Hardware: Part 1

    Thomas Dalgaty, Eiji Kawasaki, Miguel de Prado +2

    cs.AIcs.LGarXiv:2609.11281v12026
  50. Kangaroo: A Powerful Video-Language Model Supporting Long-context Video Input

    Jiajun Liu, Yibing Wang, Hanghang Ma +6

    cs.CVcs.AIcs.MMarXiv:2408.15542v12024
  51. SIM-CoT: Supervised Implicit Chain-of-Thought

    Xilin Wei, Xiaoran Liu, Yuhang Zang +5

    cs.CLcs.AIarXiv:2509.20317v22025
  52. Generative Replay Mitigates Sample Starvation in Quantum Architecture Search

    Akash Kundu, Amit Kumar Jaiswal, Sebastian Feld +1

    quant-phcs.AIcs.ETarXiv:2609.11248v12026
  53. CryptoL: Towards Scale Dominance and Physics Constraints Mitigation in Financial Multivariate Time Series Forecasting

    Yalda Taheri, Mohammad Hassan Heydari, Armon Rasooli +3

    cs.AIcs.CEcs.LGarXiv:2609.11206v12026
  54. WMPO: World Model-based Policy Optimization for Vision-Language-Action Models

    Fangqi Zhu, Zhengyang Yan, Zicong Hong +3

    cs.ROcs.AIarXiv:2511.09515v12025
  55. Training Deeper Neural Machine Translation Models with Transparent Attention

    Ankur Bapna, Mia Xu Chen, Orhan Firat +2

    cs.CLcs.AIcs.LGarXiv:1808.07561v22018
  56. VLA-0: Building State-of-the-Art VLAs with Zero Modification

    Ankit Goyal, Hugo Hadfield, Xuning Yang +2

    cs.ROcs.AIarXiv:2510.13054v12025
  57. Are We Really Doing Few-Shot Learning? A Critical Examination of Pre-Training Assumptions

    Alejandro Galan-Cuenca, Marcelo Saval-Calvo, Antonio Javier Gallego

    cs.CVcs.AIcs.LGarXiv:2609.10851v12026
  58. DriftNet: A Dual-Head Trajectory Transformer for Detecting and Localizing Prompt Injection in LLM Agents

    Asif Pinjari, Mithun Paul Saint-Germain

    cs.CRcs.AIcs.LGarXiv:2609.10892v12026
  59. TOUCAN: Synthesizing 1.5M Tool-Agentic Data from Real-World MCP Environments

    Zhangchen Xu, Adriana Meza Soria, Shawn Tan +4

    cs.LGcs.AIcs.CLarXiv:2510.01179v12025
  60. SLA: Beyond Sparsity in Diffusion Transformers via Fine-Tunable Sparse-Linear Attention

    Jintao Zhang, Haoxu Wang, Kai Jiang +10

    cs.LGcs.AIcs.CVarXiv:2509.24006v22025