Artificial Intelligence

Papers filed under cs.AI on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.

Search paper metadata (including unsummarized papers)

8,521 to 8,580 of 15,247

  1. openJiuwen: Beyond Static Harnesses for Long-Horizon Coding Agents

    openJiuwen Team, Tao Yu, Xinyu Zhang +17

    cs.AIarXiv:2608.27969v12026
  2. AERA: Adaptive Evidence Residual Allocation for Efficient Test-Time Reasoning

    Ziming Wang, Ivor Tsang, Hangwei Qian

    cs.AIarXiv:2608.27964v12026
  3. The Illusion of $\textit{What If}$: Evaluating the Breakdown of Counterfactual Reasoning in LLMs

    Yucheng Wang, Yuetian Du, Zhengyi Liu +8

    cs.AIarXiv:2608.27953v12026
  4. Cross-Session Decomposition Attacks: Scaling Risk and Intent-Aligned Retrieval Defense

    Disen Liao, Yihan Wang, Freda Shi +1

    cs.AIarXiv:2608.27945v12026
  5. Optimizing Instructions and Demonstrations for Multi-Stage Language Model Programs

    Krista Opsahl-Ong, Michael J Ryan, Josh Purtell +4

    cs.CLcs.AIcs.LGarXiv:2406.11695v22024
  6. CASTANET: Causality-Aware Spatio-Temporal Adversarial Network Using Traffic Incident Effects

    Toshiya Kitahara, Ryu Shirakami, Koh Takeuchi +1

    cs.AIarXiv:2608.27942v12026
  7. AutoAgents: A Framework for Automatic Agent Generation

    Guangyao Chen, Siwei Dong, Yu Shu +5

    cs.AIarXiv:2309.17288v32023
  8. A Deep Learning-Based Stacking Ensemble Framework for Turbofan Engine Remaining Useful Life Prediction

    Limon Bin Hossain, Md. Salehin Seyam, Md Rashedul Islam +2

    cs.AIarXiv:2608.27940v12026
  9. RoCo: Dialectic Multi-Robot Collaboration with Large Language Models

    Zhao Mandi, Shreeya Jain, Shuran Song

    cs.ROcs.AIcs.LGarXiv:2307.04738v12023
  10. From Documents to Reasoning: A Validated Synthetic Data Pipeline and Semantic-Aware Fine-Tuning for Financial Numerical Reasoning

    Lokendra Birla, Milind Savagaonkar, Visnu Srinivasan +2

    cs.AIarXiv:2608.27919v12026
  11. Resource Constraints and Performance in Agentic AI Systems

    Amaz Salman, Malka Halgamuge, Teo Susnjak

    cs.AIarXiv:2608.27886v12026
  12. CoRe-MoE: Compact Reusable MoE for Continual Multimodal Instruction Tuning

    Runze Liu, Naibin Gu, Mingxu Ai +4

    cs.AIarXiv:2608.27867v12026
  13. KLOD: Locality-Preserving Knowledge Editing via Non-Target Distribution Preservation

    Hojun Jeong, Gyunyeop Kim, Sangwoo Kang

    cs.AIarXiv:2608.27839v12026
  14. explAIner: A Visual Analytics Framework for Interactive and Explainable Machine Learning

    Thilo Spinner, Udo Schlegel, Hanna Schäfer +1

    cs.HCcs.AIcs.LGarXiv:1908.00087v22019
  15. An Empirical Evaluation of Cross-City POI Recommendation on a Large-Scale Benchmark

    Peibo Li, Yang Song, Hao Xue +2

    cs.AIcs.IRarXiv:2608.27840v12026
  16. ArtPrompt: ASCII Art-based Jailbreak Attacks against Aligned LLMs

    Fengqing Jiang, Zhangchen Xu, Luyao Niu +4

    cs.CLcs.AIarXiv:2402.11753v42024
  17. AI Alignment through a Game-theoretic Lens: A Survey

    Yanan Cai, Zhongrui Zhao, Zhigang Lu +6

    cs.AIcs.CLcs.GTarXiv:2608.27910v12026
  18. Assessing the Scalability of Biologically-Motivated Deep Learning Algorithms and Architectures

    Sergey Bartunov, Adam Santoro, Blake A. Richards +3

    cs.LGcs.AIcs.NEarXiv:1807.04587v22018
  19. ToxicChat: Unveiling Hidden Challenges of Toxicity Detection in Real-World User-AI Conversation

    Zi Lin, Zihan Wang, Yongqi Tong +4

    cs.CLcs.AIarXiv:2310.17389v12023
  20. HyQuant: Hybrid-Precision Quantization for LLM Attention

    Jiatong Ding, Bingxin Xing, Yu Zhang +9

    cs.AIarXiv:2608.27875v12026
  21. Evaluating human and LLM screening workflows in a conceptually complex scoping review: Recall--workload trade-offs and run-to-run consistency

    Nikol Figalová, Lynn Huestegge, Anne Böckler-Raettig

    cs.AIcs.HCcs.SEarXiv:2608.26885v12026
  22. Neural-Bayesian Structure Learning for Discrete Choice Modeling

    Hyunsoo Yun, Eun Hak Lee, Jiaru Zhang +2

    cs.LGcs.AIarXiv:2608.25258v12026
  23. SDXL-Lightning: Progressive Adversarial Diffusion Distillation

    Shanchuan Lin, Anran Wang, Xiao Yang

    cs.CVcs.AIcs.LGarXiv:2402.13929v32024
  24. See, Hypothesize, Validate: Multimodal Agentic Framework for Discovering Governing PDEs

    Sarang Manoj Pekhale, Amartya Roy, Rajat Sarkar +1

    cs.AIarXiv:2608.27869v12026
  25. Interpretable AI predicts a 2026 summer dry anomaly in central China

    Anran Wang, Wen Shi, Yong Luo +5

    physics.ao-phcs.AIarXiv:2608.19163v12026
  26. A Regulatory Placebo? The Systemic Failure of Mandatory GenAI Labeling

    Jingyi Chen, Chaofan Bu, Shibo Yan +1

    cs.CYcs.AIarXiv:2608.16470v12026
  27. P2E-VQ: ECG-linked representation augmentation for PPG via discrete patch retrieval

    Zhongli Wu, Zhuangzhi Gao, He Zhao +8

    cs.LGcs.AIarXiv:2608.14656v12026
  28. An analysis of incorporating an external language model into a sequence-to-sequence model

    Anjuli Kannan, Yonghui Wu, Patrick Nguyen +3

    eess.AScs.AIcs.CLarXiv:1712.01996v12017
  29. Large Language Models and their Awareness of Mechanics and Spatial Geometry

    Johannes Gerstmayr, Sebastian Weyrer, Tobias Möltner +2

    cs.AIarXiv:2608.14615v12026
  30. Artificial Intelligence Index Report 2026

    Sha Sajadieh, Loredana Fattorini, Raymond Perrault +20

    cs.AIarXiv:2606.15708v32026
    Summaries:한국어
  31. Secrets of RLHF in Large Language Models Part I: PPO

    Rui Zheng, Shihan Dou, Songyang Gao +24

    cs.CLcs.AIcs.LGarXiv:2307.04964v22023
  32. SpikeOPD: Stable On-Policy Distillation for Autoregressive Spiking Language Models

    Enqiao Lu, Xingrui Yu, Yiwei Fu +7

    cs.AIarXiv:2608.27857v12026
  33. Found-RL: foundation model-enhanced reinforcement learning for autonomous driving

    Yansong Qu, Zihao Sheng, Zilin Huang +6

    cs.AIcs.LGarXiv:2602.10458v12026
  34. Testing of Detection Tools for AI-Generated Text

    Debora Weber-Wulff, Alla Anohina-Naumeca, Sonja Bjelobaba +5

    cs.CLcs.AIcs.CYarXiv:2306.15666v22023
  35. BiomedGPT: A Generalist Vision-Language Foundation Model for Diverse Biomedical Tasks

    Kai Zhang, Rong Zhou, Eashan Adhikarla +20

    cs.CLcs.AIarXiv:2305.17100v42023
  36. From Uncertainty to Clinical Risk: Severity-Aware Conformal Planning for Interactive Medical Diagnosis

    Yue Zhou, Haiyang Zhou, Jin Zhang +4

    cs.AIarXiv:2608.27847v12026
  37. Does Localization Inform Editing? Surprising Differences in Causality-Based Localization vs. Knowledge Editing in Language Models

    Peter Hase, Mohit Bansal, Been Kim +1

    cs.LGcs.AIcs.CLarXiv:2301.04213v22023
  38. The GEM Benchmark: Natural Language Generation, its Evaluation and Metrics

    Sebastian Gehrmann, Tosin Adewumi, Karmanya Aggarwal +53

    cs.CLcs.AIcs.LGarXiv:2102.01672v32021
  39. A Survey on Ensemble Learning under the Era of Deep Learning

    Yongquan Yang, Haijun Lv, Ning Chen

    cs.LGcs.AIarXiv:2101.08387v62021
  40. Integrated Task and Motion Planning

    Caelan Reed Garrett, Rohan Chitnis, Rachel Holladay +4

    cs.ROcs.AIarXiv:2010.01083v12020
  41. Worldwide AI Ethics: a review of 200 guidelines and recommendations for AI governance

    Nicholas Kluge Corrêa, Camila Galvão, James William Santos +8

    cs.CYcs.AIarXiv:2206.11922v72022
  42. Ethical Machine Learning in Health Care

    Irene Y. Chen, Emma Pierson, Sherri Rose +3

    cs.CYcs.AIcs.LGarXiv:2009.10576v32020
  43. PyTorch-BigGraph: A Large-scale Graph Embedding System

    Adam Lerer, Ledell Wu, Jiajun Shen +4

    cs.LGcs.AIcs.DCarXiv:1903.12287v32019
  44. Explanation in Human-AI Systems: A Literature Meta-Review, Synopsis of Key Ideas and Publications, and Bibliography for Explainable AI

    Shane T. Mueller, Robert R. Hoffman, William Clancey +2

    cs.AIarXiv:1902.01876v12019
  45. A Comprehensive Survey of Data Mining-based Fraud Detection Research

    Clifton Phua, Vincent Lee, Kate Smith +1

    cs.AIcs.CEarXiv:1009.6119v12010
  46. Deep Whole-Body Control: Learning a Unified Policy for Manipulation and Locomotion

    Zipeng Fu, Xuxin Cheng, Deepak Pathak

    cs.ROcs.AIcs.CVarXiv:2210.10044v12022
  47. ReToolSQL: Agentic Reinforcement Learning for Robust Text-to-SQL

    Pratik Kakkar, Chandra Dhir, Ravi Shankar +2

    cs.AIarXiv:2608.27796v12026
  48. CURA: Certified Runtime Alarms for Computer-Use Agents

    Divake Kumar, Sina Tayebati, Devashri Naik +5

    cs.AIcs.CVcs.LGarXiv:2608.27808v12026
  49. Credo: Reusable Declarative Primitives for Agentic Workflows

    Duo Lu, Andrew Crotty, Uğur Çetintemel

    cs.AIcs.DBarXiv:2608.27790v12026
  50. RealSWE: A Compositional Evaluation of Coding Agents under Realistic User Requests

    Gyuhyeong Kim, Hyojung Gwon, Jeonghyeon Kim +2

    cs.AIcs.LGarXiv:2608.27831v12026
  51. Evidential-Based Higher-Order Set Argumentation Framework

    Shuai Tang

    cs.AImath.LOarXiv:2608.27824v12026
  52. AcCoRD: Evaluating User-Agent Collaboration Under Realistic User Preference Dynamics

    Tejas Srinivasan, Shikib Mehri, Nandita Shankar Naik +3

    cs.AIarXiv:2608.27818v12026
  53. In-context Vectors: Making In Context Learning More Effective and Controllable Through Latent Space Steering

    Sheng Liu, Haotian Ye, Lei Xing +1

    cs.LGcs.AIcs.CLarXiv:2311.06668v32023
  54. GraphDF: A Discrete Flow Model for Molecular Graph Generation

    Youzhi Luo, Keqiang Yan, Shuiwang Ji

    cs.LGcs.AIarXiv:2102.01189v22021
  55. ManiSkill3: GPU Parallelized Robotics Simulation and Rendering for Generalizable Embodied AI

    Stone Tao, Fanbo Xiang, Arth Shukla +20

    cs.ROcs.AIarXiv:2410.00425v22024
  56. Probing Perceptual Priors of MLLMs via Gibbs Sampling with Interpretable Generative Controls

    Manuel Cherep, Pattie Maes, Nikhil Singh

    cs.AIarXiv:2608.27727v12026
  57. PCFBench: A Diagnostic Benchmark for Product Carbon Footprint Estimation

    Krishna Rao, Andrew Dumit, Shaena Ulissi +7

    cs.AIarXiv:2608.27716v12026
  58. C3: Zero-shot Text-to-SQL with ChatGPT

    Xuemei Dong, Chao Zhang, Yuhang Ge +5

    cs.CLcs.AIarXiv:2307.07306v12023
  59. LongGuard: Mechanistic Analysis and Training-Free Mitigation of Long-Context Failure in Safety Guardrails

    Ziyang Chen, Xing Wu, Songlin Hu

    cs.AIarXiv:2608.27580v12026
  60. If Agents Were Angels, No Governance Would Be Necessary: Out-of-Band Policy Enforcement at a Trusted Tool Boundary

    Marc Millstone, Tyler Akidau, Johannes Brüderl +1

    cs.AIarXiv:2608.27646v12026