Artificial Intelligence

Papers filed under cs.AI on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.

Search paper metadata (including unsummarized papers)

4,141 to 4,200 of 15,395

  1. AI Literacy in K-12 and Higher Education in the Wake of Generative AI: An Integrative Review

    Xingjian Gu, Barbara J. Ericson

    cs.CYcs.AIarXiv:2503.00079v32025
  2. Kernel Language Entropy: Fine-grained Uncertainty Quantification for LLMs from Semantic Similarities

    Alexander Nikitin, Jannik Kossen, Yarin Gal +1

    cs.LGcs.AIcs.CLarXiv:2405.20003v12024
  3. Formal Policy Enforcement for Real-World Agentic Systems

    Nils Palumbo, Sarthak Choudhary, Jihye Choi +3

    cs.CRcs.AIcs.MAarXiv:2602.16708v32026
  4. Neural Language Modeling by Jointly Learning Syntax and Lexicon

    Yikang Shen, Zhouhan Lin, Chin-Wei Huang +1

    cs.CLcs.AIarXiv:1711.02013v22017
  5. MinTL: Minimalist Transfer Learning for Task-Oriented Dialogue Systems

    Zhaojiang Lin, Andrea Madotto, Genta Indra Winata +1

    cs.CLcs.AIarXiv:2009.12005v22020
  6. MedHallu: A Comprehensive Benchmark for Detecting Medical Hallucinations in Large Language Models

    Shrey Pandit, Jiawei Xu, Junyuan Hong +4

    cs.CLcs.AIcs.LGarXiv:2502.14302v12025
  7. Meta-Learning by Adjusting Priors Based on Extended PAC-Bayes Theory

    Ron Amit, Ron Meir

    stat.MLcs.AIcs.LGarXiv:1711.01244v82017
  8. A Survey on the Optimization of Large Language Model-based Agents

    Shangheng Du, Jiabao Zhao, Jinxin Shi +4

    cs.AIarXiv:2503.12434v22025
  9. R1-ShareVL: Incentivizing Reasoning Capability of Multimodal Large Language Models via Share-GRPO

    Huanjin Yao, Qixiang Yin, Jingyi Zhang +8

    cs.CVcs.AIcs.CLarXiv:2505.16673v12025
  10. Perceptive Humanoid Parkour: Chaining Dynamic Human Skills via Motion Matching

    Zhen Wu, Xiaoyu Huang, Lujie Yang +8

    cs.ROcs.AIcs.LGarXiv:2602.15827v22026
  11. The Assistant's Ideal Self

    Mert Yazan

    cs.AIarXiv:2609.00304v12026
  12. When Single-Agent with Skills Replace Multi-Agent Systems and When They Fail

    Xiaoxiao Li

    cs.AIcs.MAarXiv:2601.04748v22026
  13. BRIGHT: A Realistic and Challenging Benchmark for Reasoning-Intensive Retrieval

    Hongjin Su, Howard Yen, Mengzhou Xia +12

    cs.CLcs.AIcs.IRarXiv:2407.12883v42024
  14. Advancing Multi-Agent Systems Through Model Context Protocol: Architecture, Implementation, and Applications

    Naveen Krishnan

    cs.MAcs.AIarXiv:2504.21030v12025
  15. STAMP: Scalable Task And Model-agnostic Collaborative Perception

    Xiangbo Gao, Runsheng Xu, Jiachen Li +3

    cs.CVcs.AIcs.ROarXiv:2501.18616v12025
  16. Beyond Token Positions: Safety Alignment Across Denoising Steps in Diffusion Language Models

    Guoli Wang, Haonan Shi, Tu Ouyang +1

    cs.CLcs.AIarXiv:2609.00495v12026
  17. Operational Regimes in Non-Convex Optimization: A Multiplier-Based Taxonomy

    Seyed Mohsen Kazemi, Ali Movaghar, Shaahin hessabi

    math.OCcs.AIeess.SParXiv:2609.00471v12026
  18. Securing Agentic AI: A Comprehensive Threat Model and Mitigation Framework for Generative AI Agents

    Vineeth Sai Narajala, Om Narayan

    cs.CRcs.AIarXiv:2504.19956v22025
  19. Does RAG Know When Retrieval Is Wrong? Diagnosing Context Compliance under Knowledge Conflict

    Yihang Chen, Pin Qian, Su Wang +4

    cs.CLcs.AIarXiv:2605.14473v42026
  20. Are Language Models Models?

    Philip Resnik

    cs.CLcs.AIarXiv:2601.10421v12026
  21. The Impact of Reasoning Step Length on Large Language Models

    Mingyu Jin, Qinkai Yu, Dong Shu +5

    cs.CLcs.AIarXiv:2401.04925v42024
  22. Life-Cycle Emissions of AI Hardware: A Cradle-To-Grave Approach and Generational Trends

    Ian Schneider, Hui Xu, Stephan Benecke +4

    cs.ARcs.AIarXiv:2502.01671v12025
  23. Centering before Pruning: Lightweight Geometry Correction for Diversity-Based Visual Token Pruning in LVLMs

    Shunjie Wen, Jaeyeon Lee, Dong-Wan Choi

    cs.CVcs.AIarXiv:2608.30263v12026
  24. Parameter-Efficient Fine-Tuning for Foundation Models

    Dan Zhang, Tao Feng, Lilong Xue +3

    cs.CLcs.AIcs.LGarXiv:2501.13787v12025
  25. NVIDIA FLARE: Federated Learning from Simulation to Real-World

    Holger R. Roth, Yan Cheng, Yuhong Wen +20

    cs.LGcs.AIcs.CVarXiv:2210.13291v32022
  26. EpaCache: Error-Propagation-Aware Caching for Accelerating Diffusion-Based Visual Generation

    Yuhan Liu, Zongwei Hong, Jinglun Li +3

    cs.AIarXiv:2608.29264v12026
  27. TIPS: Turn-Level Information-Potential Reward Shaping for Search-Augmented LLMs

    Yutao Xie, Nathaniel Thomas, Nicklas Hansen +3

    cs.CLcs.AIcs.LGarXiv:2603.22293v12026
  28. Augmenting Human Performance with an XR Agent Learning from Online Behavior and BCI Evidence

    Ziheng Li, Xichen He, Haoyan Chen +10

    cs.AIcs.HCarXiv:2608.30369v12026
  29. Incompressible Knowledge Probes: Estimating Black-Box LLM Parameter Counts via Factual Capacity

    Bojie Li

    cs.LGcs.AIarXiv:2604.24827v22026
  30. School of Reward Hacks: Hacking harmless tasks generalizes to misaligned behavior in LLMs

    Mia Taylor, James Chua, Jan Betley +2

    cs.AIarXiv:2508.17511v12025
  31. Thinking with Generated Images

    Ethan Chern, Zhulin Hu, Steffi Chern +5

    cs.CVcs.AIcs.CLarXiv:2505.22525v12025
  32. Astra: A Multi-Agent System for GPU Kernel Performance Optimization

    Anjiang Wei, Tianran Sun, Yogesh Seenichamy +5

    cs.DCcs.AIcs.CLarXiv:2509.07506v22025
  33. General In-Hand Object Rotation with Vision and Touch

    Haozhi Qi, Brent Yi, Sudharshan Suresh +4

    cs.ROcs.AIcs.CVarXiv:2309.09979v22023
  34. Adapting the Interface, Not the Model: Runtime Harness Adaptation for Deterministic LLM Agents

    Tianshi Xu, Huifeng Wen, Meng Li

    cs.AIarXiv:2605.22166v22026
  35. Position: Graph Learning Will Lose Relevance Due To Poor Benchmarks

    Maya Bechler-Speicher, Ben Finkelshtein, Fabrizio Frasca +9

    cs.LGcs.AIcs.NEarXiv:2502.14546v12025
  36. Towards Efficient Large Language Model Serving: A Survey on System-Aware KV Cache Optimization

    Jiantong Jiang, Peiyu Yang, Rui Zhang +1

    cs.LGcs.AIcs.CLarXiv:2607.08057v12026
  37. A Generative Deep Learning Approach to Stochastic Downscaling of Precipitation Forecasts

    Lucy Harris, Andrew T. T. McRae, Matthew Chantry +2

    physics.ao-phcs.AIcs.CVarXiv:2204.02028v22022
  38. More Perspectives, Stronger Signals: Multi-Perspective Enhancement and Progressive Fusion for Multimodal Entity Representation Learning

    Chenyi Xiong, Yan Zhang, Jing Hu +5

    cs.AIarXiv:2608.29139v12026
  39. Agent Behavioral Contracts: Formal Specification and Runtime Enforcement for Reliable Autonomous AI Agents

    Varun Pratap Bhardwaj

    cs.AIcs.MAcs.SEarXiv:2602.22302v12026
  40. Oculi: A Conversational Agentic Platform for Automated Credit Risk Analysis

    Vennise Ho, Kristian Diana, Sandy Mourad +3

    cs.AIcs.LGarXiv:2608.28944v12026
  41. Removing RLHF Protections in GPT-4 via Fine-Tuning

    Qiusi Zhan, Richard Fang, Rohan Bindu +3

    cs.CLcs.AIarXiv:2311.05553v32023
  42. Robin: A multi-agent system for automating scientific discovery

    Ali Essam Ghareeb, Benjamin Chang, Ludovico Mitchener +7

    cs.AIcs.MAq-bio.QMarXiv:2505.13400v12025
  43. Linear Mode Connectivity in Multitask and Continual Learning

    Seyed Iman Mirzadeh, Mehrdad Farajtabar, Dilan Gorur +2

    cs.LGcs.AIcs.CVarXiv:2010.04495v12020
  44. AOI-Net: Structural Face AOI-Guided Eye-Gaze Track Representation Learning for Autism Spectrum Disorder Detection

    Zhanpei Huang, Binbin Sun, Jialiang Chen +5

    cs.CVcs.AIcs.ETarXiv:2608.29289v12026
  45. OCGQuant: Outlier-Companion Grouping for NVFP4 Quantization

    Yishan Yao, Binjun Li, Hanling Yi +5

    cs.CLcs.AIcs.LGarXiv:2609.00066v12026
  46. Unsupervised Diverse Colorization via Generative Adversarial Networks

    Yun Cao, Zhiming Zhou, Weinan Zhang +1

    cs.CVcs.AIarXiv:1702.06674v22017
  47. Doc-to-LoRA: Learning to Instantly Internalize Contexts

    Rujikorn Charakorn, Edoardo Cetin, Shinnosuke Uesaka +1

    cs.CLcs.AIarXiv:2602.15902v12026
  48. LLaVE: Large Language and Vision Embedding Models with Hardness-Weighted Contrastive Learning

    Zhibin Lan, Liqiang Niu, Fandong Meng +2

    cs.CVcs.AIcs.CLarXiv:2503.04812v22025
  49. Getting pwn'd by AI: Penetration Testing with Large Language Models

    Andreas Happe, Jürgen Cito

    cs.CLcs.AIcs.CRarXiv:2308.00121v32023
  50. ContextAgent: Context-Aware Proactive LLM Agents with Open-World Sensory Perceptions

    Bufang Yang, Lilin Xu, Liekang Zeng +7

    cs.AIcs.CLcs.HCarXiv:2505.14668v22025
  51. Test-time regression: a unifying framework for designing sequence models with associative memory

    Ke Alexander Wang, Jiaxin Shi, Emily B. Fox

    cs.LGcs.AIcs.NEarXiv:2501.12352v32025
  52. DocIntent: Answerability-Guided Agentic Restoration for Real-World Document Visual Question Answering

    Zihan Huang, Shihang Wu, Junle Liu +4

    cs.CVcs.AIarXiv:2608.29037v12026
  53. From Untrusted Input to Trusted Memory: A Systematic Study of Memory Poisoning Attacks in LLM Agents

    Pritam Dash, Tongyu Ge, Aditi Jain +2

    cs.CRcs.AIarXiv:2606.04329v22026
  54. TAPAS: Thermal- and Power-Aware Scheduling for LLM Inference in Cloud Platforms

    Jovan Stojkovic, Chaojie Zhang, Íñigo Goiri +5

    cs.DCcs.AIarXiv:2501.02600v12025
  55. Not Safe for All: Auditing the Dialect Penalty in Text-to-Image Safety Pipelines

    Minkyu Kim, Juhwan Choi, YoungBin Kim

    cs.AIarXiv:2608.29589v12026
  56. Meta Context Engineering via Agentic Skill Evolution

    Haoran Ye, Xuning He, Vincent Arak +2

    cs.AIcs.NEarXiv:2601.21557v22026
  57. Large Language Models to Enhance Bayesian Optimization

    Tennison Liu, Nicolás Astorga, Nabeel Seedat +1

    cs.LGcs.AIarXiv:2402.03921v22024
  58. PyVision: Agentic Vision with Dynamic Tooling

    Shitian Zhao, Haoquan Zhang, Shaoheng Lin +4

    cs.CLcs.AIcs.CVarXiv:2507.07998v32025
  59. Constitutional Classifiers++: Efficient Production-Grade Defenses against Universal Jailbreaks

    Hoagy Cunningham, Jerry Wei, Zihan Wang +26

    cs.CRcs.AIarXiv:2601.04603v12026
  60. UserBench: An Interactive Gym Environment for User-Centric Agents

    Cheng Qian, Zuxin Liu, Akshara Prabhakar +9

    cs.AIcs.CLcs.LGarXiv:2507.22034v12025