Artificial Intelligence

Papers filed under cs.AI on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.

Search paper metadata (including unsummarized papers)

15,361 to 15,420 of 15,427

  1. Proxy-Validated LLM UX Micro-Simulations: An Artifact-First Protocol for Early-Stage Decision Support

    Alexandre Cristovão Maiorano

    cs.HCcs.AIcs.SEarXiv:2608.13563v12026
  2. Shift Aware Transfer Learning with Adaptive Dual-Encoder Fusion for PM Forecasting in Data-Limited Environments

    Shahab Band, Hamed Mohammadi

    cs.AIarXiv:2608.14456v12026
  3. The Architect: Interactive Visualization of Deep Learning Mathematics Directly in Microsoft Excel

    Mohammad Imrul Jubair, Tom Yeh

    cs.HCcs.AIarXiv:2608.13572v12026
  4. BGA: A noise-immune neural distillation framework for malicious signature extraction in high-entropy encrypted flows

    Sheng Hong, Yixuan Huang, Weiwei Jiang +3

    cs.CRcs.AIarXiv:2608.14126v12026
  5. A Year in LLM Serving: Workload Evolution, Caching and Load-Balancing

    William Nixon, Jon Durbin, Florian Standhartinger +2

    cs.AIarXiv:2608.13573v12026
  6. Benchmarking data-driven material models on the classic Treloar dataset

    Hagen Holthusen, Moritz Flaschel, Denisa Martonová +1

    cs.AIcs.CEarXiv:2608.14063v12026
  7. Sensor-Driven Mission Synthesis for UAV/UGV Swarms: A TB-CSPN Coordination Architecture with Hardware-Enforced Safety

    Uwe M. Borghoff, Paolo Bottoni, Remo Pareschi

    cs.AIeess.SYarXiv:2608.14306v12026
  8. IterCOMP: Reasoning-aware Adaptive Prompt Compression for Multi-hop Question Answering

    JungMin Yun, YoungBin Kim

    cs.CLcs.AIarXiv:2608.13588v12026
  9. Weak-to-Strong Generalization via Direct On-Policy Distillation

    Shiyuan Feng, Huan-ang Gao, Haohan Chi +7

    cs.LGcs.AIcs.CLarXiv:2607.05394v22026
    Summaries:한국어
  10. CORE: Contrastive Reflection Enables Rapid Improvements in Reasoning

    Linas Nasvytis, Simon Jerome Han, Ben Prystawski +3

    cs.AIarXiv:2605.28742v22026
    Summaries:한국어
  11. Context Aware AI Assistant and AR Interface for Lunar Extravehicular Activity (EVA) Procedural Guidance

    Rodrigo Gallardo, Qilmeg Doudatcz, Ganit Goldstein +7

    cs.HCcs.AIarXiv:2608.13589v12026
    Summaries:한국어
  12. Agentao: A Governed Local-First Runtime for Tool-Using LLM Agents

    Bo Jin, Qiang Jiao, Xin Tong

    cs.AIcs.MAarXiv:2608.13574v12026
  13. Regime-Conditional Verification: Correctness Estimation for Adapting and Monitoring Safety Classifiers

    Thiago Sandoval, Ufuk Topcu

    cs.AIcs.CLcs.CRarXiv:2608.14089v12026
  14. Depth-Aware Sensitivity Analysis of Mixture-of-Experts Models via Magnitude-Based Expert Masking

    Pradeep Kumar Sharma, Shantanu Godbole, Hritvik Shrivastava

    cs.AIarXiv:2608.13565v12026
  15. CForce: Boosting Parallel Decoding for dLLMs via Consistency Forcing

    Yuji Ren, Chenkai Xu, Zhuocheng Gong +2

    cs.LGcs.AIcs.CLarXiv:2608.13925v12026
  16. Inducing Reward-Free Judging Rubrics that Reduce Over-Crediting in Agent Evaluation

    Darragh Quinn, David Dylan, Roisin Healy +3

    cs.AIarXiv:2608.13564v12026
  17. Don't Claim Benchmark-Oriented Optimization Improves General Coding Capability -- Diverse Evaluation Is Required

    Egor Shibaev, Vera Kudrevskaia, Timur Galimzyanov +9

    cs.LGcs.AIcs.SEarXiv:2608.13566v12026
  18. Simulation-Driven Vehicular Traffic Data Augmentation: Extending Sensor Coverage Through Virtual Sensing

    Davide Andrea Guastella, Eladio Montero Porras, Evangelos Pournaras +1

    cs.AIarXiv:2608.13993v12026
  19. Think in Latent, Explain in Language: Self-Explainable Latent Reasoning

    Dayuan Zhao, Shengcao Cao, Yu-Xiong Wang +1

    cs.CLcs.AIcs.LGarXiv:2608.13570v12026
  20. ASSERT: A Measurement Pipeline for GenAI Audits

    Riccardo Fogliato, Abhinav Palia, Xiawei Wang +11

    cs.CLcs.AIcs.CYarXiv:2608.13840v12026
  21. Reinforcement Learning-Based Production Scheduling in an Industry-Based Coating Scenario Using the Digital Model Playground

    Arne Kröger, Ralf Buschermöhle, Wilhelm Hasselbring +1

    cs.AIarXiv:2608.14122v12026
  22. Nanbeige4.2-3B on Apple Silicon: Fixing Deployment Bugs and Decreasing Looped Transformer Memory Overhead

    John T. Halloran

    cs.AIcs.LGarXiv:2608.13987v12026
  23. Not All Tokens Are Equal: Inflation-Aware Routing for Agentic LLM Systems

    Heming Fu, Shan Lin, Qianqian Xie +1

    cs.CLcs.AIarXiv:2608.13571v12026
  24. Concept Guidance: Precise, Training-Free Latent Control for Text-to-Image Generation

    Nikolai Röhrich, Isabell Hans, Felix Krause +1

    cs.CVcs.AIcs.LGarXiv:2608.14172v12026
  25. From Inaudible Inputs to Model Failures: Low-Frequency Safety Risks in LALMs

    Yuanhe Zhang, Weiliu Wang, Jie Ren +7

    cs.SDcs.AIarXiv:2608.09158v12026
  26. Overcoming Shortcut Learning in Graph Neural Networks through Active Explanation Guidance

    Taraneh Younesian, Steve Azzolin, Antonio Longa +3

    cs.LGcs.AIarXiv:2608.14121v12026
  27. Interactive Analysis of Global Explanations using Aggregated Class Activation Maps for Network Data

    Igor Cherepanov, David Sessler, Alex Ulmer +3

    cs.HCcs.AIcs.LGarXiv:2608.13575v12026
  28. Content Depth Matters in Short-Video Recommendation: Rethinking the Attention Economy

    Liwei Deng, Jing Jiang, Zhiwei Li +2

    cs.AIcs.IRarXiv:2608.13990v12026
  29. Training Fair Tabular Foundation Models

    Patrik Kenfack, Jesse C. Cresswell, Anthony L. Caterini +2

    cs.LGcs.AIarXiv:2608.14211v12026
  30. BCMT: Blockwise Causal Memory Transformer

    Rachid Arezki

    cs.CLcs.AIcs.LGarXiv:2608.13578v12026
  31. Hybrid Quantum-inspired Kolmogorov-Arnold Networks for Privacy-Aware Federated Biosignal Learning

    Chun-Hua Lin, Samuel Yen-Chi Chen, Yu-Chao Hsu +7

    cs.LGcs.AIcs.DCarXiv:2608.13914v12026
  32. From Prediction to Intervention: Personalized Meal-Level Glucose Regulation via an LLM Agent

    Mingyu Huang, Weiqing Min, Ying Jin +2

    cs.HCcs.AIcs.LGarXiv:2608.13581v12026
  33. HAM-RAG: Hierarchy-Aware Multimodal RAG for Structure-Faithful Interleaved Generation

    Yin Li, Ziyang Hu, Zhiyu Guo +7

    cs.IRcs.AIarXiv:2608.14032v12026
  34. Amplified Does Not Mean Predictive: Reasoning Behaviors in Thinking Models

    Jean de Dieu Nyandwi, Leena Mathur, Yonatan Bisk +2

    cs.CLcs.AIcs.CVarXiv:2608.13760v12026
  35. Apodex Discovery: Reality Benchmarks and Environments for Evaluating and Building Discoverative Artificial Intelligence

    Brian Wang, Bin Feng, Xiaoman Pan +26

    cs.AIarXiv:2608.11341v12026
  36. Specification-first convergence with an AI coding agent: a case study of dismantling a core architectural invariant across 189 files in a 717k-line codebase with no test oracle and no human code review

    Joel Abenhaim

    cs.SEcs.AIarXiv:2608.12440v12026
  37. TailBooster: A Dual-Layer Generative Framework for Extreme Value Augmentation with Operational Validity Enforcement

    Karim Aly, Alexei Sharpanskykh, Jacco Hoekstra

    cs.LGcs.AIarXiv:2608.11951v12026
  38. Who Speaks Matters: Authority-Aware Multi-View RAG over Italian Parliamentary Proceedings

    Mirko Tritella, Riccardo Pozzi, Matteo Palmonari

    cs.AIarXiv:2608.13410v12026
  39. SPARGen: Unifying Spatial Perception and Reasoning through Native Multimodal Generation

    Jinsheng Quan, Jianhua Li, Siyi Xie +7

    cs.CVcs.AIarXiv:2608.14138v12026
  40. Stable Miscalibration in Large Language Models: A Practical View of High-Confidence Errors

    Akira Okutomi

    cs.AIcs.CLarXiv:2608.13591v12026
  41. AI Evaluation Should Work With Humans

    Jan Kulveit, Gavin Leech, Tomáš Gavenčiak +1

    cs.AIcs.LGarXiv:2608.13577v12026
  42. DFM Mimir v1: An Open HRM Delivering Frontier Performance at 1B Parameters Using Only Permissible Post-Training Data

    Peter Schneider-Kamp, Jacob Nielsen, Gianluca Barmina +2

    cs.CLcs.AIarXiv:2608.13517v12026
  43. A Pathway to General-Purpose Scientific AI: Multimodal Comprehension of Scientific Images

    Jennifer D'Souza, Fahad Ahmed, Cecilia Andrea Bustamante Andrade +9

    cs.AIcs.CVcs.DLarXiv:2608.14075v12026
  44. Improved Large Language Diffusion Models

    Shen Nie, Qiyang Min, Shaoxuan Xu +7

    cs.CLcs.AIcs.LGarXiv:2606.25331v12026
  45. Scalable Visual Pretraining for Language Intelligence

    Yiming Zhang, Zhonghan Zhao, Wenwei Zhang +14

    cs.CVcs.AIcs.MMarXiv:2607.09657v22026
  46. Video Generation Models are General-Purpose Vision Learners

    Letian Wang, Chuhan Zhang, Rishabh Kabra +9

    cs.CVcs.AIarXiv:2607.09024v12026
  47. Self-Supervised Visual On-Policy Distillation

    Yijiang Li, Yijun Liang, Yunjie Tian +6

    cs.CVcs.AIarXiv:2608.14144v12026
    Summaries:한국어
  48. UniMoMo: Expert Merging-Based MoE Acceleration for Large Recommendation Models

    Lei Xin, Bin Gu, Peize Li +9

    cs.AIarXiv:2608.08627v12026
  49. RoFormer: Enhanced Transformer with Rotary Position Embedding

    Jianlin Su, Yu Lu, Shengfeng Pan +3

    cs.CLcs.AIcs.LGarXiv:2104.09864v52021
  50. Agent Safety Should Be a Runtime Contract

    Albus W. Ng, Yi Han, Jusheng Zhang +1

    cs.CRcs.AIarXiv:2608.11274v12026
  51. Are You Sure You're Sure? On the Impact of Instruction Tuning on Confidence and Lexical Diversity

    Irina Proskurina, Mayank Kumar, Oyindolapo O. Komolafe

    cs.CLcs.AIarXiv:2608.13430v12026
  52. Forecast Collapse in Time-Series Foundation Models

    Shu Wan, Miles Ma, Hank Zhu +4

    cs.LGcs.AIcs.CEarXiv:2608.14106v12026
  53. Activity Frames: Deterministic Screen-Activity Compilation for Agent Memory and Replay

    Nossa Iyamu

    cs.AIarXiv:2608.05784v12026
  54. MobileMem: Learning from a Year of Mobile Experiences

    Xinle Deng, Yida Xue, Xiangyuan Ru +14

    cs.AIcs.CLcs.LGarXiv:2608.13606v12026
  55. Long-Horizon-Terminal-Bench: Testing the Limits of Agents on Long-Horizon Terminal Tasks with Dense Reward-Based Grading

    Zongxia Li, Zhongzhi Li, Yucheng Shi +10

    cs.AIarXiv:2607.08964v22026
  56. SPIEval: Evaluating Large Language Models as Mobile Assistants over Scattered Personal Information

    Junjie Ye, Zhuohui Sheng, Shaofan Liu +12

    cs.CLcs.AIarXiv:2608.10692v12026
  57. SkillOpt-Lite: Better and Faster Agent Self-evolution via One Line of Vibe

    Yifei Shen, Bo Li, Xinjie Zhang

    cs.SEcs.AIcs.LGarXiv:2607.03451v12026
  58. $A^2E$ : An End-to-End Agent Auditing Engine

    Haoning Wang, Mingxun Zhang, Chenyue Yu +4

    cs.AIarXiv:2608.07346v22026
  59. Gemma 4 Technical Report

    Gemma Team, Sherif El Abd, Vaibhav Aggarwal +320

    cs.CLcs.AIarXiv:2607.02770v22026
  60. ComBodied Agents: a New Paradigm of Human-Centric Agentic AI

    Qianggang Ding, Xingyao Wang, Rui Feng +19

    cs.AIarXiv:2608.10915v22026