Artificial Intelligence

Papers filed under cs.AI on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.

Search paper metadata (including unsummarized papers)

2,041 to 2,100 of 15,246

  1. Robust Conformal Consensus: Multi-Agent LLM-as-a-Judge Interval Evaluation with Conformal Prediction

    Lihui Liu

    cs.LGcs.AIarXiv:2609.06367v12026
  2. SWE-Test: Benchmarking LLM Vulnerability Discovery via Input Prediction

    Yuanxiang Shi, Jiayi Lin, Xuanyong Lin +8

    cs.SEcs.AIarXiv:2609.06229v12026
  3. Power Mean Estimation in Stochastic Continuous Monte Carlo Tree Search

    Tuan Dam

    cs.LGcs.AIarXiv:2609.06489v12026
  4. Second-Order Smooth Planning with Optimal-Transport Bellman Smoothing

    Tuan Dam

    cs.LGcs.AIarXiv:2609.06484v12026
  5. ExBody2: Advanced Expressive Humanoid Whole-Body Control

    Mazeyu Ji, Xuanbin Peng, Fangchen Liu +4

    cs.ROcs.AIcs.LGarXiv:2412.13196v22024
  6. BioMedLM: A 2.7B Parameter Language Model Trained On Biomedical Text

    Elliot Bolton, Abhinav Venigalla, Michihiro Yasunaga +8

    cs.CLcs.AIarXiv:2403.18421v12024
  7. Collision Snapshot Guided Time-Reversed Safety-Critical Scenario Generation

    Taehyung Kim, Jongeun Choi

    cs.ROcs.AIarXiv:2609.06433v12026
  8. Model-Based Domain Generalization

    Alexander Robey, George J. Pappas, Hamed Hassani

    stat.MLcs.AIcs.LGarXiv:2102.11436v52021
  9. The Benchmark Lottery

    Mostafa Dehghani, Yi Tay, Alexey A. Gritsenko +5

    cs.LGcs.AIcs.CLarXiv:2107.07002v12021
  10. On BatchNorm Forward Modes in Value-Based Reinforcement Learning

    Daniel Palenicek, Mikael Henaff, Scott Fujimoto +1

    cs.LGcs.AIarXiv:2609.06421v12026
  11. Grounding Language Models to Images for Multimodal Inputs and Outputs

    Jing Yu Koh, Ruslan Salakhutdinov, Daniel Fried

    cs.CLcs.AIcs.CVarXiv:2301.13823v42023
  12. Parameterized and Streaming Algorithms for Euclidean Fair $k$-Center Clustering

    Zeyu Lin, Chaoqi Jia, Longkun Guo +1

    cs.LGcs.AIarXiv:2609.06384v12026
  13. Time and Activity Sequence Prediction of Business Process Instances

    Mirko Polato, Alessandro Sperduti, Andrea Burattin +1

    cs.AIarXiv:1602.07566v12016
  14. Recovering Weak Signals with Normalizing Flows

    Sarod Yatawatta

    stat.MLastro-ph.COastro-ph.IMarXiv:2609.06382v12026
  15. Exploring Model-based Planning with Policy Networks

    Tingwu Wang, Jimmy Ba

    cs.LGcs.AIcs.ROarXiv:1906.08649v12019
  16. Grounding Language with Visual Affordances over Unstructured Data

    Oier Mees, Jessica Borja-Diaz, Wolfram Burgard

    cs.ROcs.AIcs.CLarXiv:2210.01911v32022
  17. Diamond Agent: Agentic Control of Federated HPC Resources as a Service

    Haotian Xie, Junlin Chen, Mingkai Zheng +9

    cs.DCcs.AIarXiv:2609.06181v12026
  18. Investigating the Limitations of Transformers with Simple Arithmetic Tasks

    Rodrigo Nogueira, Zhiying Jiang, Jimmy Lin

    cs.CLcs.AIcs.LGarXiv:2102.13019v32021
  19. All for 1-Bit: Towards Genuine 1-Bit Post-Training Quantization for LLMs

    Zhixiong Zhao, Zukang Xu, Guangyu Sun +2

    cs.LGcs.AIarXiv:2609.06161v12026
  20. Multiple Myeloma Lesion Segmentation on Whole-Body Diffusion-Weighted Imaging via Efficient Anatomical Anticipation and Multimodal Confirmation

    Mengmeng Zhang, Shengqian Huang, Junde Zhou +12

    cs.CVcs.AIarXiv:2609.06165v12026
  21. An Attentive Inductive Bias for Sequential Recommendation beyond the Self-Attention

    Yehjin Shin, Jeongwhan Choi, Hyowon Wi +1

    cs.LGcs.AIcs.IRarXiv:2312.10325v22023
  22. Boosting Multimodal Large Language Models with Visual Tokens Withdrawal for Rapid Inference

    Zhihang Lin, Mingbao Lin, Luxi Lin +1

    cs.CVcs.AIarXiv:2405.05803v32024
  23. Boosting the Speed of Entity Alignment 10*: Dual Attention Matching Network with Normalized Hard Sample Mining

    Xin Mao, Wenting Wang, Yuanbin Wu +1

    cs.AIarXiv:2103.15452v12021
  24. Decision-Aware Suffix Prediction and Reasoning of Business Processes

    Henryk Mustroph, Stefanie Rinderle-Ma

    cs.LGcs.AIarXiv:2609.06169v12026
  25. GALIP: Generative Adversarial CLIPs for Text-to-Image Synthesis

    Ming Tao, Bing-Kun Bao, Hao Tang +1

    cs.CVcs.AIarXiv:2301.12959v12023
  26. From Splats to Silicon: Rethinking Computational Efficiency of 3DGS

    Minnan Pei, Qiwei Dong, Yihan Zhou +7

    cs.ARcs.AIcs.GRarXiv:2609.06157v12026
  27. ExpertLens: Visualizing Embedding Spaces for Post-Hoc Explainability in MoE Enhanced Retrievers

    Effrosyni Sokli, Isaac Roberts, Alexander Schulz +2

    cs.IRcs.AIarXiv:2609.06155v12026
  28. SCRIPTIOC-BENCH: A Benchmark for Recognizing Actionable Threat Intelligence from Script-Based Malware using LLMs

    Hanna Kim, Jian Cui, Minkyoo Song +4

    cs.CRcs.AIarXiv:2609.06149v12026
  29. What the Window Does Not Contain: Auditing Provenance in a Document-Grounded Instability Benchmark

    Seyed Mosayeb Alam

    cs.CLcs.AIcs.LGarXiv:2609.06147v12026
  30. ExecCritic: Learn to Test, Test to Improve for Coding Agents

    Leitian Tao, Baolin Peng, Haorui Wang +7

    cs.AIcs.CLcs.SEarXiv:2609.09133v12026
  31. Co-Evolving Harnesses and Models: On-Policy Correction Helps Weaker Models Catch Up Where Imitation Fails

    Zhou Yu, Bin Bi, Shiva Kumar Pentyala +8

    cs.AIarXiv:2609.09134v12026
  32. MeClear: Cooperative Game-Theoretic Attribution and Risk-Aware Memory Clearance for Long-Horizon LLM Agents

    Boyu Yang, Jiazheng Sun, Zilong Lu +3

    cs.AIcs.SEarXiv:2609.09115v12026
  33. Message Passing for Hyper-Relational Knowledge Graphs

    Mikhail Galkin, Priyansh Trivedi, Gaurav Maheshwari +2

    cs.LGcs.AIcs.CLarXiv:2009.10847v12020
  34. A Generalization of Amari's Bayesian Duality

    Mohammad Emtiyaz Khan, Thomas Möllenhoff

    cs.AIcs.LGstat.MLarXiv:2609.09126v12026
  35. Semantic Grouping Network for Video Captioning

    Hobin Ryu, Sunghun Kang, Haeyong Kang +1

    cs.CVcs.AIarXiv:2102.00831v22021
  36. AGSA-Net: Abundance-Guided Self-Attention Network for Spectral Unmixing-Aware Hyperspectral Remote Sensing Image Classification

    Nafisa Anjum, Satavisa Dey Borno, Ananna Saha +4

    cs.CVcs.AIcs.LGarXiv:2609.06359v12026
  37. Linear Algebra Foundations of Efficient Attention: A Phase Reversal in Rank Collapse Under SVD Compression

    Anjaneya Teja Sarma Kalvakolanu

    cs.LGcs.AIcs.NEarXiv:2609.06341v12026
  38. SIDE: Sensor Impersonation Detection at the Edge via Sequence Prediction

    Nahom Birhan

    cs.CRcs.AIarXiv:2609.06271v12026
  39. It is Not Yet Another Tool: Creating and Deploying an Agentic AI Companion in a Security Operations Center

    Kritan Banstola, Faayed Al Faisal, Duy Dao +3

    cs.CRcs.AIarXiv:2609.06250v12026
  40. VDiff-Bench: A Challenging Benchmark for Fine-Grained Image Difference Identification

    Yixin Wan, Tianle Zheng, Kai-Wei Chang

    cs.CVcs.AIcs.CLarXiv:2609.06245v12026
  41. Debiasing Graph Neural Networks via Learning Disentangled Causal Substructure

    Shaohua Fan, Xiao Wang, Yanhu Mo +2

    cs.LGcs.AIarXiv:2209.14107v12022
  42. A Data-Driven Framework for Identifying and Prioritizing RPA Opportunities in Healthcare Processes

    Maria Alejandra Gomez, Juan Manuel Castillo

    cs.AIcs.CLarXiv:2609.09137v12026
  43. TransEdge: Translating Relation-contextualized Embeddings for Knowledge Graphs

    Zequn Sun, Jiacheng Huang, Wei Hu +3

    cs.AIcs.CLcs.LGarXiv:2004.13579v12020
  44. No Metrics Are Perfect: Adversarial Reward Learning for Visual Storytelling

    Xin Wang, Wenhu Chen, Yuan-Fang Wang +1

    cs.CLcs.AIcs.CVarXiv:1804.09160v22018
  45. Benchmarking Complex Instruction-Following with Multiple Constraints Composition

    Bosi Wen, Pei Ke, Xiaotao Gu +11

    cs.CLcs.AIarXiv:2407.03978v32024
  46. PlannerForge: LLM Agents for Scenario-Based Testing of Motion Planners in Autonomous Driving

    Yuan Gao, Sebastian Müller, Mattia Piccinini +5

    cs.AIcs.CLcs.ROarXiv:2609.08965v12026
  47. Robotic Telekinesis: Learning a Robotic Hand Imitator by Watching Humans on Youtube

    Aravind Sivakumar, Kenneth Shaw, Deepak Pathak

    cs.ROcs.AIcs.CVarXiv:2202.10448v22022
  48. Everything in Moderation: Per-Domain Coverage Optima and Alignment-Resistant Domain Gaps in Multi-Domain Mid-Training

    Yunpeng Xu, Kun Zheng

    cs.AIarXiv:2609.09081v12026
  49. A review of Generative Adversarial Networks (GANs) and its applications in a wide variety of disciplines -- From Medical to Remote Sensing

    Ankan Dash, Junyi Ye, Guiling Wang

    cs.LGcs.AIcs.CVarXiv:2110.01442v12021
  50. Random vector functional link neural network based ensemble deep learning for short-term load forecasting

    Ruobin Gao, Liang Du, P. N. Suganthan +2

    cs.LGcs.AIeess.SParXiv:2107.14385v12021
  51. RL-RRT: Kinodynamic Motion Planning via Learning Reachability Estimators from RL Policies

    Hao-Tien Lewis Chiang, Jasmine Hsu, Marek Fiser +2

    cs.ROcs.AIcs.LGarXiv:1907.04799v22019
  52. SAEScientist-Bench: Can AI Agents Conduct Autonomous SAE Interpretability Research?

    Yuqiao Tan, Shizhu He, Jun Zhao +1

    cs.AIcs.CLcs.LGarXiv:2609.09113v12026
  53. Learning to Control Self-Assembling Morphologies: A Study of Generalization via Modularity

    Deepak Pathak, Chris Lu, Trevor Darrell +2

    cs.LGcs.AIcs.CVarXiv:1902.05546v22019
  54. Keyformer: KV Cache Reduction through Key Tokens Selection for Efficient Generative Inference

    Muhammad Adnan, Akhil Arunkumar, Gaurav Jain +3

    cs.LGcs.AIcs.ARarXiv:2403.09054v22024
  55. Good Pretraining, Bad SFT: Checkpoint Quality Across the Training Stack

    Sohir Maskey, Philipp Scholl, Jonas Knupp +2

    cs.AIcs.CLarXiv:2609.08966v12026
  56. Time-Varying Data as Sheaves: an Invitation to Narratives

    Wilmer Leal, Benjamin Merlin Bumpus, Jana K. Nickel +3

    cs.AIcs.MAeess.SYarXiv:2609.09056v12026
  57. Answer-Distribution Trajectories: A Stochastic-Dynamics View of LLM Reasoning

    Mar Gonzàlez I Català, Haitz Sáez de Ocáriz Borde, Davide Murari +3

    cs.AIcs.CLcs.ITarXiv:2609.09030v12026
  58. The Surprising Effectiveness of Approximate Value Iteration in Self-Play

    Raphael Boige, Amine Boumaza, Bruno Scherrer

    cs.AIarXiv:2609.09094v12026
  59. Deposon: An Auditable, Conservation-Guaranteed, Game-Theoretically Tested Scattering Layer over LLM Reasoning Paths

    Qihao Yuan

    cs.AIcs.LGarXiv:2609.09001v12026
  60. SkillAdam: Stable and Efficient Skill Evolution for Agents

    Gaoyuan Li, Meihao Fan, Yizhe Liu +7

    cs.AIarXiv:2609.08944v12026