Artificial Intelligence
Papers filed under cs.AI on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.
Search paper metadata (including unsummarized papers)
2,041 to 2,100 of 15,246
Robust Conformal Consensus: Multi-Agent LLM-as-a-Judge Interval Evaluation with Conformal Prediction
Lihui Liu
cs.LGcs.AIarXiv:2609.06367v12026SWE-Test: Benchmarking LLM Vulnerability Discovery via Input Prediction
Yuanxiang Shi, Jiayi Lin, Xuanyong Lin +8
cs.SEcs.AIarXiv:2609.06229v12026Power Mean Estimation in Stochastic Continuous Monte Carlo Tree Search
Tuan Dam
cs.LGcs.AIarXiv:2609.06489v12026Second-Order Smooth Planning with Optimal-Transport Bellman Smoothing
Tuan Dam
cs.LGcs.AIarXiv:2609.06484v12026ExBody2: Advanced Expressive Humanoid Whole-Body Control
Mazeyu Ji, Xuanbin Peng, Fangchen Liu +4
cs.ROcs.AIcs.LGarXiv:2412.13196v22024BioMedLM: A 2.7B Parameter Language Model Trained On Biomedical Text
Elliot Bolton, Abhinav Venigalla, Michihiro Yasunaga +8
cs.CLcs.AIarXiv:2403.18421v12024Collision Snapshot Guided Time-Reversed Safety-Critical Scenario Generation
Taehyung Kim, Jongeun Choi
cs.ROcs.AIarXiv:2609.06433v12026Model-Based Domain Generalization
Alexander Robey, George J. Pappas, Hamed Hassani
stat.MLcs.AIcs.LGarXiv:2102.11436v52021The Benchmark Lottery
Mostafa Dehghani, Yi Tay, Alexey A. Gritsenko +5
cs.LGcs.AIcs.CLarXiv:2107.07002v12021On BatchNorm Forward Modes in Value-Based Reinforcement Learning
Daniel Palenicek, Mikael Henaff, Scott Fujimoto +1
cs.LGcs.AIarXiv:2609.06421v12026Grounding Language Models to Images for Multimodal Inputs and Outputs
Jing Yu Koh, Ruslan Salakhutdinov, Daniel Fried
cs.CLcs.AIcs.CVarXiv:2301.13823v42023Parameterized and Streaming Algorithms for Euclidean Fair $k$-Center Clustering
Zeyu Lin, Chaoqi Jia, Longkun Guo +1
cs.LGcs.AIarXiv:2609.06384v12026Time and Activity Sequence Prediction of Business Process Instances
Mirko Polato, Alessandro Sperduti, Andrea Burattin +1
cs.AIarXiv:1602.07566v12016Recovering Weak Signals with Normalizing Flows
Sarod Yatawatta
stat.MLastro-ph.COastro-ph.IMarXiv:2609.06382v12026Exploring Model-based Planning with Policy Networks
Tingwu Wang, Jimmy Ba
cs.LGcs.AIcs.ROarXiv:1906.08649v12019Grounding Language with Visual Affordances over Unstructured Data
Oier Mees, Jessica Borja-Diaz, Wolfram Burgard
cs.ROcs.AIcs.CLarXiv:2210.01911v32022Diamond Agent: Agentic Control of Federated HPC Resources as a Service
Haotian Xie, Junlin Chen, Mingkai Zheng +9
cs.DCcs.AIarXiv:2609.06181v12026Investigating the Limitations of Transformers with Simple Arithmetic Tasks
Rodrigo Nogueira, Zhiying Jiang, Jimmy Lin
cs.CLcs.AIcs.LGarXiv:2102.13019v32021All for 1-Bit: Towards Genuine 1-Bit Post-Training Quantization for LLMs
Zhixiong Zhao, Zukang Xu, Guangyu Sun +2
cs.LGcs.AIarXiv:2609.06161v12026Multiple Myeloma Lesion Segmentation on Whole-Body Diffusion-Weighted Imaging via Efficient Anatomical Anticipation and Multimodal Confirmation
Mengmeng Zhang, Shengqian Huang, Junde Zhou +12
cs.CVcs.AIarXiv:2609.06165v12026An Attentive Inductive Bias for Sequential Recommendation beyond the Self-Attention
Yehjin Shin, Jeongwhan Choi, Hyowon Wi +1
cs.LGcs.AIcs.IRarXiv:2312.10325v22023Boosting Multimodal Large Language Models with Visual Tokens Withdrawal for Rapid Inference
Zhihang Lin, Mingbao Lin, Luxi Lin +1
cs.CVcs.AIarXiv:2405.05803v32024Boosting the Speed of Entity Alignment 10*: Dual Attention Matching Network with Normalized Hard Sample Mining
Xin Mao, Wenting Wang, Yuanbin Wu +1
cs.AIarXiv:2103.15452v12021Decision-Aware Suffix Prediction and Reasoning of Business Processes
Henryk Mustroph, Stefanie Rinderle-Ma
cs.LGcs.AIarXiv:2609.06169v12026GALIP: Generative Adversarial CLIPs for Text-to-Image Synthesis
Ming Tao, Bing-Kun Bao, Hao Tang +1
cs.CVcs.AIarXiv:2301.12959v12023From Splats to Silicon: Rethinking Computational Efficiency of 3DGS
Minnan Pei, Qiwei Dong, Yihan Zhou +7
cs.ARcs.AIcs.GRarXiv:2609.06157v12026ExpertLens: Visualizing Embedding Spaces for Post-Hoc Explainability in MoE Enhanced Retrievers
Effrosyni Sokli, Isaac Roberts, Alexander Schulz +2
cs.IRcs.AIarXiv:2609.06155v12026SCRIPTIOC-BENCH: A Benchmark for Recognizing Actionable Threat Intelligence from Script-Based Malware using LLMs
Hanna Kim, Jian Cui, Minkyoo Song +4
cs.CRcs.AIarXiv:2609.06149v12026What the Window Does Not Contain: Auditing Provenance in a Document-Grounded Instability Benchmark
Seyed Mosayeb Alam
cs.CLcs.AIcs.LGarXiv:2609.06147v12026ExecCritic: Learn to Test, Test to Improve for Coding Agents
Leitian Tao, Baolin Peng, Haorui Wang +7
cs.AIcs.CLcs.SEarXiv:2609.09133v12026Co-Evolving Harnesses and Models: On-Policy Correction Helps Weaker Models Catch Up Where Imitation Fails
Zhou Yu, Bin Bi, Shiva Kumar Pentyala +8
cs.AIarXiv:2609.09134v12026MeClear: Cooperative Game-Theoretic Attribution and Risk-Aware Memory Clearance for Long-Horizon LLM Agents
Boyu Yang, Jiazheng Sun, Zilong Lu +3
cs.AIcs.SEarXiv:2609.09115v12026Message Passing for Hyper-Relational Knowledge Graphs
Mikhail Galkin, Priyansh Trivedi, Gaurav Maheshwari +2
cs.LGcs.AIcs.CLarXiv:2009.10847v12020A Generalization of Amari's Bayesian Duality
Mohammad Emtiyaz Khan, Thomas Möllenhoff
cs.AIcs.LGstat.MLarXiv:2609.09126v12026Semantic Grouping Network for Video Captioning
Hobin Ryu, Sunghun Kang, Haeyong Kang +1
cs.CVcs.AIarXiv:2102.00831v22021AGSA-Net: Abundance-Guided Self-Attention Network for Spectral Unmixing-Aware Hyperspectral Remote Sensing Image Classification
Nafisa Anjum, Satavisa Dey Borno, Ananna Saha +4
cs.CVcs.AIcs.LGarXiv:2609.06359v12026Linear Algebra Foundations of Efficient Attention: A Phase Reversal in Rank Collapse Under SVD Compression
Anjaneya Teja Sarma Kalvakolanu
cs.LGcs.AIcs.NEarXiv:2609.06341v12026SIDE: Sensor Impersonation Detection at the Edge via Sequence Prediction
Nahom Birhan
cs.CRcs.AIarXiv:2609.06271v12026It is Not Yet Another Tool: Creating and Deploying an Agentic AI Companion in a Security Operations Center
Kritan Banstola, Faayed Al Faisal, Duy Dao +3
cs.CRcs.AIarXiv:2609.06250v12026VDiff-Bench: A Challenging Benchmark for Fine-Grained Image Difference Identification
Yixin Wan, Tianle Zheng, Kai-Wei Chang
cs.CVcs.AIcs.CLarXiv:2609.06245v12026Debiasing Graph Neural Networks via Learning Disentangled Causal Substructure
Shaohua Fan, Xiao Wang, Yanhu Mo +2
cs.LGcs.AIarXiv:2209.14107v12022A Data-Driven Framework for Identifying and Prioritizing RPA Opportunities in Healthcare Processes
Maria Alejandra Gomez, Juan Manuel Castillo
cs.AIcs.CLarXiv:2609.09137v12026TransEdge: Translating Relation-contextualized Embeddings for Knowledge Graphs
Zequn Sun, Jiacheng Huang, Wei Hu +3
cs.AIcs.CLcs.LGarXiv:2004.13579v12020No Metrics Are Perfect: Adversarial Reward Learning for Visual Storytelling
Xin Wang, Wenhu Chen, Yuan-Fang Wang +1
cs.CLcs.AIcs.CVarXiv:1804.09160v22018Benchmarking Complex Instruction-Following with Multiple Constraints Composition
Bosi Wen, Pei Ke, Xiaotao Gu +11
cs.CLcs.AIarXiv:2407.03978v32024PlannerForge: LLM Agents for Scenario-Based Testing of Motion Planners in Autonomous Driving
Yuan Gao, Sebastian Müller, Mattia Piccinini +5
cs.AIcs.CLcs.ROarXiv:2609.08965v12026Robotic Telekinesis: Learning a Robotic Hand Imitator by Watching Humans on Youtube
Aravind Sivakumar, Kenneth Shaw, Deepak Pathak
cs.ROcs.AIcs.CVarXiv:2202.10448v22022Everything in Moderation: Per-Domain Coverage Optima and Alignment-Resistant Domain Gaps in Multi-Domain Mid-Training
Yunpeng Xu, Kun Zheng
cs.AIarXiv:2609.09081v12026A review of Generative Adversarial Networks (GANs) and its applications in a wide variety of disciplines -- From Medical to Remote Sensing
Ankan Dash, Junyi Ye, Guiling Wang
cs.LGcs.AIcs.CVarXiv:2110.01442v12021Random vector functional link neural network based ensemble deep learning for short-term load forecasting
Ruobin Gao, Liang Du, P. N. Suganthan +2
cs.LGcs.AIeess.SParXiv:2107.14385v12021RL-RRT: Kinodynamic Motion Planning via Learning Reachability Estimators from RL Policies
Hao-Tien Lewis Chiang, Jasmine Hsu, Marek Fiser +2
cs.ROcs.AIcs.LGarXiv:1907.04799v22019SAEScientist-Bench: Can AI Agents Conduct Autonomous SAE Interpretability Research?
Yuqiao Tan, Shizhu He, Jun Zhao +1
cs.AIcs.CLcs.LGarXiv:2609.09113v12026Learning to Control Self-Assembling Morphologies: A Study of Generalization via Modularity
Deepak Pathak, Chris Lu, Trevor Darrell +2
cs.LGcs.AIcs.CVarXiv:1902.05546v22019Keyformer: KV Cache Reduction through Key Tokens Selection for Efficient Generative Inference
Muhammad Adnan, Akhil Arunkumar, Gaurav Jain +3
cs.LGcs.AIcs.ARarXiv:2403.09054v22024Good Pretraining, Bad SFT: Checkpoint Quality Across the Training Stack
Sohir Maskey, Philipp Scholl, Jonas Knupp +2
cs.AIcs.CLarXiv:2609.08966v12026Time-Varying Data as Sheaves: an Invitation to Narratives
Wilmer Leal, Benjamin Merlin Bumpus, Jana K. Nickel +3
cs.AIcs.MAeess.SYarXiv:2609.09056v12026Answer-Distribution Trajectories: A Stochastic-Dynamics View of LLM Reasoning
Mar Gonzàlez I Català, Haitz Sáez de Ocáriz Borde, Davide Murari +3
cs.AIcs.CLcs.ITarXiv:2609.09030v12026The Surprising Effectiveness of Approximate Value Iteration in Self-Play
Raphael Boige, Amine Boumaza, Bruno Scherrer
cs.AIarXiv:2609.09094v12026Deposon: An Auditable, Conservation-Guaranteed, Game-Theoretically Tested Scattering Layer over LLM Reasoning Paths
Qihao Yuan
cs.AIcs.LGarXiv:2609.09001v12026SkillAdam: Stable and Efficient Skill Evolution for Agents
Gaoyuan Li, Meihao Fan, Yizhe Liu +7
cs.AIarXiv:2609.08944v12026