Every paper with a summary
Every arXiv paper Paperlayer has summarized so far. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.
Search paper metadata (including unsummarized papers)
59,221 to 59,280 of 61,289
Identifying Confusion Trends in Concept-based XAI for Multi-Label Classification
Haadia Amjad, Ronald Tetzlaff
cs.CVcs.AIarXiv:2608.15731v12026RRFC: Recursive Refinement via Feedback Conditioning for Iterative Image-to-Image Generation
Kareem Hassani, Chaymaa Abbas, Hadi Al Mubasher +1
cs.CVcs.AIarXiv:2608.15694v12026Not All Attention Is Equal: A Quantitative Survey of the EEI Trade-off
Aditya Singh
cs.LGcs.AIarXiv:2608.15459v12026Solvable Sokoban Without a Solver via Diffusion
Sina Baghal
cs.AIcs.GTcs.LGarXiv:2608.15958v12026PandasCorpus: A Resource of Real-World Pandas Workflows and Usage Patterns
Syrym Abdikhan, Mazhar Hameed
cs.SEcs.LGarXiv:2608.14742v12026EMASAM: a Computationally Efficient Sharpness-Aware Minimization via EMA-Guided Perturbations
Tanapat Ratchatorn, Masayuki Tanaka
cs.LGcs.CVarXiv:2608.15105v12026SAGA: Structure-Attended Generative Action Embedding Model that encodes Multi-Surface User Action Sequences
Tsz Fung Pang, Po Jen Chen, Nimish Ronghe +2
cs.LGcs.IRarXiv:2608.15429v12026Gated Against One Model, Open to the Next: Option-Only Solvability in Legal Multiple-Choice Benchmarks
Volodymyr Ovcharov
cs.CLcs.AIcs.CYarXiv:2608.15428v12026Chameleon: An Adaptive AI-Driven Honeypot Architecture Using Threat-Calibrated Particle Swarm Optimization and Semantic Deception Rapidly-Exploring Random Trees
Rohit Swami, Tushar Singh, Akash Warde +1
cs.CRcs.AIcs.NEarXiv:2608.15407v12026WeSCE: A Benchmark for Measuring Security Drift in LLM-Driven Code Editing
Zhiyu Zhang, Tingyue Wen, Senke Sun +2
cs.CRcs.AIcs.SEarXiv:2608.15092v12026M-LINKX: Multiview Graph Learning for Brain Cognitive Disease Detection
An Phan, Yufei Jin, Xingquan Zhu
cs.LGarXiv:2608.14847v12026Distinguishing AI-Generated Music from Edited Audio as a Hard-Negative Robustness Task
Alexandru-Stefan Morosanu, Valerian Cecan, Stefan-Daniel Achirei +1
cs.SDcs.AIcs.LGarXiv:2608.14916v12026Adaptive Volumetric Mechanical Property Fields Invariant to Resolution
Rishit Dagli, Donglai Xiang, Vismay Modi +4
cs.CVcs.LGcs.ROarXiv:2606.18231v12026Trust the Right Teacher: Quality-Aware Self-Distillation for GUI Grounding
Jingyuan Huang, Zuming Huang, Yucheng Shi +4
cs.AIarXiv:2606.18101v22026OGX: An Open-Source, Vendor-Neutral Generative AI Application Server
Francisco Javier Arceo, Sébastien Han, Matthew Farrellee +8
cs.AIcs.IRarXiv:2608.14580v12026Robusto-2: Benchmarking Humans & VLMs for Autonomous Driving in Lima & New York City
Adrian Cespedes, Marcelo Chincha, Dunant Cusipuma +3
cs.CVcs.AIcs.ROarXiv:2606.20980v12026From Trainee to Trainer: LLM-Designed Training Environment for RL with Multi-Agent Reasoning
Chao Chen, Chengzu Li, Zhiwei Li +2
cs.CLarXiv:2606.17682v12026LegalHalluLens: Typed Hallucination Auditing and Calibrated Multi-Agent Debate for Trustworthy Legal AI
Lalit Yadav, Akshaj Gurugubelli
cs.AIcs.CLcs.LGarXiv:2606.18021v12026Does VLA Even Know the Basics? Measuring Commonsense and World Knowledge Retention in Vision-Language-Action Models
Nikita Kachaev, Andrey Moskalenko, Matvey Skripkin +10
cs.LGcs.ROarXiv:2606.19297v12026FAPO: Fully Automated Prompt Optimization of Multi-Step LLM Pipelines
Paul Kassianik, Baturay Saglam, Huaibo Zhao +4
cs.SEcs.AIarXiv:2606.19605v22026Freeing the Law with LOCUS: A Local Ordinance Corpus for the United States
Denis Peskoff, Joe Barrow, Christopher Vu +1
cs.CLcs.CYcs.LGarXiv:2606.19334v12026Beyond Reward Engineering: A Data Recipe for Long-Context Reinforcement Learning
Xiaoyue Xu, Sikui Zhang, Xiaorong Wang +2
cs.CLcs.AIarXiv:2606.18831v12026Configurable Clinical Information Extraction with Agentic RAG: What Works, What Breaks, and Why
Osman Alperen Çinar-Koraş, Marie Bauer, Sameh Khattab +7
cs.AIarXiv:2606.19602v12026LooseControlVideo: Directorial Video Control using Spatial Blocking
Shariq Farooq Bhat, Niloy J. Mitra, Kalyan Sunkavalli
cs.CVarXiv:2606.19495v12026StylisticBias: A Few Human Visual Cues Drive Most Social Biases in MLLMs
Shaghayegh Kolli, Timo Cavelius, Nafiseh Nikeghbal +2
cs.CLcs.CVarXiv:2606.20527v12026OpenRath: Session-Centered Runtime State for Agent Systems
Fukang Wen, Zhijie Wang, Ruilin Xu
cs.SEcs.PLarXiv:2606.19409v12026Object-Centric Residual RL for Zero-Shot Sim-to-Real VLA Enhancement
Kinam Kim, Namiko Saito, Heecheol Kim +3
cs.ROarXiv:2606.18953v12026Rethinking Shrinkage Bias in LLM FP4 Pretraining: Geometric Origin, Systemic Impact, and UFP4 Recipe
Qian Zhao, Kunlong Chen, Changxin Tian +9
cs.AIarXiv:2606.20381v12026JanusMesh: Fast and Zero-Shot 3D Visual Illusion Generation via Cross-Space Denoising
Siang-Ling Zhang, Huai-Hsun Cheng, Tsung-Ju Yang +1
cs.CVarXiv:2606.20563v12026CogniRoute: Learning to Route Social Evidence in Omni-Modal Models
Yifan Shen, Pei Tian, Xinzhuo Li +8
cs.CVarXiv:2606.20970v12026How Post-Training Shapes Biological Reasoning Models
Lukas Fesser, Hanlin Zhang, Michelle M. Li +5
cs.LGq-bio.QMarXiv:2606.16517v22026Characterization of Thermal Systems from Noisy and Low-resolution Measurements Using Dynamic Mode Decomposition
M. E. P. Silva, L. S. Araujo, F. T. Colombo +2
physics.comp-phcs.LGeess.SParXiv:2608.14581v12026Variable-Width Transformers
Zhaofeng Wu, Oliver Sieberling, Shawn Tan +3
cs.CLarXiv:2606.18246v12026Reinforcing Dual-Path Reasoning in Spatial Vision Language Models
Yatai Ji, An-Chieh Cheng, Yang Fu +13
cs.CVcs.AIarXiv:2606.17539v12026GeneralVLA-2: Geometry-Aware Reconstruction and Governed Memory for Robot Planning
Haoyu Wang, Guoqing Ma, Zeyu Zhang +3
cs.CVcs.ROarXiv:2606.17480v12026MCompassRAG: Topic Metadata as a Semantic Compass for Paragraph-Level Retrieval
Amirhossein Abaskohi, Raymond Li, Gaetano Cimino +3
cs.CLcs.IRarXiv:2606.18508v12026SproutRAG: Attention-Guided Tree Search with Progressive Embeddings for Long-Document RAG
Amirhossein Abaskohi, Issam H. Laradji, Peter West +1
cs.CLcs.IRarXiv:2606.18381v12026Evaluating the impact of adversarial traffic patterns on vanet communication using veins simulation
Henry Agyapong
cs.NIcs.LGarXiv:2608.14583v12026Morpheus: A Morphology-Aware Neural Tokenizer and Word Embedder for Turkish
Tolga Şakar
cs.CLcs.AIarXiv:2606.18717v12026Sumi: Open Uniform Diffusion Language Model from Scratch
Mengyu Ye, Keito Kudo, Wataru Ikeda +3
cs.CLcs.LGarXiv:2606.19005v12026Beyond the Current Observation: Evaluating Multimodal Large Language Models in Controllable Non-Markov Games
Shengyuan Ding, Xilin Wei, Xinyu Fang +4
cs.CVarXiv:2606.19338v12026PerceptionDLM: Parallel Region Perception with Multimodal Diffusion Language Models
Yueyi Sun, Yuhao Wang, Jason Li +8
cs.CVcs.AIcs.CLarXiv:2606.19534v12026Moebius: 0.2B Lightweight Image Inpainting Framework with 10B-Level Performance
Kangsheng Duan, Ziyang Xu, Wenyu Liu +3
cs.CVarXiv:2606.19195v12026HumanScale: Egocentric Human Video Can Outperform Real-Robot Data for Embodied Pretraining
Juncheng Ma, Jianxin Bi, Yufan Deng +19
cs.CVarXiv:2606.20521v12026DF3DV-1K: A Large-Scale Dataset and Benchmark for Distractor-Free Novel View Synthesis
Cheng-You Lu, Yi-Shan Hung, Wei-Ling Chi +6
cs.CVcs.AIarXiv:2604.13416v32026Beyond Static Leaderboards: Predictive Validity for the Evaluation of LLM Agents
Dhaval C. Patel, Kaoutar El Maghraoui, Shuxin Lin +58
cs.AIarXiv:2606.19704v12026When, Where, and How: Adaptive Binning for Tabular Self-Supervised Learning
Daehwan Kim, Haejun Chung, Ikbeom Jang
cs.LGcs.AIarXiv:2606.19827v12026Toward Parking Spot Occupancy Recognition: A Self-Supervised Approach
Luan Marko Kujavski, Rayson Laroca, Paulo Lisboa de Almeida
cs.CVarXiv:2606.20886v12026Grouped Query Experts: Mixture-of-Experts on GQA Self-Attention
Vishesh Tripathi, Abhay Kumar
cs.LGarXiv:2606.20945v22026QG-MIL: A Gated Transformer Aggregator for Domain-Agnostic Multiple Instance Learning in Medical Imaging
Luca Zedda, Davide Antonio Mura, Cecilia Di Ruberto +4
cs.CVarXiv:2606.20027v12026HydraHead: From Head-Level Functional Heterogeneity to Specialized Attention Hybridization
Zhentao Tan, Wei Chen, Jingyi Shen +4
cs.CLarXiv:2606.20097v12026An Exploratory Case Study of LLM-Assisted Refactoring and Gameplay Feature Generation in an Endless Runner Game
Jan Wunderlich, Markus Kleffmann, Sebastian Lempert
cs.SEcs.AIarXiv:2606.21171v12026EgoSteer: A Full-Stack System Towards Steerable Dexterous Manipulation from Egocentric Videos
Yifan Zhong, Zhang Chen, Tianrui Guan +13
cs.ROarXiv:2607.09701v12026Causal Discovery in the Era of Agents
Yujia Zheng, Vishal Verma, Mantej Gill +3
cs.AIcs.LGcs.SEarXiv:2606.23608v12026Arbor: Explicit Geometric Conditioning for Controllable 3D Asset Generation
Jan-Niklas Dihlmann, Andreas Engelhardt, Simon Donne +2
cs.CVcs.GRarXiv:2606.23514v12026Vera: A Layered Diffusion Model for Content-Preserving Video Editing
Hongkai Zheng, Ta-Ying Cheng, Benjamin Klein +2
cs.CVarXiv:2606.23610v12026EnterpriseClawBench: Benchmarking Agents from Real Workplace Sessions
Jincheng Zhong, Weizhi Wang, Che Jiang +5
cs.CLcs.SEarXiv:2606.23654v12026PhoneBuddy: Training Open Models for Agentic Phone Use
Zhengyang Tang, Xin Lai, Pengyuan Lyu +23
cs.CLcs.AIarXiv:2606.23049v22026ChartWalker: Benchmarking the Cross-Chart RAG Task with Hierarchical Knowledge Graphs
Ning Tang, Chenghan Xie, Hanyang Yuan +6
cs.IRarXiv:2606.23997v12026VeriEvol: Scaling Multimodal Mathematical Reasoning via Verifiable Evol-Instruct
Haoling Li, Kai Zheng, Jie Wu +4
cs.AIcs.CLcs.CVarXiv:2606.23543v12026