Every paper with a summary
Every arXiv paper Paperlayer has summarized so far. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.
Search paper metadata (including unsummarized papers)
59,941 to 60,000 of 61,302
Macaron-A2UI: A Model for Generative UI in Personal Agents
Fancy Kong, Congjie Zheng, Murphy Zhuang +8
cs.HCarXiv:2605.24830v12026Hallucinations Undermine Trust; Metacognition is a Way Forward
Gal Yona, Mor Geva, Yossi Matias
cs.CLarXiv:2605.01428v12026Auto-Rubric as Reward: From Implicit Preferences to Explicit Multimodal Generative Criteria
Juanxi Tian, Fengyuan Liu, Jiaming Han +6
cs.AIarXiv:2605.08354v12026AeroCopilotBench: A Two-Tier Benchmark for Evaluating LLM Agents as Aviation Copilots in an Interactive Virtual Cockpit Environment
Yuchen Yuan, Zhenghuang Wu, Yuangan Li +2
cs.AIarXiv:2608.16349v12026Towards On-Policy Data Evolution for Visual-Native Multimodal Deep Search Agents
Shijue Huang, Hangyu Guo, Guanting Dong +8
cs.CLarXiv:2605.10832v22026Baseline-Relative Counterfactual Refinement for Bit-Aware Visual Token Communication
Jia Guo, Xiaohan Zhao, Changwang Liu +4
cs.AIarXiv:2608.16192v12026OpenSTBench: Beyond Semantic Evaluation for Speech Translation
Yanjie An, Yuxiang Zhao, Yichi Zhang +5
eess.AScs.AIarXiv:2605.30792v12026Comprehensive Benchmarking of Deep Learning Architectures for Lung Cancer Histopathology
Hadi Hasan, Safaa Salman, Lama Sleem +2
cs.CVcs.AIarXiv:2608.15915v12026Don't Drop the BATON: Long-Horizon Robot Manipulation via Agentic Subtask Exploration and Transition-aware Memory
Bingxin Xu, Yuzhang Shang, Emilio Ferrara
cs.ROcs.AIcs.CVarXiv:2608.16889v12026Persistent Visual Memory: Sustaining Perception for Deep Generation in LVLMs
Siyuan Huang, Xiaoye Qu, Yafu Li +6
cs.CVcs.AIarXiv:2605.00814v22026WildTableBench: Benchmarking Multimodal Foundation Models on Table Understanding In the Wild
Junzhe Huang, Xiaoxiao Sun, Yan Yang +6
cs.CVarXiv:2605.01018v22026Missing Old Logits in Asynchronous Agentic RL: Semantic Mismatch and Repair Methods for Off-Policy Correction
Zhong Guan, Yongjian Guo, Haoran Sun +5
cs.LGcs.AIarXiv:2605.12070v22026Asymmetric Flow Models
Hansheng Chen, Jan Ackermann, Minseo Kim +2
cs.CVarXiv:2605.12964v22026SkCC: Portable and Secure Skill Compilation for Cross-Framework LLM Agents
Yipeng Ouyang, Yi Xiao, Yuhao Gu +1
cs.CRcs.AIarXiv:2605.03353v42026Generative Quantum-inspired Kolmogorov-Arnold Eigensolver
Yu-Cheng Lin, Yu-Chao Hsu, I-Shan Tsai +9
quant-phcs.LGarXiv:2605.04604v12026HAGE: Harnessing Agentic Memory via RL-Driven Weighted Graph Evolution
Dongming Jiang, Yi Li, Guanpeng Li +2
cs.AIarXiv:2605.09942v12026OmniHumanoid: Streaming Cross-Embodiment Video Generation with Paired-Free Adaptation
Yiren Song, Xiyao Deng, Pei Yang +2
cs.CVarXiv:2605.12038v12026Pion: A Spectrum-Preserving Optimizer via Orthogonal Equivalence Transformation
Kexuan Shi, Hanxuan Li, Zeju Qiu +3
cs.LGstat.MLarXiv:2605.12492v12026Diagnosing Dense Same-Class Attribute Misbinding in Large Vision-Language Models
Yuanzhi Xu, Qian Gao, Jun Fan +4
cs.CVcs.AIarXiv:2608.16805v12026JoyAI-Image: Awaking Spatial Intelligence in Unified Multimodal Understanding and Generation
Lin Song, Wenbo Li, Guoqing Ma +16
cs.GRcs.AIcs.CLarXiv:2605.04128v22026Agent-BRACE: Decoupling Beliefs from Actions in Long-Horizon Tasks via Verbalized State Uncertainty
Joykirat Singh, Zaid Khan, Archiki Prasad +5
cs.CLcs.AIarXiv:2605.11436v12026Steering the Flow: Inverting Face Recognition Models via Gradient-Guided Flow Matching
Ye Lu, Shen Wang, Zhaoyang Zhang +4
cs.CVcs.AIcs.CRarXiv:2608.16791v12026MISA: Mixture of Indexer Sparse Attention for Long-Context LLM Inference
Ruijie Zhou, Fanxu Meng, Yufei Xu +4
cs.LGcs.AIarXiv:2605.07363v12026jina-embeddings-v5-omni: Geometry-preserving Embeddings via Locked Aligned Towers
Florian Hönicke, Michael Günther, Andreas Koukounas +3
cs.CLarXiv:2605.08384v42026HarmTrace: Anchor-Calibrated Decoupled Optimization for Fine-Grained Target Identification in Harmful Memes
Yujia Li, Yiqun Zhang, Zihan Cheng +7
cs.CVcs.AIarXiv:2608.16622v12026AMPLIFAI: A Multiphase CT Dataset for Benchmarking Clinical Reasoning in LI-RADS Assessment of Liver Lesions
Pranav Kulkarni, Nikhil Shah, Amritansh Suryavanshi +9
cs.CVcs.LGarXiv:2608.14778v12026Learning from Language Feedback via Variational Policy Distillation
Yang Li, Erik Nijkamp, Semih Yavuz +1
cs.LGarXiv:2605.15113v22026Graph Machine Learning: An Opportunity for Power Systems
Martin Sadric, Sebastian Pütz, Christian Nauck +4
cs.LGcs.AIcs.CEarXiv:2608.16494v12026Native Audio-Visual Alignment for Generation
Longbin Ji, Guan Wang, Xuan Wei +6
cs.CVarXiv:2605.30073v12026RagGAD: Rationale-Aware Conditional Gaussian Mixture Normalizing Flow for Unsupervised Graph Anomaly Detection
Junxin Lu, Jing Zhao, Shiliang Sun
cs.LGcs.AIarXiv:2608.16018v12026GEO-Flag: Detecting and Measuring GEO-Optimized Web Content
Junjie Chu, Ye Leng, Mingjie Li +3
cs.LGcs.CRcs.IRarXiv:2608.16824v12026SAUL: Sharpness-Aware Augmented-Lagrangian Unlearning
Jaewan Choi, Junyoung Yang, Sangdon Park
cs.LGarXiv:2608.16249v12026ATLAS: Scaffold-Free Algorithm Synthesis by LLMs via Embedding-Guided Quality-Diversity Search
Danial Yazdani, Mohammad Nabi Omidvar, Yuan Sun +2
cs.AIcs.NEarXiv:2608.15546v22026Beyond Individual Intelligence: Surveying Collaboration, Failure Attribution, and Self-Evolution in LLM-based Multi-Agent Systems
Shihao Qi, Jie Ma, Rui Xing +15
cs.AIarXiv:2605.14892v22026Model-Adaptive Tool Necessity Reveals the Knowing-Doing Gap in LLM Tool Use
Yize Cheng, Chenrui Fan, Mahdi JafariRaviz +2
cs.AIarXiv:2605.14038v22026Delta Attention Residuals
Cheng Luo, Zefan Cai, Junjie Hu
cs.LGcs.CVarXiv:2605.18855v12026Summaries:한국어AgentFugue: Agent Scaling for Long-Horizon Tasks through Collective Reasoning
Yuyang Hu, Hongjin Qian, Shuting Wang +5
cs.AIcs.CLarXiv:2605.24486v12026Time-Aware Validation of Machine Learning Fuel Consumption Models: Evidence from 1\,Hz Operational Data, CCGS \textit{Sir Wilfrid Laurier}
Samarasimha Reddy Chittamuru, Ayhan Akinturk, Allison Kennedy +2
cs.LGarXiv:2608.16833v12026Towards Reasonable Molecular Structure Elucidation from Infrared Spectroscopy with Chemical Feedback
Yusen Tan, Hongyu Zhan, Hai-tao Yu +3
cs.LGarXiv:2608.16082v12026Geometry-Aware Image Flow Matching
Junho Lee, Kwanseok Kim, Joonseok Lee
cs.CVarXiv:2605.25294v12026The Ethical Decision Head: Operationalizing Normative Ethics in Autonomous Vehicles via Reinforcement Learning from Human Feedback
Thomas Mbrice, Ammar Ali, Sami Mian +5
cs.LGarXiv:2608.16710v12026Contrastive Energy Fields for Inference-Time Procedure Planning in Instructional Videos
Mohamed Afham, Christoph Reich, Oliver Hahn +2
cs.CVcs.AIarXiv:2608.16457v12026Graph Neural Assisted Actor-Critic for Latency-Efficient Edge Vision System
Alam Noor, Luis Almeida, Kai Li +3
cs.CVcs.AIarXiv:2608.16142v12026MUSE: An Interactive Meta-Agent for Understanding and Steering LLM-powered Data Science Systems
Wei-Hao Chen, Weixi Tong, Yuan Tian +2
cs.HCcs.AIarXiv:2608.16181v12026AutoLab: Can Frontier Models Solve Long-Horizon Auto Research and Engineering Tasks?
Zhangchen Xu, Junda Chen, Yue Huang +16
cs.AIcs.LGarXiv:2606.05080v12026Online Skill Learning for Web Agents via State-Grounded Dynamic Retrieval
Jiaxi Li, Ke Deng, Yun Wang +5
cs.AIarXiv:2606.04391v12026CrevasseSeg: A Label-Efficient UAV Crevasse Segmentation Framework
Steven Wallace, William D. Harcourt, Richard Hann +3
cs.LGarXiv:2608.15790v22026SparDA: Sparse Decoupled Attention for Efficient Long-Context LLM Inference
Yaosheng Fu, Guangxuan Xiao, Xin Dong +2
cs.CLcs.LGarXiv:2606.04511v12026P3D-Bench: Benchmarking MLLMs for Parametric 3D Generation and Structural Reasoning
Yikang Yang, Zhanpeng Hu, Youtian Lin +5
cs.CVarXiv:2606.11152v22026A Cognitively Motivated Multidimensional Framework for Evaluating Metaphor Explanations
Ana Naveriani, Jakob Suchan, Stefano Zoia +3
cs.CLcs.AIarXiv:2608.15828v12026Boosting Omni-Modal Language Models: Staged Post-Training with Visually Debiased Evaluation
Che Liu, Lichao Ma, Xiangyu Tony Zhang +4
cs.MMcs.AIcs.CVarXiv:2605.12034v22026CiteVQA: Benchmarking Evidence Attribution for Trustworthy Document Intelligence
Dongsheng Ma, Jiayu Li, Zhengren Wang +8
cs.CLcs.CVarXiv:2605.12882v12026KVServe: Service-Aware KV Cache Compression for Communication-Efficient Disaggregated LLM Serving
Zedong Liu, Xinyang Ma, Dejun Luo +9
cs.DCcs.AIcs.NIarXiv:2605.13734v12026Toward AI-Friendly Cartography: Understanding How Color Design Influences Foundation Model Spatial Reasoning on Sequential Choropleth Maps
Yonghe Sun, Zhenjia Liu, Hua Liao +4
cs.AIarXiv:2608.15736v12026Large language model-assisted discovery of cohorts from scientific literature
Moritz Sturm, Lisa M. Berg, Inken Berg +6
cs.IRcs.CLarXiv:2608.15909v12026FashionChameleon: Towards Real-Time and Interactive Human-Garment Video Customization
Quanjian Song, Yefeng Shen, Mengting Chen +5
cs.CVarXiv:2605.15824v22026Look Before You Leap: Autonomous Exploration for LLM Agents
Ziang Ye, Wentao Shi, Yuxin Liu +6
cs.AIcs.CLarXiv:2605.16143v12026Iterative Self-Learning for Expressive Text-to-Speech Synthesis
Nicholas Sanders, Gustav Eje Henter, Simon King +1
eess.AScs.CLcs.SDarXiv:2608.15910v12026RoPE Distinguishes Neither Positions Nor Tokens in Long Contexts, Provably
Yufeng Du, Phillip Harris, Minyang Tian +5
cs.CLcs.AIcs.LGarXiv:2605.15514v12026Full Attention Strikes Back: Transferring Full Attention into Sparse within Hundred Training Steps
Yanke Zhou, Yiduo Li, Hanlin Tang +6
cs.CLcs.AIarXiv:2605.16928v22026