Every paper with a summary

Every arXiv paper Paperlayer has summarized so far. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.

Search paper metadata (including unsummarized papers)

59,941 to 60,000 of 61,302

  1. Macaron-A2UI: A Model for Generative UI in Personal Agents

    Fancy Kong, Congjie Zheng, Murphy Zhuang +8

    cs.HCarXiv:2605.24830v12026
  2. Hallucinations Undermine Trust; Metacognition is a Way Forward

    Gal Yona, Mor Geva, Yossi Matias

    cs.CLarXiv:2605.01428v12026
  3. Auto-Rubric as Reward: From Implicit Preferences to Explicit Multimodal Generative Criteria

    Juanxi Tian, Fengyuan Liu, Jiaming Han +6

    cs.AIarXiv:2605.08354v12026
  4. AeroCopilotBench: A Two-Tier Benchmark for Evaluating LLM Agents as Aviation Copilots in an Interactive Virtual Cockpit Environment

    Yuchen Yuan, Zhenghuang Wu, Yuangan Li +2

    cs.AIarXiv:2608.16349v12026
  5. Towards On-Policy Data Evolution for Visual-Native Multimodal Deep Search Agents

    Shijue Huang, Hangyu Guo, Guanting Dong +8

    cs.CLarXiv:2605.10832v22026
  6. Baseline-Relative Counterfactual Refinement for Bit-Aware Visual Token Communication

    Jia Guo, Xiaohan Zhao, Changwang Liu +4

    cs.AIarXiv:2608.16192v12026
  7. OpenSTBench: Beyond Semantic Evaluation for Speech Translation

    Yanjie An, Yuxiang Zhao, Yichi Zhang +5

    eess.AScs.AIarXiv:2605.30792v12026
  8. Comprehensive Benchmarking of Deep Learning Architectures for Lung Cancer Histopathology

    Hadi Hasan, Safaa Salman, Lama Sleem +2

    cs.CVcs.AIarXiv:2608.15915v12026
  9. Don't Drop the BATON: Long-Horizon Robot Manipulation via Agentic Subtask Exploration and Transition-aware Memory

    Bingxin Xu, Yuzhang Shang, Emilio Ferrara

    cs.ROcs.AIcs.CVarXiv:2608.16889v12026
  10. Persistent Visual Memory: Sustaining Perception for Deep Generation in LVLMs

    Siyuan Huang, Xiaoye Qu, Yafu Li +6

    cs.CVcs.AIarXiv:2605.00814v22026
  11. WildTableBench: Benchmarking Multimodal Foundation Models on Table Understanding In the Wild

    Junzhe Huang, Xiaoxiao Sun, Yan Yang +6

    cs.CVarXiv:2605.01018v22026
  12. Missing Old Logits in Asynchronous Agentic RL: Semantic Mismatch and Repair Methods for Off-Policy Correction

    Zhong Guan, Yongjian Guo, Haoran Sun +5

    cs.LGcs.AIarXiv:2605.12070v22026
  13. Asymmetric Flow Models

    Hansheng Chen, Jan Ackermann, Minseo Kim +2

    cs.CVarXiv:2605.12964v22026
  14. SkCC: Portable and Secure Skill Compilation for Cross-Framework LLM Agents

    Yipeng Ouyang, Yi Xiao, Yuhao Gu +1

    cs.CRcs.AIarXiv:2605.03353v42026
  15. Generative Quantum-inspired Kolmogorov-Arnold Eigensolver

    Yu-Cheng Lin, Yu-Chao Hsu, I-Shan Tsai +9

    quant-phcs.LGarXiv:2605.04604v12026
  16. HAGE: Harnessing Agentic Memory via RL-Driven Weighted Graph Evolution

    Dongming Jiang, Yi Li, Guanpeng Li +2

    cs.AIarXiv:2605.09942v12026
  17. OmniHumanoid: Streaming Cross-Embodiment Video Generation with Paired-Free Adaptation

    Yiren Song, Xiyao Deng, Pei Yang +2

    cs.CVarXiv:2605.12038v12026
  18. Pion: A Spectrum-Preserving Optimizer via Orthogonal Equivalence Transformation

    Kexuan Shi, Hanxuan Li, Zeju Qiu +3

    cs.LGstat.MLarXiv:2605.12492v12026
  19. Diagnosing Dense Same-Class Attribute Misbinding in Large Vision-Language Models

    Yuanzhi Xu, Qian Gao, Jun Fan +4

    cs.CVcs.AIarXiv:2608.16805v12026
  20. JoyAI-Image: Awaking Spatial Intelligence in Unified Multimodal Understanding and Generation

    Lin Song, Wenbo Li, Guoqing Ma +16

    cs.GRcs.AIcs.CLarXiv:2605.04128v22026
  21. Agent-BRACE: Decoupling Beliefs from Actions in Long-Horizon Tasks via Verbalized State Uncertainty

    Joykirat Singh, Zaid Khan, Archiki Prasad +5

    cs.CLcs.AIarXiv:2605.11436v12026
  22. Steering the Flow: Inverting Face Recognition Models via Gradient-Guided Flow Matching

    Ye Lu, Shen Wang, Zhaoyang Zhang +4

    cs.CVcs.AIcs.CRarXiv:2608.16791v12026
  23. MISA: Mixture of Indexer Sparse Attention for Long-Context LLM Inference

    Ruijie Zhou, Fanxu Meng, Yufei Xu +4

    cs.LGcs.AIarXiv:2605.07363v12026
  24. jina-embeddings-v5-omni: Geometry-preserving Embeddings via Locked Aligned Towers

    Florian Hönicke, Michael Günther, Andreas Koukounas +3

    cs.CLarXiv:2605.08384v42026
  25. HarmTrace: Anchor-Calibrated Decoupled Optimization for Fine-Grained Target Identification in Harmful Memes

    Yujia Li, Yiqun Zhang, Zihan Cheng +7

    cs.CVcs.AIarXiv:2608.16622v12026
  26. AMPLIFAI: A Multiphase CT Dataset for Benchmarking Clinical Reasoning in LI-RADS Assessment of Liver Lesions

    Pranav Kulkarni, Nikhil Shah, Amritansh Suryavanshi +9

    cs.CVcs.LGarXiv:2608.14778v12026
  27. Learning from Language Feedback via Variational Policy Distillation

    Yang Li, Erik Nijkamp, Semih Yavuz +1

    cs.LGarXiv:2605.15113v22026
  28. Graph Machine Learning: An Opportunity for Power Systems

    Martin Sadric, Sebastian Pütz, Christian Nauck +4

    cs.LGcs.AIcs.CEarXiv:2608.16494v12026
  29. Native Audio-Visual Alignment for Generation

    Longbin Ji, Guan Wang, Xuan Wei +6

    cs.CVarXiv:2605.30073v12026
  30. RagGAD: Rationale-Aware Conditional Gaussian Mixture Normalizing Flow for Unsupervised Graph Anomaly Detection

    Junxin Lu, Jing Zhao, Shiliang Sun

    cs.LGcs.AIarXiv:2608.16018v12026
  31. GEO-Flag: Detecting and Measuring GEO-Optimized Web Content

    Junjie Chu, Ye Leng, Mingjie Li +3

    cs.LGcs.CRcs.IRarXiv:2608.16824v12026
  32. SAUL: Sharpness-Aware Augmented-Lagrangian Unlearning

    Jaewan Choi, Junyoung Yang, Sangdon Park

    cs.LGarXiv:2608.16249v12026
  33. ATLAS: Scaffold-Free Algorithm Synthesis by LLMs via Embedding-Guided Quality-Diversity Search

    Danial Yazdani, Mohammad Nabi Omidvar, Yuan Sun +2

    cs.AIcs.NEarXiv:2608.15546v22026
  34. Beyond Individual Intelligence: Surveying Collaboration, Failure Attribution, and Self-Evolution in LLM-based Multi-Agent Systems

    Shihao Qi, Jie Ma, Rui Xing +15

    cs.AIarXiv:2605.14892v22026
  35. Model-Adaptive Tool Necessity Reveals the Knowing-Doing Gap in LLM Tool Use

    Yize Cheng, Chenrui Fan, Mahdi JafariRaviz +2

    cs.AIarXiv:2605.14038v22026
  36. Delta Attention Residuals

    Cheng Luo, Zefan Cai, Junjie Hu

    cs.LGcs.CVarXiv:2605.18855v12026
    Summaries:한국어
  37. AgentFugue: Agent Scaling for Long-Horizon Tasks through Collective Reasoning

    Yuyang Hu, Hongjin Qian, Shuting Wang +5

    cs.AIcs.CLarXiv:2605.24486v12026
  38. Time-Aware Validation of Machine Learning Fuel Consumption Models: Evidence from 1\,Hz Operational Data, CCGS \textit{Sir Wilfrid Laurier}

    Samarasimha Reddy Chittamuru, Ayhan Akinturk, Allison Kennedy +2

    cs.LGarXiv:2608.16833v12026
  39. Towards Reasonable Molecular Structure Elucidation from Infrared Spectroscopy with Chemical Feedback

    Yusen Tan, Hongyu Zhan, Hai-tao Yu +3

    cs.LGarXiv:2608.16082v12026
  40. Geometry-Aware Image Flow Matching

    Junho Lee, Kwanseok Kim, Joonseok Lee

    cs.CVarXiv:2605.25294v12026
  41. The Ethical Decision Head: Operationalizing Normative Ethics in Autonomous Vehicles via Reinforcement Learning from Human Feedback

    Thomas Mbrice, Ammar Ali, Sami Mian +5

    cs.LGarXiv:2608.16710v12026
  42. Contrastive Energy Fields for Inference-Time Procedure Planning in Instructional Videos

    Mohamed Afham, Christoph Reich, Oliver Hahn +2

    cs.CVcs.AIarXiv:2608.16457v12026
  43. Graph Neural Assisted Actor-Critic for Latency-Efficient Edge Vision System

    Alam Noor, Luis Almeida, Kai Li +3

    cs.CVcs.AIarXiv:2608.16142v12026
  44. MUSE: An Interactive Meta-Agent for Understanding and Steering LLM-powered Data Science Systems

    Wei-Hao Chen, Weixi Tong, Yuan Tian +2

    cs.HCcs.AIarXiv:2608.16181v12026
  45. AutoLab: Can Frontier Models Solve Long-Horizon Auto Research and Engineering Tasks?

    Zhangchen Xu, Junda Chen, Yue Huang +16

    cs.AIcs.LGarXiv:2606.05080v12026
  46. Online Skill Learning for Web Agents via State-Grounded Dynamic Retrieval

    Jiaxi Li, Ke Deng, Yun Wang +5

    cs.AIarXiv:2606.04391v12026
  47. CrevasseSeg: A Label-Efficient UAV Crevasse Segmentation Framework

    Steven Wallace, William D. Harcourt, Richard Hann +3

    cs.LGarXiv:2608.15790v22026
  48. SparDA: Sparse Decoupled Attention for Efficient Long-Context LLM Inference

    Yaosheng Fu, Guangxuan Xiao, Xin Dong +2

    cs.CLcs.LGarXiv:2606.04511v12026
  49. P3D-Bench: Benchmarking MLLMs for Parametric 3D Generation and Structural Reasoning

    Yikang Yang, Zhanpeng Hu, Youtian Lin +5

    cs.CVarXiv:2606.11152v22026
  50. A Cognitively Motivated Multidimensional Framework for Evaluating Metaphor Explanations

    Ana Naveriani, Jakob Suchan, Stefano Zoia +3

    cs.CLcs.AIarXiv:2608.15828v12026
  51. Boosting Omni-Modal Language Models: Staged Post-Training with Visually Debiased Evaluation

    Che Liu, Lichao Ma, Xiangyu Tony Zhang +4

    cs.MMcs.AIcs.CVarXiv:2605.12034v22026
  52. CiteVQA: Benchmarking Evidence Attribution for Trustworthy Document Intelligence

    Dongsheng Ma, Jiayu Li, Zhengren Wang +8

    cs.CLcs.CVarXiv:2605.12882v12026
  53. KVServe: Service-Aware KV Cache Compression for Communication-Efficient Disaggregated LLM Serving

    Zedong Liu, Xinyang Ma, Dejun Luo +9

    cs.DCcs.AIcs.NIarXiv:2605.13734v12026
  54. Toward AI-Friendly Cartography: Understanding How Color Design Influences Foundation Model Spatial Reasoning on Sequential Choropleth Maps

    Yonghe Sun, Zhenjia Liu, Hua Liao +4

    cs.AIarXiv:2608.15736v12026
  55. Large language model-assisted discovery of cohorts from scientific literature

    Moritz Sturm, Lisa M. Berg, Inken Berg +6

    cs.IRcs.CLarXiv:2608.15909v12026
  56. FashionChameleon: Towards Real-Time and Interactive Human-Garment Video Customization

    Quanjian Song, Yefeng Shen, Mengting Chen +5

    cs.CVarXiv:2605.15824v22026
  57. Look Before You Leap: Autonomous Exploration for LLM Agents

    Ziang Ye, Wentao Shi, Yuxin Liu +6

    cs.AIcs.CLarXiv:2605.16143v12026
  58. Iterative Self-Learning for Expressive Text-to-Speech Synthesis

    Nicholas Sanders, Gustav Eje Henter, Simon King +1

    eess.AScs.CLcs.SDarXiv:2608.15910v12026
  59. RoPE Distinguishes Neither Positions Nor Tokens in Long Contexts, Provably

    Yufeng Du, Phillip Harris, Minyang Tian +5

    cs.CLcs.AIcs.LGarXiv:2605.15514v12026
  60. Full Attention Strikes Back: Transferring Full Attention into Sparse within Hundred Training Steps

    Yanke Zhou, Yiduo Li, Hanlin Tang +6

    cs.CLcs.AIarXiv:2605.16928v22026