Every paper with a summary

Every arXiv paper Paperlayer has summarized so far. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.

Search paper metadata (including unsummarized papers)

25,861 to 25,920 of 61,351

  1. SpatialVLA: Exploring Spatial Representations for Visual-Language-Action Model

    Delin Qu, Haoming Song, Qizhi Chen +8

    cs.ROcs.AIarXiv:2501.15830v52025
  2. A Simple Neural Attentive Meta-Learner

    Nikhil Mishra, Mostafa Rohaninejad, Xi Chen +1

    cs.AIcs.LGcs.NEarXiv:1707.03141v32017
  3. Deciding superellipticity and computing the Weierstrass normal form

    T. Shaska

    math.AGcs.SCmath.NTarXiv:2609.00672v12026
  4. Use HiResCAM instead of Grad-CAM for faithful explanations of convolutional neural networks

    Rachel Lea Draelos, Lawrence Carin

    eess.IVcs.CVcs.LGarXiv:2011.08891v42020
  5. UniVLA: Learning to Act Anywhere with Task-centric Latent Actions

    Qingwen Bu, Yanting Yang, Jisong Cai +5

    cs.ROcs.AIcs.LGarXiv:2505.06111v32025
  6. GenScale: A Benchmark for Relative Object Scale in Image Generation and Editing

    Lingxiao Li, Max Whitton, Ledell Wu +1

    cs.CVarXiv:2609.00525v12026
  7. EarthLD: Towards Unified Open-World Landslide Understanding via Vision-Language Guided Diffusion Models

    Yuanchao Su, Lianru Gao, Mengying Jiang +3

    cs.CVarXiv:2609.00712v12026
  8. Exploring and Unleashing the Power of Large Language Models in Automated Code Translation

    Zhen Yang, Fang Liu, Zhongxing Yu +7

    cs.SEcs.AIarXiv:2404.14646v22024
  9. SFAD: Speculative Factuality-Aware Decoding

    Guanqiao Chen, Di Wang, Lijie Hu

    cs.CLarXiv:2609.00796v22026
  10. Okutama-Action: An Aerial View Video Dataset for Concurrent Human Action Detection

    Mohammadamin Barekatain, Miquel Martí, Hsueh-Fu Shih +4

    cs.CVarXiv:1706.03038v22017
  11. SiGMa: Simple Greedy Matching for Aligning Large Knowledge Bases

    Simon Lacoste-Julien, Konstantina Palla, Alex Davies +3

    cs.AIcs.DBcs.IRarXiv:1207.4525v12012
  12. SimpleRL-Zoo: Investigating and Taming Zero Reinforcement Learning for Open Base Models in the Wild

    Weihao Zeng, Yuzhen Huang, Qian Liu +4

    cs.LGcs.AIcs.CLarXiv:2503.18892v32025
  13. Scaling Near-Optimal SFT-RL Annotation Budget Allocation from Small to Large LLMs

    Jingtan Wang, Arun Verma, Xiaoqiang Lin +4

    cs.CLcs.AIcs.LGarXiv:2609.01573v12026
  14. Intriguing Properties of Contrastive Losses

    Ting Chen, Calvin Luo, Lala Li

    cs.LGcs.AIcs.CVarXiv:2011.02803v32020
  15. GSVA: Generalized Segmentation via Multimodal Large Language Models

    Zhuofan Xia, Dongchen Han, Yizeng Han +3

    cs.CVarXiv:2312.10103v32023
  16. Debug like a Human: A Large Language Model Debugger via Verifying Runtime Execution Step-by-step

    Li Zhong, Zilong Wang, Jingbo Shang

    cs.SEcs.AIcs.CLarXiv:2402.16906v62024
  17. GEPA: Reflective Prompt Evolution Can Outperform Reinforcement Learning

    Lakshya A Agrawal, Shangyin Tan, Dilara Soylu +14

    cs.CLcs.AIcs.LGarXiv:2507.19457v22025
  18. MobileVLM V2: Faster and Stronger Baseline for Vision Language Model

    Xiangxiang Chu, Limeng Qiao, Xinyu Zhang +8

    cs.CVcs.AIarXiv:2402.03766v12024
  19. BigEarthNet-MM: A Large Scale Multi-Modal Multi-Label Benchmark Archive for Remote Sensing Image Classification and Retrieval

    Gencer Sumbul, Arne de Wall, Tristan Kreuziger +6

    cs.CVarXiv:2105.07921v22021
  20. A Study of Conditional Diffusion Models for Open-Loop Control under Dry Friction and Stiction

    Eric Aislan Antonelo

    cs.LGarXiv:2609.01756v12026
  21. Efficiently Approximating the Minimum-Volume Bounding Box of a Point Set in Three Dimensions

    Gill Barequet, Sariel Har-Peled

    cs.CGarXiv:2512.12391v12025
  22. Sequence-to-Sequence Knowledge Graph Completion and Question Answering

    Apoorv Saxena, Adrian Kochsiek, Rainer Gemulla

    cs.CLcs.LGarXiv:2203.10321v12022
  23. Model Context Protocol (MCP): Landscape, Security Threats, and Future Research Directions

    Xinyi Hou, Yanjie Zhao, Shenao Wang +1

    cs.CRcs.AIarXiv:2503.23278v32025
  24. Super-Resolution Delay-Doppler Estimation for OFDM Passive Radar

    Le Zheng, Xiaodong Wang

    cs.ITarXiv:1610.04218v22016
  25. LeWorldModel: Stable End-to-End Joint-Embedding Predictive Architecture from Pixels

    Lucas Maes, Quentin Le Lidec, Damien Scieur +2

    cs.LGcs.AIarXiv:2603.19312v32026
  26. Graph Convolution for Multimodal Information Extraction from Visually Rich Documents

    Xiaojing Liu, Feiyu Gao, Qiong Zhang +1

    cs.IRcs.CVcs.LGarXiv:1903.11279v12019
  27. HealthBench: Evaluating Large Language Models Towards Improved Human Health

    Rahul K. Arora, Jason Wei, Rebecca Soskin Hicks +9

    cs.CLarXiv:2505.08775v12025
  28. TSLANet: Rethinking Transformers for Time Series Representation Learning

    Emadeldeen Eldele, Mohamed Ragab, Zhenghua Chen +2

    cs.LGstat.MLarXiv:2404.08472v22024
  29. Step1X-Edit: A Practical Framework for General Image Editing

    Shiyu Liu, Yucheng Han, Peng Xing +21

    cs.CVarXiv:2504.17761v52025
  30. Global Self-Attention as a Replacement for Graph Convolution

    Md Shamim Hussain, Mohammed J. Zaki, Dharmashankar Subramanian

    cs.LGarXiv:2108.03348v32021
  31. ASTEC -- the Aarhus STellar Evolution Code

    J. Christensen-Dalsgaard

    astro-pharXiv:0710.3114v12007
  32. LIMO: Less is More for Reasoning

    Yixin Ye, Zhen Huang, Yang Xiao +3

    cs.CLcs.AIarXiv:2502.03387v32025
  33. Utility Optimal Scheduling in Energy Harvesting Networks

    Longbo Huang, Michael J. Neely

    math.OCarXiv:1012.1945v12010
  34. Smart Contracts Claimed Vulnerable by the CVE Database, with Labels and Source Locations

    Monika di Angelo, Gernot Salzer

    cs.CRcs.SEarXiv:2609.01186v12026
  35. Consistency Policy: Accelerated Visuomotor Policies via Consistency Distillation

    Aaditya Prasad, Kevin Lin, Jimmy Wu +2

    cs.ROcs.AIarXiv:2405.07503v22024
  36. YOLOv13: Real-Time Object Detection with Hypergraph-Enhanced Adaptive Visual Perception

    Mengqi Lei, Siqi Li, Yihong Wu +7

    cs.CVarXiv:2506.17733v22025
  37. Dynamics Based 3D Skeletal Hand Tracking

    Stan Melax, Leonid Keselman, Sterling Orsten

    cs.CVcs.GRarXiv:1705.07640v12017
  38. VideoLLaMA 3: Frontier Multimodal Foundation Models for Image and Video Understanding

    Boqiang Zhang, Kehan Li, Zesen Cheng +12

    cs.CVarXiv:2501.13106v42025
  39. ChronoNet: A Deep Recurrent Neural Network for Abnormal EEG Identification

    Subhrajit Roy, Isabell Kiral-Kornek, Stefan Harrer

    eess.SPcs.LGarXiv:1802.00308v22018
  40. Residual Attention: A Simple but Effective Method for Multi-Label Recognition

    Ke Zhu, Jianxin Wu

    cs.CVarXiv:2108.02456v22021
  41. Fast-WAM: Do World Action Models Need Test-time Future Imagination?

    Tianyuan Yuan, Zibin Dong, Yicheng Liu +1

    cs.CVcs.AIarXiv:2603.16666v22026
  42. Fast-dLLM: Training-free Acceleration of Diffusion LLM by Enabling KV Cache and Parallel Decoding

    Chengyue Wu, Hao Zhang, Shuchen Xue +6

    cs.CLarXiv:2505.22618v32025
  43. Penalized Composite Quasi-Likelihood for Ultrahigh-Dimensional Variable Selection

    Jelena Bradic, Jianqing Fan, Weiwei Wang

    stat.MEmath.STarXiv:0912.5200v22009
  44. DoFE: Domain-oriented Feature Embedding for Generalizable Fundus Image Segmentation on Unseen Datasets

    Shujun Wang, Lequan Yu, Kang Li +3

    cs.CVarXiv:2010.06208v12020
  45. The Entropy Mechanism of Reinforcement Learning for Reasoning Language Models

    Ganqu Cui, Yuchen Zhang, Jiacheng Chen +14

    cs.LGcs.AIcs.CLarXiv:2505.22617v12025
  46. MedGemma Technical Report

    Andrew Sellergren, Sahar Kazemzadeh, Tiam Jaroensri +78

    cs.AIcs.CLcs.CVarXiv:2507.05201v42025
  47. Escaping Redundant Reasoning: Structure-Aware Search for Inference-Time LLMs

    Lu Cheng

    cs.AIarXiv:2609.00738v12026
  48. Vision-R1: Incentivizing Reasoning Capability in Multimodal Large Language Models

    Wenxuan Huang, Bohan Jia, Zijie Zhai +7

    cs.CVcs.AIcs.CLarXiv:2503.06749v42025
  49. Grounded Human-Object Interaction Hotspots from Video

    Tushar Nagarajan, Christoph Feichtenhofer, Kristen Grauman

    cs.CVarXiv:1812.04558v22018
  50. Simple linear attention language models balance the recall-throughput tradeoff

    Simran Arora, Sabri Eyuboglu, Michael Zhang +6

    cs.CLcs.LGarXiv:2402.18668v22024
  51. FAST: Efficient Action Tokenization for Vision-Language-Action Models

    Karl Pertsch, Kyle Stachowicz, Brian Ichter +6

    cs.ROcs.LGarXiv:2501.09747v12025
  52. FUSE: An Evaluating Framework for Dangerous Capabilities of LLMs

    Zhengyi Jin, Ru Zhang, Xiao Chen +5

    cs.AIarXiv:2609.02168v12026
  53. Social Learning and Distributed Hypothesis Testing

    Anusha Lalitha, Tara Javidi, Anand Sarwate

    math.STcs.ITmath.OCarXiv:1410.4307v52014
    Summaries:한국어
  54. Probabilistic Tools for the Analysis of Randomized Optimization Heuristics

    Benjamin Doerr

    cs.DScs.DMcs.NEarXiv:1801.06733v62018
  55. Why Do Multi-Agent LLM Systems Fail?

    Mert Cemri, Melissa Z. Pan, Shuyi Yang +10

    cs.AIarXiv:2503.13657v32025
  56. Comprehensive Graph-conditional Similarity Preserving Network for Unsupervised Cross-modal Hashing

    Jun Yu, Hao Zhou, Yibing Zhan +1

    cs.IRcs.CVarXiv:2012.13538v12020
  57. "It's a Fair Game", or Is It? Examining How Users Navigate Disclosure Risks and Benefits When Using LLM-Based Conversational Agents

    Zhiping Zhang, Michelle Jia, Hao-Ping Lee +5

    cs.HCcs.AIcs.CRarXiv:2309.11653v22023
  58. Understanding metric-related pitfalls in image analysis validation

    Annika Reinke, Minu D. Tizabi, Michael Baumgartner +75

    cs.CVarXiv:2302.01790v42023
  59. Beyond the 80/20 Rule: High-Entropy Minority Tokens Drive Effective Reinforcement Learning for LLM Reasoning

    Shenzhi Wang, Le Yu, Chang Gao +15

    cs.CLcs.AIcs.LGarXiv:2506.01939v22025
  60. Optimus: Organizing Sentences via Pre-trained Modeling of a Latent Space

    Chunyuan Li, Xiang Gao, Yuan Li +4

    cs.CLcs.LGstat.MLarXiv:2004.04092v42020