Every paper with a summary
Every arXiv paper Paperlayer has summarized so far. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.
Search paper metadata (including unsummarized papers)
25,861 to 25,920 of 61,351
SpatialVLA: Exploring Spatial Representations for Visual-Language-Action Model
Delin Qu, Haoming Song, Qizhi Chen +8
cs.ROcs.AIarXiv:2501.15830v52025A Simple Neural Attentive Meta-Learner
Nikhil Mishra, Mostafa Rohaninejad, Xi Chen +1
cs.AIcs.LGcs.NEarXiv:1707.03141v32017Deciding superellipticity and computing the Weierstrass normal form
T. Shaska
math.AGcs.SCmath.NTarXiv:2609.00672v12026Use HiResCAM instead of Grad-CAM for faithful explanations of convolutional neural networks
Rachel Lea Draelos, Lawrence Carin
eess.IVcs.CVcs.LGarXiv:2011.08891v42020UniVLA: Learning to Act Anywhere with Task-centric Latent Actions
Qingwen Bu, Yanting Yang, Jisong Cai +5
cs.ROcs.AIcs.LGarXiv:2505.06111v32025GenScale: A Benchmark for Relative Object Scale in Image Generation and Editing
Lingxiao Li, Max Whitton, Ledell Wu +1
cs.CVarXiv:2609.00525v12026EarthLD: Towards Unified Open-World Landslide Understanding via Vision-Language Guided Diffusion Models
Yuanchao Su, Lianru Gao, Mengying Jiang +3
cs.CVarXiv:2609.00712v12026Exploring and Unleashing the Power of Large Language Models in Automated Code Translation
Zhen Yang, Fang Liu, Zhongxing Yu +7
cs.SEcs.AIarXiv:2404.14646v22024SFAD: Speculative Factuality-Aware Decoding
Guanqiao Chen, Di Wang, Lijie Hu
cs.CLarXiv:2609.00796v22026Okutama-Action: An Aerial View Video Dataset for Concurrent Human Action Detection
Mohammadamin Barekatain, Miquel Martí, Hsueh-Fu Shih +4
cs.CVarXiv:1706.03038v22017SiGMa: Simple Greedy Matching for Aligning Large Knowledge Bases
Simon Lacoste-Julien, Konstantina Palla, Alex Davies +3
cs.AIcs.DBcs.IRarXiv:1207.4525v12012SimpleRL-Zoo: Investigating and Taming Zero Reinforcement Learning for Open Base Models in the Wild
Weihao Zeng, Yuzhen Huang, Qian Liu +4
cs.LGcs.AIcs.CLarXiv:2503.18892v32025Scaling Near-Optimal SFT-RL Annotation Budget Allocation from Small to Large LLMs
Jingtan Wang, Arun Verma, Xiaoqiang Lin +4
cs.CLcs.AIcs.LGarXiv:2609.01573v12026Intriguing Properties of Contrastive Losses
Ting Chen, Calvin Luo, Lala Li
cs.LGcs.AIcs.CVarXiv:2011.02803v32020GSVA: Generalized Segmentation via Multimodal Large Language Models
Zhuofan Xia, Dongchen Han, Yizeng Han +3
cs.CVarXiv:2312.10103v32023Debug like a Human: A Large Language Model Debugger via Verifying Runtime Execution Step-by-step
Li Zhong, Zilong Wang, Jingbo Shang
cs.SEcs.AIcs.CLarXiv:2402.16906v62024GEPA: Reflective Prompt Evolution Can Outperform Reinforcement Learning
Lakshya A Agrawal, Shangyin Tan, Dilara Soylu +14
cs.CLcs.AIcs.LGarXiv:2507.19457v22025MobileVLM V2: Faster and Stronger Baseline for Vision Language Model
Xiangxiang Chu, Limeng Qiao, Xinyu Zhang +8
cs.CVcs.AIarXiv:2402.03766v12024BigEarthNet-MM: A Large Scale Multi-Modal Multi-Label Benchmark Archive for Remote Sensing Image Classification and Retrieval
Gencer Sumbul, Arne de Wall, Tristan Kreuziger +6
cs.CVarXiv:2105.07921v22021A Study of Conditional Diffusion Models for Open-Loop Control under Dry Friction and Stiction
Eric Aislan Antonelo
cs.LGarXiv:2609.01756v12026Efficiently Approximating the Minimum-Volume Bounding Box of a Point Set in Three Dimensions
Gill Barequet, Sariel Har-Peled
cs.CGarXiv:2512.12391v12025Sequence-to-Sequence Knowledge Graph Completion and Question Answering
Apoorv Saxena, Adrian Kochsiek, Rainer Gemulla
cs.CLcs.LGarXiv:2203.10321v12022Model Context Protocol (MCP): Landscape, Security Threats, and Future Research Directions
Xinyi Hou, Yanjie Zhao, Shenao Wang +1
cs.CRcs.AIarXiv:2503.23278v32025Super-Resolution Delay-Doppler Estimation for OFDM Passive Radar
Le Zheng, Xiaodong Wang
cs.ITarXiv:1610.04218v22016LeWorldModel: Stable End-to-End Joint-Embedding Predictive Architecture from Pixels
Lucas Maes, Quentin Le Lidec, Damien Scieur +2
cs.LGcs.AIarXiv:2603.19312v32026Graph Convolution for Multimodal Information Extraction from Visually Rich Documents
Xiaojing Liu, Feiyu Gao, Qiong Zhang +1
cs.IRcs.CVcs.LGarXiv:1903.11279v12019HealthBench: Evaluating Large Language Models Towards Improved Human Health
Rahul K. Arora, Jason Wei, Rebecca Soskin Hicks +9
cs.CLarXiv:2505.08775v12025TSLANet: Rethinking Transformers for Time Series Representation Learning
Emadeldeen Eldele, Mohamed Ragab, Zhenghua Chen +2
cs.LGstat.MLarXiv:2404.08472v22024Step1X-Edit: A Practical Framework for General Image Editing
Shiyu Liu, Yucheng Han, Peng Xing +21
cs.CVarXiv:2504.17761v52025Global Self-Attention as a Replacement for Graph Convolution
Md Shamim Hussain, Mohammed J. Zaki, Dharmashankar Subramanian
cs.LGarXiv:2108.03348v32021ASTEC -- the Aarhus STellar Evolution Code
J. Christensen-Dalsgaard
astro-pharXiv:0710.3114v12007LIMO: Less is More for Reasoning
Yixin Ye, Zhen Huang, Yang Xiao +3
cs.CLcs.AIarXiv:2502.03387v32025Utility Optimal Scheduling in Energy Harvesting Networks
Longbo Huang, Michael J. Neely
math.OCarXiv:1012.1945v12010Smart Contracts Claimed Vulnerable by the CVE Database, with Labels and Source Locations
Monika di Angelo, Gernot Salzer
cs.CRcs.SEarXiv:2609.01186v12026Consistency Policy: Accelerated Visuomotor Policies via Consistency Distillation
Aaditya Prasad, Kevin Lin, Jimmy Wu +2
cs.ROcs.AIarXiv:2405.07503v22024YOLOv13: Real-Time Object Detection with Hypergraph-Enhanced Adaptive Visual Perception
Mengqi Lei, Siqi Li, Yihong Wu +7
cs.CVarXiv:2506.17733v22025Dynamics Based 3D Skeletal Hand Tracking
Stan Melax, Leonid Keselman, Sterling Orsten
cs.CVcs.GRarXiv:1705.07640v12017VideoLLaMA 3: Frontier Multimodal Foundation Models for Image and Video Understanding
Boqiang Zhang, Kehan Li, Zesen Cheng +12
cs.CVarXiv:2501.13106v42025ChronoNet: A Deep Recurrent Neural Network for Abnormal EEG Identification
Subhrajit Roy, Isabell Kiral-Kornek, Stefan Harrer
eess.SPcs.LGarXiv:1802.00308v22018Residual Attention: A Simple but Effective Method for Multi-Label Recognition
Ke Zhu, Jianxin Wu
cs.CVarXiv:2108.02456v22021Fast-WAM: Do World Action Models Need Test-time Future Imagination?
Tianyuan Yuan, Zibin Dong, Yicheng Liu +1
cs.CVcs.AIarXiv:2603.16666v22026Fast-dLLM: Training-free Acceleration of Diffusion LLM by Enabling KV Cache and Parallel Decoding
Chengyue Wu, Hao Zhang, Shuchen Xue +6
cs.CLarXiv:2505.22618v32025Penalized Composite Quasi-Likelihood for Ultrahigh-Dimensional Variable Selection
Jelena Bradic, Jianqing Fan, Weiwei Wang
stat.MEmath.STarXiv:0912.5200v22009DoFE: Domain-oriented Feature Embedding for Generalizable Fundus Image Segmentation on Unseen Datasets
Shujun Wang, Lequan Yu, Kang Li +3
cs.CVarXiv:2010.06208v12020The Entropy Mechanism of Reinforcement Learning for Reasoning Language Models
Ganqu Cui, Yuchen Zhang, Jiacheng Chen +14
cs.LGcs.AIcs.CLarXiv:2505.22617v12025MedGemma Technical Report
Andrew Sellergren, Sahar Kazemzadeh, Tiam Jaroensri +78
cs.AIcs.CLcs.CVarXiv:2507.05201v42025Escaping Redundant Reasoning: Structure-Aware Search for Inference-Time LLMs
Lu Cheng
cs.AIarXiv:2609.00738v12026Vision-R1: Incentivizing Reasoning Capability in Multimodal Large Language Models
Wenxuan Huang, Bohan Jia, Zijie Zhai +7
cs.CVcs.AIcs.CLarXiv:2503.06749v42025Grounded Human-Object Interaction Hotspots from Video
Tushar Nagarajan, Christoph Feichtenhofer, Kristen Grauman
cs.CVarXiv:1812.04558v22018Simple linear attention language models balance the recall-throughput tradeoff
Simran Arora, Sabri Eyuboglu, Michael Zhang +6
cs.CLcs.LGarXiv:2402.18668v22024FAST: Efficient Action Tokenization for Vision-Language-Action Models
Karl Pertsch, Kyle Stachowicz, Brian Ichter +6
cs.ROcs.LGarXiv:2501.09747v12025FUSE: An Evaluating Framework for Dangerous Capabilities of LLMs
Zhengyi Jin, Ru Zhang, Xiao Chen +5
cs.AIarXiv:2609.02168v12026Social Learning and Distributed Hypothesis Testing
Anusha Lalitha, Tara Javidi, Anand Sarwate
math.STcs.ITmath.OCarXiv:1410.4307v52014Summaries:한국어Probabilistic Tools for the Analysis of Randomized Optimization Heuristics
Benjamin Doerr
cs.DScs.DMcs.NEarXiv:1801.06733v62018Why Do Multi-Agent LLM Systems Fail?
Mert Cemri, Melissa Z. Pan, Shuyi Yang +10
cs.AIarXiv:2503.13657v32025Comprehensive Graph-conditional Similarity Preserving Network for Unsupervised Cross-modal Hashing
Jun Yu, Hao Zhou, Yibing Zhan +1
cs.IRcs.CVarXiv:2012.13538v12020"It's a Fair Game", or Is It? Examining How Users Navigate Disclosure Risks and Benefits When Using LLM-Based Conversational Agents
Zhiping Zhang, Michelle Jia, Hao-Ping Lee +5
cs.HCcs.AIcs.CRarXiv:2309.11653v22023Understanding metric-related pitfalls in image analysis validation
Annika Reinke, Minu D. Tizabi, Michael Baumgartner +75
cs.CVarXiv:2302.01790v42023Beyond the 80/20 Rule: High-Entropy Minority Tokens Drive Effective Reinforcement Learning for LLM Reasoning
Shenzhi Wang, Le Yu, Chang Gao +15
cs.CLcs.AIcs.LGarXiv:2506.01939v22025Optimus: Organizing Sentences via Pre-trained Modeling of a Latent Space
Chunyuan Li, Xiang Gao, Yuan Li +4
cs.CLcs.LGstat.MLarXiv:2004.04092v42020