Every paper with a summary

Every arXiv paper Paperlayer has summarized so far. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.

Search paper metadata (including unsummarized papers)

25,561 to 25,620 of 61,100

  1. Text Segmentation as a Supervised Learning Task

    Omri Koshorek, Adir Cohen, Noam Mor +2

    cs.CLarXiv:1803.09337v12018
  2. Projective Affine Body Dynamics for Multibody Systems

    Zimeng Ye, Xiaowei He, Yuzhong Guo +3

    cs.GRarXiv:2609.02675v12026
  3. Deep Learning Approaches on Image Captioning: A Review

    Taraneh Ghandi, Hamidreza Pourreza, Hamidreza Mahyar

    cs.CVarXiv:2201.12944v52022
  4. Inversion by Direct Iteration: An Alternative to Denoising Diffusion for Image Restoration

    Mauricio Delbracio, Peyman Milanfar

    eess.IVcs.CVcs.LGarXiv:2303.11435v52023
  5. Age of Information in Random Access Channels

    Xingran Chen, Konstantinos Gatsis, Hamed Hassani +1

    cs.NIcs.ITarXiv:1912.01473v62019
  6. Agents That Model Agents: Five Principles Toward a Theory of Mind for 6G Networks

    Hatim Chergui, Carolina Fernández-Martínez, Mehdi Bennis +1

    cs.NIcs.AIcs.MAarXiv:2609.01779v12026
  7. SCOP: Scientific Control for Reliable Neural Network Pruning

    Yehui Tang, Yunhe Wang, Yixing Xu +4

    cs.CVcs.LGarXiv:2010.10732v22020
  8. The Ceiling Is in the Channel: Auditing Learner Gaps and Measurement Frontiers in Clinical Prediction

    Sayeed Shafayet Chowdhury, Nusrat Jahan, Snehasis Mukhopadhyay +2

    cs.AIarXiv:2609.01909v12026
  9. SpatialVLA: Exploring Spatial Representations for Visual-Language-Action Model

    Delin Qu, Haoming Song, Qizhi Chen +8

    cs.ROcs.AIarXiv:2501.15830v52025
  10. A Simple Neural Attentive Meta-Learner

    Nikhil Mishra, Mostafa Rohaninejad, Xi Chen +1

    cs.AIcs.LGcs.NEarXiv:1707.03141v32017
  11. Deciding superellipticity and computing the Weierstrass normal form

    T. Shaska

    math.AGcs.SCmath.NTarXiv:2609.00672v12026
  12. Use HiResCAM instead of Grad-CAM for faithful explanations of convolutional neural networks

    Rachel Lea Draelos, Lawrence Carin

    eess.IVcs.CVcs.LGarXiv:2011.08891v42020
  13. UniVLA: Learning to Act Anywhere with Task-centric Latent Actions

    Qingwen Bu, Yanting Yang, Jisong Cai +5

    cs.ROcs.AIcs.LGarXiv:2505.06111v32025
  14. GenScale: A Benchmark for Relative Object Scale in Image Generation and Editing

    Lingxiao Li, Max Whitton, Ledell Wu +1

    cs.CVarXiv:2609.00525v12026
  15. EarthLD: Towards Unified Open-World Landslide Understanding via Vision-Language Guided Diffusion Models

    Yuanchao Su, Lianru Gao, Mengying Jiang +3

    cs.CVarXiv:2609.00712v12026
  16. Exploring and Unleashing the Power of Large Language Models in Automated Code Translation

    Zhen Yang, Fang Liu, Zhongxing Yu +7

    cs.SEcs.AIarXiv:2404.14646v22024
  17. SFAD: Speculative Factuality-Aware Decoding

    Guanqiao Chen, Di Wang, Lijie Hu

    cs.CLarXiv:2609.00796v22026
  18. Okutama-Action: An Aerial View Video Dataset for Concurrent Human Action Detection

    Mohammadamin Barekatain, Miquel Martí, Hsueh-Fu Shih +4

    cs.CVarXiv:1706.03038v22017
  19. SiGMa: Simple Greedy Matching for Aligning Large Knowledge Bases

    Simon Lacoste-Julien, Konstantina Palla, Alex Davies +3

    cs.AIcs.DBcs.IRarXiv:1207.4525v12012
  20. SimpleRL-Zoo: Investigating and Taming Zero Reinforcement Learning for Open Base Models in the Wild

    Weihao Zeng, Yuzhen Huang, Qian Liu +4

    cs.LGcs.AIcs.CLarXiv:2503.18892v32025
  21. Scaling Near-Optimal SFT-RL Annotation Budget Allocation from Small to Large LLMs

    Jingtan Wang, Arun Verma, Xiaoqiang Lin +4

    cs.CLcs.AIcs.LGarXiv:2609.01573v12026
  22. Intriguing Properties of Contrastive Losses

    Ting Chen, Calvin Luo, Lala Li

    cs.LGcs.AIcs.CVarXiv:2011.02803v32020
  23. GSVA: Generalized Segmentation via Multimodal Large Language Models

    Zhuofan Xia, Dongchen Han, Yizeng Han +3

    cs.CVarXiv:2312.10103v32023
  24. Debug like a Human: A Large Language Model Debugger via Verifying Runtime Execution Step-by-step

    Li Zhong, Zilong Wang, Jingbo Shang

    cs.SEcs.AIcs.CLarXiv:2402.16906v62024
  25. GEPA: Reflective Prompt Evolution Can Outperform Reinforcement Learning

    Lakshya A Agrawal, Shangyin Tan, Dilara Soylu +14

    cs.CLcs.AIcs.LGarXiv:2507.19457v22025
  26. MobileVLM V2: Faster and Stronger Baseline for Vision Language Model

    Xiangxiang Chu, Limeng Qiao, Xinyu Zhang +8

    cs.CVcs.AIarXiv:2402.03766v12024
  27. BigEarthNet-MM: A Large Scale Multi-Modal Multi-Label Benchmark Archive for Remote Sensing Image Classification and Retrieval

    Gencer Sumbul, Arne de Wall, Tristan Kreuziger +6

    cs.CVarXiv:2105.07921v22021
  28. A Study of Conditional Diffusion Models for Open-Loop Control under Dry Friction and Stiction

    Eric Aislan Antonelo

    cs.LGarXiv:2609.01756v12026
  29. Efficiently Approximating the Minimum-Volume Bounding Box of a Point Set in Three Dimensions

    Gill Barequet, Sariel Har-Peled

    cs.CGarXiv:2512.12391v12025
  30. Sequence-to-Sequence Knowledge Graph Completion and Question Answering

    Apoorv Saxena, Adrian Kochsiek, Rainer Gemulla

    cs.CLcs.LGarXiv:2203.10321v12022
  31. Model Context Protocol (MCP): Landscape, Security Threats, and Future Research Directions

    Xinyi Hou, Yanjie Zhao, Shenao Wang +1

    cs.CRcs.AIarXiv:2503.23278v32025
  32. Super-Resolution Delay-Doppler Estimation for OFDM Passive Radar

    Le Zheng, Xiaodong Wang

    cs.ITarXiv:1610.04218v22016
  33. LeWorldModel: Stable End-to-End Joint-Embedding Predictive Architecture from Pixels

    Lucas Maes, Quentin Le Lidec, Damien Scieur +2

    cs.LGcs.AIarXiv:2603.19312v32026
  34. Graph Convolution for Multimodal Information Extraction from Visually Rich Documents

    Xiaojing Liu, Feiyu Gao, Qiong Zhang +1

    cs.IRcs.CVcs.LGarXiv:1903.11279v12019
  35. HealthBench: Evaluating Large Language Models Towards Improved Human Health

    Rahul K. Arora, Jason Wei, Rebecca Soskin Hicks +9

    cs.CLarXiv:2505.08775v12025
  36. TSLANet: Rethinking Transformers for Time Series Representation Learning

    Emadeldeen Eldele, Mohamed Ragab, Zhenghua Chen +2

    cs.LGstat.MLarXiv:2404.08472v22024
  37. Step1X-Edit: A Practical Framework for General Image Editing

    Shiyu Liu, Yucheng Han, Peng Xing +21

    cs.CVarXiv:2504.17761v52025
  38. Global Self-Attention as a Replacement for Graph Convolution

    Md Shamim Hussain, Mohammed J. Zaki, Dharmashankar Subramanian

    cs.LGarXiv:2108.03348v32021
  39. ASTEC -- the Aarhus STellar Evolution Code

    J. Christensen-Dalsgaard

    astro-pharXiv:0710.3114v12007
  40. LIMO: Less is More for Reasoning

    Yixin Ye, Zhen Huang, Yang Xiao +3

    cs.CLcs.AIarXiv:2502.03387v32025
  41. Utility Optimal Scheduling in Energy Harvesting Networks

    Longbo Huang, Michael J. Neely

    math.OCarXiv:1012.1945v12010
  42. Smart Contracts Claimed Vulnerable by the CVE Database, with Labels and Source Locations

    Monika di Angelo, Gernot Salzer

    cs.CRcs.SEarXiv:2609.01186v12026
  43. Consistency Policy: Accelerated Visuomotor Policies via Consistency Distillation

    Aaditya Prasad, Kevin Lin, Jimmy Wu +2

    cs.ROcs.AIarXiv:2405.07503v22024
  44. YOLOv13: Real-Time Object Detection with Hypergraph-Enhanced Adaptive Visual Perception

    Mengqi Lei, Siqi Li, Yihong Wu +7

    cs.CVarXiv:2506.17733v22025
  45. Dynamics Based 3D Skeletal Hand Tracking

    Stan Melax, Leonid Keselman, Sterling Orsten

    cs.CVcs.GRarXiv:1705.07640v12017
  46. VideoLLaMA 3: Frontier Multimodal Foundation Models for Image and Video Understanding

    Boqiang Zhang, Kehan Li, Zesen Cheng +12

    cs.CVarXiv:2501.13106v42025
  47. ChronoNet: A Deep Recurrent Neural Network for Abnormal EEG Identification

    Subhrajit Roy, Isabell Kiral-Kornek, Stefan Harrer

    eess.SPcs.LGarXiv:1802.00308v22018
  48. Residual Attention: A Simple but Effective Method for Multi-Label Recognition

    Ke Zhu, Jianxin Wu

    cs.CVarXiv:2108.02456v22021
  49. Fast-WAM: Do World Action Models Need Test-time Future Imagination?

    Tianyuan Yuan, Zibin Dong, Yicheng Liu +1

    cs.CVcs.AIarXiv:2603.16666v22026
  50. Fast-dLLM: Training-free Acceleration of Diffusion LLM by Enabling KV Cache and Parallel Decoding

    Chengyue Wu, Hao Zhang, Shuchen Xue +6

    cs.CLarXiv:2505.22618v32025
  51. Penalized Composite Quasi-Likelihood for Ultrahigh-Dimensional Variable Selection

    Jelena Bradic, Jianqing Fan, Weiwei Wang

    stat.MEmath.STarXiv:0912.5200v22009
  52. DoFE: Domain-oriented Feature Embedding for Generalizable Fundus Image Segmentation on Unseen Datasets

    Shujun Wang, Lequan Yu, Kang Li +3

    cs.CVarXiv:2010.06208v12020
  53. The Entropy Mechanism of Reinforcement Learning for Reasoning Language Models

    Ganqu Cui, Yuchen Zhang, Jiacheng Chen +14

    cs.LGcs.AIcs.CLarXiv:2505.22617v12025
  54. MedGemma Technical Report

    Andrew Sellergren, Sahar Kazemzadeh, Tiam Jaroensri +78

    cs.AIcs.CLcs.CVarXiv:2507.05201v42025
  55. Escaping Redundant Reasoning: Structure-Aware Search for Inference-Time LLMs

    Lu Cheng

    cs.AIarXiv:2609.00738v12026
  56. Vision-R1: Incentivizing Reasoning Capability in Multimodal Large Language Models

    Wenxuan Huang, Bohan Jia, Zijie Zhai +7

    cs.CVcs.AIcs.CLarXiv:2503.06749v42025
  57. Grounded Human-Object Interaction Hotspots from Video

    Tushar Nagarajan, Christoph Feichtenhofer, Kristen Grauman

    cs.CVarXiv:1812.04558v22018
  58. Simple linear attention language models balance the recall-throughput tradeoff

    Simran Arora, Sabri Eyuboglu, Michael Zhang +6

    cs.CLcs.LGarXiv:2402.18668v22024
  59. FAST: Efficient Action Tokenization for Vision-Language-Action Models

    Karl Pertsch, Kyle Stachowicz, Brian Ichter +6

    cs.ROcs.LGarXiv:2501.09747v12025
  60. FUSE: An Evaluating Framework for Dangerous Capabilities of LLMs

    Zhengyi Jin, Ru Zhang, Xiao Chen +5

    cs.AIarXiv:2609.02168v12026