Every paper with a summary
Every arXiv paper Paperlayer has summarized so far. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.
Search paper metadata (including unsummarized papers)
58,021 to 58,080 of 61,190
RoBERTa: A Robustly Optimized BERT Pretraining Approach
Yinhan Liu, Myle Ott, Naman Goyal +7
cs.CLarXiv:1907.11692v12019FCOS: Fully Convolutional One-Stage Object Detection
Zhi Tian, Chunhua Shen, Hao Chen +1
cs.CVarXiv:1904.01355v52019nuScenes: A multimodal dataset for autonomous driving
Holger Caesar, Varun Bankiti, Alex H. Lang +7
cs.LGcs.CVcs.ROarXiv:1903.11027v52019Parameter-Efficient Transfer Learning for NLP
Neil Houlsby, Andrei Giurgiu, Stanislaw Jastrzebski +5
cs.LGcs.CLstat.MLarXiv:1902.00751v22019Random Erasing Data Augmentation
Zhun Zhong, Liang Zheng, Guoliang Kang +2
cs.CVarXiv:1708.04896v22017On Calibration of Modern Neural Networks
Chuan Guo, Geoff Pleiss, Yu Sun +1
cs.LGarXiv:1706.04599v22017Asynchronous Methods for Deep Reinforcement Learning
Volodymyr Mnih, Adrià Puigdomènech Badia, Mehdi Mirza +5
cs.LGarXiv:1602.01783v22016Striving for Simplicity: The All Convolutional Net
Jost Tobias Springenberg, Alexey Dosovitskiy, Thomas Brox +1
cs.LGcs.CVcs.NEarXiv:1412.6806v32014MuSS: A Large-Scale Dataset and Cinematic Narrative Benchmark for Multi-Shot Subject-to-Video Generation
Haojie Zhang, Di Wu, Bingyan Liu +5
cs.CVarXiv:2604.23789v32026Summaries:한국어Prompt-Level Distillation: A Non-Parametric Alternative to Model Fine-Tuning for Efficient Reasoning
Sanket Badhe, Deep Shah
cs.CLcs.IRarXiv:2602.21103v22026Diffusion-Pretrained Dense and Contextual Embeddings
Sedigheh Eslami, Maksim Gaiduk, Markus Krimmel +3
cs.LGcs.CLcs.IRarXiv:2602.11151v22026Summaries:한국어FollowUpBot: An LLM-Based Conversational Robot for Automatic Postoperative Follow-up
Chen Chen, Jianing Yin, Jiannong Cao +5
cs.HCcs.CLcs.ROarXiv:2507.15502v12025DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models
Zhihong Shao, Peiyi Wang, Qihao Zhu +8
cs.CLcs.AIcs.LGarXiv:2402.03300v32024Qwen Technical Report
Jinze Bai, Shuai Bai, Yunfei Chu +45
cs.CLarXiv:2309.16609v12023RT-2: Vision-Language-Action Models Transfer Web Knowledge to Robotic Control
Anthony Brohan, Noah Brown, Justice Carbajal +51
cs.ROcs.CLcs.CVarXiv:2307.15818v12023DETRs Beat YOLOs on Real-time Object Detection
Yian Zhao, Wenyu Lv, Shangliang Xu +5
cs.CVarXiv:2304.08069v32023Grounding DINO: Marrying DINO with Grounded Pre-Training for Open-Set Object Detection
Shilong Liu, Zhaoyang Zeng, Tianhe Ren +9
cs.CVarXiv:2303.05499v52023Large Language Models Encode Clinical Knowledge
Karan Singhal, Shekoofeh Azizi, Tao Tu +27
cs.CLarXiv:2212.13138v12022Robust Speech Recognition via Large-Scale Weak Supervision
Alec Radford, Jong Wook Kim, Tao Xu +3
eess.AScs.CLcs.LGarXiv:2212.04356v12022A Time Series is Worth 64 Words: Long-term Forecasting with Transformers
Yuqi Nie, Nam H. Nguyen, Phanwadee Sinthong +1
cs.LGcs.AIarXiv:2211.14730v22022Large Language Models are Zero-Shot Reasoners
Takeshi Kojima, Shixiang Shane Gu, Machel Reid +2
cs.CLcs.AIcs.LGarXiv:2205.11916v42022Training a Helpful and Harmless Assistant with Reinforcement Learning from Human Feedback
Yuntao Bai, Andy Jones, Kamal Ndousse +28
cs.CLcs.LGarXiv:2204.05862v12022TruthfulQA: Measuring How Models Mimic Human Falsehoods
Stephanie Lin, Jacob Hilton, Owain Evans
cs.CLcs.AIcs.CYarXiv:2109.07958v22021Swin-Unet: Unet-like Pure Transformer for Medical Image Segmentation
Hu Cao, Yueyue Wang, Joy Chen +4
eess.IVcs.CVarXiv:2105.05537v12021Barlow Twins: Self-Supervised Learning via Redundancy Reduction
Jure Zbontar, Li Jing, Ishan Misra +2
cs.CVcs.AIcs.LGarXiv:2103.03230v32021Learning Transferable Visual Models From Natural Language Supervision
Alec Radford, Jong Wook Kim, Chris Hallacy +9
cs.CVcs.LGarXiv:2103.00020v12021Summaries:한국어Evaluation: from precision, recall and F-measure to ROC, informedness, markedness and correlation
David M. W. Powers
cs.LGstat.MEstat.MLarXiv:2010.16061v12020Supervised Contrastive Learning
Prannay Khosla, Piotr Teterwak, Chen Wang +6
cs.LGcs.CVstat.MLarXiv:2004.11362v52020Don't Stop Pretraining: Adapt Language Models to Domains and Tasks
Suchin Gururangan, Ana Marasović, Swabha Swayamdipta +4
cs.CLcs.LGarXiv:2004.10964v32020Dense Passage Retrieval for Open-Domain Question Answering
Vladimir Karpukhin, Barlas Oğuz, Sewon Min +5
cs.CLarXiv:2004.04906v32020A Simple Framework for Contrastive Learning of Visual Representations
Ting Chen, Simon Kornblith, Mohammad Norouzi +1
cs.LGcs.CVstat.MLarXiv:2002.05709v32020Momentum Contrast for Unsupervised Visual Representation Learning
Kaiming He, Haoqi Fan, Yuxin Wu +2
cs.CVarXiv:1911.05722v32019Generative Modeling by Estimating Gradients of the Data Distribution
Yang Song, Stefano Ermon
cs.LGstat.MLarXiv:1907.05600v32019BERT: Pre-training of Deep Bidirectional Transformers for Language Understanding
Jacob Devlin, Ming-Wei Chang, Kenton Lee +1
cs.CLarXiv:1810.04805v22018Intelligent Reflecting Surface Enhanced Wireless Network via Joint Active and Passive Beamforming
Qingqing Wu, Rui Zhang
cs.ITmath.OCarXiv:1810.03961v42018Deep learning in agriculture: A survey
Andreas Kamilaris, Francesc X. Prenafeta-Boldu
cs.LGcs.CVstat.MLarXiv:1807.11809v12018CBAM: Convolutional Block Attention Module
Sanghyun Woo, Jongchan Park, Joon-Young Lee +1
cs.CVarXiv:1807.06521v22018Glow: Generative Flow with Invertible 1x1 Convolutions
Diederik P. Kingma, Prafulla Dhariwal
stat.MLcs.AIcs.LGarXiv:1807.03039v22018BDD100K: A Diverse Driving Dataset for Heterogeneous Multitask Learning
Fisher Yu, Haofeng Chen, Xin Wang +5
cs.CVarXiv:1805.04687v22018An Empirical Evaluation of Generic Convolutional and Recurrent Networks for Sequence Modeling
Shaojie Bai, J. Zico Kolter, Vladlen Koltun
cs.LGcs.AIcs.CLarXiv:1803.01271v22018Soft Actor-Critic: Off-Policy Maximum Entropy Deep Reinforcement Learning with a Stochastic Actor
Tuomas Haarnoja, Aurick Zhou, Pieter Abbeel +1
cs.LGcs.AIstat.MLarXiv:1801.01290v22018Convolutional 2D Knowledge Graph Embeddings
Tim Dettmers, Pasquale Minervini, Pontus Stenetorp +1
cs.LGarXiv:1707.01476v62017SphereFace: Deep Hypersphere Embedding for Face Recognition
Weiyang Liu, Yandong Wen, Zhiding Yu +3
cs.CVarXiv:1704.08063v42017Domain Randomization for Transferring Deep Neural Networks from Simulation to the Real World
Josh Tobin, Rachel Fong, Alex Ray +3
cs.ROcs.LGarXiv:1703.06907v12017Model-Agnostic Meta-Learning for Fast Adaptation of Deep Networks
Chelsea Finn, Pieter Abbeel, Sergey Levine
cs.LGcs.AIcs.CVarXiv:1703.03400v32017On Large-Batch Training for Deep Learning: Generalization Gap and Sharp Minima
Nitish Shirish Keskar, Dheevatsa Mudigere, Jorge Nocedal +2
cs.LGmath.OCarXiv:1609.04836v22016Fully-Convolutional Siamese Networks for Object Tracking
Luca Bertinetto, Jack Valmadre, João F. Henriques +2
cs.CVarXiv:1606.09549v32016Joint Face Detection and Alignment using Multi-task Cascaded Convolutional Networks
Kaipeng Zhang, Zhanpeng Zhang, Zhifeng Li +1
cs.CVarXiv:1604.02878v12016XGBoost: A Scalable Tree Boosting System
Tianqi Chen, Carlos Guestrin
cs.LGarXiv:1603.02754v32016Summaries:한국어Prioritized Experience Replay
Tom Schaul, John Quan, Ioannis Antonoglou +1
cs.LGarXiv:1511.05952v42015Deep Unsupervised Learning using Nonequilibrium Thermodynamics
Jascha Sohl-Dickstein, Eric A. Weiss, Niru Maheswaranathan +1
cs.LGcond-mat.dis-nnq-bio.NCarXiv:1503.03585v82015Deep Neural Networks are Easily Fooled: High Confidence Predictions for Unrecognizable Images
Anh Nguyen, Jason Yosinski, Jeff Clune
cs.CVcs.AIcs.NEarXiv:1412.1897v42014Julia: A Fresh Approach to Numerical Computing
Jeff Bezanson, Alan Edelman, Stefan Karpinski +1
cs.MSarXiv:1411.1607v42014FastJet user manual
Matteo Cacciari, Gavin P. Salam, Gregory Soyez
hep-phhep-exarXiv:1111.6097v12011MadGraph 5 : Going Beyond
Johan Alwall, Michel Herquet, Fabio Maltoni +2
hep-pharXiv:1106.0522v12011A Tutorial on Spectral Clustering
Ulrike von Luxburg
cs.DScs.LGarXiv:0711.0189v12007S-Bus: Automatic Read-Set Reconstruction for Multi-Agent LLM State Coordination
Sajjad Khan
cs.LGcs.AIcs.DCarXiv:2605.17076v22026DeepTCM1.0: A Multi-Expert AI Agent for Deciphering Mechanisms of Chinese Herbal Formulae Based on General Large Language Models
Wenxin Duan, Hanwei Wang, Zhongying Peng +4
cs.CLcs.AIarXiv:2608.18103v12026Decision-Metric Alignment in Latent World Models: Diagnostics and Action-Conditioned Objectives for MPC Planning
Jiawei Wang, Ke Rui, Yushen Zuo +2
cs.LGcs.CVarXiv:2608.18746v12026H2R-Bench: Benchmarking Human-to-Robot Manipulation Video Generation in World Models
Dingyi Rong, Yue Shi, Chaofan Ma +6
cs.ROcs.CVarXiv:2608.13049v12026