Every paper with a summary
Every arXiv paper Paperlayer has summarized so far. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.
Search paper metadata (including unsummarized papers)
55,861 to 55,920 of 61,207
Uncovering Entity Identity Confusion in Multimodal Knowledge Editing
Shu Wu, Xiaotian Ye, Xinyu Mou +3
cs.CLcs.CVarXiv:2605.06096v12026Stabilizing Off-Policy Q-Learning via Bootstrapping Error Reduction
Aviral Kumar, Justin Fu, George Tucker +1
cs.LGstat.MLarXiv:1906.00949v22019AllenNLP: A Deep Semantic Natural Language Processing Platform
Matt Gardner, Joel Grus, Mark Neumann +6
cs.CLarXiv:1803.07640v22018How Contextual are Contextualized Word Representations? Comparing the Geometry of BERT, ELMo, and GPT-2 Embeddings
Kawin Ethayarajh
cs.CLarXiv:1909.00512v12019Why we (usually) don't have to worry about multiple comparisons
Andrew Gelman, Jennifer Hill, Masanao Yajima
stat.APstat.MEarXiv:0907.2478v12009Meta-learners for Estimating Heterogeneous Treatment Effects using Machine Learning
Sören R. Künzel, Jasjeet S. Sekhon, Peter J. Bickel +1
math.STstat.MEarXiv:1706.03461v62017Efficient Large-Scale Language Model Training on GPU Clusters Using Megatron-LM
Deepak Narayanan, Mohammad Shoeybi, Jared Casper +9
cs.CLcs.DCarXiv:2104.04473v52021Deep Learning based Recommender System: A Survey and New Perspectives
Shuai Zhang, Lina Yao, Aixin Sun +1
cs.IRarXiv:1707.07435v72017Planning with Diffusion for Flexible Behavior Synthesis
Michael Janner, Yilun Du, Joshua B. Tenenbaum +1
cs.LGcs.AIarXiv:2205.09991v22022Model-Driven Development of Complex Software: A Research Roadmap
Robert France, Bernhard Rumpe
cs.SEarXiv:1409.6620v12014Visual Saliency Based on Multiscale Deep Features
Guanbin Li, Yizhou Yu
cs.CVarXiv:1503.08663v32015Language-agnostic BERT Sentence Embedding
Fangxiaoyu Feng, Yinfei Yang, Daniel Cer +2
cs.CLarXiv:2007.01852v22020Deep Biaffine Attention for Neural Dependency Parsing
Timothy Dozat, Christopher D. Manning
cs.CLcs.NEarXiv:1611.01734v32016A Survey on Object Detection in Optical Remote Sensing Images
Gong Cheng, Junwei Han
cs.CVarXiv:1603.06201v22016DROP: A Reading Comprehension Benchmark Requiring Discrete Reasoning Over Paragraphs
Dheeru Dua, Yizhong Wang, Pradeep Dasigi +3
cs.CLarXiv:1903.00161v22019Universally Sloppy Parameter Sensitivities in Systems Biology
Ryan N. Gutenkunst, Joshua J. Waterfall, Fergal P. Casey +3
q-bio.QMq-bio.MNarXiv:q-bio/0701039v32007Learning Sparse Neural Networks through $L_0$ Regularization
Christos Louizos, Max Welling, Diederik P. Kingma
stat.MLcs.LGarXiv:1712.01312v22017Adversarially Learned Inference
Vincent Dumoulin, Ishmael Belghazi, Ben Poole +4
stat.MLcs.LGarXiv:1606.00704v32016Twins: Revisiting the Design of Spatial Attention in Vision Transformers
Xiangxiang Chu, Zhi Tian, Yuqing Wang +5
cs.CVcs.AIcs.LGarXiv:2104.13840v42021Quantifying Attention Flow in Transformers
Samira Abnar, Willem Zuidema
cs.LGcs.AIcs.CLarXiv:2005.00928v22020Self-Attention Graph Pooling
Junhyun Lee, Inyeop Lee, Jaewoo Kang
cs.LGstat.MLarXiv:1904.08082v42019SimLex-999: Evaluating Semantic Models with (Genuine) Similarity Estimation
Felix Hill, Roi Reichart, Anna Korhonen
cs.CLarXiv:1408.3456v12014Learning Discriminative Model Prediction for Tracking
Goutam Bhat, Martin Danelljan, Luc Van Gool +1
cs.CVarXiv:1904.07220v22019FastText.zip: Compressing text classification models
Armand Joulin, Edouard Grave, Piotr Bojanowski +3
cs.CLcs.LGarXiv:1612.03651v12016Mind2Web: Towards a Generalist Agent for the Web
Xiang Deng, Yu Gu, Boyuan Zheng +5
cs.CLarXiv:2306.06070v32023API design for machine learning software: experiences from the scikit-learn project
Lars Buitinck, Gilles Louppe, Mathieu Blondel +12
cs.LGcs.MSarXiv:1309.0238v12013The Option-Critic Architecture
Pierre-Luc Bacon, Jean Harb, Doina Precup
cs.AIarXiv:1609.05140v22016ATOM: Accurate Tracking by Overlap Maximization
Martin Danelljan, Goutam Bhat, Fahad Shahbaz Khan +1
cs.CVarXiv:1811.07628v22018Adafactor: Adaptive Learning Rates with Sublinear Memory Cost
Noam Shazeer, Mitchell Stern
cs.LGcs.AIstat.MLarXiv:1804.04235v12018Atlas: Few-shot Learning with Retrieval Augmented Language Models
Gautier Izacard, Patrick Lewis, Maria Lomeli +7
cs.CLarXiv:2208.03299v32022On Lattices, Learning with Errors, Random Linear Codes, and Cryptography
Oded Regev
cs.CRcs.CCquant-pharXiv:2401.03703v12024SummaRuNNer: A Recurrent Neural Network based Sequence Model for Extractive Summarization of Documents
Ramesh Nallapati, Feifei Zhai, Bowen Zhou
cs.CLarXiv:1611.04230v12016AtlasNet: A Papier-Mâché Approach to Learning 3D Surface Generation
Thibault Groueix, Matthew Fisher, Vladimir G. Kim +2
cs.CVarXiv:1802.05384v32018FlexMatch: Boosting Semi-Supervised Learning with Curriculum Pseudo Labeling
Bowen Zhang, Yidong Wang, Wenxin Hou +4
cs.LGcs.CVarXiv:2110.08263v32021End-to-End Incremental Learning
Francisco M. Castro, Manuel J. Marín-Jiménez, Nicolás Guil +2
cs.CVarXiv:1807.09536v22018signSGD: Compressed Optimisation for Non-Convex Problems
Jeremy Bernstein, Yu-Xiang Wang, Kamyar Azizzadenesheli +1
cs.LGcs.DCmath.OCarXiv:1802.04434v32018Simultaneous Detection and Segmentation
Bharath Hariharan, Pablo Arbeláez, Ross Girshick +1
cs.CVarXiv:1407.1808v12014Train Short, Test Long: Attention with Linear Biases Enables Input Length Extrapolation
Ofir Press, Noah A. Smith, Mike Lewis
cs.CLarXiv:2108.12409v22021Simple Copy-Paste is a Strong Data Augmentation Method for Instance Segmentation
Golnaz Ghiasi, Yin Cui, Aravind Srinivas +5
cs.CVarXiv:2012.07177v22020Advances in Pre-Training Distributed Word Representations
Tomas Mikolov, Edouard Grave, Piotr Bojanowski +2
cs.CLarXiv:1712.09405v12017The Zwicky Transient Facility: System Overview, Performance, and First Results
Eric C. Bellm, Shrinivas R. Kulkarni, Matthew J. Graham +113
astro-ph.IMarXiv:1902.01932v12019Relation Networks for Object Detection
Han Hu, Jiayuan Gu, Zheng Zhang +2
cs.CVarXiv:1711.11575v22017From Captions to Visual Concepts and Back
Hao Fang, Saurabh Gupta, Forrest Iandola +9
cs.CVcs.CLarXiv:1411.4952v32014Learning to Prompt for Continual Learning
Zifeng Wang, Zizhao Zhang, Chen-Yu Lee +7
cs.LGcs.CVarXiv:2112.08654v22021Deep Interest Evolution Network for Click-Through Rate Prediction
Guorui Zhou, Na Mou, Ying Fan +5
stat.MLcs.IRcs.LGarXiv:1809.03672v52018Social networks that matter: Twitter under the microscope
Bernardo A. Huberman, Daniel M. Romero, Fang Wu
cs.CYphysics.soc-pharXiv:0812.1045v12008Multi-Task Deep Neural Networks for Natural Language Understanding
Xiaodong Liu, Pengcheng He, Weizhu Chen +1
cs.CLarXiv:1901.11504v22019Understanding the limits of LoRaWAN
Ferran Adelantado, Xavier Vilajosana, Pere Tuset-Peiro +3
cs.NIarXiv:1607.08011v22016A Survey on Gas Sensing Technology
Xiao Liu, Sitian Cheng, Hong Liu +3
physics.ins-detarXiv:1305.7427v12013Learning Fine-grained Image Similarity with Deep Ranking
Jiang Wang, Yang song, Thomas Leung +5
cs.CVarXiv:1404.4661v12014Fast Byte Latent Transformer
Julie Kallini, Artidoro Pagnoni, Tomasz Limisiewicz +5
cs.CLcs.AIcs.LGarXiv:2605.08044v12026Going deeper with Image Transformers
Hugo Touvron, Matthieu Cord, Alexandre Sablayrolles +2
cs.CVarXiv:2103.17239v22021Open-vocabulary Object Detection via Vision and Language Knowledge Distillation
Xiuye Gu, Tsung-Yi Lin, Weicheng Kuo +1
cs.CVcs.AIcs.LGarXiv:2104.13921v32021Few-Shot Learning with Graph Neural Networks
Victor Garcia, Joan Bruna
stat.MLcs.LGarXiv:1711.04043v32017Stand-Alone Self-Attention in Vision Models
Prajit Ramachandran, Niki Parmar, Ashish Vaswani +3
cs.CVarXiv:1906.05909v12019code2vec: Learning Distributed Representations of Code
Uri Alon, Meital Zilberstein, Omer Levy +1
cs.LGcs.AIcs.PLarXiv:1803.09473v52018A tale of two databases: The use of Web of Science and Scopus in academic papers
Junwen Zhu, Weishu Liu
cs.DLarXiv:2002.02608v12020A quantile-based g-computation approach to addressing the effects of exposure mixtures
Alexander P. Keil, Jessie P. Buckley, Katie M. OBrien +3
stat.MEarXiv:1902.04200v42019Do not copy and paste! Rewriting strategies for code retrieval
Andrea Gurioli, Federico Pennino, Maurizio Gabbrielli
cs.SEcs.AIarXiv:2605.08299v12026Convolutional Neural Network Architectures for Matching Natural Language Sentences
Baotian Hu, Zhengdong Lu, Hang Li +1
cs.CLcs.LGcs.NEarXiv:1503.03244v12015