Every paper with a summary

Every arXiv paper Paperlayer has summarized so far. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.

Search paper metadata (including unsummarized papers)

55,861 to 55,920 of 61,207

  1. Uncovering Entity Identity Confusion in Multimodal Knowledge Editing

    Shu Wu, Xiaotian Ye, Xinyu Mou +3

    cs.CLcs.CVarXiv:2605.06096v12026
  2. Stabilizing Off-Policy Q-Learning via Bootstrapping Error Reduction

    Aviral Kumar, Justin Fu, George Tucker +1

    cs.LGstat.MLarXiv:1906.00949v22019
  3. AllenNLP: A Deep Semantic Natural Language Processing Platform

    Matt Gardner, Joel Grus, Mark Neumann +6

    cs.CLarXiv:1803.07640v22018
  4. How Contextual are Contextualized Word Representations? Comparing the Geometry of BERT, ELMo, and GPT-2 Embeddings

    Kawin Ethayarajh

    cs.CLarXiv:1909.00512v12019
  5. Why we (usually) don't have to worry about multiple comparisons

    Andrew Gelman, Jennifer Hill, Masanao Yajima

    stat.APstat.MEarXiv:0907.2478v12009
  6. Meta-learners for Estimating Heterogeneous Treatment Effects using Machine Learning

    Sören R. Künzel, Jasjeet S. Sekhon, Peter J. Bickel +1

    math.STstat.MEarXiv:1706.03461v62017
  7. Efficient Large-Scale Language Model Training on GPU Clusters Using Megatron-LM

    Deepak Narayanan, Mohammad Shoeybi, Jared Casper +9

    cs.CLcs.DCarXiv:2104.04473v52021
  8. Deep Learning based Recommender System: A Survey and New Perspectives

    Shuai Zhang, Lina Yao, Aixin Sun +1

    cs.IRarXiv:1707.07435v72017
  9. Planning with Diffusion for Flexible Behavior Synthesis

    Michael Janner, Yilun Du, Joshua B. Tenenbaum +1

    cs.LGcs.AIarXiv:2205.09991v22022
  10. Model-Driven Development of Complex Software: A Research Roadmap

    Robert France, Bernhard Rumpe

    cs.SEarXiv:1409.6620v12014
  11. Visual Saliency Based on Multiscale Deep Features

    Guanbin Li, Yizhou Yu

    cs.CVarXiv:1503.08663v32015
  12. Language-agnostic BERT Sentence Embedding

    Fangxiaoyu Feng, Yinfei Yang, Daniel Cer +2

    cs.CLarXiv:2007.01852v22020
  13. Deep Biaffine Attention for Neural Dependency Parsing

    Timothy Dozat, Christopher D. Manning

    cs.CLcs.NEarXiv:1611.01734v32016
  14. A Survey on Object Detection in Optical Remote Sensing Images

    Gong Cheng, Junwei Han

    cs.CVarXiv:1603.06201v22016
  15. DROP: A Reading Comprehension Benchmark Requiring Discrete Reasoning Over Paragraphs

    Dheeru Dua, Yizhong Wang, Pradeep Dasigi +3

    cs.CLarXiv:1903.00161v22019
  16. Universally Sloppy Parameter Sensitivities in Systems Biology

    Ryan N. Gutenkunst, Joshua J. Waterfall, Fergal P. Casey +3

    q-bio.QMq-bio.MNarXiv:q-bio/0701039v32007
  17. Learning Sparse Neural Networks through $L_0$ Regularization

    Christos Louizos, Max Welling, Diederik P. Kingma

    stat.MLcs.LGarXiv:1712.01312v22017
  18. Adversarially Learned Inference

    Vincent Dumoulin, Ishmael Belghazi, Ben Poole +4

    stat.MLcs.LGarXiv:1606.00704v32016
  19. Twins: Revisiting the Design of Spatial Attention in Vision Transformers

    Xiangxiang Chu, Zhi Tian, Yuqing Wang +5

    cs.CVcs.AIcs.LGarXiv:2104.13840v42021
  20. Quantifying Attention Flow in Transformers

    Samira Abnar, Willem Zuidema

    cs.LGcs.AIcs.CLarXiv:2005.00928v22020
  21. Self-Attention Graph Pooling

    Junhyun Lee, Inyeop Lee, Jaewoo Kang

    cs.LGstat.MLarXiv:1904.08082v42019
  22. SimLex-999: Evaluating Semantic Models with (Genuine) Similarity Estimation

    Felix Hill, Roi Reichart, Anna Korhonen

    cs.CLarXiv:1408.3456v12014
  23. Learning Discriminative Model Prediction for Tracking

    Goutam Bhat, Martin Danelljan, Luc Van Gool +1

    cs.CVarXiv:1904.07220v22019
  24. FastText.zip: Compressing text classification models

    Armand Joulin, Edouard Grave, Piotr Bojanowski +3

    cs.CLcs.LGarXiv:1612.03651v12016
  25. Mind2Web: Towards a Generalist Agent for the Web

    Xiang Deng, Yu Gu, Boyuan Zheng +5

    cs.CLarXiv:2306.06070v32023
  26. API design for machine learning software: experiences from the scikit-learn project

    Lars Buitinck, Gilles Louppe, Mathieu Blondel +12

    cs.LGcs.MSarXiv:1309.0238v12013
  27. The Option-Critic Architecture

    Pierre-Luc Bacon, Jean Harb, Doina Precup

    cs.AIarXiv:1609.05140v22016
  28. ATOM: Accurate Tracking by Overlap Maximization

    Martin Danelljan, Goutam Bhat, Fahad Shahbaz Khan +1

    cs.CVarXiv:1811.07628v22018
  29. Adafactor: Adaptive Learning Rates with Sublinear Memory Cost

    Noam Shazeer, Mitchell Stern

    cs.LGcs.AIstat.MLarXiv:1804.04235v12018
  30. Atlas: Few-shot Learning with Retrieval Augmented Language Models

    Gautier Izacard, Patrick Lewis, Maria Lomeli +7

    cs.CLarXiv:2208.03299v32022
  31. On Lattices, Learning with Errors, Random Linear Codes, and Cryptography

    Oded Regev

    cs.CRcs.CCquant-pharXiv:2401.03703v12024
  32. SummaRuNNer: A Recurrent Neural Network based Sequence Model for Extractive Summarization of Documents

    Ramesh Nallapati, Feifei Zhai, Bowen Zhou

    cs.CLarXiv:1611.04230v12016
  33. AtlasNet: A Papier-Mâché Approach to Learning 3D Surface Generation

    Thibault Groueix, Matthew Fisher, Vladimir G. Kim +2

    cs.CVarXiv:1802.05384v32018
  34. FlexMatch: Boosting Semi-Supervised Learning with Curriculum Pseudo Labeling

    Bowen Zhang, Yidong Wang, Wenxin Hou +4

    cs.LGcs.CVarXiv:2110.08263v32021
  35. End-to-End Incremental Learning

    Francisco M. Castro, Manuel J. Marín-Jiménez, Nicolás Guil +2

    cs.CVarXiv:1807.09536v22018
  36. signSGD: Compressed Optimisation for Non-Convex Problems

    Jeremy Bernstein, Yu-Xiang Wang, Kamyar Azizzadenesheli +1

    cs.LGcs.DCmath.OCarXiv:1802.04434v32018
  37. Simultaneous Detection and Segmentation

    Bharath Hariharan, Pablo Arbeláez, Ross Girshick +1

    cs.CVarXiv:1407.1808v12014
  38. Train Short, Test Long: Attention with Linear Biases Enables Input Length Extrapolation

    Ofir Press, Noah A. Smith, Mike Lewis

    cs.CLarXiv:2108.12409v22021
  39. Simple Copy-Paste is a Strong Data Augmentation Method for Instance Segmentation

    Golnaz Ghiasi, Yin Cui, Aravind Srinivas +5

    cs.CVarXiv:2012.07177v22020
  40. Advances in Pre-Training Distributed Word Representations

    Tomas Mikolov, Edouard Grave, Piotr Bojanowski +2

    cs.CLarXiv:1712.09405v12017
  41. The Zwicky Transient Facility: System Overview, Performance, and First Results

    Eric C. Bellm, Shrinivas R. Kulkarni, Matthew J. Graham +113

    astro-ph.IMarXiv:1902.01932v12019
  42. Relation Networks for Object Detection

    Han Hu, Jiayuan Gu, Zheng Zhang +2

    cs.CVarXiv:1711.11575v22017
  43. From Captions to Visual Concepts and Back

    Hao Fang, Saurabh Gupta, Forrest Iandola +9

    cs.CVcs.CLarXiv:1411.4952v32014
  44. Learning to Prompt for Continual Learning

    Zifeng Wang, Zizhao Zhang, Chen-Yu Lee +7

    cs.LGcs.CVarXiv:2112.08654v22021
  45. Deep Interest Evolution Network for Click-Through Rate Prediction

    Guorui Zhou, Na Mou, Ying Fan +5

    stat.MLcs.IRcs.LGarXiv:1809.03672v52018
  46. Social networks that matter: Twitter under the microscope

    Bernardo A. Huberman, Daniel M. Romero, Fang Wu

    cs.CYphysics.soc-pharXiv:0812.1045v12008
  47. Multi-Task Deep Neural Networks for Natural Language Understanding

    Xiaodong Liu, Pengcheng He, Weizhu Chen +1

    cs.CLarXiv:1901.11504v22019
  48. Understanding the limits of LoRaWAN

    Ferran Adelantado, Xavier Vilajosana, Pere Tuset-Peiro +3

    cs.NIarXiv:1607.08011v22016
  49. A Survey on Gas Sensing Technology

    Xiao Liu, Sitian Cheng, Hong Liu +3

    physics.ins-detarXiv:1305.7427v12013
  50. Learning Fine-grained Image Similarity with Deep Ranking

    Jiang Wang, Yang song, Thomas Leung +5

    cs.CVarXiv:1404.4661v12014
  51. Fast Byte Latent Transformer

    Julie Kallini, Artidoro Pagnoni, Tomasz Limisiewicz +5

    cs.CLcs.AIcs.LGarXiv:2605.08044v12026
  52. Going deeper with Image Transformers

    Hugo Touvron, Matthieu Cord, Alexandre Sablayrolles +2

    cs.CVarXiv:2103.17239v22021
  53. Open-vocabulary Object Detection via Vision and Language Knowledge Distillation

    Xiuye Gu, Tsung-Yi Lin, Weicheng Kuo +1

    cs.CVcs.AIcs.LGarXiv:2104.13921v32021
  54. Few-Shot Learning with Graph Neural Networks

    Victor Garcia, Joan Bruna

    stat.MLcs.LGarXiv:1711.04043v32017
  55. Stand-Alone Self-Attention in Vision Models

    Prajit Ramachandran, Niki Parmar, Ashish Vaswani +3

    cs.CVarXiv:1906.05909v12019
  56. code2vec: Learning Distributed Representations of Code

    Uri Alon, Meital Zilberstein, Omer Levy +1

    cs.LGcs.AIcs.PLarXiv:1803.09473v52018
  57. A tale of two databases: The use of Web of Science and Scopus in academic papers

    Junwen Zhu, Weishu Liu

    cs.DLarXiv:2002.02608v12020
  58. A quantile-based g-computation approach to addressing the effects of exposure mixtures

    Alexander P. Keil, Jessie P. Buckley, Katie M. OBrien +3

    stat.MEarXiv:1902.04200v42019
  59. Do not copy and paste! Rewriting strategies for code retrieval

    Andrea Gurioli, Federico Pennino, Maurizio Gabbrielli

    cs.SEcs.AIarXiv:2605.08299v12026
  60. Convolutional Neural Network Architectures for Matching Natural Language Sentences

    Baotian Hu, Zhengdong Lu, Hang Li +1

    cs.CLcs.LGcs.NEarXiv:1503.03244v12015