Every paper with a summary

Every arXiv paper Paperlayer has summarized so far. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.

Search paper metadata (including unsummarized papers)

57,121 to 57,180 of 61,090

  1. U$^2$-Net: Going Deeper with Nested U-Structure for Salient Object Detection

    Xuebin Qin, Zichen Zhang, Chenyang Huang +3

    cs.CVarXiv:2005.09007v32020
  2. ArcANE: Do Role-Playing Language Agents Stay in Character at the Right Time?

    Woojung Song, Nalim Kim, Sangjun Song +3

    cs.CLcs.AIarXiv:2606.05553v12026
  3. Google's Multilingual Neural Machine Translation System: Enabling Zero-Shot Translation

    Melvin Johnson, Mike Schuster, Quoc V. Le +9

    cs.CLcs.AIarXiv:1611.04558v22016
  4. Federated Optimization: Distributed Machine Learning for On-Device Intelligence

    Jakub Konečný, H. Brendan McMahan, Daniel Ramage +1

    cs.LGarXiv:1610.02527v12016
  5. Autoencoding beyond pixels using a learned similarity metric

    Anders Boesen Lindbo Larsen, Søren Kaae Sønderby, Hugo Larochelle +1

    cs.LGcs.CVstat.MLarXiv:1512.09300v22015
  6. SubtleMemory: A Benchmark for Fine-Grained Relational Memory Discrimination in Long-Horizon AI Agents

    Wenxuan Wang, Haoyu Sun, Fukuan Hou +4

    cs.AIcs.CLarXiv:2606.05761v22026
  7. Can Spatiotemporal 3D CNNs Retrace the History of 2D CNNs and ImageNet?

    Kensho Hara, Hirokatsu Kataoka, Yutaka Satoh

    cs.CVarXiv:1711.09577v22017
  8. Learning Dexterous In-Hand Manipulation

    OpenAI, Marcin Andrychowicz, Bowen Baker +14

    cs.LGcs.AIcs.ROarXiv:1808.00177v52018
  9. Robots Need More than VLA and World Models

    Elis Karcini, Faisal Mehrban, Quang Nguyen +6

    cs.ROarXiv:2606.06556v12026
  10. Learning Hand-Eye Coordination for Robotic Grasping with Deep Learning and Large-Scale Data Collection

    Sergey Levine, Peter Pastor, Alex Krizhevsky +1

    cs.LGcs.AIcs.CVarXiv:1603.02199v42016
  11. Deep Speech: Scaling up end-to-end speech recognition

    Awni Hannun, Carl Case, Jared Casper +8

    cs.CLcs.LGcs.NEarXiv:1412.5567v22014
  12. AsyncWebRL: Efficient Asynchronous Reinforcement Learning for Multi-Step Visual Web Agents

    Hao Bai, Rui Yang, Chenlu Ye +3

    cs.LGarXiv:2606.05597v32026
  13. A Survey of Indoor Localization Systems and Technologies

    Faheem Zafari, Athanasios Gkelias, Kin Leung

    cs.NIarXiv:1709.01015v32017
  14. Multi-Stage Progressive Image Restoration

    Syed Waqas Zamir, Aditya Arora, Salman Khan +4

    cs.CVarXiv:2102.02808v22021
  15. GShard: Scaling Giant Models with Conditional Computation and Automatic Sharding

    Dmitry Lepikhin, HyoukJoong Lee, Yuanzhong Xu +6

    cs.CLcs.LGstat.MLarXiv:2006.16668v12020
  16. RhymeFlow: Training-Free Acceleration for Video Generation with Asynchronous Denoising Flow Scheduling

    Chensheng Dai, Shengjun Zhang, Yifan Li +3

    cs.CVarXiv:2606.06309v12026
  17. Explaining Explanations: An Overview of Interpretability of Machine Learning

    Leilani H. Gilpin, David Bau, Ben Z. Yuan +3

    cs.AIcs.LGstat.MLarXiv:1806.00069v32018
  18. Beyond Alignment: Value Diversity as a Collective Property in Multicultural Agent Systems

    Shaoyang Xu, Jingshen Zhang, Long P. Hoang +2

    cs.CLcs.CYarXiv:2606.05985v12026
  19. IR3DE: A Linear Router for Large Language Models

    Eros Fanì, Oğuzhan Ersoy

    cs.CLcs.LGarXiv:2606.06098v12026
  20. MetaGPT: Meta Programming for A Multi-Agent Collaborative Framework

    Sirui Hong, Mingchen Zhuge, Jiaqi Chen +12

    cs.AIcs.MAarXiv:2308.00352v72023
  21. AID: A Benchmark Dataset for Performance Evaluation of Aerial Scene Classification

    Gui-Song Xia, Jingwen Hu, Fan Hu +4

    cs.CVarXiv:1608.05167v12016
  22. Efficient Streaming Language Models with Attention Sinks

    Guangxuan Xiao, Yuandong Tian, Beidi Chen +2

    cs.CLcs.AIarXiv:2309.17453v42023
  23. Can You Trust Your Model's Uncertainty? Evaluating Predictive Uncertainty Under Dataset Shift

    Yaniv Ovadia, Emily Fertig, Jie Ren +6

    stat.MLcs.LGarXiv:1906.02530v22019
  24. Towards Principled Methods for Training Generative Adversarial Networks

    Martin Arjovsky, Léon Bottou

    stat.MLcs.LGarXiv:1701.04862v12017
  25. Uformer: A General U-Shaped Transformer for Image Restoration

    Zhendong Wang, Xiaodong Cun, Jianmin Bao +3

    cs.CVarXiv:2106.03106v22021
  26. Face2Face: Real-time Face Capture and Reenactment of RGB Videos

    Justus Thies, Michael Zollhöfer, Marc Stamminger +2

    cs.CVarXiv:2007.14808v12020
  27. Understanding the Effective Receptive Field in Deep Convolutional Neural Networks

    Wenjie Luo, Yujia Li, Raquel Urtasun +1

    cs.CVcs.AIcs.LGarXiv:1701.04128v22017
  28. DoReFa-Net: Training Low Bitwidth Convolutional Neural Networks with Low Bitwidth Gradients

    Shuchang Zhou, Yuxin Wu, Zekun Ni +3

    cs.NEcs.LGarXiv:1606.06160v32016
  29. Reproducing GW150914: the first observation of gravitational waves from a binary black hole merger

    Duncan A. Brown, Karan Vahi, Michela Taufer +2

    cs.DCastro-ph.IMgr-qcarXiv:2010.07244v22020
  30. KGAT: Knowledge Graph Attention Network for Recommendation

    Xiang Wang, Xiangnan He, Yixin Cao +2

    cs.LGcs.IRstat.MLarXiv:1905.07854v22019
  31. Palette: Image-to-Image Diffusion Models

    Chitwan Saharia, William Chan, Huiwen Chang +5

    cs.CVcs.LGarXiv:2111.05826v22021
  32. Deep Learning Face Representation by Joint Identification-Verification

    Yi Sun, Xiaogang Wang, Xiaoou Tang

    cs.CVarXiv:1406.4773v12014
  33. Zero-Reference Deep Curve Estimation for Low-Light Image Enhancement

    Chunle Guo, Chongyi Li, Jichang Guo +4

    cs.CVarXiv:2001.06826v22020
  34. Targeted Backdoor Attacks on Deep Learning Systems Using Data Poisoning

    Xinyun Chen, Chang Liu, Bo Li +2

    cs.CRcs.LGarXiv:1712.05526v12017
  35. An Overview of Recent Progress in the Study of Distributed Multi-agent Coordination

    Yongcan Cao, Wenwu Yu, Wei Ren +1

    math.OCarXiv:1207.3231v22012
  36. Federated Learning in Mobile Edge Networks: A Comprehensive Survey

    Wei Yang Bryan Lim, Nguyen Cong Luong, Dinh Thai Hoang +5

    cs.NIeess.SParXiv:1909.11875v22019
  37. A Data-Driven Approximation of the Koopman Operator: Extending Dynamic Mode Decomposition

    Matthew O. Williams, Ioannis G. Kevrekidis, Clarence W. Rowley

    math.DSarXiv:1408.4408v12014
  38. Multimodal Transformer for Unaligned Multimodal Language Sequences

    Yao-Hung Hubert Tsai, Shaojie Bai, Paul Pu Liang +3

    cs.CLarXiv:1906.00295v12019
  39. Evasion Attacks against Machine Learning at Test Time

    Battista Biggio, Igino Corona, Davide Maiorca +5

    cs.CRcs.LGarXiv:1708.06131v12017
  40. MXNet: A Flexible and Efficient Machine Learning Library for Heterogeneous Distributed Systems

    Tianqi Chen, Mu Li, Yutian Li +7

    cs.DCcs.LGcs.MSarXiv:1512.01274v12015
  41. DeepXDE: A deep learning library for solving differential equations

    Lu Lu, Xuhui Meng, Zhiping Mao +1

    cs.LGphysics.comp-phstat.MLarXiv:1907.04502v22019
  42. Deep Generative Image Models using a Laplacian Pyramid of Adversarial Networks

    Emily Denton, Soumith Chintala, Arthur Szlam +1

    cs.CVarXiv:1506.05751v12015
  43. BadNets: Identifying Vulnerabilities in the Machine Learning Model Supply Chain

    Tianyu Gu, Brendan Dolan-Gavitt, Siddharth Garg

    cs.CRcs.LGarXiv:1708.06733v22017
  44. FEVER: a large-scale dataset for Fact Extraction and VERification

    James Thorne, Andreas Vlachos, Christos Christodoulopoulos +1

    cs.CLarXiv:1803.05355v32018
  45. Playing for Data: Ground Truth from Computer Games

    Stephan R. Richter, Vibhav Vineet, Stefan Roth +1

    cs.CVarXiv:1608.02192v12016
  46. VisualBERT: A Simple and Performant Baseline for Vision and Language

    Liunian Harold Li, Mark Yatskar, Da Yin +2

    cs.CVcs.CLcs.LGarXiv:1908.03557v12019
  47. Deep Variational Information Bottleneck

    Alexander A. Alemi, Ian Fischer, Joshua V. Dillon +1

    cs.LGcs.ITarXiv:1612.00410v72016
  48. Plenoxels: Radiance Fields without Neural Networks

    Alex Yu, Sara Fridovich-Keil, Matthew Tancik +3

    cs.CVcs.GRarXiv:2112.05131v12021
  49. Personalized Top-N Sequential Recommendation via Convolutional Sequence Embedding

    Jiaxi Tang, Ke Wang

    cs.IRcs.LGarXiv:1809.07426v12018
  50. RAPPOR: Randomized Aggregatable Privacy-Preserving Ordinal Response

    Úlfar Erlingsson, Vasyl Pihur, Aleksandra Korolova

    cs.CRarXiv:1407.6981v22014
  51. Learning Convolutional Neural Networks for Graphs

    Mathias Niepert, Mohamed Ahmed, Konstantin Kutzkov

    cs.LGcs.AIstat.MLarXiv:1605.05273v42016
  52. Pre-Trained Image Processing Transformer

    Hanting Chen, Yunhe Wang, Tianyu Guo +7

    cs.CVcs.LGarXiv:2012.00364v42020
  53. Image Inpainting for Irregular Holes Using Partial Convolutions

    Guilin Liu, Fitsum A. Reda, Kevin J. Shih +3

    cs.CVarXiv:1804.07723v22018
  54. Do ImageNet Classifiers Generalize to ImageNet?

    Benjamin Recht, Rebecca Roelofs, Ludwig Schmidt +1

    cs.CVcs.LGstat.MLarXiv:1902.10811v22019
  55. Gemma 2: Improving Open Language Models at a Practical Size

    Gemma Team, Morgane Riviere, Shreya Pathak +195

    cs.CLcs.AIarXiv:2408.00118v32024
  56. Building high-level features using large scale unsupervised learning

    Quoc V. Le, Marc'Aurelio Ranzato, Rajat Monga +5

    cs.LGarXiv:1112.6209v52011
  57. A Structured Self-attentive Sentence Embedding

    Zhouhan Lin, Minwei Feng, Cicero Nogueira dos Santos +4

    cs.CLcs.AIcs.LGarXiv:1703.03130v12017
  58. PV-RCNN: Point-Voxel Feature Set Abstraction for 3D Object Detection

    Shaoshuai Shi, Chaoxu Guo, Li Jiang +4

    cs.CVcs.LGeess.IVarXiv:1912.13192v22019
  59. Deep Learning for 3D Point Clouds: A Survey

    Yulan Guo, Hanyun Wang, Qingyong Hu +3

    cs.CVcs.LGcs.ROarXiv:1912.12033v22019
  60. Wavelets on Graphs via Spectral Graph Theory

    David K Hammond, Pierre Vandergheynst, Rémi Gribonval

    math.FAcs.ITarXiv:0912.3848v12009