Every paper with a summary

Every arXiv paper Paperlayer has summarized so far. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.

Search paper metadata (including unsummarized papers)

56,461 to 56,520 of 61,098

  1. Robot Operating System 2: Design, Architecture, and Uses In The Wild

    Steve Macenski, Tully Foote, Brian Gerkey +2

    cs.ROarXiv:2211.07752v12022
  2. EAST: An Efficient and Accurate Scene Text Detector

    Xinyu Zhou, Cong Yao, He Wen +4

    cs.CVarXiv:1704.03155v22017
  3. Balancing Fidelity and Diversity in Diffusion Models via Symmetric Attention Decomposition: Hopfield Perspective

    Hyunmin Cho, Woo Kyoung Han, Kyong Hwan Jin

    cs.LGcs.AIarXiv:2605.27476v12026
  4. BEVFusion: Multi-Task Multi-Sensor Fusion with Unified Bird's-Eye View Representation

    Zhijian Liu, Haotian Tang, Alexander Amini +4

    cs.CVarXiv:2205.13542v32022
  5. JLT: Clean-Latent Prediction in Latent Diffusion Transformers

    Funing Fu, Tenghui Wang, Guanyu Zhou +2

    cs.CVcs.LGarXiv:2605.27102v22026
  6. Approximate Nearest Neighbor Negative Contrastive Learning for Dense Text Retrieval

    Lee Xiong, Chenyan Xiong, Ye Li +5

    cs.IRcs.CLcs.LGarXiv:2007.00808v22020
  7. GQA: Training Generalized Multi-Query Transformer Models from Multi-Head Checkpoints

    Joshua Ainslie, James Lee-Thorp, Michiel de Jong +3

    cs.CLcs.LGarXiv:2305.13245v32023
  8. BatteryMFormer: Multi-level Learning for Battery Degradation Trajectory Forecasting

    Ruifeng Tan, Jintao Dong, Weixiang Hong +3

    cs.AIarXiv:2605.27044v22026
  9. Contrastive Learning for Unpaired Image-to-Image Translation

    Taesung Park, Alexei A. Efros, Richard Zhang +1

    cs.CVcs.LGarXiv:2007.15651v32020
  10. A SIDARTHE Model of COVID-19 Epidemic in Italy

    Giulia Giordano, Franco Blanchini, Raffaele Bruno +5

    q-bio.PEeess.SYmath.DSarXiv:2003.09861v12020
  11. CoAtNet: Marrying Convolution and Attention for All Data Sizes

    Zihang Dai, Hanxiao Liu, Quoc V. Le +1

    cs.CVcs.LGarXiv:2106.04803v22021
  12. AgensFlow: A Coordination-Policy Substrate for Multi-Agent Systems

    Nicole Koenigstein

    cs.MAcs.AIcs.LGarXiv:2605.27466v12026
  13. How far are we from solving the 2D & 3D Face Alignment problem? (and a dataset of 230,000 3D facial landmarks)

    Adrian Bulat, Georgios Tzimiropoulos

    cs.CVcs.LGarXiv:1703.07332v32017
  14. MentorNet: Learning Data-Driven Curriculum for Very Deep Neural Networks on Corrupted Labels

    Lu Jiang, Zhengyuan Zhou, Thomas Leung +2

    cs.CVarXiv:1712.05055v22017
  15. Cascaded Diffusion Models for High Fidelity Image Generation

    Jonathan Ho, Chitwan Saharia, William Chan +3

    cs.CVcs.AIcs.LGarXiv:2106.15282v32021
  16. Model soups: averaging weights of multiple fine-tuned models improves accuracy without increasing inference time

    Mitchell Wortsman, Gabriel Ilharco, Samir Yitzhak Gadre +8

    cs.LGcs.CLcs.CVarXiv:2203.05482v32022
  17. The Roadmap to 6G -- AI Empowered Wireless Networks

    Khaled B. Letaief, Wei Chen, Yuanming Shi +2

    cs.NIcs.LGarXiv:1904.11686v22019
  18. ChatGPT Outperforms Crowd-Workers for Text-Annotation Tasks

    Fabrizio Gilardi, Meysam Alizadeh, Maël Kubli

    cs.CLcs.CYarXiv:2303.15056v22023
  19. Junction Tree Variational Autoencoder for Molecular Graph Generation

    Wengong Jin, Regina Barzilay, Tommi Jaakkola

    cs.LGcs.NEstat.MLarXiv:1802.04364v42018
  20. MELD: A Multimodal Multi-Party Dataset for Emotion Recognition in Conversations

    Soujanya Poria, Devamanyu Hazarika, Navonil Majumder +3

    cs.CLarXiv:1810.02508v62018
  21. M3-Embedding: Multi-Linguality, Multi-Functionality, Multi-Granularity Text Embeddings Through Self-Knowledge Distillation

    Jianlv Chen, Shitao Xiao, Peitian Zhang +3

    cs.CLcs.AIcs.LGarXiv:2402.03216v52024
  22. Pre-trained Models for Natural Language Processing: A Survey

    Xipeng Qiu, Tianxiang Sun, Yige Xu +3

    cs.CLcs.LGarXiv:2003.08271v42020
  23. Deep Facial Expression Recognition: A Survey

    Shan Li, Weihong Deng

    cs.CVarXiv:1804.08348v22018
  24. BiSeNet V2: Bilateral Network with Guided Aggregation for Real-time Semantic Segmentation

    Changqian Yu, Changxin Gao, Jingbo Wang +3

    cs.CVarXiv:2004.02147v12020
  25. DeblurGAN: Blind Motion Deblurring Using Conditional Adversarial Networks

    Orest Kupyn, Volodymyr Budzan, Mykola Mykhailych +2

    cs.CVarXiv:1711.07064v42017
  26. Multi-Agent Reinforcement Learning: A Selective Overview of Theories and Algorithms

    Kaiqing Zhang, Zhuoran Yang, Tamer Başar

    cs.LGcs.AIcs.MAarXiv:1911.10635v22019
  27. A simple neural network module for relational reasoning

    Adam Santoro, David Raposo, David G. T. Barrett +4

    cs.CLcs.LGarXiv:1706.01427v12017
  28. word2vec Explained: deriving Mikolov et al.'s negative-sampling word-embedding method

    Yoav Goldberg, Omer Levy

    cs.CLcs.LGstat.MLarXiv:1402.3722v12014
  29. Making Deep Neural Networks Robust to Label Noise: a Loss Correction Approach

    Giorgio Patrini, Alessandro Rozza, Aditya Menon +2

    stat.MLcs.LGarXiv:1609.03683v22016
  30. Wild Patterns: Ten Years After the Rise of Adversarial Machine Learning

    Battista Biggio, Fabio Roli

    cs.CVcs.CRcs.GTarXiv:1712.03141v22017
  31. Towards the Development of Realistic Botnet Dataset in the Internet of Things for Network Forensic Analytics: Bot-IoT Dataset

    Nickolaos Koroniotis, Nour Moustafa, Elena Sitnikova +1

    cs.CRarXiv:1811.00701v12018
  32. A Theoretically Grounded Application of Dropout in Recurrent Neural Networks

    Yarin Gal, Zoubin Ghahramani

    stat.MLarXiv:1512.05287v52015
  33. Understanding intermediate layers using linear classifier probes

    Guillaume Alain, Yoshua Bengio

    stat.MLcs.LGarXiv:1610.01644v42016
  34. Learning Quadrupedal Locomotion over Challenging Terrain

    Joonho Lee, Jemin Hwangbo, Lorenz Wellhausen +2

    cs.ROcs.LGeess.SYarXiv:2010.11251v12020
  35. ImageBind: One Embedding Space To Bind Them All

    Rohit Girdhar, Alaaeldin El-Nouby, Zhuang Liu +4

    cs.CVcs.AIcs.LGarXiv:2305.05665v22023
  36. IP-Adapter: Text Compatible Image Prompt Adapter for Text-to-Image Diffusion Models

    Hu Ye, Jun Zhang, Sibo Liu +2

    cs.CVcs.AIarXiv:2308.06721v12023
  37. MVSNet: Depth Inference for Unstructured Multi-view Stereo

    Yao Yao, Zixin Luo, Shiwei Li +2

    cs.CVarXiv:1804.02505v22018
  38. A guide to convolution arithmetic for deep learning

    Vincent Dumoulin, Francesco Visin

    stat.MLcs.LGcs.NEarXiv:1603.07285v22016
  39. Multiscale Vision Transformers

    Haoqi Fan, Bo Xiong, Karttikeya Mangalam +4

    cs.CVcs.AIcs.LGarXiv:2104.11227v12021
  40. Expanding Performance Boundaries of Open-Source Multimodal Models with Model, Data, and Test-Time Scaling

    Zhe Chen, Weiyun Wang, Yue Cao +39

    cs.CVarXiv:2412.05271v52024
  41. SDR - half-baked or well done?

    Jonathan Le Roux, Scott Wisdom, Hakan Erdogan +1

    cs.SDeess.ASarXiv:1811.02508v12018
  42. Learning to Reweight Examples for Robust Deep Learning

    Mengye Ren, Wenyuan Zeng, Bin Yang +1

    cs.LGstat.MLarXiv:1803.09050v32018
  43. Generative Adversarial Network in Medical Imaging: A Review

    Xin Yi, Ekta Walia, Paul Babyn

    cs.CVcs.LGarXiv:1809.07294v42018
  44. Deep Gradient Compression: Reducing the Communication Bandwidth for Distributed Training

    Yujun Lin, Song Han, Huizi Mao +2

    cs.CVcs.DCcs.LGarXiv:1712.01887v32017
  45. Quantum repeaters based on atomic ensembles and linear optics

    Nicolas Sangouard, Christoph Simon, Hugues de Riedmatten +1

    quant-pharXiv:0906.2699v22009
  46. Recursive Partitioning for Heterogeneous Causal Effects

    Susan Athey, Guido Imbens

    stat.MLecon.EMarXiv:1504.01132v32015
  47. Networks beyond pairwise interactions: structure and dynamics

    Federico Battiston, Giulia Cencetti, Iacopo Iacopini +5

    physics.soc-phcond-mat.dis-nncs.SIarXiv:2006.01764v12020
  48. MINE: Mutual Information Neural Estimation

    Mohamed Ishmael Belghazi, Aristide Baratin, Sai Rajeswar +4

    cs.LGstat.MLarXiv:1801.04062v52018
  49. Differentially Private Federated Learning: A Client Level Perspective

    Robin C. Geyer, Tassilo Klein, Moin Nabi

    cs.CRcs.LGstat.MLarXiv:1712.07557v22017
  50. Predicting Positive and Negative Links in Online Social Networks

    Jure Leskovec, Daniel Huttenlocher, Jon Kleinberg

    physics.soc-phcs.AIcs.CYarXiv:1003.2429v12010
  51. ChatGLM: A Family of Large Language Models from GLM-130B to GLM-4 All Tools

    Team GLM, :, Aohan Zeng +56

    cs.CLarXiv:2406.12793v22024
  52. Deep learning for universal linear embeddings of nonlinear dynamics

    Bethany Lusch, J. Nathan Kutz, Steven L. Brunton

    math.DScs.LGstat.MLarXiv:1712.09707v22017
  53. Coupled Generative Adversarial Networks

    Ming-Yu Liu, Oncel Tuzel

    cs.CVarXiv:1606.07536v22016
  54. MUSAN: A Music, Speech, and Noise Corpus

    David Snyder, Guoguo Chen, Daniel Povey

    cs.SDarXiv:1510.08484v12015
  55. Mitigating Unwanted Biases with Adversarial Learning

    Brian Hu Zhang, Blake Lemoine, Margaret Mitchell

    cs.LGcs.AIcs.CYarXiv:1801.07593v12018
  56. Video-LLaVA: Learning United Visual Representation by Alignment Before Projection

    Bin Lin, Yang Ye, Bin Zhu +4

    cs.CVarXiv:2311.10122v32023
  57. "Liar, Liar Pants on Fire": A New Benchmark Dataset for Fake News Detection

    William Yang Wang

    cs.CLcs.CYarXiv:1705.00648v12017
  58. COCO-Stuff: Thing and Stuff Classes in Context

    Holger Caesar, Jasper Uijlings, Vittorio Ferrari

    cs.CVarXiv:1612.03716v42016
  59. Segmentation Transformer: Object-Contextual Representations for Semantic Segmentation

    Yuhui Yuan, Xiaokang Chen, Xilin Chen +1

    cs.CVarXiv:1909.11065v62019
  60. LinkNet: Exploiting Encoder Representations for Efficient Semantic Segmentation

    Abhishek Chaurasia, Eugenio Culurciello

    cs.CVcs.LGarXiv:1707.03718v12017