Every paper with a summary
Every arXiv paper Paperlayer has summarized so far. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.
Search paper metadata (including unsummarized papers)
56,461 to 56,520 of 61,098
Robot Operating System 2: Design, Architecture, and Uses In The Wild
Steve Macenski, Tully Foote, Brian Gerkey +2
cs.ROarXiv:2211.07752v12022EAST: An Efficient and Accurate Scene Text Detector
Xinyu Zhou, Cong Yao, He Wen +4
cs.CVarXiv:1704.03155v22017Balancing Fidelity and Diversity in Diffusion Models via Symmetric Attention Decomposition: Hopfield Perspective
Hyunmin Cho, Woo Kyoung Han, Kyong Hwan Jin
cs.LGcs.AIarXiv:2605.27476v12026BEVFusion: Multi-Task Multi-Sensor Fusion with Unified Bird's-Eye View Representation
Zhijian Liu, Haotian Tang, Alexander Amini +4
cs.CVarXiv:2205.13542v32022JLT: Clean-Latent Prediction in Latent Diffusion Transformers
Funing Fu, Tenghui Wang, Guanyu Zhou +2
cs.CVcs.LGarXiv:2605.27102v22026Approximate Nearest Neighbor Negative Contrastive Learning for Dense Text Retrieval
Lee Xiong, Chenyan Xiong, Ye Li +5
cs.IRcs.CLcs.LGarXiv:2007.00808v22020GQA: Training Generalized Multi-Query Transformer Models from Multi-Head Checkpoints
Joshua Ainslie, James Lee-Thorp, Michiel de Jong +3
cs.CLcs.LGarXiv:2305.13245v32023BatteryMFormer: Multi-level Learning for Battery Degradation Trajectory Forecasting
Ruifeng Tan, Jintao Dong, Weixiang Hong +3
cs.AIarXiv:2605.27044v22026Contrastive Learning for Unpaired Image-to-Image Translation
Taesung Park, Alexei A. Efros, Richard Zhang +1
cs.CVcs.LGarXiv:2007.15651v32020A SIDARTHE Model of COVID-19 Epidemic in Italy
Giulia Giordano, Franco Blanchini, Raffaele Bruno +5
q-bio.PEeess.SYmath.DSarXiv:2003.09861v12020CoAtNet: Marrying Convolution and Attention for All Data Sizes
Zihang Dai, Hanxiao Liu, Quoc V. Le +1
cs.CVcs.LGarXiv:2106.04803v22021AgensFlow: A Coordination-Policy Substrate for Multi-Agent Systems
Nicole Koenigstein
cs.MAcs.AIcs.LGarXiv:2605.27466v12026How far are we from solving the 2D & 3D Face Alignment problem? (and a dataset of 230,000 3D facial landmarks)
Adrian Bulat, Georgios Tzimiropoulos
cs.CVcs.LGarXiv:1703.07332v32017MentorNet: Learning Data-Driven Curriculum for Very Deep Neural Networks on Corrupted Labels
Lu Jiang, Zhengyuan Zhou, Thomas Leung +2
cs.CVarXiv:1712.05055v22017Cascaded Diffusion Models for High Fidelity Image Generation
Jonathan Ho, Chitwan Saharia, William Chan +3
cs.CVcs.AIcs.LGarXiv:2106.15282v32021Model soups: averaging weights of multiple fine-tuned models improves accuracy without increasing inference time
Mitchell Wortsman, Gabriel Ilharco, Samir Yitzhak Gadre +8
cs.LGcs.CLcs.CVarXiv:2203.05482v32022The Roadmap to 6G -- AI Empowered Wireless Networks
Khaled B. Letaief, Wei Chen, Yuanming Shi +2
cs.NIcs.LGarXiv:1904.11686v22019ChatGPT Outperforms Crowd-Workers for Text-Annotation Tasks
Fabrizio Gilardi, Meysam Alizadeh, Maël Kubli
cs.CLcs.CYarXiv:2303.15056v22023Junction Tree Variational Autoencoder for Molecular Graph Generation
Wengong Jin, Regina Barzilay, Tommi Jaakkola
cs.LGcs.NEstat.MLarXiv:1802.04364v42018MELD: A Multimodal Multi-Party Dataset for Emotion Recognition in Conversations
Soujanya Poria, Devamanyu Hazarika, Navonil Majumder +3
cs.CLarXiv:1810.02508v62018M3-Embedding: Multi-Linguality, Multi-Functionality, Multi-Granularity Text Embeddings Through Self-Knowledge Distillation
Jianlv Chen, Shitao Xiao, Peitian Zhang +3
cs.CLcs.AIcs.LGarXiv:2402.03216v52024Pre-trained Models for Natural Language Processing: A Survey
Xipeng Qiu, Tianxiang Sun, Yige Xu +3
cs.CLcs.LGarXiv:2003.08271v42020Deep Facial Expression Recognition: A Survey
Shan Li, Weihong Deng
cs.CVarXiv:1804.08348v22018BiSeNet V2: Bilateral Network with Guided Aggregation for Real-time Semantic Segmentation
Changqian Yu, Changxin Gao, Jingbo Wang +3
cs.CVarXiv:2004.02147v12020DeblurGAN: Blind Motion Deblurring Using Conditional Adversarial Networks
Orest Kupyn, Volodymyr Budzan, Mykola Mykhailych +2
cs.CVarXiv:1711.07064v42017Multi-Agent Reinforcement Learning: A Selective Overview of Theories and Algorithms
Kaiqing Zhang, Zhuoran Yang, Tamer Başar
cs.LGcs.AIcs.MAarXiv:1911.10635v22019A simple neural network module for relational reasoning
Adam Santoro, David Raposo, David G. T. Barrett +4
cs.CLcs.LGarXiv:1706.01427v12017word2vec Explained: deriving Mikolov et al.'s negative-sampling word-embedding method
Yoav Goldberg, Omer Levy
cs.CLcs.LGstat.MLarXiv:1402.3722v12014Making Deep Neural Networks Robust to Label Noise: a Loss Correction Approach
Giorgio Patrini, Alessandro Rozza, Aditya Menon +2
stat.MLcs.LGarXiv:1609.03683v22016Wild Patterns: Ten Years After the Rise of Adversarial Machine Learning
Battista Biggio, Fabio Roli
cs.CVcs.CRcs.GTarXiv:1712.03141v22017Towards the Development of Realistic Botnet Dataset in the Internet of Things for Network Forensic Analytics: Bot-IoT Dataset
Nickolaos Koroniotis, Nour Moustafa, Elena Sitnikova +1
cs.CRarXiv:1811.00701v12018A Theoretically Grounded Application of Dropout in Recurrent Neural Networks
Yarin Gal, Zoubin Ghahramani
stat.MLarXiv:1512.05287v52015Understanding intermediate layers using linear classifier probes
Guillaume Alain, Yoshua Bengio
stat.MLcs.LGarXiv:1610.01644v42016Learning Quadrupedal Locomotion over Challenging Terrain
Joonho Lee, Jemin Hwangbo, Lorenz Wellhausen +2
cs.ROcs.LGeess.SYarXiv:2010.11251v12020ImageBind: One Embedding Space To Bind Them All
Rohit Girdhar, Alaaeldin El-Nouby, Zhuang Liu +4
cs.CVcs.AIcs.LGarXiv:2305.05665v22023IP-Adapter: Text Compatible Image Prompt Adapter for Text-to-Image Diffusion Models
Hu Ye, Jun Zhang, Sibo Liu +2
cs.CVcs.AIarXiv:2308.06721v12023MVSNet: Depth Inference for Unstructured Multi-view Stereo
Yao Yao, Zixin Luo, Shiwei Li +2
cs.CVarXiv:1804.02505v22018A guide to convolution arithmetic for deep learning
Vincent Dumoulin, Francesco Visin
stat.MLcs.LGcs.NEarXiv:1603.07285v22016Multiscale Vision Transformers
Haoqi Fan, Bo Xiong, Karttikeya Mangalam +4
cs.CVcs.AIcs.LGarXiv:2104.11227v12021Expanding Performance Boundaries of Open-Source Multimodal Models with Model, Data, and Test-Time Scaling
Zhe Chen, Weiyun Wang, Yue Cao +39
cs.CVarXiv:2412.05271v52024SDR - half-baked or well done?
Jonathan Le Roux, Scott Wisdom, Hakan Erdogan +1
cs.SDeess.ASarXiv:1811.02508v12018Learning to Reweight Examples for Robust Deep Learning
Mengye Ren, Wenyuan Zeng, Bin Yang +1
cs.LGstat.MLarXiv:1803.09050v32018Generative Adversarial Network in Medical Imaging: A Review
Xin Yi, Ekta Walia, Paul Babyn
cs.CVcs.LGarXiv:1809.07294v42018Deep Gradient Compression: Reducing the Communication Bandwidth for Distributed Training
Yujun Lin, Song Han, Huizi Mao +2
cs.CVcs.DCcs.LGarXiv:1712.01887v32017Quantum repeaters based on atomic ensembles and linear optics
Nicolas Sangouard, Christoph Simon, Hugues de Riedmatten +1
quant-pharXiv:0906.2699v22009Recursive Partitioning for Heterogeneous Causal Effects
Susan Athey, Guido Imbens
stat.MLecon.EMarXiv:1504.01132v32015Networks beyond pairwise interactions: structure and dynamics
Federico Battiston, Giulia Cencetti, Iacopo Iacopini +5
physics.soc-phcond-mat.dis-nncs.SIarXiv:2006.01764v12020MINE: Mutual Information Neural Estimation
Mohamed Ishmael Belghazi, Aristide Baratin, Sai Rajeswar +4
cs.LGstat.MLarXiv:1801.04062v52018Differentially Private Federated Learning: A Client Level Perspective
Robin C. Geyer, Tassilo Klein, Moin Nabi
cs.CRcs.LGstat.MLarXiv:1712.07557v22017Predicting Positive and Negative Links in Online Social Networks
Jure Leskovec, Daniel Huttenlocher, Jon Kleinberg
physics.soc-phcs.AIcs.CYarXiv:1003.2429v12010ChatGLM: A Family of Large Language Models from GLM-130B to GLM-4 All Tools
Team GLM, :, Aohan Zeng +56
cs.CLarXiv:2406.12793v22024Deep learning for universal linear embeddings of nonlinear dynamics
Bethany Lusch, J. Nathan Kutz, Steven L. Brunton
math.DScs.LGstat.MLarXiv:1712.09707v22017Coupled Generative Adversarial Networks
Ming-Yu Liu, Oncel Tuzel
cs.CVarXiv:1606.07536v22016MUSAN: A Music, Speech, and Noise Corpus
David Snyder, Guoguo Chen, Daniel Povey
cs.SDarXiv:1510.08484v12015Mitigating Unwanted Biases with Adversarial Learning
Brian Hu Zhang, Blake Lemoine, Margaret Mitchell
cs.LGcs.AIcs.CYarXiv:1801.07593v12018Video-LLaVA: Learning United Visual Representation by Alignment Before Projection
Bin Lin, Yang Ye, Bin Zhu +4
cs.CVarXiv:2311.10122v32023"Liar, Liar Pants on Fire": A New Benchmark Dataset for Fake News Detection
William Yang Wang
cs.CLcs.CYarXiv:1705.00648v12017COCO-Stuff: Thing and Stuff Classes in Context
Holger Caesar, Jasper Uijlings, Vittorio Ferrari
cs.CVarXiv:1612.03716v42016Segmentation Transformer: Object-Contextual Representations for Semantic Segmentation
Yuhui Yuan, Xiaokang Chen, Xilin Chen +1
cs.CVarXiv:1909.11065v62019LinkNet: Exploiting Encoder Representations for Efficient Semantic Segmentation
Abhishek Chaurasia, Eugenio Culurciello
cs.CVcs.LGarXiv:1707.03718v12017