Every paper with a summary
Every arXiv paper Paperlayer has summarized so far. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.
Search paper metadata (including unsummarized papers)
45,061 to 45,120 of 61,134
LanguageBind: Extending Video-Language Pretraining to N-modality by Language-based Semantic Alignment
Bin Zhu, Bin Lin, Munan Ning +11
cs.CVcs.AIarXiv:2310.01852v72023XCOPA: A Multilingual Dataset for Causal Commonsense Reasoning
Edoardo Maria Ponti, Goran Glavaš, Olga Majewska +3
cs.CLarXiv:2005.00333v22020PatchmatchNet: Learned Multi-View Patchmatch Stereo
Fangjinhua Wang, Silvano Galliani, Christoph Vogel +2
cs.CVarXiv:2012.01411v12020PRNet: Self-Supervised Learning for Partial-to-Partial Registration
Yue Wang, Justin M. Solomon
cs.LGstat.MLarXiv:1910.12240v22019You may implement this later: Cofunctors as partial implementations
Vincent Wang-Maścianica
cs.PLcs.LOarXiv:2608.27180v12026PMLB: A Large Benchmark Suite for Machine Learning Evaluation and Comparison
Randal S. Olson, William La Cava, Patryk Orzechowski +2
cs.LGcs.AIarXiv:1703.00512v12017GCAN: Graph-aware Co-Attention Networks for Explainable Fake News Detection on Social Media
Yi-Ju Lu, Cheng-Te Li
cs.CLcs.LGstat.MLarXiv:2004.11648v12020Deep MANTA: A Coarse-to-fine Many-Task Network for joint 2D and 3D vehicle analysis from monocular image
Florian Chabot, Mohamed Chaouch, Jaonary Rabarisoa +2
cs.CVarXiv:1703.07570v12017Investigation of Prediction Accuracy, Sensitivity, and Parameter Stability of Large-Scale Propagation Path Loss Models for 5G Wireless Communications
Shu Sun, Theodore S. Rappaport, Timothy A. Thomas +6
cs.ITarXiv:1603.04404v82016Learning to Prompt for Open-Vocabulary Object Detection with Vision-Language Model
Yu Du, Fangyun Wei, Zihe Zhang +3
cs.CVarXiv:2203.14940v12022CAT3D: Create Anything in 3D with Multi-View Diffusion Models
Ruiqi Gao, Aleksander Holynski, Philipp Henzler +5
cs.CVarXiv:2405.10314v12024Wayformer: Motion Forecasting via Simple & Efficient Attention Networks
Nigamaa Nayakanti, Rami Al-Rfou, Aurick Zhou +3
cs.CVarXiv:2207.05844v12022Active Object Localization with Deep Reinforcement Learning
Juan C. Caicedo, Svetlana Lazebnik
cs.CVarXiv:1511.06015v12015Neighbor2Neighbor: Self-Supervised Denoising from Single Noisy Images
Tao Huang, Songjiang Li, Xu Jia +2
eess.IVcs.CVarXiv:2101.02824v32021ScanQA: 3D Question Answering for Spatial Scene Understanding
Daichi Azuma, Taiki Miyanishi, Shuhei Kurita +1
cs.CVarXiv:2112.10482v32021A Gentle Introduction to Deep Learning in Medical Image Processing
Andreas Maier, Christopher Syben, Tobias Lasser +1
cs.CVarXiv:1810.05401v22018Capture, Learning, and Synthesis of 3D Speaking Styles
Daniel Cudeiro, Timo Bolkart, Cassidy Laidlaw +2
cs.CVarXiv:1905.03079v12019Conditional contraction coefficients and their applications to quantum networks
Christoph Hirche, Ian George, Theshani Nuradha +1
quant-phcs.ITarXiv:2608.27171v12026Normalization Techniques in Training DNNs: Methodology, Analysis and Application
Lei Huang, Jie Qin, Yi Zhou +3
cs.LGcs.CVstat.MLarXiv:2009.12836v12020Visual Question Answering: A Survey of Methods and Datasets
Qi Wu, Damien Teney, Peng Wang +3
cs.CVarXiv:1607.05910v12016Modeling Spatial-Temporal Clues in a Hybrid Deep Learning Framework for Video Classification
Zuxuan Wu, Xi Wang, Yu-Gang Jiang +2
cs.CVcs.MMarXiv:1504.01561v12015Monkey: Image Resolution and Text Label Are Important Things for Large Multi-modal Models
Zhang Li, Biao Yang, Qiang Liu +6
cs.CVcs.AIcs.CLarXiv:2311.06607v42023Amnesiac Machine Learning
Laura Graves, Vineel Nagisetty, Vijay Ganesh
cs.LGcs.AIcs.CRarXiv:2010.10981v12020Sparsity information and regularization in the horseshoe and other shrinkage priors
Juho Piironen, Aki Vehtari
stat.MEarXiv:1707.01694v12017A Dynamic Likelihood Approach to Filtering for Advection-Diffusion Dynamics
Johannes Krotz, Juan M. Restrepo, Jorge Ramirez
math.DScs.LGmath.STarXiv:2406.06837v22024Deep Industrial Image Anomaly Detection: A Survey
Jiaqi Liu, Guoyang Xie, Jinbao Wang +4
cs.CVarXiv:2301.11514v52023Diffusion Strategies Outperform Consensus Strategies for Distributed Estimation over Adaptive Networks
Sheng-Yuan Tu, Ali H. Sayed
cs.ITcs.SIarXiv:1205.3993v22012TOPIQ: Statistical Error Propagation for Quantity-of-Interest Prediction under Lossy Compression
Youyuan Liu, Bo Jiang, Taolue Yang +3
cs.PFcs.ITmath.NAarXiv:2608.26912v12026LoRA+: Efficient Low Rank Adaptation of Large Models
Soufiane Hayou, Nikhil Ghosh, Bin Yu
cs.LGcs.AIcs.CLarXiv:2402.12354v22024Multi-scale Dynamic Graph Convolutional Network for Hyperspectral Image Classification
Sheng Wan, Chen Gong, Ping Zhong +3
eess.IVcs.LGstat.MLarXiv:1905.06133v12019Anchor-free Oriented Proposal Generator for Object Detection
Gong Cheng, Jiabao Wang, Ke Li +4
cs.CVarXiv:2110.01931v22021Transformers as Soft Reasoners over Language
Peter Clark, Oyvind Tafjord, Kyle Richardson
cs.CLcs.AIarXiv:2002.05867v22020Large Pose 3D Face Reconstruction from a Single Image via Direct Volumetric CNN Regression
Aaron S. Jackson, Adrian Bulat, Vasileios Argyriou +1
cs.CVarXiv:1703.07834v22017ActionVLAD: Learning spatio-temporal aggregation for action classification
Rohit Girdhar, Deva Ramanan, Abhinav Gupta +2
cs.CVarXiv:1704.02895v12017Occ3D: A Large-Scale 3D Occupancy Prediction Benchmark for Autonomous Driving
Xiaoyu Tian, Tao Jiang, Longfei Yun +5
cs.CVarXiv:2304.14365v32023End-to-end Symmetry Preserving Inter-atomic Potential Energy Model for Finite and Extended Systems
Linfeng Zhang, Jiequn Han, Han Wang +3
physics.comp-phcond-mat.mtrl-sciphysics.chem-pharXiv:1805.09003v22018Analysis of Explainers of Black Box Deep Neural Networks for Computer Vision: A Survey
Vanessa Buhrmester, David Münch, Michael Arens
cs.AIcs.CVarXiv:1911.12116v12019A randomized algorithm for principal component analysis
Vladimir Rokhlin, Arthur Szlam, Mark Tygert
stat.COarXiv:0809.2274v42008Textual Explanations for Self-Driving Vehicles
Jinkyu Kim, Anna Rohrbach, Trevor Darrell +2
cs.CVarXiv:1807.11546v12018A Deep Learning Approach to Structured Signal Recovery
Ali Mousavi, Ankit B. Patel, Richard G. Baraniuk
cs.LGstat.MLarXiv:1508.04065v12015Knowledge Distillation Driven Semantic NOMA with GAN Refinement for 6G Robotic Vehicle Networks
Qifei Wang, Zhen Gao, Li Qiao +4
cs.ITcs.CVeess.IVarXiv:2608.27198v12026Further Optimal Regret Bounds for Thompson Sampling
Shipra Agrawal, Navin Goyal
cs.LGcs.DSstat.MLarXiv:1209.3353v12012D3VO: Deep Depth, Deep Pose and Deep Uncertainty for Monocular Visual Odometry
Nan Yang, Lukas von Stumberg, Rui Wang +1
cs.CVcs.AIarXiv:2003.01060v22020Heavy Rain Image Restoration: Integrating Physics Model and Conditional Adversarial Learning
Ruotent Li, Loong Fah Cheong, Robby T. Tan
cs.CVarXiv:1904.05050v12019University-1652: A Multi-view Multi-source Benchmark for Drone-based Geo-localization
Zhedong Zheng, Yunchao Wei, Yi Yang
cs.CVarXiv:2002.12186v22020Energy-Neutral Coverage Optimization by Joint Deployment and Scheduling in Ambient IoT Devices with Directional Sensing
David E. Ruíz-Guirola, Samuel Montejo-Sánchez, Richard Demo Souza +1
eess.SYcs.ITarXiv:2608.26944v12026A Simple Method for Commonsense Reasoning
Trieu H. Trinh, Quoc V. Le
cs.AIcs.CLcs.LGarXiv:1806.02847v22018ActBERT: Learning Global-Local Video-Text Representations
Linchao Zhu, Yi Yang
cs.CVarXiv:2011.07231v12020Noise or Signal: The Role of Image Backgrounds in Object Recognition
Kai Xiao, Logan Engstrom, Andrew Ilyas +1
cs.CVcs.LGarXiv:2006.09994v12020Size Bounds for CQs Under Acyclic Constraints
Stefan Mengel, Andrei Romashchenko
cs.DBcs.ITarXiv:2608.26775v12026V2V-PoseNet: Voxel-to-Voxel Prediction Network for Accurate 3D Hand and Human Pose Estimation from a Single Depth Map
Gyeongsik Moon, Ju Yong Chang, Kyoung Mu Lee
cs.CVarXiv:1711.07399v32017Improved Techniques for Training Consistency Models
Yang Song, Prafulla Dhariwal
cs.LGarXiv:2310.14189v12023Physics-guided Convolutional Neural Network (PhyCNN) for Data-driven Seismic Response Modeling
Ruiyang Zhang, Yang Liu, Hao Sun
eess.SPcs.CEarXiv:1909.08118v12019Limits of modularity maximization in community detection
Andrea Lancichinetti, Santo Fortunato
physics.soc-phcs.SIarXiv:1107.1155v22011(Sequential) Joint Detection and Estimation: Classic Results and New Directions
Dominik Reinhard, Abdelhak M. Zoubir
eess.SPcs.ITmath.STarXiv:2608.26278v12026YOLOP: You Only Look Once for Panoptic Driving Perception
Dong Wu, Manwen Liao, Weitian Zhang +4
cs.CVarXiv:2108.11250v72021The INTERSPEECH 2020 Deep Noise Suppression Challenge: Datasets, Subjective Testing Framework, and Challenge Results
Chandan K. A. Reddy, Vishak Gopal, Ross Cutler +10
eess.AScs.LGcs.SDarXiv:2005.13981v32020Deterministic Identification over Additive Gaussian Channels
Jonathan E. W. Huffmann, Holger Boche
cs.ITarXiv:2608.27243v12026The Instruction Hierarchy: Training LLMs to Prioritize Privileged Instructions
Eric Wallace, Kai Xiao, Reimar Leike +3
cs.CRcs.CLcs.LGarXiv:2404.13208v12024A Comprehensive Review of Deep Learning Applications in Hydrology and Water Resources
Muhammed Sit, Bekir Z. Demiray, Zhongrun Xiang +3
physics.geo-phcs.LGstat.MLarXiv:2007.12269v12020