Every paper with a summary

Every arXiv paper Paperlayer has summarized so far. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.

Search paper metadata (including unsummarized papers)

45,061 to 45,120 of 61,134

  1. LanguageBind: Extending Video-Language Pretraining to N-modality by Language-based Semantic Alignment

    Bin Zhu, Bin Lin, Munan Ning +11

    cs.CVcs.AIarXiv:2310.01852v72023
  2. XCOPA: A Multilingual Dataset for Causal Commonsense Reasoning

    Edoardo Maria Ponti, Goran Glavaš, Olga Majewska +3

    cs.CLarXiv:2005.00333v22020
  3. PatchmatchNet: Learned Multi-View Patchmatch Stereo

    Fangjinhua Wang, Silvano Galliani, Christoph Vogel +2

    cs.CVarXiv:2012.01411v12020
  4. PRNet: Self-Supervised Learning for Partial-to-Partial Registration

    Yue Wang, Justin M. Solomon

    cs.LGstat.MLarXiv:1910.12240v22019
  5. You may implement this later: Cofunctors as partial implementations

    Vincent Wang-Maścianica

    cs.PLcs.LOarXiv:2608.27180v12026
  6. PMLB: A Large Benchmark Suite for Machine Learning Evaluation and Comparison

    Randal S. Olson, William La Cava, Patryk Orzechowski +2

    cs.LGcs.AIarXiv:1703.00512v12017
  7. GCAN: Graph-aware Co-Attention Networks for Explainable Fake News Detection on Social Media

    Yi-Ju Lu, Cheng-Te Li

    cs.CLcs.LGstat.MLarXiv:2004.11648v12020
  8. Deep MANTA: A Coarse-to-fine Many-Task Network for joint 2D and 3D vehicle analysis from monocular image

    Florian Chabot, Mohamed Chaouch, Jaonary Rabarisoa +2

    cs.CVarXiv:1703.07570v12017
  9. Investigation of Prediction Accuracy, Sensitivity, and Parameter Stability of Large-Scale Propagation Path Loss Models for 5G Wireless Communications

    Shu Sun, Theodore S. Rappaport, Timothy A. Thomas +6

    cs.ITarXiv:1603.04404v82016
  10. Learning to Prompt for Open-Vocabulary Object Detection with Vision-Language Model

    Yu Du, Fangyun Wei, Zihe Zhang +3

    cs.CVarXiv:2203.14940v12022
  11. CAT3D: Create Anything in 3D with Multi-View Diffusion Models

    Ruiqi Gao, Aleksander Holynski, Philipp Henzler +5

    cs.CVarXiv:2405.10314v12024
  12. Wayformer: Motion Forecasting via Simple & Efficient Attention Networks

    Nigamaa Nayakanti, Rami Al-Rfou, Aurick Zhou +3

    cs.CVarXiv:2207.05844v12022
  13. Active Object Localization with Deep Reinforcement Learning

    Juan C. Caicedo, Svetlana Lazebnik

    cs.CVarXiv:1511.06015v12015
  14. Neighbor2Neighbor: Self-Supervised Denoising from Single Noisy Images

    Tao Huang, Songjiang Li, Xu Jia +2

    eess.IVcs.CVarXiv:2101.02824v32021
  15. ScanQA: 3D Question Answering for Spatial Scene Understanding

    Daichi Azuma, Taiki Miyanishi, Shuhei Kurita +1

    cs.CVarXiv:2112.10482v32021
  16. A Gentle Introduction to Deep Learning in Medical Image Processing

    Andreas Maier, Christopher Syben, Tobias Lasser +1

    cs.CVarXiv:1810.05401v22018
  17. Capture, Learning, and Synthesis of 3D Speaking Styles

    Daniel Cudeiro, Timo Bolkart, Cassidy Laidlaw +2

    cs.CVarXiv:1905.03079v12019
  18. Conditional contraction coefficients and their applications to quantum networks

    Christoph Hirche, Ian George, Theshani Nuradha +1

    quant-phcs.ITarXiv:2608.27171v12026
  19. Normalization Techniques in Training DNNs: Methodology, Analysis and Application

    Lei Huang, Jie Qin, Yi Zhou +3

    cs.LGcs.CVstat.MLarXiv:2009.12836v12020
  20. Visual Question Answering: A Survey of Methods and Datasets

    Qi Wu, Damien Teney, Peng Wang +3

    cs.CVarXiv:1607.05910v12016
  21. Modeling Spatial-Temporal Clues in a Hybrid Deep Learning Framework for Video Classification

    Zuxuan Wu, Xi Wang, Yu-Gang Jiang +2

    cs.CVcs.MMarXiv:1504.01561v12015
  22. Monkey: Image Resolution and Text Label Are Important Things for Large Multi-modal Models

    Zhang Li, Biao Yang, Qiang Liu +6

    cs.CVcs.AIcs.CLarXiv:2311.06607v42023
  23. Amnesiac Machine Learning

    Laura Graves, Vineel Nagisetty, Vijay Ganesh

    cs.LGcs.AIcs.CRarXiv:2010.10981v12020
  24. Sparsity information and regularization in the horseshoe and other shrinkage priors

    Juho Piironen, Aki Vehtari

    stat.MEarXiv:1707.01694v12017
  25. A Dynamic Likelihood Approach to Filtering for Advection-Diffusion Dynamics

    Johannes Krotz, Juan M. Restrepo, Jorge Ramirez

    math.DScs.LGmath.STarXiv:2406.06837v22024
  26. Deep Industrial Image Anomaly Detection: A Survey

    Jiaqi Liu, Guoyang Xie, Jinbao Wang +4

    cs.CVarXiv:2301.11514v52023
  27. Diffusion Strategies Outperform Consensus Strategies for Distributed Estimation over Adaptive Networks

    Sheng-Yuan Tu, Ali H. Sayed

    cs.ITcs.SIarXiv:1205.3993v22012
  28. TOPIQ: Statistical Error Propagation for Quantity-of-Interest Prediction under Lossy Compression

    Youyuan Liu, Bo Jiang, Taolue Yang +3

    cs.PFcs.ITmath.NAarXiv:2608.26912v12026
  29. LoRA+: Efficient Low Rank Adaptation of Large Models

    Soufiane Hayou, Nikhil Ghosh, Bin Yu

    cs.LGcs.AIcs.CLarXiv:2402.12354v22024
  30. Multi-scale Dynamic Graph Convolutional Network for Hyperspectral Image Classification

    Sheng Wan, Chen Gong, Ping Zhong +3

    eess.IVcs.LGstat.MLarXiv:1905.06133v12019
  31. Anchor-free Oriented Proposal Generator for Object Detection

    Gong Cheng, Jiabao Wang, Ke Li +4

    cs.CVarXiv:2110.01931v22021
  32. Transformers as Soft Reasoners over Language

    Peter Clark, Oyvind Tafjord, Kyle Richardson

    cs.CLcs.AIarXiv:2002.05867v22020
  33. Large Pose 3D Face Reconstruction from a Single Image via Direct Volumetric CNN Regression

    Aaron S. Jackson, Adrian Bulat, Vasileios Argyriou +1

    cs.CVarXiv:1703.07834v22017
  34. ActionVLAD: Learning spatio-temporal aggregation for action classification

    Rohit Girdhar, Deva Ramanan, Abhinav Gupta +2

    cs.CVarXiv:1704.02895v12017
  35. Occ3D: A Large-Scale 3D Occupancy Prediction Benchmark for Autonomous Driving

    Xiaoyu Tian, Tao Jiang, Longfei Yun +5

    cs.CVarXiv:2304.14365v32023
  36. End-to-end Symmetry Preserving Inter-atomic Potential Energy Model for Finite and Extended Systems

    Linfeng Zhang, Jiequn Han, Han Wang +3

    physics.comp-phcond-mat.mtrl-sciphysics.chem-pharXiv:1805.09003v22018
  37. Analysis of Explainers of Black Box Deep Neural Networks for Computer Vision: A Survey

    Vanessa Buhrmester, David Münch, Michael Arens

    cs.AIcs.CVarXiv:1911.12116v12019
  38. A randomized algorithm for principal component analysis

    Vladimir Rokhlin, Arthur Szlam, Mark Tygert

    stat.COarXiv:0809.2274v42008
  39. Textual Explanations for Self-Driving Vehicles

    Jinkyu Kim, Anna Rohrbach, Trevor Darrell +2

    cs.CVarXiv:1807.11546v12018
  40. A Deep Learning Approach to Structured Signal Recovery

    Ali Mousavi, Ankit B. Patel, Richard G. Baraniuk

    cs.LGstat.MLarXiv:1508.04065v12015
  41. Knowledge Distillation Driven Semantic NOMA with GAN Refinement for 6G Robotic Vehicle Networks

    Qifei Wang, Zhen Gao, Li Qiao +4

    cs.ITcs.CVeess.IVarXiv:2608.27198v12026
  42. Further Optimal Regret Bounds for Thompson Sampling

    Shipra Agrawal, Navin Goyal

    cs.LGcs.DSstat.MLarXiv:1209.3353v12012
  43. D3VO: Deep Depth, Deep Pose and Deep Uncertainty for Monocular Visual Odometry

    Nan Yang, Lukas von Stumberg, Rui Wang +1

    cs.CVcs.AIarXiv:2003.01060v22020
  44. Heavy Rain Image Restoration: Integrating Physics Model and Conditional Adversarial Learning

    Ruotent Li, Loong Fah Cheong, Robby T. Tan

    cs.CVarXiv:1904.05050v12019
  45. University-1652: A Multi-view Multi-source Benchmark for Drone-based Geo-localization

    Zhedong Zheng, Yunchao Wei, Yi Yang

    cs.CVarXiv:2002.12186v22020
  46. Energy-Neutral Coverage Optimization by Joint Deployment and Scheduling in Ambient IoT Devices with Directional Sensing

    David E. Ruíz-Guirola, Samuel Montejo-Sánchez, Richard Demo Souza +1

    eess.SYcs.ITarXiv:2608.26944v12026
  47. A Simple Method for Commonsense Reasoning

    Trieu H. Trinh, Quoc V. Le

    cs.AIcs.CLcs.LGarXiv:1806.02847v22018
  48. ActBERT: Learning Global-Local Video-Text Representations

    Linchao Zhu, Yi Yang

    cs.CVarXiv:2011.07231v12020
  49. Noise or Signal: The Role of Image Backgrounds in Object Recognition

    Kai Xiao, Logan Engstrom, Andrew Ilyas +1

    cs.CVcs.LGarXiv:2006.09994v12020
  50. Size Bounds for CQs Under Acyclic Constraints

    Stefan Mengel, Andrei Romashchenko

    cs.DBcs.ITarXiv:2608.26775v12026
  51. V2V-PoseNet: Voxel-to-Voxel Prediction Network for Accurate 3D Hand and Human Pose Estimation from a Single Depth Map

    Gyeongsik Moon, Ju Yong Chang, Kyoung Mu Lee

    cs.CVarXiv:1711.07399v32017
  52. Improved Techniques for Training Consistency Models

    Yang Song, Prafulla Dhariwal

    cs.LGarXiv:2310.14189v12023
  53. Physics-guided Convolutional Neural Network (PhyCNN) for Data-driven Seismic Response Modeling

    Ruiyang Zhang, Yang Liu, Hao Sun

    eess.SPcs.CEarXiv:1909.08118v12019
  54. Limits of modularity maximization in community detection

    Andrea Lancichinetti, Santo Fortunato

    physics.soc-phcs.SIarXiv:1107.1155v22011
  55. (Sequential) Joint Detection and Estimation: Classic Results and New Directions

    Dominik Reinhard, Abdelhak M. Zoubir

    eess.SPcs.ITmath.STarXiv:2608.26278v12026
  56. YOLOP: You Only Look Once for Panoptic Driving Perception

    Dong Wu, Manwen Liao, Weitian Zhang +4

    cs.CVarXiv:2108.11250v72021
  57. The INTERSPEECH 2020 Deep Noise Suppression Challenge: Datasets, Subjective Testing Framework, and Challenge Results

    Chandan K. A. Reddy, Vishak Gopal, Ross Cutler +10

    eess.AScs.LGcs.SDarXiv:2005.13981v32020
  58. Deterministic Identification over Additive Gaussian Channels

    Jonathan E. W. Huffmann, Holger Boche

    cs.ITarXiv:2608.27243v12026
  59. The Instruction Hierarchy: Training LLMs to Prioritize Privileged Instructions

    Eric Wallace, Kai Xiao, Reimar Leike +3

    cs.CRcs.CLcs.LGarXiv:2404.13208v12024
  60. A Comprehensive Review of Deep Learning Applications in Hydrology and Water Resources

    Muhammed Sit, Bekir Z. Demiray, Zhongrun Xiang +3

    physics.geo-phcs.LGstat.MLarXiv:2007.12269v12020