Computer Vision and Pattern Recognition

Papers filed under cs.CV on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.

Search paper metadata (including unsummarized papers)

12,061 to 12,120 of 18,867

  1. Sharp U-Net: Depthwise Convolutional Network for Biomedical Image Segmentation

    Hasib Zunair, A. Ben Hamza

    eess.IVcs.CVarXiv:2107.12461v12021
  2. Progressive Domain Adaptation for Object Detection

    Han-Kai Hsu, Chun-Han Yao, Yi-Hsuan Tsai +4

    cs.CVarXiv:1910.11319v12019
  3. VectorMapNet: End-to-end Vectorized HD Map Learning

    Yicheng Liu, Tianyuan Yuan, Yue Wang +2

    cs.CVcs.ROarXiv:2206.08920v62022
  4. Modeling the Background for Incremental Learning in Semantic Segmentation

    Fabio Cermelli, Massimiliano Mancini, Samuel Rota Bulò +2

    cs.CVarXiv:2002.00718v22020
  5. Flexible Diffusion Modeling of Long Videos

    William Harvey, Saeid Naderiparizi, Vaden Masrani +2

    cs.CVcs.LGarXiv:2205.11495v32022
  6. In-Place Activated BatchNorm for Memory-Optimized Training of DNNs

    Samuel Rota Bulò, Lorenzo Porzi, Peter Kontschieder

    cs.CVarXiv:1712.02616v32017
  7. Appearance-and-Relation Networks for Video Classification

    Limin Wang, Wei Li, Wen Li +1

    cs.CVarXiv:1711.09125v22017
  8. UniFormer: Unified Transformer for Efficient Spatiotemporal Representation Learning

    Kunchang Li, Yali Wang, Peng Gao +4

    cs.CVarXiv:2201.04676v32022
  9. Skeleton-Based Action Recognition with Spatial Reasoning and Temporal Stack Learning

    Chenyang Si, Ya Jing, Wei Wang +2

    cs.CVarXiv:1805.02335v22018
  10. Morphing and Sampling Network for Dense Point Cloud Completion

    Minghua Liu, Lu Sheng, Sheng Yang +2

    cs.CVarXiv:1912.00280v12019
  11. Building a Large Scale Dataset for Image Emotion Recognition: The Fine Print and The Benchmark

    Quanzeng You, Jiebo Luo, Hailin Jin +1

    cs.AIcs.CVarXiv:1605.02677v12016
  12. Residual Networks of Residual Networks: Multilevel Residual Networks

    Ke Zhang, Miao Sun, Tony X. Han +3

    cs.CVarXiv:1608.02908v22016
  13. Video Summarization with Attention-Based Encoder-Decoder Networks

    Zhong Ji, Kailin Xiong, Yanwei Pang +1

    cs.CVarXiv:1708.09545v22017
  14. Learning Two-View Correspondences and Geometry Using Order-Aware Network

    Jiahui Zhang, Dawei Sun, Zixin Luo +6

    cs.CVcs.CGcs.LGarXiv:1908.04964v12019
  15. Deep Learning Based Brain Tumor Segmentation: A Survey

    Zhihua Liu, Lei Tong, Zheheng Jiang +6

    eess.IVcs.CVarXiv:2007.09479v32020
  16. Self-supervised Learning in Remote Sensing: A Review

    Yi Wang, Conrad M Albrecht, Nassim Ait Ali Braham +2

    cs.CVarXiv:2206.13188v22022
  17. NuScenes-QA: A Multi-modal Visual Question Answering Benchmark for Autonomous Driving Scenario

    Tianwen Qian, Jingjing Chen, Linhai Zhuo +2

    cs.CVarXiv:2305.14836v22023
  18. Neural Nearest Neighbors Networks

    Tobias Plötz, Stefan Roth

    cs.CVcs.LGarXiv:1810.12575v12018
  19. DreamSim: Learning New Dimensions of Human Visual Similarity using Synthetic Data

    Stephanie Fu, Netanel Tamir, Shobhita Sundaram +4

    cs.CVcs.LGarXiv:2306.09344v32023
  20. Camera Distance-aware Top-down Approach for 3D Multi-person Pose Estimation from a Single RGB Image

    Gyeongsik Moon, Ju Yong Chang, Kyoung Mu Lee

    cs.CVarXiv:1907.11346v22019
  21. RoadTracer: Automatic Extraction of Road Networks from Aerial Images

    Favyen Bastani, Songtao He, Sofiane Abbar +5

    cs.CVarXiv:1802.03680v22018
  22. Classification of Hyperspectral and LiDAR Data Using Coupled CNNs

    Renlong Hang, Zhu Li, Pedram Ghamisi +3

    cs.CVeess.IVarXiv:2002.01144v12020
  23. Low-rank Bilinear Pooling for Fine-Grained Classification

    Shu Kong, Charless Fowlkes

    cs.CVarXiv:1611.05109v22016
  24. Compressed Video Action Recognition

    Chao-Yuan Wu, Manzil Zaheer, Hexiang Hu +3

    cs.CVarXiv:1712.00636v22017
  25. Neural Prototype Trees for Interpretable Fine-grained Image Recognition

    Meike Nauta, Ron van Bree, Christin Seifert

    cs.CVcs.AIcs.LGarXiv:2012.02046v22020
  26. Uncertainty-Aware Blind Image Quality Assessment in the Laboratory and Wild

    Weixia Zhang, Kede Ma, Guangtao Zhai +1

    cs.CVcs.LGcs.MMarXiv:2005.13983v62020
  27. Disentangled Non-Local Neural Networks

    Minghao Yin, Zhuliang Yao, Yue Cao +4

    cs.CVcs.CLcs.LGarXiv:2006.06668v22020
  28. How Neural Networks Extrapolate: From Feedforward to Graph Neural Networks

    Keyulu Xu, Mozhi Zhang, Jingling Li +3

    cs.LGcs.AIcs.CVarXiv:2009.11848v52020
  29. Stereo DSO: Large-Scale Direct Sparse Visual Odometry with Stereo Cameras

    Rui Wang, Martin Schwörer, Daniel Cremers

    cs.CVarXiv:1708.07878v12017
  30. Designing Deep Networks for Surface Normal Estimation

    Xiaolong Wang, David F. Fouhey, Abhinav Gupta

    cs.CVarXiv:1411.4958v12014
  31. Multi-task Collaborative Network for Joint Referring Expression Comprehension and Segmentation

    Gen Luo, Yiyi Zhou, Xiaoshuai Sun +4

    cs.CVarXiv:2003.08813v12020
  32. Appearance-Based Loop Closure Detection for Online Large-Scale and Long-Term Operation

    Mathieu Labbé, François Michaud

    cs.ROcs.CVarXiv:2407.15304v12024
  33. Dual Motion GAN for Future-Flow Embedded Video Prediction

    Xiaodan Liang, Lisa Lee, Wei Dai +1

    cs.CVarXiv:1708.00284v22017
  34. TOPIQ: A Top-down Approach from Semantics to Distortions for Image Quality Assessment

    Chaofeng Chen, Jiadi Mo, Jingwen Hou +5

    cs.CVarXiv:2308.03060v12023
  35. Plug-and-Play Priors for Bright Field Electron Tomography and Sparse Interpolation

    Suhas Sreehari, S. V. Venkatakrishnan, Brendt Wohlberg +3

    cs.CVeess.IVarXiv:1512.07331v12015
  36. Probabilistic Face Embeddings

    Yichun Shi, Anil K. Jain

    cs.CVarXiv:1904.09658v42019
  37. FaceScape: a Large-scale High Quality 3D Face Dataset and Detailed Riggable 3D Face Prediction

    Haotian Yang, Hao Zhu, Yanru Wang +4

    cs.CVarXiv:2003.13989v32020
  38. Hough-CNN: Deep Learning for Segmentation of Deep Brain Regions in MRI and Ultrasound

    Fausto Milletari, Seyed-Ahmad Ahmadi, Christine Kroll +8

    cs.CVarXiv:1601.07014v32016
  39. PolyGen: An Autoregressive Generative Model of 3D Meshes

    Charlie Nash, Yaroslav Ganin, S. M. Ali Eslami +1

    cs.GRcs.CVcs.LGarXiv:2002.10880v12020
  40. HP-GAN: Probabilistic 3D human motion prediction via GAN

    Emad Barsoum, John Kender, Zicheng Liu

    cs.CVcs.AIcs.HCarXiv:1711.09561v12017
  41. Variational Denoising Network: Toward Blind Noise Modeling and Removal

    Zongsheng Yue, Hongwei Yong, Qian Zhao +2

    cs.CVarXiv:1908.11314v52019
  42. Can Deep Learning Outperform Modern Commercial CT Image Reconstruction Methods?

    Hongming Shan, Atul Padole, Fatemeh Homayounieh +5

    cs.CVphysics.med-pharXiv:1811.03691v12018
  43. UC-Net: Uncertainty Inspired RGB-D Saliency Detection via Conditional Variational Autoencoders

    Jing Zhang, Deng-Ping Fan, Yuchao Dai +4

    cs.CVarXiv:2004.05763v12020
  44. Synergistic Image and Feature Adaptation: Towards Cross-Modality Domain Adaptation for Medical Image Segmentation

    Cheng Chen, Qi Dou, Hao Chen +2

    cs.CVarXiv:1901.08211v42019
  45. DecideNet: Counting Varying Density Crowds Through Attention Guided Detection and Density Estimation

    Jiang Liu, Chenqiang Gao, Deyu Meng +1

    cs.CVarXiv:1712.06679v22017
  46. YOLOv6 v3.0: A Full-Scale Reloading

    Chuyi Li, Lulu Li, Yifei Geng +6

    cs.CVarXiv:2301.05586v12023
  47. TAM: Temporal Adaptive Module for Video Recognition

    Zhaoyang Liu, Limin Wang, Wayne Wu +2

    cs.CVarXiv:2005.06803v32020
  48. Not All Points Are Equal: Learning Highly Efficient Point-based Detectors for 3D LiDAR Point Clouds

    Yifan Zhang, Qingyong Hu, Guoquan Xu +3

    cs.CVcs.ROarXiv:2203.11139v12022
  49. SegStereo: Exploiting Semantic Information for Disparity Estimation

    Guorun Yang, Hengshuang Zhao, Jianping Shi +2

    cs.CVarXiv:1807.11699v12018
  50. Target-Aware Deep Tracking

    Xin Li, Chao Ma, Baoyuan Wu +2

    cs.CVarXiv:1904.01772v12019
  51. Look into Person: Joint Body Parsing & Pose Estimation Network and A New Benchmark

    Xiaodan Liang, Ke Gong, Xiaohui Shen +1

    cs.CVarXiv:1804.01984v12018
  52. Object-Part Attention Model for Fine-grained Image Classification

    Yuxin Peng, Xiangteng He, Junjie Zhao

    cs.CVarXiv:1704.01740v22017
  53. The More You Know: Using Knowledge Graphs for Image Classification

    Kenneth Marino, Ruslan Salakhutdinov, Abhinav Gupta

    cs.CVarXiv:1612.04844v22016
  54. Representation Learning by Learning to Count

    Mehdi Noroozi, Hamed Pirsiavash, Paolo Favaro

    cs.CVarXiv:1708.06734v12017
  55. PseudoSeg: Designing Pseudo Labels for Semantic Segmentation

    Yuliang Zou, Zizhao Zhang, Han Zhang +4

    cs.CVarXiv:2010.09713v22020
  56. Universal Correspondence Network

    Christopher B. Choy, JunYoung Gwak, Silvio Savarese +1

    cs.CVarXiv:1606.03558v32016
  57. Do We Really Need to Collect Millions of Faces for Effective Face Recognition?

    Iacopo Masi, Anh Tuan Tran, Jatuporn Toy Leksut +2

    cs.CVarXiv:1603.07057v22016
  58. Training-Free Layout Control with Cross-Attention Guidance

    Minghao Chen, Iro Laina, Andrea Vedaldi

    cs.CVarXiv:2304.03373v22023
  59. Safety-Enhanced Autonomous Driving Using Interpretable Sensor Fusion Transformer

    Hao Shao, Letian Wang, RuoBing Chen +2

    cs.CVcs.AIcs.LGarXiv:2207.14024v52022
  60. PaMIR: Parametric Model-Conditioned Implicit Representation for Image-based Human Reconstruction

    Zerong Zheng, Tao Yu, Yebin Liu +1

    cs.CVarXiv:2007.03858v22020