Computer Vision and Pattern Recognition

Papers filed under cs.CV on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.

Search paper metadata (including unsummarized papers)

11,461 to 11,520 of 18,959

  1. Generative Novel View Synthesis with 3D-Aware Diffusion Models

    Eric R. Chan, Koki Nagano, Matthew A. Chan +7

    cs.CVcs.AIcs.GRarXiv:2304.02602v12023
  2. Deep Adversarial Training for Multi-Organ Nuclei Segmentation in Histopathology Images

    Faisal Mahmood, Daniel Borders, Richard Chen +4

    cs.CVarXiv:1810.00236v22018
  3. Perceptual Quality Prediction on Authentically Distorted Images Using a Bag of Features Approach

    Deepti Ghadiyaram, Alan C. Bovik

    cs.CVarXiv:1609.04757v12016
  4. How Well Do Self-Supervised Models Transfer?

    Linus Ericsson, Henry Gouk, Timothy M. Hospedales

    cs.CVarXiv:2011.13377v22020
  5. Multispectral Fusion for Object Detection with Cyclic Fuse-and-Refine Blocks

    Heng Zhang, Elisa Fromont, Sébastien Lefevre +1

    cs.CVarXiv:2009.12664v12020
  6. Gaussian Opacity Fields: Efficient Adaptive Surface Reconstruction in Unbounded Scenes

    Zehao Yu, Torsten Sattler, Andreas Geiger

    cs.CVarXiv:2404.10772v22024
  7. StainGAN: Stain Style Transfer for Digital Histological Images

    M Tarek Shaban, Christoph Baur, Nassir Navab +1

    cs.CVarXiv:1804.01601v12018
  8. Improving Multispectral Pedestrian Detection by Addressing Modality Imbalance Problems

    Kailai Zhou, Linsen Chen, Xun Cao

    cs.CVarXiv:2008.03043v22020
  9. BOP Challenge 2020 on 6D Object Localization

    Tomas Hodan, Martin Sundermeyer, Bertram Drost +5

    cs.CVcs.GRcs.LGarXiv:2009.07378v22020
  10. Asynchronous, Photometric Feature Tracking using Events and Frames

    Daniel Gehrig, Henri Rebecq, Guillermo Gallego +1

    cs.CVcs.ROarXiv:1807.09713v12018
  11. Deep Neural Networks for Anatomical Brain Segmentation

    Alexandre de Brebisson, Giovanni Montana

    cs.CVcs.LGstat.AParXiv:1502.02445v22015
  12. Self-supervised Learning of Motion Capture

    Hsiao-Yu Fish Tung, Hsiao-Wei Tung, Ersin Yumer +1

    cs.CVarXiv:1712.01337v12017
  13. Key.Net: Keypoint Detection by Handcrafted and Learned CNN Filters

    Axel Barroso-Laguna, Edgar Riba, Daniel Ponsa +1

    cs.CVarXiv:1904.00889v32019
  14. CelebV-HQ: A Large-Scale Video Facial Attributes Dataset

    Hao Zhu, Wayne Wu, Wentao Zhu +5

    cs.CVarXiv:2207.12393v12022
  15. LongVILA: Scaling Long-Context Visual Language Models for Long Videos

    Yukang Chen, Fuzhao Xue, Dacheng Li +15

    cs.CVcs.CLarXiv:2408.10188v62024
  16. Modeling and Propagating CNNs in a Tree Structure for Visual Tracking

    Hyeonseob Nam, Mooyeol Baek, Bohyung Han

    cs.CVarXiv:1608.07242v12016
  17. Graduated Non-Convexity for Robust Spatial Perception: From Non-Minimal Solvers to Global Outlier Rejection

    Heng Yang, Pasquale Antonante, Vasileios Tzoumas +1

    cs.CVcs.ROmath.OCarXiv:1909.08605v42019
  18. Inverse Rendering for Complex Indoor Scenes: Shape, Spatially-Varying Lighting and SVBRDF from a Single Image

    Zhengqin Li, Mohammad Shafiei, Ravi Ramamoorthi +2

    cs.CVarXiv:1905.02722v12019
  19. Generalized Source-free Domain Adaptation

    Shiqi Yang, Yaxing Wang, Joost van de Weijer +2

    cs.CVarXiv:2108.01614v22021
  20. Aggregated Contextual Transformations for High-Resolution Image Inpainting

    Yanhong Zeng, Jianlong Fu, Hongyang Chao +1

    cs.CVarXiv:2104.01431v12021
  21. On Regularized Losses for Weakly-supervised CNN Segmentation

    Meng Tang, Federico Perazzi, Abdelaziz Djelouah +3

    cs.CVarXiv:1803.09569v22018
  22. Cross-view Semantic Segmentation for Sensing Surroundings

    Bowen Pan, Jiankai Sun, Ho Yin Tiga Leung +2

    cs.CVeess.IVarXiv:1906.03560v32019
  23. A Comprehensive Analysis of Deep Regression

    Stéphane Lathuilière, Pablo Mesejo, Xavier Alameda-Pineda +1

    cs.CVarXiv:1803.08450v32018
  24. On the uncertainty of self-supervised monocular depth estimation

    Matteo Poggi, Filippo Aleotti, Fabio Tosi +1

    cs.CVarXiv:2005.06209v12020
  25. Large-scale Multi-Modal Pre-trained Models: A Comprehensive Survey

    Xiao Wang, Guangyao Chen, Guangwu Qian +5

    cs.CVcs.AIcs.MMarXiv:2302.10035v32023
  26. Diffusion Probabilistic Modeling for Video Generation

    Ruihan Yang, Prakhar Srivastava, Stephan Mandt

    cs.CVcs.LGstat.MLarXiv:2203.09481v52022
  27. Towards Dropout Training for Convolutional Neural Networks

    Haibing Wu, Xiaodong Gu

    cs.LGcs.CVcs.NEarXiv:1512.00242v12015
  28. Unsupervised Cross-Modality Domain Adaptation of ConvNets for Biomedical Image Segmentations with Adversarial Loss

    Qi Dou, Cheng Ouyang, Cheng Chen +2

    cs.CVarXiv:1804.10916v22018
  29. ScanComplete: Large-Scale Scene Completion and Semantic Segmentation for 3D Scans

    Angela Dai, Daniel Ritchie, Martin Bokeloh +3

    cs.CVarXiv:1712.10215v22017
  30. Alpha-IoU: A Family of Power Intersection over Union Losses for Bounding Box Regression

    Jiabo He, Sarah Erfani, Xingjun Ma +3

    cs.CVarXiv:2110.13675v22021
  31. Adversarial Active Learning for Deep Networks: a Margin Based Approach

    Melanie Ducoffe, Frederic Precioso

    cs.LGcs.CVstat.MLarXiv:1802.09841v12018
  32. Spatially-Attentive Patch-Hierarchical Network for Adaptive Motion Deblurring

    Maitreya Suin, Kuldeep Purohit, A. N. Rajagopalan

    cs.CVeess.IVarXiv:2004.05343v12020
  33. ZSON: Zero-Shot Object-Goal Navigation using Multimodal Goal Embeddings

    Arjun Majumdar, Gunjan Aggarwal, Bhavika Devnani +2

    cs.CVcs.LGcs.ROarXiv:2206.12403v22022
  34. Disentangling Adversarial Robustness and Generalization

    David Stutz, Matthias Hein, Bernt Schiele

    cs.CVcs.CRcs.LGarXiv:1812.00740v22018
  35. Source-Free Domain Adaptation for Semantic Segmentation

    Yuang Liu, Wei Zhang, Jun Wang

    cs.CVarXiv:2103.16372v12021
  36. Automatic Spatially-aware Fashion Concept Discovery

    Xintong Han, Zuxuan Wu, Phoenix X. Huang +5

    cs.CVarXiv:1708.01311v12017
  37. Towards Large-Pose Face Frontalization in the Wild

    Xi Yin, Xiang Yu, Kihyuk Sohn +2

    cs.CVarXiv:1704.06244v32017
  38. Neural Video Compression with Diverse Contexts

    Jiahao Li, Bin Li, Yan Lu

    eess.IVcs.CVcs.MMarXiv:2302.14402v32023
  39. On the Challenges and Perspectives of Foundation Models for Medical Image Analysis

    Shaoting Zhang, Dimitris Metaxas

    eess.IVcs.CVarXiv:2306.05705v22023
  40. Dynamic Refinement Network for Oriented and Densely Packed Object Detection

    Xingjia Pan, Yuqiang Ren, Kekai Sheng +5

    cs.CVarXiv:2005.09973v22020
  41. U-Net Transformer: Self and Cross Attention for Medical Image Segmentation

    Olivier Petit, Nicolas Thome, Clément Rambour +1

    eess.IVcs.CVarXiv:2103.06104v22021
  42. Jointly Attentive Spatial-Temporal Pooling Networks for Video-based Person Re-Identification

    Shuangjie Xu, Yu Cheng, Kang Gu +3

    cs.CVcs.LGstat.MLarXiv:1708.02286v22017
  43. Deep Supervised Discrete Hashing

    Qi Li, Zhenan Sun, Ran He +1

    cs.CVarXiv:1705.10999v22017
  44. The DeepFake Detection Challenge (DFDC) Dataset

    Brian Dolhansky, Joanna Bitton, Ben Pflaum +4

    cs.CVcs.LGarXiv:2006.07397v42020
  45. LasHeR: A Large-scale High-diversity Benchmark for RGBT Tracking

    Chenglong Li, Wanlin Xue, Yaqing Jia +4

    cs.CVarXiv:2104.13202v22021
  46. Ambient Sound Provides Supervision for Visual Learning

    Andrew Owens, Jiajun Wu, Josh H. McDermott +2

    cs.CVarXiv:1608.07017v22016
  47. When Face Recognition Meets with Deep Learning: an Evaluation of Convolutional Neural Networks for Face Recognition

    Guosheng Hu, Yongxin Yang, Dong Yi +4

    cs.CVcs.LGcs.NEarXiv:1504.02351v12015
  48. STAR: Sparse Trained Articulated Human Body Regressor

    Ahmed A. A. Osman, Timo Bolkart, Michael J. Black

    cs.CVarXiv:2008.08535v12020
  49. Recurrent Instance Segmentation

    Bernardino Romera-Paredes, Philip H. S. Torr

    cs.CVcs.AIarXiv:1511.08250v32015
  50. UniRepLKNet: A Universal Perception Large-Kernel ConvNet for Audio, Video, Point Cloud, Time-Series and Image Recognition

    Xiaohan Ding, Yiyuan Zhang, Yixiao Ge +4

    cs.CVcs.AIcs.LGarXiv:2311.15599v22023
  51. GPT4Tools: Teaching Large Language Model to Use Tools via Self-instruction

    Rui Yang, Lin Song, Yanwei Li +4

    cs.CVcs.CLarXiv:2305.18752v12023
  52. MonoPair: Monocular 3D Object Detection Using Pairwise Spatial Relationships

    Yongjian Chen, Lei Tai, Kai Sun +1

    cs.CVcs.AIcs.LGarXiv:2003.00504v12020
  53. AANet: Attribute Attention Network for Person Re-Identifications

    Chiat-Pin Tay, Sharmili Roy, Kim-Hui Yap

    cs.CVarXiv:1912.09021v12019
  54. Perceive Where to Focus: Learning Visibility-aware Part-level Features for Partial Person Re-identification

    Yifan Sun, Qin Xu, Yali Li +4

    cs.CVarXiv:1904.00537v12019
  55. PPDM: Parallel Point Detection and Matching for Real-time Human-Object Interaction Detection

    Yue Liao, Si Liu, Fei Wang +3

    cs.CVarXiv:1912.12898v32019
  56. LViT: Language meets Vision Transformer in Medical Image Segmentation

    Zihan Li, Yunxiang Li, Qingde Li +6

    cs.CVarXiv:2206.14718v42022
  57. Learning Visual Clothing Style with Heterogeneous Dyadic Co-occurrences

    Andreas Veit, Balazs Kovacs, Sean Bell +3

    cs.CVarXiv:1509.07473v12015
  58. Person Re-Identification by Camera Correlation Aware Feature Augmentation

    Ying-Cong Chen, Xiatian Zhu, Wei-Shi Zheng +1

    cs.CVarXiv:1703.08837v12017
  59. Multi-Agent Diverse Generative Adversarial Networks

    Arnab Ghosh, Viveka Kulharia, Vinay Namboodiri +2

    cs.CVcs.AIcs.GRarXiv:1704.02906v32017
  60. Deep Alignment Network: A convolutional neural network for robust face alignment

    Marek Kowalski, Jacek Naruniec, Tomasz Trzcinski

    cs.CVarXiv:1706.01789v22017