Computer Vision and Pattern Recognition
Papers filed under cs.CV on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.
Search paper metadata (including unsummarized papers)
11,461 to 11,520 of 18,959
Generative Novel View Synthesis with 3D-Aware Diffusion Models
Eric R. Chan, Koki Nagano, Matthew A. Chan +7
cs.CVcs.AIcs.GRarXiv:2304.02602v12023Deep Adversarial Training for Multi-Organ Nuclei Segmentation in Histopathology Images
Faisal Mahmood, Daniel Borders, Richard Chen +4
cs.CVarXiv:1810.00236v22018Perceptual Quality Prediction on Authentically Distorted Images Using a Bag of Features Approach
Deepti Ghadiyaram, Alan C. Bovik
cs.CVarXiv:1609.04757v12016How Well Do Self-Supervised Models Transfer?
Linus Ericsson, Henry Gouk, Timothy M. Hospedales
cs.CVarXiv:2011.13377v22020Multispectral Fusion for Object Detection with Cyclic Fuse-and-Refine Blocks
Heng Zhang, Elisa Fromont, Sébastien Lefevre +1
cs.CVarXiv:2009.12664v12020Gaussian Opacity Fields: Efficient Adaptive Surface Reconstruction in Unbounded Scenes
Zehao Yu, Torsten Sattler, Andreas Geiger
cs.CVarXiv:2404.10772v22024StainGAN: Stain Style Transfer for Digital Histological Images
M Tarek Shaban, Christoph Baur, Nassir Navab +1
cs.CVarXiv:1804.01601v12018Improving Multispectral Pedestrian Detection by Addressing Modality Imbalance Problems
Kailai Zhou, Linsen Chen, Xun Cao
cs.CVarXiv:2008.03043v22020BOP Challenge 2020 on 6D Object Localization
Tomas Hodan, Martin Sundermeyer, Bertram Drost +5
cs.CVcs.GRcs.LGarXiv:2009.07378v22020Asynchronous, Photometric Feature Tracking using Events and Frames
Daniel Gehrig, Henri Rebecq, Guillermo Gallego +1
cs.CVcs.ROarXiv:1807.09713v12018Deep Neural Networks for Anatomical Brain Segmentation
Alexandre de Brebisson, Giovanni Montana
cs.CVcs.LGstat.AParXiv:1502.02445v22015Self-supervised Learning of Motion Capture
Hsiao-Yu Fish Tung, Hsiao-Wei Tung, Ersin Yumer +1
cs.CVarXiv:1712.01337v12017Key.Net: Keypoint Detection by Handcrafted and Learned CNN Filters
Axel Barroso-Laguna, Edgar Riba, Daniel Ponsa +1
cs.CVarXiv:1904.00889v32019CelebV-HQ: A Large-Scale Video Facial Attributes Dataset
Hao Zhu, Wayne Wu, Wentao Zhu +5
cs.CVarXiv:2207.12393v12022LongVILA: Scaling Long-Context Visual Language Models for Long Videos
Yukang Chen, Fuzhao Xue, Dacheng Li +15
cs.CVcs.CLarXiv:2408.10188v62024Modeling and Propagating CNNs in a Tree Structure for Visual Tracking
Hyeonseob Nam, Mooyeol Baek, Bohyung Han
cs.CVarXiv:1608.07242v12016Graduated Non-Convexity for Robust Spatial Perception: From Non-Minimal Solvers to Global Outlier Rejection
Heng Yang, Pasquale Antonante, Vasileios Tzoumas +1
cs.CVcs.ROmath.OCarXiv:1909.08605v42019Inverse Rendering for Complex Indoor Scenes: Shape, Spatially-Varying Lighting and SVBRDF from a Single Image
Zhengqin Li, Mohammad Shafiei, Ravi Ramamoorthi +2
cs.CVarXiv:1905.02722v12019Generalized Source-free Domain Adaptation
Shiqi Yang, Yaxing Wang, Joost van de Weijer +2
cs.CVarXiv:2108.01614v22021Aggregated Contextual Transformations for High-Resolution Image Inpainting
Yanhong Zeng, Jianlong Fu, Hongyang Chao +1
cs.CVarXiv:2104.01431v12021On Regularized Losses for Weakly-supervised CNN Segmentation
Meng Tang, Federico Perazzi, Abdelaziz Djelouah +3
cs.CVarXiv:1803.09569v22018Cross-view Semantic Segmentation for Sensing Surroundings
Bowen Pan, Jiankai Sun, Ho Yin Tiga Leung +2
cs.CVeess.IVarXiv:1906.03560v32019A Comprehensive Analysis of Deep Regression
Stéphane Lathuilière, Pablo Mesejo, Xavier Alameda-Pineda +1
cs.CVarXiv:1803.08450v32018On the uncertainty of self-supervised monocular depth estimation
Matteo Poggi, Filippo Aleotti, Fabio Tosi +1
cs.CVarXiv:2005.06209v12020Large-scale Multi-Modal Pre-trained Models: A Comprehensive Survey
Xiao Wang, Guangyao Chen, Guangwu Qian +5
cs.CVcs.AIcs.MMarXiv:2302.10035v32023Diffusion Probabilistic Modeling for Video Generation
Ruihan Yang, Prakhar Srivastava, Stephan Mandt
cs.CVcs.LGstat.MLarXiv:2203.09481v52022Towards Dropout Training for Convolutional Neural Networks
Haibing Wu, Xiaodong Gu
cs.LGcs.CVcs.NEarXiv:1512.00242v12015Unsupervised Cross-Modality Domain Adaptation of ConvNets for Biomedical Image Segmentations with Adversarial Loss
Qi Dou, Cheng Ouyang, Cheng Chen +2
cs.CVarXiv:1804.10916v22018ScanComplete: Large-Scale Scene Completion and Semantic Segmentation for 3D Scans
Angela Dai, Daniel Ritchie, Martin Bokeloh +3
cs.CVarXiv:1712.10215v22017Alpha-IoU: A Family of Power Intersection over Union Losses for Bounding Box Regression
Jiabo He, Sarah Erfani, Xingjun Ma +3
cs.CVarXiv:2110.13675v22021Adversarial Active Learning for Deep Networks: a Margin Based Approach
Melanie Ducoffe, Frederic Precioso
cs.LGcs.CVstat.MLarXiv:1802.09841v12018Spatially-Attentive Patch-Hierarchical Network for Adaptive Motion Deblurring
Maitreya Suin, Kuldeep Purohit, A. N. Rajagopalan
cs.CVeess.IVarXiv:2004.05343v12020ZSON: Zero-Shot Object-Goal Navigation using Multimodal Goal Embeddings
Arjun Majumdar, Gunjan Aggarwal, Bhavika Devnani +2
cs.CVcs.LGcs.ROarXiv:2206.12403v22022Disentangling Adversarial Robustness and Generalization
David Stutz, Matthias Hein, Bernt Schiele
cs.CVcs.CRcs.LGarXiv:1812.00740v22018Source-Free Domain Adaptation for Semantic Segmentation
Yuang Liu, Wei Zhang, Jun Wang
cs.CVarXiv:2103.16372v12021Automatic Spatially-aware Fashion Concept Discovery
Xintong Han, Zuxuan Wu, Phoenix X. Huang +5
cs.CVarXiv:1708.01311v12017Towards Large-Pose Face Frontalization in the Wild
Xi Yin, Xiang Yu, Kihyuk Sohn +2
cs.CVarXiv:1704.06244v32017Neural Video Compression with Diverse Contexts
Jiahao Li, Bin Li, Yan Lu
eess.IVcs.CVcs.MMarXiv:2302.14402v32023On the Challenges and Perspectives of Foundation Models for Medical Image Analysis
Shaoting Zhang, Dimitris Metaxas
eess.IVcs.CVarXiv:2306.05705v22023Dynamic Refinement Network for Oriented and Densely Packed Object Detection
Xingjia Pan, Yuqiang Ren, Kekai Sheng +5
cs.CVarXiv:2005.09973v22020U-Net Transformer: Self and Cross Attention for Medical Image Segmentation
Olivier Petit, Nicolas Thome, Clément Rambour +1
eess.IVcs.CVarXiv:2103.06104v22021Jointly Attentive Spatial-Temporal Pooling Networks for Video-based Person Re-Identification
Shuangjie Xu, Yu Cheng, Kang Gu +3
cs.CVcs.LGstat.MLarXiv:1708.02286v22017Deep Supervised Discrete Hashing
Qi Li, Zhenan Sun, Ran He +1
cs.CVarXiv:1705.10999v22017The DeepFake Detection Challenge (DFDC) Dataset
Brian Dolhansky, Joanna Bitton, Ben Pflaum +4
cs.CVcs.LGarXiv:2006.07397v42020LasHeR: A Large-scale High-diversity Benchmark for RGBT Tracking
Chenglong Li, Wanlin Xue, Yaqing Jia +4
cs.CVarXiv:2104.13202v22021Ambient Sound Provides Supervision for Visual Learning
Andrew Owens, Jiajun Wu, Josh H. McDermott +2
cs.CVarXiv:1608.07017v22016When Face Recognition Meets with Deep Learning: an Evaluation of Convolutional Neural Networks for Face Recognition
Guosheng Hu, Yongxin Yang, Dong Yi +4
cs.CVcs.LGcs.NEarXiv:1504.02351v12015STAR: Sparse Trained Articulated Human Body Regressor
Ahmed A. A. Osman, Timo Bolkart, Michael J. Black
cs.CVarXiv:2008.08535v12020Recurrent Instance Segmentation
Bernardino Romera-Paredes, Philip H. S. Torr
cs.CVcs.AIarXiv:1511.08250v32015UniRepLKNet: A Universal Perception Large-Kernel ConvNet for Audio, Video, Point Cloud, Time-Series and Image Recognition
Xiaohan Ding, Yiyuan Zhang, Yixiao Ge +4
cs.CVcs.AIcs.LGarXiv:2311.15599v22023GPT4Tools: Teaching Large Language Model to Use Tools via Self-instruction
Rui Yang, Lin Song, Yanwei Li +4
cs.CVcs.CLarXiv:2305.18752v12023MonoPair: Monocular 3D Object Detection Using Pairwise Spatial Relationships
Yongjian Chen, Lei Tai, Kai Sun +1
cs.CVcs.AIcs.LGarXiv:2003.00504v12020AANet: Attribute Attention Network for Person Re-Identifications
Chiat-Pin Tay, Sharmili Roy, Kim-Hui Yap
cs.CVarXiv:1912.09021v12019Perceive Where to Focus: Learning Visibility-aware Part-level Features for Partial Person Re-identification
Yifan Sun, Qin Xu, Yali Li +4
cs.CVarXiv:1904.00537v12019PPDM: Parallel Point Detection and Matching for Real-time Human-Object Interaction Detection
Yue Liao, Si Liu, Fei Wang +3
cs.CVarXiv:1912.12898v32019LViT: Language meets Vision Transformer in Medical Image Segmentation
Zihan Li, Yunxiang Li, Qingde Li +6
cs.CVarXiv:2206.14718v42022Learning Visual Clothing Style with Heterogeneous Dyadic Co-occurrences
Andreas Veit, Balazs Kovacs, Sean Bell +3
cs.CVarXiv:1509.07473v12015Person Re-Identification by Camera Correlation Aware Feature Augmentation
Ying-Cong Chen, Xiatian Zhu, Wei-Shi Zheng +1
cs.CVarXiv:1703.08837v12017Multi-Agent Diverse Generative Adversarial Networks
Arnab Ghosh, Viveka Kulharia, Vinay Namboodiri +2
cs.CVcs.AIcs.GRarXiv:1704.02906v32017Deep Alignment Network: A convolutional neural network for robust face alignment
Marek Kowalski, Jacek Naruniec, Tomasz Trzcinski
cs.CVarXiv:1706.01789v22017