Computer Vision and Pattern Recognition
Papers filed under cs.CV on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.
Search paper metadata (including unsummarized papers)
8,101 to 8,160 of 18,849
Contextual Encoder-Decoder Network for Visual Saliency Prediction
Alexander Kroner, Mario Senden, Kurt Driessens +1
cs.CVarXiv:1902.06634v42019Learning in an Uncertain World: Representing Ambiguity Through Multiple Hypotheses
Christian Rupprecht, Iro Laina, Robert DiPietro +4
cs.CVarXiv:1612.00197v32016A-Lamp: Adaptive Layout-Aware Multi-Patch Deep Convolutional Neural Network for Photo Aesthetic Assessment
Shuang Ma, Jing Liu, Chang Wen Chen
cs.CVarXiv:1704.00248v12017From Synthetic to Real: Image Dehazing Collaborating with Unlabeled Real Data
Ye Liu, Lei Zhu, Shunda Pei +5
cs.CVarXiv:2108.02934v12021FlowVVTON: Flow-Guided Mask-Free Video Virtual Try-On
Shengyao Chen, Xianbing Sun, Liqing Zhang +1
cs.CVarXiv:2608.30450v12026PRISM: Predictive Recomposition via Semantic Latent Decomposition for View-invariant Video Representation Learning
Youngchae Chee, Hosu Lee, Sungjune Park +2
cs.CVcs.AIarXiv:2608.30388v12026Weakly Supervised Video Moment Retrieval From Text Queries
Niluthpol Chowdhury Mithun, Sujoy Paul, Amit K. Roy-Chowdhury
cs.CVcs.MMarXiv:1904.03282v22019You Only Need Adversarial Supervision for Semantic Image Synthesis
Vadim Sushko, Edgar Schönfeld, Dan Zhang +3
cs.CVcs.LGeess.IVarXiv:2012.04781v32020Differentiable Rendering: A Survey
Hiroharu Kato, Deniz Beker, Mihai Morariu +4
cs.CVcs.GRarXiv:2006.12057v22020A Taxonomy of Deep Convolutional Neural Nets for Computer Vision
Suraj Srinivas, Ravi Kiran Sarvadevabhatla, Konda Reddy Mopuri +3
cs.CVcs.LGcs.MMarXiv:1601.06615v12016REVISE: A Tool for Measuring and Mitigating Bias in Visual Datasets
Angelina Wang, Alexander Liu, Ryan Zhang +6
cs.CVarXiv:2004.07999v42020The Devil is in the Tails: Fine-grained Classification in the Wild
Grant Van Horn, Pietro Perona
cs.CVarXiv:1709.01450v12017Doc-REFRAG: Rethinking Multimodal Document Retrieval-Augmented Generation
Ruofan Hu, Shengyang Xu, Minjie Hong +5
cs.IRcs.CVarXiv:2608.30163v12026DeRF: Decomposed Radiance Fields
Daniel Rebain, Wei Jiang, Soroosh Yazdani +3
cs.CVcs.GRarXiv:2011.12490v12020Overcoming Catastrophic Forgetting in Incremental Few-Shot Learning by Finding Flat Minima
Guangyuan Shi, Jiaxin Chen, Wenlong Zhang +2
cs.LGcs.CVarXiv:2111.01549v22021DALES: A Large-scale Aerial LiDAR Data Set for Semantic Segmentation
Nina Varney, Vijayan K. Asari, Quinn Graehling
cs.CVcs.LGstat.MLarXiv:2004.11985v12020Direction-aware Spatial Context Features for Shadow Detection
Xiaowei Hu, Lei Zhu, Chi-Wing Fu +2
cs.CVarXiv:1712.04142v22017Neural Rendering for Stereo 3D Reconstruction of Deformable Tissues in Robotic Surgery
Yuehao Wang, Yonghao Long, Siu Hin Fan +1
cs.CVarXiv:2206.15255v12022Everybody Tracking Every Body
Daeyun Shin, Yunhan Zhao, Shu Kong +2
cs.CVarXiv:2608.29927v12026Underwater Optical Image Processing: A Comprehensive Review
Huimin Lu, Yujie Li, Yudong Zhang +3
cs.CVarXiv:1702.03600v12017Boundary-aware Context Neural Network for Medical Image Segmentation
Ruxin Wang, Shuyuan Chen, Chaojie Ji +2
eess.IVcs.CVarXiv:2005.00966v12020Adapting Without Gradients: Affine Statistics Transport and What Its Certificate Can Tell You
Salim Khazem, Ibrahim Mohamed Serouis
cs.LGcs.AIcs.CVarXiv:2609.00374v12026InfraOcc: An Infrastructure Occupancy Benchmark with Static-to-Dynamic Reasoning
Lei Yang, Xiaokai Bai, Boqi Li +8
cs.CVarXiv:2608.30657v12026DSR -- A dual subspace re-projection network for surface anomaly detection
Vitjan Zavrtanik, Matej Kristan, Danijel Skočaj
cs.CVarXiv:2208.01521v22022Chained Multi-stream Networks Exploiting Pose, Motion, and Appearance for Action Classification and Detection
Mohammadreza Zolfaghari, Gabriel L. Oliveira, Nima Sedaghat +1
cs.CVcs.AIcs.HCarXiv:1704.00616v22017What Do Compressed Deep Neural Networks Forget?
Sara Hooker, Aaron Courville, Gregory Clark +2
cs.LGcs.AIcs.CVarXiv:1911.05248v32019Saliency Detection for Stereoscopic Images Based on Depth Confidence Analysis and Multiple Cues Fusion
Runmin Cong, Jianjun Lei, Changqing Zhang +3
cs.CVarXiv:1710.05174v12017Dynamic Hub-and-Spoke Memory for Streaming Video Understanding
Xinru Jiang, Lin Zhao, Xi Xiao +7
cs.CVarXiv:2608.30294v12026Scene recognition with CNNs: objects, scales and dataset bias
Luis Herranz, Shuqiang Jiang, Xiangyang Li
cs.CVarXiv:1801.06867v12018Common pitfalls and recommendations for using machine learning to detect and prognosticate for COVID-19 using chest radiographs and CT scans
Michael Roberts, Derek Driggs, Matthew Thorpe +13
cs.LGcs.CVeess.IVarXiv:2008.06388v42020Scalable Sparse Subspace Clustering by Orthogonal Matching Pursuit
Chong You, Daniel P. Robinson, Rene Vidal
cs.CVcs.LGstat.MLarXiv:1507.01238v32015Efficient Deformable ConvNets: Rethinking Dynamic and Sparse Operator for Vision Applications
Yuwen Xiong, Zhiqi Li, Yuntao Chen +10
cs.CVarXiv:2401.06197v12024Between-class Learning for Image Classification
Yuji Tokozume, Yoshitaka Ushiku, Tatsuya Harada
cs.LGcs.CVstat.MLarXiv:1711.10284v22017SceneFormer: Indoor Scene Generation with Transformers
Xinpeng Wang, Chandan Yeshwanth, Matthias Nießner
cs.CVarXiv:2012.09793v22020Deep Ranking for Person Re-identification via Joint Representation Learning
Shi-Zhe Chen, Chun-Chao Guo, Jian-Huang Lai
cs.CVarXiv:1505.06821v22015Iterative Visual Reasoning Beyond Convolutions
Xinlei Chen, Li-Jia Li, Li Fei-Fei +1
cs.CVarXiv:1803.11189v12018RetinaTrack: Online Single Stage Joint Detection and Tracking
Zhichao Lu, Vivek Rathod, Ronny Votel +1
cs.CVcs.LGeess.IVarXiv:2003.13870v12020Swin3D: A Pretrained Transformer Backbone for 3D Indoor Scene Understanding
Yu-Qi Yang, Yu-Xiao Guo, Jian-Yu Xiong +5
cs.CVarXiv:2304.06906v32023Sparse 3D convolutional neural networks
Ben Graham
cs.CVarXiv:1505.02890v22015Detecting Deep-Fake Videos from Appearance and Behavior
Shruti Agarwal, Tarek El-Gaaly, Hany Farid +1
cs.CVcs.LGcs.MMarXiv:2004.14491v12020Cascaded Boundary Regression for Temporal Action Detection
Jiyang Gao, Zhenheng Yang, Ram Nevatia
cs.CVarXiv:1705.01180v12017SRNet: Improving Generalization in 3D Human Pose Estimation with a Split-and-Recombine Approach
Ailing Zeng, Xiao Sun, Fuyang Huang +3
cs.CVarXiv:2007.09389v12020CTformer: Convolution-free Token2Token Dilated Vision Transformer for Low-dose CT Denoising
Dayang Wang, Fenglei Fan, Zhan Wu +3
eess.IVcs.CVarXiv:2202.13517v12022Wavelet-based Fourier Information Interaction with Frequency Diffusion Adjustment for Underwater Image Restoration
Chen Zhao, Weiling Cai, Chenyu Dong +1
cs.CVarXiv:2311.16845v12023GridFormer: Residual Dense Transformer with Grid Structure for Image Restoration in Adverse Weather Conditions
Tao Wang, Kaihao Zhang, Ziqian Shao +6
cs.CVarXiv:2305.17863v22023End-to-End Object Detection with Adaptive Clustering Transformer
Minghang Zheng, Peng Gao, Renrui Zhang +4
cs.CVarXiv:2011.09315v22020Boosting Contrastive Self-Supervised Learning with False Negative Cancellation
Tri Huynh, Simon Kornblith, Matthew R. Walter +2
cs.CVcs.LGarXiv:2011.11765v22020Learning Normalized Inputs for Iterative Estimation in Medical Image Segmentation
Michal Drozdzal, Gabriel Chartrand, Eugene Vorontsov +6
cs.CVarXiv:1702.05174v12017Transparency by Design: Closing the Gap Between Performance and Interpretability in Visual Reasoning
David Mascharka, Philip Tran, Ryan Soklaski +1
cs.CVarXiv:1803.05268v22018Self-supervised Spatio-temporal Representation Learning for Videos by Predicting Motion and Appearance Statistics
Jiangliu Wang, Jianbo Jiao, Linchao Bao +3
cs.CVarXiv:1904.03597v12019Efficient Pipeline for Camera Trap Image Review
Sara Beery, Dan Morris, Siyu Yang
cs.CVarXiv:1907.06772v12019High-Resolution Virtual Try-On with Misalignment and Occlusion-Handled Conditions
Sangyun Lee, Gyojung Gu, Sunghyun Park +2
cs.CVcs.AIarXiv:2206.14180v22022Summaries:한국어Semantic Video CNNs through Representation Warping
Raghudeep Gadde, Varun Jampani, Peter V. Gehler
cs.CVarXiv:1708.03088v12017Sparse Upcycling: Training Mixture-of-Experts from Dense Checkpoints
Aran Komatsuzaki, Joan Puigcerver, James Lee-Thorp +6
cs.LGcs.CLcs.CVarXiv:2212.05055v22022Learning to Count Objects in Natural Images for Visual Question Answering
Yan Zhang, Jonathon Hare, Adam Prügel-Bennett
cs.CVcs.CLarXiv:1802.05766v12018Gradient Centralization: A New Optimization Technique for Deep Neural Networks
Hongwei Yong, Jianqiang Huang, Xiansheng Hua +1
cs.CVarXiv:2004.01461v22020Dimensionality Reduction on SPD Manifolds: The Emergence of Geometry-Aware Methods
Mehrtash Harandi, Mathieu Salzmann, Richard Hartley
cs.CVarXiv:1605.06182v12016Semantic Adversarial Examples
Hossein Hosseini, Radha Poovendran
cs.CVcs.AIcs.LGarXiv:1804.00499v12018Text-Adaptive Generative Adversarial Networks: Manipulating Images with Natural Language
Seonghyeon Nam, Yunji Kim, Seon Joo Kim
cs.CVarXiv:1810.11919v22018ForgeryNet: A Versatile Benchmark for Comprehensive Forgery Analysis
Yinan He, Bei Gan, Siyu Chen +6
cs.CVcs.LGarXiv:2103.05630v22021