Computer Vision and Pattern Recognition

Papers filed under cs.CV on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.

Search paper metadata (including unsummarized papers)

8,101 to 8,160 of 18,849

  1. Contextual Encoder-Decoder Network for Visual Saliency Prediction

    Alexander Kroner, Mario Senden, Kurt Driessens +1

    cs.CVarXiv:1902.06634v42019
  2. Learning in an Uncertain World: Representing Ambiguity Through Multiple Hypotheses

    Christian Rupprecht, Iro Laina, Robert DiPietro +4

    cs.CVarXiv:1612.00197v32016
  3. A-Lamp: Adaptive Layout-Aware Multi-Patch Deep Convolutional Neural Network for Photo Aesthetic Assessment

    Shuang Ma, Jing Liu, Chang Wen Chen

    cs.CVarXiv:1704.00248v12017
  4. From Synthetic to Real: Image Dehazing Collaborating with Unlabeled Real Data

    Ye Liu, Lei Zhu, Shunda Pei +5

    cs.CVarXiv:2108.02934v12021
  5. FlowVVTON: Flow-Guided Mask-Free Video Virtual Try-On

    Shengyao Chen, Xianbing Sun, Liqing Zhang +1

    cs.CVarXiv:2608.30450v12026
  6. PRISM: Predictive Recomposition via Semantic Latent Decomposition for View-invariant Video Representation Learning

    Youngchae Chee, Hosu Lee, Sungjune Park +2

    cs.CVcs.AIarXiv:2608.30388v12026
  7. Weakly Supervised Video Moment Retrieval From Text Queries

    Niluthpol Chowdhury Mithun, Sujoy Paul, Amit K. Roy-Chowdhury

    cs.CVcs.MMarXiv:1904.03282v22019
  8. You Only Need Adversarial Supervision for Semantic Image Synthesis

    Vadim Sushko, Edgar Schönfeld, Dan Zhang +3

    cs.CVcs.LGeess.IVarXiv:2012.04781v32020
  9. Differentiable Rendering: A Survey

    Hiroharu Kato, Deniz Beker, Mihai Morariu +4

    cs.CVcs.GRarXiv:2006.12057v22020
  10. A Taxonomy of Deep Convolutional Neural Nets for Computer Vision

    Suraj Srinivas, Ravi Kiran Sarvadevabhatla, Konda Reddy Mopuri +3

    cs.CVcs.LGcs.MMarXiv:1601.06615v12016
  11. REVISE: A Tool for Measuring and Mitigating Bias in Visual Datasets

    Angelina Wang, Alexander Liu, Ryan Zhang +6

    cs.CVarXiv:2004.07999v42020
  12. The Devil is in the Tails: Fine-grained Classification in the Wild

    Grant Van Horn, Pietro Perona

    cs.CVarXiv:1709.01450v12017
  13. Doc-REFRAG: Rethinking Multimodal Document Retrieval-Augmented Generation

    Ruofan Hu, Shengyang Xu, Minjie Hong +5

    cs.IRcs.CVarXiv:2608.30163v12026
  14. DeRF: Decomposed Radiance Fields

    Daniel Rebain, Wei Jiang, Soroosh Yazdani +3

    cs.CVcs.GRarXiv:2011.12490v12020
  15. Overcoming Catastrophic Forgetting in Incremental Few-Shot Learning by Finding Flat Minima

    Guangyuan Shi, Jiaxin Chen, Wenlong Zhang +2

    cs.LGcs.CVarXiv:2111.01549v22021
  16. DALES: A Large-scale Aerial LiDAR Data Set for Semantic Segmentation

    Nina Varney, Vijayan K. Asari, Quinn Graehling

    cs.CVcs.LGstat.MLarXiv:2004.11985v12020
  17. Direction-aware Spatial Context Features for Shadow Detection

    Xiaowei Hu, Lei Zhu, Chi-Wing Fu +2

    cs.CVarXiv:1712.04142v22017
  18. Neural Rendering for Stereo 3D Reconstruction of Deformable Tissues in Robotic Surgery

    Yuehao Wang, Yonghao Long, Siu Hin Fan +1

    cs.CVarXiv:2206.15255v12022
  19. Everybody Tracking Every Body

    Daeyun Shin, Yunhan Zhao, Shu Kong +2

    cs.CVarXiv:2608.29927v12026
  20. Underwater Optical Image Processing: A Comprehensive Review

    Huimin Lu, Yujie Li, Yudong Zhang +3

    cs.CVarXiv:1702.03600v12017
  21. Boundary-aware Context Neural Network for Medical Image Segmentation

    Ruxin Wang, Shuyuan Chen, Chaojie Ji +2

    eess.IVcs.CVarXiv:2005.00966v12020
  22. Adapting Without Gradients: Affine Statistics Transport and What Its Certificate Can Tell You

    Salim Khazem, Ibrahim Mohamed Serouis

    cs.LGcs.AIcs.CVarXiv:2609.00374v12026
  23. InfraOcc: An Infrastructure Occupancy Benchmark with Static-to-Dynamic Reasoning

    Lei Yang, Xiaokai Bai, Boqi Li +8

    cs.CVarXiv:2608.30657v12026
  24. DSR -- A dual subspace re-projection network for surface anomaly detection

    Vitjan Zavrtanik, Matej Kristan, Danijel Skočaj

    cs.CVarXiv:2208.01521v22022
  25. Chained Multi-stream Networks Exploiting Pose, Motion, and Appearance for Action Classification and Detection

    Mohammadreza Zolfaghari, Gabriel L. Oliveira, Nima Sedaghat +1

    cs.CVcs.AIcs.HCarXiv:1704.00616v22017
  26. What Do Compressed Deep Neural Networks Forget?

    Sara Hooker, Aaron Courville, Gregory Clark +2

    cs.LGcs.AIcs.CVarXiv:1911.05248v32019
  27. Saliency Detection for Stereoscopic Images Based on Depth Confidence Analysis and Multiple Cues Fusion

    Runmin Cong, Jianjun Lei, Changqing Zhang +3

    cs.CVarXiv:1710.05174v12017
  28. Dynamic Hub-and-Spoke Memory for Streaming Video Understanding

    Xinru Jiang, Lin Zhao, Xi Xiao +7

    cs.CVarXiv:2608.30294v12026
  29. Scene recognition with CNNs: objects, scales and dataset bias

    Luis Herranz, Shuqiang Jiang, Xiangyang Li

    cs.CVarXiv:1801.06867v12018
  30. Common pitfalls and recommendations for using machine learning to detect and prognosticate for COVID-19 using chest radiographs and CT scans

    Michael Roberts, Derek Driggs, Matthew Thorpe +13

    cs.LGcs.CVeess.IVarXiv:2008.06388v42020
  31. Scalable Sparse Subspace Clustering by Orthogonal Matching Pursuit

    Chong You, Daniel P. Robinson, Rene Vidal

    cs.CVcs.LGstat.MLarXiv:1507.01238v32015
  32. Efficient Deformable ConvNets: Rethinking Dynamic and Sparse Operator for Vision Applications

    Yuwen Xiong, Zhiqi Li, Yuntao Chen +10

    cs.CVarXiv:2401.06197v12024
  33. Between-class Learning for Image Classification

    Yuji Tokozume, Yoshitaka Ushiku, Tatsuya Harada

    cs.LGcs.CVstat.MLarXiv:1711.10284v22017
  34. SceneFormer: Indoor Scene Generation with Transformers

    Xinpeng Wang, Chandan Yeshwanth, Matthias Nießner

    cs.CVarXiv:2012.09793v22020
  35. Deep Ranking for Person Re-identification via Joint Representation Learning

    Shi-Zhe Chen, Chun-Chao Guo, Jian-Huang Lai

    cs.CVarXiv:1505.06821v22015
  36. Iterative Visual Reasoning Beyond Convolutions

    Xinlei Chen, Li-Jia Li, Li Fei-Fei +1

    cs.CVarXiv:1803.11189v12018
  37. RetinaTrack: Online Single Stage Joint Detection and Tracking

    Zhichao Lu, Vivek Rathod, Ronny Votel +1

    cs.CVcs.LGeess.IVarXiv:2003.13870v12020
  38. Swin3D: A Pretrained Transformer Backbone for 3D Indoor Scene Understanding

    Yu-Qi Yang, Yu-Xiao Guo, Jian-Yu Xiong +5

    cs.CVarXiv:2304.06906v32023
  39. Sparse 3D convolutional neural networks

    Ben Graham

    cs.CVarXiv:1505.02890v22015
  40. Detecting Deep-Fake Videos from Appearance and Behavior

    Shruti Agarwal, Tarek El-Gaaly, Hany Farid +1

    cs.CVcs.LGcs.MMarXiv:2004.14491v12020
  41. Cascaded Boundary Regression for Temporal Action Detection

    Jiyang Gao, Zhenheng Yang, Ram Nevatia

    cs.CVarXiv:1705.01180v12017
  42. SRNet: Improving Generalization in 3D Human Pose Estimation with a Split-and-Recombine Approach

    Ailing Zeng, Xiao Sun, Fuyang Huang +3

    cs.CVarXiv:2007.09389v12020
  43. CTformer: Convolution-free Token2Token Dilated Vision Transformer for Low-dose CT Denoising

    Dayang Wang, Fenglei Fan, Zhan Wu +3

    eess.IVcs.CVarXiv:2202.13517v12022
  44. Wavelet-based Fourier Information Interaction with Frequency Diffusion Adjustment for Underwater Image Restoration

    Chen Zhao, Weiling Cai, Chenyu Dong +1

    cs.CVarXiv:2311.16845v12023
  45. GridFormer: Residual Dense Transformer with Grid Structure for Image Restoration in Adverse Weather Conditions

    Tao Wang, Kaihao Zhang, Ziqian Shao +6

    cs.CVarXiv:2305.17863v22023
  46. End-to-End Object Detection with Adaptive Clustering Transformer

    Minghang Zheng, Peng Gao, Renrui Zhang +4

    cs.CVarXiv:2011.09315v22020
  47. Boosting Contrastive Self-Supervised Learning with False Negative Cancellation

    Tri Huynh, Simon Kornblith, Matthew R. Walter +2

    cs.CVcs.LGarXiv:2011.11765v22020
  48. Learning Normalized Inputs for Iterative Estimation in Medical Image Segmentation

    Michal Drozdzal, Gabriel Chartrand, Eugene Vorontsov +6

    cs.CVarXiv:1702.05174v12017
  49. Transparency by Design: Closing the Gap Between Performance and Interpretability in Visual Reasoning

    David Mascharka, Philip Tran, Ryan Soklaski +1

    cs.CVarXiv:1803.05268v22018
  50. Self-supervised Spatio-temporal Representation Learning for Videos by Predicting Motion and Appearance Statistics

    Jiangliu Wang, Jianbo Jiao, Linchao Bao +3

    cs.CVarXiv:1904.03597v12019
  51. Efficient Pipeline for Camera Trap Image Review

    Sara Beery, Dan Morris, Siyu Yang

    cs.CVarXiv:1907.06772v12019
  52. High-Resolution Virtual Try-On with Misalignment and Occlusion-Handled Conditions

    Sangyun Lee, Gyojung Gu, Sunghyun Park +2

    cs.CVcs.AIarXiv:2206.14180v22022
    Summaries:한국어
  53. Semantic Video CNNs through Representation Warping

    Raghudeep Gadde, Varun Jampani, Peter V. Gehler

    cs.CVarXiv:1708.03088v12017
  54. Sparse Upcycling: Training Mixture-of-Experts from Dense Checkpoints

    Aran Komatsuzaki, Joan Puigcerver, James Lee-Thorp +6

    cs.LGcs.CLcs.CVarXiv:2212.05055v22022
  55. Learning to Count Objects in Natural Images for Visual Question Answering

    Yan Zhang, Jonathon Hare, Adam Prügel-Bennett

    cs.CVcs.CLarXiv:1802.05766v12018
  56. Gradient Centralization: A New Optimization Technique for Deep Neural Networks

    Hongwei Yong, Jianqiang Huang, Xiansheng Hua +1

    cs.CVarXiv:2004.01461v22020
  57. Dimensionality Reduction on SPD Manifolds: The Emergence of Geometry-Aware Methods

    Mehrtash Harandi, Mathieu Salzmann, Richard Hartley

    cs.CVarXiv:1605.06182v12016
  58. Semantic Adversarial Examples

    Hossein Hosseini, Radha Poovendran

    cs.CVcs.AIcs.LGarXiv:1804.00499v12018
  59. Text-Adaptive Generative Adversarial Networks: Manipulating Images with Natural Language

    Seonghyeon Nam, Yunji Kim, Seon Joo Kim

    cs.CVarXiv:1810.11919v22018
  60. ForgeryNet: A Versatile Benchmark for Comprehensive Forgery Analysis

    Yinan He, Bei Gan, Siyu Chen +6

    cs.CVcs.LGarXiv:2103.05630v22021