Computer Vision and Pattern Recognition

Papers filed under cs.CV on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.

Search paper metadata (including unsummarized papers)

13,381 to 13,440 of 18,849

  1. Differential Recurrent Neural Networks for Action Recognition

    Vivek Veeriah, Naifan Zhuang, Guo-Jun Qi

    cs.CVarXiv:1504.06678v12015
  2. The TUM VI Benchmark for Evaluating Visual-Inertial Odometry

    David Schubert, Thore Goll, Nikolaus Demmel +3

    cs.CVcs.ROarXiv:1804.06120v32018
  3. Lesion Border Detection in Dermoscopy Images

    M. Emre Celebi, Hitoshi Iyatomi, Gerald Schaefer +1

    cs.CVarXiv:1011.0640v12010
  4. MedViT: A Robust Vision Transformer for Generalized Medical Image Classification

    Omid Nejati Manzari, Hamid Ahmadabadi, Hossein Kashiani +2

    cs.CVarXiv:2302.09462v12023
  5. DrivingGaussian: Composite Gaussian Splatting for Surrounding Dynamic Autonomous Driving Scenes

    Xiaoyu Zhou, Zhiwei Lin, Xiaojun Shan +3

    cs.CVarXiv:2312.07920v32023
  6. Three-Phase Scribble-Adaptive Curriculum Learning for autoPETV Grand Challenge

    Libo Zhang

    cs.CVarXiv:2608.22096v12026
  7. Billion-scale semi-supervised learning for image classification

    I. Zeki Yalniz, Hervé Jégou, Kan Chen +2

    cs.CVarXiv:1905.00546v12019
  8. Focal Frequency Loss for Image Reconstruction and Synthesis

    Liming Jiang, Bo Dai, Wayne Wu +1

    cs.CVcs.LGeess.IVarXiv:2012.12821v32020
  9. Dataset Condensation with Distribution Matching

    Bo Zhao, Hakan Bilen

    cs.LGcs.CVarXiv:2110.04181v32021
  10. FastFlow: Unsupervised Anomaly Detection and Localization via 2D Normalizing Flows

    Jiawei Yu, Ye Zheng, Xiang Wang +4

    cs.CVarXiv:2111.07677v22021
  11. Focusing Attention: Towards Accurate Text Recognition in Natural Images

    Zhanzhan Cheng, Fan Bai, Yunlu Xu +3

    cs.CVarXiv:1709.02054v32017
  12. Inversion-Based Style Transfer with Diffusion Models

    Yuxin Zhang, Nisha Huang, Fan Tang +4

    cs.CVcs.GRarXiv:2211.13203v32022
  13. Analog Bits: Generating Discrete Data using Diffusion Models with Self-Conditioning

    Ting Chen, Ruixiang Zhang, Geoffrey Hinton

    cs.CVcs.AIcs.CLarXiv:2208.04202v22022
  14. ELITE: Encoding Visual Concepts into Textual Embeddings for Customized Text-to-Image Generation

    Yuxiang Wei, Yabo Zhang, Zhilong Ji +3

    cs.CVarXiv:2302.13848v22023
  15. Poly Kernel Inception Network for Remote Sensing Detection

    Xinhao Cai, Qiuxia Lai, Yuwei Wang +3

    cs.CVarXiv:2403.06258v22024
  16. MambaVision: A Hybrid Mamba-Transformer Vision Backbone

    Ali Hatamizadeh, Jan Kautz

    cs.CVarXiv:2407.08083v22024
  17. Uni-ControlNet: All-in-One Control to Text-to-Image Diffusion Models

    Shihao Zhao, Dongdong Chen, Yen-Chun Chen +4

    cs.CVcs.GRarXiv:2305.16322v32023
  18. One Thousand and One Hours: Self-driving Motion Prediction Dataset

    John Houston, Guido Zuidhof, Luca Bergamini +6

    cs.CVcs.LGcs.ROarXiv:2006.14480v22020
  19. Quasi-Dense Similarity Learning for Multiple Object Tracking

    Jiangmiao Pang, Linlu Qiu, Xia Li +4

    cs.CVcs.LGarXiv:2006.06664v42020
  20. Open-Vocabulary Panoptic Segmentation with Text-to-Image Diffusion Models

    Jiarui Xu, Sifei Liu, Arash Vahdat +3

    cs.CVarXiv:2303.04803v42023
  21. View Adaptive Neural Networks for High Performance Skeleton-based Human Action Recognition

    Pengfei Zhang, Cuiling Lan, Junliang Xing +3

    cs.CVarXiv:1804.07453v32018
  22. EPNet: Enhancing Point Features with Image Semantics for 3D Object Detection

    Tengteng Huang, Zhe Liu, Xiwu Chen +1

    cs.CVarXiv:2007.08856v12020
  23. Generic Attention-model Explainability for Interpreting Bi-Modal and Encoder-Decoder Transformers

    Hila Chefer, Shir Gur, Lior Wolf

    cs.CVcs.LGarXiv:2103.15679v12021
  24. Learning to Navigate for Fine-grained Classification

    Ze Yang, Tiange Luo, Dong Wang +3

    cs.CVarXiv:1809.00287v12018
  25. Last Layer Re-Training is Sufficient for Robustness to Spurious Correlations

    Polina Kirichenko, Pavel Izmailov, Andrew Gordon Wilson

    cs.LGcs.CVstat.MLarXiv:2204.02937v22022
  26. SA-UNet: Spatial Attention U-Net for Retinal Vessel Segmentation

    Changlu Guo, Márton Szemenyei, Yugen Yi +3

    eess.IVcs.CVarXiv:2004.03696v32020
  27. COVID-ResNet: A Deep Learning Framework for Screening of COVID19 from Radiographs

    Muhammad Farooq, Abdul Hafeez

    eess.IVcs.CVcs.LGarXiv:2003.14395v12020
  28. CiUNet: A Hybrid Swin-CNN UNet for Medical Image Segmentation

    Bin Dong, Jinghong Chen

    eess.IVcs.CVarXiv:2608.22281v12026
  29. Unsupervised Learning of Probabilistic Diffeomorphic Registration for Images and Surfaces

    Adrian V. Dalca, Guha Balakrishnan, John Guttag +1

    cs.CVcs.GRarXiv:1903.03545v22019
  30. Two-branch Recurrent Network for Isolating Deepfakes in Videos

    Iacopo Masi, Aditya Killekar, Royston Marian Mascarenhas +2

    cs.CVcs.CYcs.LGarXiv:2008.03412v32020
  31. Motion Transformer with Global Intention Localization and Local Movement Refinement

    Shaoshuai Shi, Li Jiang, Dengxin Dai +1

    cs.CVarXiv:2209.13508v22022
  32. Events-to-Video: Bringing Modern Computer Vision to Event Cameras

    Henri Rebecq, René Ranftl, Vladlen Koltun +1

    cs.CVarXiv:1904.08298v12019
  33. Rainbow Memory: Continual Learning with a Memory of Diverse Samples

    Jihwan Bang, Heesu Kim, YoungJoon Yoo +2

    cs.CVcs.LGarXiv:2103.17230v12021
  34. When Test-Time Adaptation Helps, Harms, or Becomes Inactive: A Condition-Level Study on CIFAR-10-C

    Sreeja Guha Majumdar, Aratrika Saha

    cs.LGcs.CVarXiv:2608.22233v12026
  35. FreeMatch: Self-adaptive Thresholding for Semi-supervised Learning

    Yidong Wang, Hao Chen, Qiang Heng +9

    cs.LGcs.CVarXiv:2205.07246v32022
  36. Balanced Multimodal Learning via On-the-fly Gradient Modulation

    Xiaokang Peng, Yake Wei, Andong Deng +2

    cs.CVcs.AIarXiv:2203.15332v12022
  37. Slicing Aided Hyper Inference and Fine-tuning for Small Object Detection

    Fatih Cagatay Akyon, Sinan Onur Altinuc, Alptekin Temizel

    cs.CVcs.LGarXiv:2202.06934v52022
  38. Shallow and Deep Convolutional Networks for Saliency Prediction

    Junting Pan, Kevin McGuinness, Elisa Sayrol +2

    cs.CVcs.LGarXiv:1603.00845v12016
  39. Convolutional neural networks with low-rank regularization

    Cheng Tai, Tong Xiao, Yi Zhang +2

    cs.LGcs.CVstat.MLarXiv:1511.06067v32015
  40. Deep Autoencoding Models for Unsupervised Anomaly Segmentation in Brain MR Images

    Christoph Baur, Benedikt Wiestler, Shadi Albarqouni +1

    cs.CVarXiv:1804.04488v12018
  41. Improved Image Captioning via Policy Gradient optimization of SPIDEr

    Siqi Liu, Zhenhai Zhu, Ning Ye +2

    cs.CVcs.CLarXiv:1612.00370v42016
  42. Do Adversarially Robust ImageNet Models Transfer Better?

    Hadi Salman, Andrew Ilyas, Logan Engstrom +2

    cs.CVcs.LGstat.MLarXiv:2007.08489v22020
  43. Multi-Agent Cooperation and the Emergence of (Natural) Language

    Angeliki Lazaridou, Alexander Peysakhovich, Marco Baroni

    cs.CLcs.CVcs.GTarXiv:1612.07182v22016
  44. Learning feed-forward one-shot learners

    Luca Bertinetto, João F. Henriques, Jack Valmadre +2

    cs.CVcs.LGarXiv:1606.05233v12016
  45. Radiomics strategies for risk assessment of tumour failure in head-and-neck cancer

    Martin Vallières, Emily Kay-Rivest, Léo Jean Perrin +9

    cs.CVarXiv:1703.08516v12017
  46. How Far are We from Solving Pedestrian Detection?

    Shanshan Zhang, Rodrigo Benenson, Mohamed Omran +2

    cs.CVarXiv:1602.01237v22016
  47. On the Integration of Self-Attention and Convolution

    Xuran Pan, Chunjiang Ge, Rui Lu +4

    cs.CVarXiv:2111.14556v22021
  48. Does a Modern-Handwriting Warm-Up Help Historical Arabic OCR? A Reproducible, Compute-Matched Evaluation on Muharaf and KHATT

    Sumaih Almarshad, Maram Alamri, Dona Aloraini +4

    cs.LGcs.CVarXiv:2608.22316v12026
  49. Deep Parametric Continuous Convolutional Neural Networks

    Shenlong Wang, Simon Suo, Wei-Chiu Ma +2

    cs.CVcs.AIcs.LGarXiv:2101.06742v12021
  50. No Fear of Heterogeneity: Classifier Calibration for Federated Learning with Non-IID Data

    Mi Luo, Fei Chen, Dapeng Hu +3

    cs.LGcs.CVcs.DCarXiv:2106.05001v22021
  51. GETNET: A General End-to-end Two-dimensional CNN Framework for Hyperspectral Image Change Detection

    Qi Wang, Zhenghang Yuan, Qian Du +1

    cs.CVeess.IVarXiv:1905.01662v12019
  52. SpaceNet: A Remote Sensing Dataset and Challenge Series

    Adam Van Etten, Dave Lindenbaum, Todd M. Bacastow

    cs.CVarXiv:1807.01232v32018
  53. Compact 3D Gaussian Representation for Radiance Field

    Joo Chan Lee, Daniel Rho, Xiangyu Sun +2

    cs.CVcs.GRarXiv:2311.13681v22023
  54. Global Context-Aware Progressive Aggregation Network for Salient Object Detection

    Zuyao Chen, Qianqian Xu, Runmin Cong +1

    cs.CVarXiv:2003.00651v12020
  55. Disparities in Dermatology AI Performance on a Diverse, Curated Clinical Image Set

    Roxana Daneshjou, Kailas Vodrahalli, Roberto A Novoa +16

    eess.IVcs.AIcs.CVarXiv:2203.08807v12022
  56. Transformation Consistent Self-ensembling Model for Semi-supervised Medical Image Segmentation

    Xiaomeng Li, Lequan Yu, Hao Chen +3

    cs.CVarXiv:1903.00348v32019
  57. VoxelNeXt: Fully Sparse VoxelNet for 3D Object Detection and Tracking

    Yukang Chen, Jianhui Liu, Xiangyu Zhang +2

    cs.CVarXiv:2303.11301v12023
  58. Efficient and Accurate Arbitrary-Shaped Text Detection with Pixel Aggregation Network

    Wenhai Wang, Enze Xie, Xiaoge Song +5

    cs.CVarXiv:1908.05900v22019
  59. Partial Adversarial Domain Adaptation

    Zhangjie Cao, Lijia Ma, Mingsheng Long +1

    cs.CVarXiv:1808.04205v12018
  60. Toward Characteristic-Preserving Image-based Virtual Try-On Network

    Bochao Wang, Huabin Zheng, Xiaodan Liang +3

    cs.CVarXiv:1807.07688v32018