Computer Vision and Pattern Recognition

Papers filed under cs.CV on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.

Search paper metadata (including unsummarized papers)

6,721 to 6,780 of 18,866

  1. DUNet: A deformable network for retinal vessel segmentation

    Qiangguo Jin, Zhaopeng Meng, Tuan D. Pham +3

    cs.CVarXiv:1811.01206v12018
  2. Weisfeiler and Leman Go Neural: Higher-order Graph Neural Networks

    Christopher Morris, Martin Ritzert, Matthias Fey +4

    cs.LGcs.AIcs.CVarXiv:1810.02244v52018
  3. VoxelMorph: A Learning Framework for Deformable Medical Image Registration

    Guha Balakrishnan, Amy Zhao, Mert R. Sabuncu +2

    cs.CVarXiv:1809.05231v32018
  4. Transfer between Modalities with MetaQueries

    Xichen Pan, Satya Narayan Shukla, Aashu Singh +9

    cs.CVarXiv:2504.06256v12025
  5. CT Super-resolution GAN Constrained by the Identical, Residual, and Cycle Learning Ensemble(GAN-CIRCLE)

    Chenyu You, Guang Li, Yi Zhang +9

    eess.IVcs.CVcs.LGarXiv:1808.04256v32018
  6. Video Depth Anything: Consistent Depth Estimation for Super-Long Videos

    Sili Chen, Hengkai Guo, Shengnan Zhu +4

    cs.CVcs.AIarXiv:2501.12375v32025
  7. Deep Learning for Single Image Super-Resolution: A Brief Review

    Wenming Yang, Xuechen Zhang, Yapeng Tian +2

    cs.CVarXiv:1808.03344v32018
  8. Object Detection with Deep Learning: A Review

    Zhong-Qiu Zhao, Peng Zheng, Shou-tao Xu +1

    cs.CVarXiv:1807.05511v22018
  9. Deep Face Recognition: A Survey

    Mei Wang, Weihong Deng

    cs.CVarXiv:1804.06655v92018
  10. VLocNet++: Deep Multitask Learning for Semantic Visual Localization and Odometry

    Noha Radwan, Abhinav Valada, Wolfram Burgard

    cs.ROcs.CVarXiv:1804.08366v62018
  11. Tri-MipRF: Tri-Mip Representation for Efficient Anti-Aliasing Neural Radiance Fields

    Wenbo Hu, Yuling Wang, Lin Ma +4

    cs.CVcs.AIcs.GRarXiv:2307.11335v12023
  12. Direct Sparse Visual-Inertial Odometry using Dynamic Marginalization

    Lukas von Stumberg, Vladyslav Usenko, Daniel Cremers

    cs.CVcs.ROarXiv:1804.05625v12018
  13. A Recurrent CNN for Automatic Detection and Classification of Coronary Artery Plaque and Stenosis in Coronary CT Angiography

    Majd Zreik, Robbert W. van Hamersvelt, Jelmer M. Wolterink +3

    cs.CVarXiv:1804.04360v42018
  14. Matrix-game 2.0: An open-source real-time and streaming interactive world model

    Xianglong He, Chunli Peng, Zexiang Liu +16

    cs.CVarXiv:2508.13009v42025
  15. Demystifying Parallel and Distributed Deep Learning: An In-Depth Concurrency Analysis

    Tal Ben-Nun, Torsten Hoefler

    cs.LGcs.CVcs.DCarXiv:1802.09941v22018
  16. A DIRT-T Approach to Unsupervised Domain Adaptation

    Rui Shu, Hung H. Bui, Hirokazu Narui +1

    stat.MLcs.CVcs.LGarXiv:1802.08735v22018
  17. VideoChat-R1: Enhancing Spatio-Temporal Perception via Reinforcement Fine-Tuning

    Xinhao Li, Ziang Yan, Desen Meng +7

    cs.CVarXiv:2504.06958v52025
  18. Towards End-to-End Lane Detection: an Instance Segmentation Approach

    Davy Neven, Bert De Brabandere, Stamatios Georgoulis +2

    cs.CVarXiv:1802.05591v12018
  19. SelfLift: Accelerating Few-Step Diffusion via Self-Recovering Resolution Transition

    Tingyan Wen, Chenqian Yan, Xurui Peng +4

    cs.CVarXiv:2609.02036v12026
  20. Dynamic Graph CNN for Learning on Point Clouds

    Yue Wang, Yongbin Sun, Ziwei Liu +3

    cs.CVarXiv:1801.07829v22018
  21. Optimal ANN-SNN Conversion for Fast and Accurate Inference in Deep Spiking Neural Networks

    Jianhao Ding, Zhaofei Yu, Yonghong Tian +1

    cs.NEcs.AIcs.CVarXiv:2105.11654v12021
  22. TextBoxes++: A Single-Shot Oriented Scene Text Detector

    Minghui Liao, Baoguang Shi, Xiang Bai

    cs.CVarXiv:1801.02765v32018
  23. Threat of Adversarial Attacks on Deep Learning in Computer Vision: A Survey

    Naveed Akhtar, Ajmal Mian

    cs.CVarXiv:1801.00553v32018
  24. MoDL: Model Based Deep Learning Architecture for Inverse Problems

    Hemant Kumar Aggarwal, Merry P. Mani, Mathews Jacob

    cs.CVarXiv:1712.02862v42017
  25. SqueezeSeg: Convolutional Neural Nets with Recurrent CRF for Real-Time Road-Object Segmentation from 3D LiDAR Point Cloud

    Bichen Wu, Alvin Wan, Xiangyu Yue +1

    cs.CVarXiv:1710.07368v12017
  26. DeepSeek-OCR: Contexts Optical Compression

    Haoran Wei, Yaofeng Sun, Yukun Li

    cs.CVarXiv:2510.18234v12025
  27. Interactive Medical Image Segmentation using Deep Learning with Image-specific Fine-tuning

    Guotai Wang, Wenqi Li, Maria A. Zuluaga +8

    cs.CVarXiv:1710.04043v12017
  28. End-to-end Driving via Conditional Imitation Learning

    Felipe Codevilla, Matthias Müller, Antonio López +2

    cs.ROcs.CVcs.LGarXiv:1710.02410v22017
  29. VBench-2.0: Advancing Video Generation Benchmark Suite for Intrinsic Faithfulness

    Dian Zheng, Ziqi Huang, Hongbo Liu +9

    cs.CVarXiv:2503.21755v22025
  30. Label Refinery: Improving ImageNet Classification through Label Progression

    Hessam Bagherinezhad, Maxwell Horton, Mohammad Rastegari +1

    cs.CVarXiv:1805.02641v12018
  31. Matterport3D: Learning from RGB-D Data in Indoor Environments

    Angel Chang, Angela Dai, Thomas Funkhouser +6

    cs.CVarXiv:1709.06158v12017
  32. EuroSAT: A Novel Dataset and Deep Learning Benchmark for Land Use and Land Cover Classification

    Patrick Helber, Benjamin Bischke, Andreas Dengel +1

    cs.CVcs.LGarXiv:1709.00029v22017
  33. Hyperspectral Image Restoration via Total Variation Regularized Low-rank Tensor Decomposition

    Yao Wang, Jiangjun Peng, Qian Zhao +3

    cs.CVarXiv:1707.02477v12017
  34. Cross-Model Distillation of a Human-Pose Foundation Model from Unannotated Infant Video for Markerless 3D Pose Estimation

    R. James Cotton, Divya Joshi, Colleen Peyton

    cs.CVarXiv:2609.01840v12026
  35. DocHop: Benchmarking Out-of-domain Multi-hop Reasoning in Information-Dense Documents

    Zhuoran Yu, Le Thien Phuc Nguyen, Jaden Park +5

    cs.AIcs.CVcs.LGarXiv:2609.02059v12026
  36. IoU-aware Single-stage Object Detector for Accurate Localization

    Shengkai Wu, Xiaoping Li, Xinggang Wang

    cs.CVarXiv:1912.05992v42019
  37. NormFace: L2 Hypersphere Embedding for Face Verification

    Feng Wang, Xiang Xiang, Jian Cheng +1

    cs.CVarXiv:1704.06369v42017
  38. PointNet++: Deep Hierarchical Feature Learning on Point Sets in a Metric Space

    Charles R. Qi, Li Yi, Hao Su +1

    cs.CVarXiv:1706.02413v12017
  39. Convolutional Neural Networks for Medical Image Analysis: Full Training or Fine Tuning?

    Nima Tajbakhsh, Jae Y. Shin, Suryakanth R. Gurudu +4

    cs.CVcs.LGarXiv:1706.00712v12017
  40. Efficient Processing of Deep Neural Networks: A Tutorial and Survey

    Vivienne Sze, Yu-Hsin Chen, Tien-Ju Yang +1

    cs.CVarXiv:1703.09039v22017
  41. Reweighted Infrared Patch-Tensor Model With Both Non-Local and Local Priors for Single-Frame Small Target Detection

    Yimian Dai, Yiquan Wu

    cs.CVarXiv:1703.09157v12017
  42. Bidirectional-Convolutional LSTM Based Spectral-Spatial Feature Learning for Hyperspectral Image Classification

    Qingshan Liu, Feng Zhou, Renlong Hang +1

    cs.CVarXiv:1703.07910v12017
  43. Arbitrary-Oriented Scene Text Detection via Rotation Proposals

    Jianqi Ma, Weiyuan Shao, Hao Ye +4

    cs.CVarXiv:1703.01086v32017
  44. Soft + Hardwired Attention: An LSTM Framework for Human Trajectory Prediction and Abnormal Event Detection

    Tharindu Fernando, Simon Denman, Sridha Sridharan +1

    cs.CVcs.NEarXiv:1702.05552v12017
  45. A deep learning model integrating FCNNs and CRFs for brain tumor segmentation

    Xiaomei Zhao, Yihong Wu, Guidong Song +3

    cs.CVarXiv:1702.04528v32017
  46. Discriminative Correlation Filter with Channel and Spatial Reliability

    Alan Lukežič, Tomáš Vojíř, Luka Čehovin +2

    cs.CVarXiv:1611.08461v32016
  47. Deeply supervised salient object detection with short connections

    Qibin Hou, Ming-Ming Cheng, Xiao-Wei Hu +3

    cs.CVarXiv:1611.04849v42016
  48. ORB-SLAM2: an Open-Source SLAM System for Monocular, Stereo and RGB-D Cameras

    Raul Mur-Artal, Juan D. Tardos

    cs.ROcs.CVarXiv:1610.06475v22016
  49. Visual-Inertial Monocular SLAM with Map Reuse

    Raul Mur-Artal, Juan D. Tardos

    cs.ROcs.CVarXiv:1610.05949v22016
  50. A Survey of Multi-View Representation Learning

    Yingming Li, Ming Yang, Zhongfei Zhang

    cs.LGcs.CVcs.IRarXiv:1610.01206v52016
  51. Deep Visual Foresight for Planning Robot Motion

    Chelsea Finn, Sergey Levine

    cs.LGcs.AIcs.CVarXiv:1610.00696v22016
  52. Clearing the Skies: A deep network architecture for single-image rain removal

    Xueyang Fu, Jiabin Huang, Xinghao Ding +2

    cs.CVarXiv:1609.02087v22016
  53. Towards Evaluating the Robustness of Neural Networks

    Nicholas Carlini, David Wagner

    cs.CRcs.CVarXiv:1608.04644v22016
  54. SIFT Meets CNN: A Decade Survey of Instance Retrieval

    Liang Zheng, Yi Yang, Qi Tian

    cs.CVarXiv:1608.01807v22016
  55. Adversarial examples in the physical world

    Alexey Kurakin, Ian Goodfellow, Samy Bengio

    cs.CVcs.CRcs.LGarXiv:1607.02533v42016
  56. DeepLab: Semantic Image Segmentation with Deep Convolutional Nets, Atrous Convolution, and Fully Connected CRFs

    Liang-Chieh Chen, George Papandreou, Iasonas Kokkinos +2

    cs.CVarXiv:1606.00915v22016
  57. Estimating Depth from Monocular Images as Classification Using Deep Fully Convolutional Residual Networks

    Yuanzhouhan Cao, Zifeng Wu, Chunhua Shen

    cs.CVarXiv:1605.02305v32016
  58. Plug-and-Play ADMM for Image Restoration: Fixed Point Convergence and Applications

    Stanley H. Chan, Xiran Wang, Omar A. Elgendy

    cs.CVarXiv:1605.01710v22016
  59. Going Deeper with Contextual CNN for Hyperspectral Image Classification

    Hyungtae Lee, Heesung Kwon

    cs.CVcs.LGarXiv:1604.03519v32016
  60. A survey of sparse representation: algorithms and applications

    Zheng Zhang, Yong Xu, Jian Yang +2

    cs.CVcs.LGarXiv:1602.07017v12016