Computer Vision and Pattern Recognition
Papers filed under cs.CV on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.
Search paper metadata (including unsummarized papers)
6,721 to 6,780 of 18,866
DUNet: A deformable network for retinal vessel segmentation
Qiangguo Jin, Zhaopeng Meng, Tuan D. Pham +3
cs.CVarXiv:1811.01206v12018Weisfeiler and Leman Go Neural: Higher-order Graph Neural Networks
Christopher Morris, Martin Ritzert, Matthias Fey +4
cs.LGcs.AIcs.CVarXiv:1810.02244v52018VoxelMorph: A Learning Framework for Deformable Medical Image Registration
Guha Balakrishnan, Amy Zhao, Mert R. Sabuncu +2
cs.CVarXiv:1809.05231v32018Transfer between Modalities with MetaQueries
Xichen Pan, Satya Narayan Shukla, Aashu Singh +9
cs.CVarXiv:2504.06256v12025CT Super-resolution GAN Constrained by the Identical, Residual, and Cycle Learning Ensemble(GAN-CIRCLE)
Chenyu You, Guang Li, Yi Zhang +9
eess.IVcs.CVcs.LGarXiv:1808.04256v32018Video Depth Anything: Consistent Depth Estimation for Super-Long Videos
Sili Chen, Hengkai Guo, Shengnan Zhu +4
cs.CVcs.AIarXiv:2501.12375v32025Deep Learning for Single Image Super-Resolution: A Brief Review
Wenming Yang, Xuechen Zhang, Yapeng Tian +2
cs.CVarXiv:1808.03344v32018Object Detection with Deep Learning: A Review
Zhong-Qiu Zhao, Peng Zheng, Shou-tao Xu +1
cs.CVarXiv:1807.05511v22018Deep Face Recognition: A Survey
Mei Wang, Weihong Deng
cs.CVarXiv:1804.06655v92018VLocNet++: Deep Multitask Learning for Semantic Visual Localization and Odometry
Noha Radwan, Abhinav Valada, Wolfram Burgard
cs.ROcs.CVarXiv:1804.08366v62018Tri-MipRF: Tri-Mip Representation for Efficient Anti-Aliasing Neural Radiance Fields
Wenbo Hu, Yuling Wang, Lin Ma +4
cs.CVcs.AIcs.GRarXiv:2307.11335v12023Direct Sparse Visual-Inertial Odometry using Dynamic Marginalization
Lukas von Stumberg, Vladyslav Usenko, Daniel Cremers
cs.CVcs.ROarXiv:1804.05625v12018A Recurrent CNN for Automatic Detection and Classification of Coronary Artery Plaque and Stenosis in Coronary CT Angiography
Majd Zreik, Robbert W. van Hamersvelt, Jelmer M. Wolterink +3
cs.CVarXiv:1804.04360v42018Matrix-game 2.0: An open-source real-time and streaming interactive world model
Xianglong He, Chunli Peng, Zexiang Liu +16
cs.CVarXiv:2508.13009v42025Demystifying Parallel and Distributed Deep Learning: An In-Depth Concurrency Analysis
Tal Ben-Nun, Torsten Hoefler
cs.LGcs.CVcs.DCarXiv:1802.09941v22018A DIRT-T Approach to Unsupervised Domain Adaptation
Rui Shu, Hung H. Bui, Hirokazu Narui +1
stat.MLcs.CVcs.LGarXiv:1802.08735v22018VideoChat-R1: Enhancing Spatio-Temporal Perception via Reinforcement Fine-Tuning
Xinhao Li, Ziang Yan, Desen Meng +7
cs.CVarXiv:2504.06958v52025Towards End-to-End Lane Detection: an Instance Segmentation Approach
Davy Neven, Bert De Brabandere, Stamatios Georgoulis +2
cs.CVarXiv:1802.05591v12018SelfLift: Accelerating Few-Step Diffusion via Self-Recovering Resolution Transition
Tingyan Wen, Chenqian Yan, Xurui Peng +4
cs.CVarXiv:2609.02036v12026Dynamic Graph CNN for Learning on Point Clouds
Yue Wang, Yongbin Sun, Ziwei Liu +3
cs.CVarXiv:1801.07829v22018Optimal ANN-SNN Conversion for Fast and Accurate Inference in Deep Spiking Neural Networks
Jianhao Ding, Zhaofei Yu, Yonghong Tian +1
cs.NEcs.AIcs.CVarXiv:2105.11654v12021TextBoxes++: A Single-Shot Oriented Scene Text Detector
Minghui Liao, Baoguang Shi, Xiang Bai
cs.CVarXiv:1801.02765v32018Threat of Adversarial Attacks on Deep Learning in Computer Vision: A Survey
Naveed Akhtar, Ajmal Mian
cs.CVarXiv:1801.00553v32018MoDL: Model Based Deep Learning Architecture for Inverse Problems
Hemant Kumar Aggarwal, Merry P. Mani, Mathews Jacob
cs.CVarXiv:1712.02862v42017SqueezeSeg: Convolutional Neural Nets with Recurrent CRF for Real-Time Road-Object Segmentation from 3D LiDAR Point Cloud
Bichen Wu, Alvin Wan, Xiangyu Yue +1
cs.CVarXiv:1710.07368v12017DeepSeek-OCR: Contexts Optical Compression
Haoran Wei, Yaofeng Sun, Yukun Li
cs.CVarXiv:2510.18234v12025Interactive Medical Image Segmentation using Deep Learning with Image-specific Fine-tuning
Guotai Wang, Wenqi Li, Maria A. Zuluaga +8
cs.CVarXiv:1710.04043v12017End-to-end Driving via Conditional Imitation Learning
Felipe Codevilla, Matthias Müller, Antonio López +2
cs.ROcs.CVcs.LGarXiv:1710.02410v22017VBench-2.0: Advancing Video Generation Benchmark Suite for Intrinsic Faithfulness
Dian Zheng, Ziqi Huang, Hongbo Liu +9
cs.CVarXiv:2503.21755v22025Label Refinery: Improving ImageNet Classification through Label Progression
Hessam Bagherinezhad, Maxwell Horton, Mohammad Rastegari +1
cs.CVarXiv:1805.02641v12018Matterport3D: Learning from RGB-D Data in Indoor Environments
Angel Chang, Angela Dai, Thomas Funkhouser +6
cs.CVarXiv:1709.06158v12017EuroSAT: A Novel Dataset and Deep Learning Benchmark for Land Use and Land Cover Classification
Patrick Helber, Benjamin Bischke, Andreas Dengel +1
cs.CVcs.LGarXiv:1709.00029v22017Hyperspectral Image Restoration via Total Variation Regularized Low-rank Tensor Decomposition
Yao Wang, Jiangjun Peng, Qian Zhao +3
cs.CVarXiv:1707.02477v12017Cross-Model Distillation of a Human-Pose Foundation Model from Unannotated Infant Video for Markerless 3D Pose Estimation
R. James Cotton, Divya Joshi, Colleen Peyton
cs.CVarXiv:2609.01840v12026DocHop: Benchmarking Out-of-domain Multi-hop Reasoning in Information-Dense Documents
Zhuoran Yu, Le Thien Phuc Nguyen, Jaden Park +5
cs.AIcs.CVcs.LGarXiv:2609.02059v12026IoU-aware Single-stage Object Detector for Accurate Localization
Shengkai Wu, Xiaoping Li, Xinggang Wang
cs.CVarXiv:1912.05992v42019NormFace: L2 Hypersphere Embedding for Face Verification
Feng Wang, Xiang Xiang, Jian Cheng +1
cs.CVarXiv:1704.06369v42017PointNet++: Deep Hierarchical Feature Learning on Point Sets in a Metric Space
Charles R. Qi, Li Yi, Hao Su +1
cs.CVarXiv:1706.02413v12017Convolutional Neural Networks for Medical Image Analysis: Full Training or Fine Tuning?
Nima Tajbakhsh, Jae Y. Shin, Suryakanth R. Gurudu +4
cs.CVcs.LGarXiv:1706.00712v12017Efficient Processing of Deep Neural Networks: A Tutorial and Survey
Vivienne Sze, Yu-Hsin Chen, Tien-Ju Yang +1
cs.CVarXiv:1703.09039v22017Reweighted Infrared Patch-Tensor Model With Both Non-Local and Local Priors for Single-Frame Small Target Detection
Yimian Dai, Yiquan Wu
cs.CVarXiv:1703.09157v12017Bidirectional-Convolutional LSTM Based Spectral-Spatial Feature Learning for Hyperspectral Image Classification
Qingshan Liu, Feng Zhou, Renlong Hang +1
cs.CVarXiv:1703.07910v12017Arbitrary-Oriented Scene Text Detection via Rotation Proposals
Jianqi Ma, Weiyuan Shao, Hao Ye +4
cs.CVarXiv:1703.01086v32017Soft + Hardwired Attention: An LSTM Framework for Human Trajectory Prediction and Abnormal Event Detection
Tharindu Fernando, Simon Denman, Sridha Sridharan +1
cs.CVcs.NEarXiv:1702.05552v12017A deep learning model integrating FCNNs and CRFs for brain tumor segmentation
Xiaomei Zhao, Yihong Wu, Guidong Song +3
cs.CVarXiv:1702.04528v32017Discriminative Correlation Filter with Channel and Spatial Reliability
Alan Lukežič, Tomáš Vojíř, Luka Čehovin +2
cs.CVarXiv:1611.08461v32016Deeply supervised salient object detection with short connections
Qibin Hou, Ming-Ming Cheng, Xiao-Wei Hu +3
cs.CVarXiv:1611.04849v42016ORB-SLAM2: an Open-Source SLAM System for Monocular, Stereo and RGB-D Cameras
Raul Mur-Artal, Juan D. Tardos
cs.ROcs.CVarXiv:1610.06475v22016Visual-Inertial Monocular SLAM with Map Reuse
Raul Mur-Artal, Juan D. Tardos
cs.ROcs.CVarXiv:1610.05949v22016A Survey of Multi-View Representation Learning
Yingming Li, Ming Yang, Zhongfei Zhang
cs.LGcs.CVcs.IRarXiv:1610.01206v52016Deep Visual Foresight for Planning Robot Motion
Chelsea Finn, Sergey Levine
cs.LGcs.AIcs.CVarXiv:1610.00696v22016Clearing the Skies: A deep network architecture for single-image rain removal
Xueyang Fu, Jiabin Huang, Xinghao Ding +2
cs.CVarXiv:1609.02087v22016Towards Evaluating the Robustness of Neural Networks
Nicholas Carlini, David Wagner
cs.CRcs.CVarXiv:1608.04644v22016SIFT Meets CNN: A Decade Survey of Instance Retrieval
Liang Zheng, Yi Yang, Qi Tian
cs.CVarXiv:1608.01807v22016Adversarial examples in the physical world
Alexey Kurakin, Ian Goodfellow, Samy Bengio
cs.CVcs.CRcs.LGarXiv:1607.02533v42016DeepLab: Semantic Image Segmentation with Deep Convolutional Nets, Atrous Convolution, and Fully Connected CRFs
Liang-Chieh Chen, George Papandreou, Iasonas Kokkinos +2
cs.CVarXiv:1606.00915v22016Estimating Depth from Monocular Images as Classification Using Deep Fully Convolutional Residual Networks
Yuanzhouhan Cao, Zifeng Wu, Chunhua Shen
cs.CVarXiv:1605.02305v32016Plug-and-Play ADMM for Image Restoration: Fixed Point Convergence and Applications
Stanley H. Chan, Xiran Wang, Omar A. Elgendy
cs.CVarXiv:1605.01710v22016Going Deeper with Contextual CNN for Hyperspectral Image Classification
Hyungtae Lee, Heesung Kwon
cs.CVcs.LGarXiv:1604.03519v32016A survey of sparse representation: algorithms and applications
Zheng Zhang, Yong Xu, Jian Yang +2
cs.CVcs.LGarXiv:1602.07017v12016