Computer Vision and Pattern Recognition
Papers filed under cs.CV on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.
Search paper metadata (including unsummarized papers)
1,081 to 1,140 of 18,817
Learning Inverse Depth Regression for Multi-View Stereo with Correlation Cost Volume
Qingshan Xu, Wenbing Tao
cs.CVarXiv:1912.11746v12019LoFT: Parameter-Efficient Fine-Tuning for Long-tailed Semi-Supervised Learning in Open-World Scenarios
Zhiyuan Huang, Jiahao Chen, Bing Su
cs.LGcs.CVarXiv:2509.09926v52025Hybrid Noise Removal in Hyperspectral Imagery With a Spatial-Spectral Gradient Network
Qiang Zhang, Qiangqiang Yuan, Jie Li +3
cs.CVarXiv:1810.00495v32018High-Speed Tracking with Kernelized Correlation Filters
João F. Henriques, Rui Caseiro, Pedro Martins +1
cs.CVarXiv:1404.7584v32014Learning to Predict Visual Attributes in the Wild
Khoi Pham, Kushal Kafle, Zhe Lin +4
cs.CVarXiv:2106.09707v12021Rethinking Differentiable Search for Mixed-Precision Neural Networks
Zhaowei Cai, Nuno Vasconcelos
cs.LGcs.CVcs.NEarXiv:2004.05795v12020Dexterous Manipulation Policies from RGB Human Videos via 3D Hand-Object Trajectory Reconstruction
Hongyi Chen, Tony Dong, Tiancheng Wu +7
cs.ROcs.CVarXiv:2602.09013v22026Early- and in-season crop type mapping without current-year ground truth: generating labels from historical information via a topology-based approach
Chenxi Lin, Liheng Zhong, Xiao-Peng Song +3
cs.CVcs.LGarXiv:2110.10275v12021Modality-aware Mutual Learning for Multi-modal Medical Image Segmentation
Yao Zhang, Jiawei Yang, Jiang Tian +4
eess.IVcs.CVarXiv:2107.09842v12021Fixed Anchors Are Not Enough: Dynamic Retrieval and Persistent Homology for Dataset Distillation
Muquan Li, Hang Gou, Yingyi Ma +3
cs.CVarXiv:2602.24144v32026Event Enhanced High-Quality Image Recovery
Bishan Wang, Jingwei He, Lei Yu +2
cs.CVarXiv:2007.08336v12020Paint3D: Paint Anything 3D with Lighting-Less Texture Diffusion Models
Xianfang Zeng, Xin Chen, Zhongqi Qi +6
cs.CVarXiv:2312.13913v22023Faster ILOD: Incremental Learning for Object Detectors based on Faster RCNN
Can Peng, Kun Zhao, Brian C. Lovell
cs.CVarXiv:2003.03901v22020Dexterous Imitation Made Easy: A Learning-Based Framework for Efficient Dexterous Manipulation
Sridhar Pandian Arunachalam, Sneha Silwal, Ben Evans +1
cs.ROcs.AIcs.CVarXiv:2203.13251v12022CoRA-NAS: Coarse Ranking and Anchor-Residual Refinement for Neural Architecture Search
Yifan Yang, Zhaoyan Wang, Zheng Gao +2
cs.LGcs.CVarXiv:2609.11884v12026Auto-Encoder Guided GAN for Chinese Calligraphy Synthesis
Pengyuan Lyu, Xiang Bai, Cong Yao +3
cs.CVarXiv:1706.08789v12017Fully-automated Body Composition Analysis in Routine CT Imaging Using 3D Semantic Segmentation Convolutional Neural Networks
Sven Koitka, Lennard Kroll, Eugen Malamutmann +2
eess.IVcs.CVarXiv:2002.10776v12020Deep Learning for Micro-expression Recognition: A Survey
Yante Li, Jinsheng Wei, Yang Liu +2
cs.CVcs.HCarXiv:2107.02823v52021AugMax: Adversarial Composition of Random Augmentations for Robust Training
Haotao Wang, Chaowei Xiao, Jean Kossaifi +3
cs.CVcs.LGarXiv:2110.13771v32021Optically lightweight tracking of objects around a corner
Jonathan Klein, Christoph Peters, Jaime Martín +2
cs.CVcs.GRphysics.opticsarXiv:1606.01873v12016Segmentation-Aware Convolutional Networks Using Local Attention Masks
Adam W. Harley, Konstantinos G. Derpanis, Iasonas Kokkinos
cs.CVarXiv:1708.04607v12017Detecting Photoshopped Faces by Scripting Photoshop
Sheng-Yu Wang, Oliver Wang, Andrew Owens +2
cs.CVarXiv:1906.05856v22019Segmentation of Roots in Soil with U-Net
Abraham George Smith, Jens Petersen, Raghavendra Selvan +1
cs.CVarXiv:1902.11050v22019Customizable Architecture Search for Semantic Segmentation
Yiheng Zhang, Zhaofan Qiu, Jingen Liu +3
cs.CVarXiv:1908.09550v12019LaneAF: Robust Multi-Lane Detection with Affinity Fields
Hala Abualsaud, Sean Liu, David Lu +3
cs.CVcs.ROarXiv:2103.12040v42021Predicting with Confidence on Unseen Distributions
Devin Guillory, Vaishaal Shankar, Sayna Ebrahimi +2
cs.LGcs.CVstat.MLarXiv:2107.03315v22021MixerCSeg: An Efficient Mixer Architecture for Crack Segmentation via Decoupled Mamba Attention
Zilong Zhao, Zhengming Ding, Pei Niu +2
cs.CVcs.AIarXiv:2603.01361v12026Fast Spatially-Varying Indoor Lighting Estimation
Mathieu Garon, Kalyan Sunkavalli, Sunil Hadap +2
cs.CVarXiv:1906.03799v12019SROBB: Targeted Perceptual Loss for Single Image Super-Resolution
Mohammad Saeed Rad, Behzad Bozorgtabar, Urs-Viktor Marti +3
cs.CVarXiv:1908.07222v12019TokenGS: Decoupling 3D Gaussian Prediction from Pixels with Learnable Tokens
Jiawei Ren, Michal Jan Tyszkiewicz, Jiahui Huang +1
cs.CVarXiv:2604.15239v12026Rethinking Zero-Shot Learning: A Conditional Visual Classification Perspective
Kai Li, Martin Renqiang Min, Yun Fu
cs.CVarXiv:1909.05995v22019SiamAPN++: Siamese Attentional Aggregation Network for Real-Time UAV Tracking
Ziang Cao, Changhong Fu, Junjie Ye +2
cs.CVarXiv:2106.08816v22021A Survey on Graph-Based Deep Learning for Computational Histopathology
David Ahmedt-Aristizabal, Mohammad Ali Armin, Simon Denman +2
cs.LGcs.CVq-bio.TOarXiv:2107.00272v22021Synthetic CT Generation from MRI using 3D Transformer-based Denoising Diffusion Model
Shaoyan Pan, Elham Abouei, Jacob Wynne +10
eess.IVcs.CVarXiv:2305.19467v12023Joint Weakly and Semi-Supervised Deep Learning for Localization and Classification of Masses in Breast Ultrasound Images
Seung Yeon Shin, Soochahn Lee, Il Dong Yun +2
cs.CVarXiv:1710.03778v22017MetaAnchor: Learning to Detect Objects with Customized Anchors
Tong Yang, Xiangyu Zhang, Zeming Li +2
cs.CVarXiv:1807.00980v22018Generative Models as a Data Source for Multiview Representation Learning
Ali Jahanian, Xavier Puig, Yonglong Tian +1
cs.CVarXiv:2106.05258v32021Compression of Deep Learning Models for Text: A Survey
Manish Gupta, Puneet Agrawal
cs.CLcs.AIcs.CVarXiv:2008.05221v42020Top-down Visual Saliency Guided by Captions
Vasili Ramanishka, Abir Das, Jianming Zhang +1
cs.CVarXiv:1612.07360v22016A Comprehensive Survey on Segment Anything Model for Vision and Beyond
Chunhui Zhang, Li Liu, Yawen Cui +4
cs.CVcs.AIarXiv:2305.08196v22023Backbones-Review: Feature Extraction Networks for Deep Learning and Deep Reinforcement Learning Approaches
Omar Elharrouss, Younes Akbari, Noor Almaadeed +1
cs.CVarXiv:2206.08016v12022ReconX: Reconstruct Any Scene from Sparse Views with Video Diffusion Model
Fangfu Liu, Wenqiang Sun, Hanyang Wang +5
cs.CVcs.AIcs.GRarXiv:2408.16767v42024Photon counting compressive depth mapping
Gregory A. Howland, Daniel J. Lum, Matthew R. Ware +1
physics.opticscs.CVarXiv:1309.4385v12013A Simple Baseline for Semi-supervised Semantic Segmentation with Strong Data Augmentation
Jianlong Yuan, Yifan Liu, Chunhua Shen +2
cs.CVarXiv:2104.07256v42021Combining Multiple Feature Extraction Techniques for Handwritten Devnagari Character Recognition
Sandhya Arora, Debotosh Bhattacharjee, Mita Nasipuri +2
cs.CVcs.AIarXiv:1005.4032v120103D Semi-Supervised Learning with Uncertainty-Aware Multi-View Co-Training
Yingda Xia, Fengze Liu, Dong Yang +6
cs.CVarXiv:1811.12506v22018OpenMix: Reviving Known Knowledge for Discovering Novel Visual Categories in An Open World
Zhun Zhong, Linchao Zhu, Zhiming Luo +3
cs.CVarXiv:2004.05551v12020Face Morphing Attack Generation & Detection: A Comprehensive Survey
Sushma Venkatesh, Raghavendra Ramachandra, Kiran Raja +1
cs.CVcs.CRcs.CYarXiv:2011.02045v12020SAM on Medical Images: A Comprehensive Study on Three Prompt Modes
Dongjie Cheng, Ziyuan Qin, Zekun Jiang +3
cs.CVcs.AIarXiv:2305.00035v12023Deep supervision with additional labels for retinal vessel segmentation task
Yishuo Zhang, Albert C. S. Chung
cs.CVarXiv:1806.02132v32018Classification of EEG-Based Brain Connectivity Networks in Schizophrenia Using a Multi-Domain Connectome Convolutional Neural Network
Chun-Ren Phang, Chee-Ming Ting, Fuad Noman +1
cs.LGcs.CVq-bio.NCarXiv:1903.08858v12019GAN Memory with No Forgetting
Yulai Cong, Miaoyun Zhao, Jianqiao Li +2
cs.CVcs.LGarXiv:2006.07543v22020DeXpression: Deep Convolutional Neural Network for Expression Recognition
Peter Burkert, Felix Trier, Muhammad Zeshan Afzal +2
cs.CVcs.LGarXiv:1509.05371v22015ProtoPShare: Prototype Sharing for Interpretable Image Classification and Similarity Discovery
Dawid Rymarczyk, Łukasz Struski, Jacek Tabor +1
cs.CVcs.AIcs.LGarXiv:2011.14340v12020Beyond accuracy: quantifying trial-by-trial behaviour of CNNs and humans by measuring error consistency
Robert Geirhos, Kristof Meding, Felix A. Wichmann
cs.CVcs.LGq-bio.NCarXiv:2006.16736v32020CALM: Conditional Adversarial Latent Models for Directable Virtual Characters
Chen Tessler, Yoni Kasten, Yunrong Guo +3
cs.CVcs.AIcs.ROarXiv:2305.02195v12023Hybrid Convolutional and Attention Network for Hyperspectral Image Denoising
Shuai Hu, Feng Gao, Xiaowei Zhou +2
eess.IVcs.CVarXiv:2403.10067v12024Connecting Look and Feel: Associating the visual and tactile properties of physical materials
Wenzhen Yuan, Shaoxiong Wang, Siyuan Dong +1
cs.CVarXiv:1704.03822v12017Interactive Sketch & Fill: Multiclass Sketch-to-Image Translation
Arnab Ghosh, Richard Zhang, Puneet K. Dokania +4
cs.CVcs.LGeess.IVarXiv:1909.11081v22019SIZER: A Dataset and Model for Parsing 3D Clothing and Learning Size Sensitive 3D Clothing
Garvita Tiwari, Bharat Lal Bhatnagar, Tony Tung +1
cs.CVarXiv:2007.11610v12020