Computer Vision and Pattern Recognition
Papers filed under cs.CV on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.
Search paper metadata (including unsummarized papers)
2,461 to 2,520 of 18,855
Few-Example Object Detection with Model Communication
Xuanyi Dong, Liang Zheng, Fan Ma +2
cs.CVarXiv:1706.08249v82017Intra-Retinal Layer Segmentation of 3D Optical Coherence Tomography Using Coarse Grained Diffusion Map
Raheleh Kafieh, Hossein Rabbani, Michael D. Abramoff +1
cs.CVarXiv:1210.0310v22012Visual Explanations From Deep 3D Convolutional Neural Networks for Alzheimer's Disease Classification
Chengliang Yang, Anand Rangarajan, Sanjay Ranka
cs.CVcs.AIcs.LGarXiv:1803.02544v32018CoreDiff: Contextual Error-Modulated Generalized Diffusion Model for Low-Dose CT Denoising and Generalization
Qi Gao, Zilong Li, Junping Zhang +2
eess.IVcs.CVcs.LGarXiv:2304.01814v22023On the generalization of GAN image forensics
Xinsheng Xuan, Bo Peng, Wei Wang +1
cs.CVcs.LGstat.MLarXiv:1902.11153v22019Machine Vision for Natural Gas Methane Emissions Detection Using an Infrared Camera
Jingfan Wang, Lyne P. Tchapmi, Arvind P. Ravikumara +5
cs.CVcs.LGeess.IVarXiv:1904.08500v12019In-context learning enables multimodal large language models to classify cancer pathology images
Dyke Ferber, Georg Wölflein, Isabella C. Wiest +8
cs.CVarXiv:2403.07407v12024Deep Learning-Based Autonomous Driving Systems: A Survey of Attacks and Defenses
Yao Deng, Tiehua Zhang, Guannan Lou +3
cs.LGcs.CRcs.CVarXiv:2104.01789v22021Layer-Wise Gate-Controlled Prompt Truncation in a Multimodal Chest X-Ray Classifier
Jingtao Lei, Hongji Li, Dexiang Shu
cs.LGcs.AIcs.CVarXiv:2609.06590v12026Adversarial Attacks Beyond the Image Space
Xiaohui Zeng, Chenxi Liu, Yu-Siang Wang +5
cs.CVarXiv:1711.07183v62017Phonocardiographic Sensing using Deep Learning for Abnormal Heartbeat Detection
Siddique Latif, Muhammad Usman, Rajib Rana +1
cs.CVarXiv:1801.08322v42018Reading Decoder Trajectories: Training-Free Counterfactual Query-Trajectory Reliability for Small-Object Detection
Zhaoning Shi, Bo Ma
cs.CVcs.AIarXiv:2609.06581v12026LargeKernel3D: Scaling up Kernels in 3D Sparse CNNs
Yukang Chen, Jianhui Liu, Xiangyu Zhang +2
cs.CVcs.LGarXiv:2206.10555v22022OracleZoom: On-Policy Self-Distillation Inspired Reference-Constrained Recursive Image Super Resolution
Shubhashis Roy Dipta, Sourajit Saha, Shaswati Saha +1
cs.CVcs.AIcs.CLarXiv:2609.06490v12026Fully Convolutional One-Stage 3D Object Detection on LiDAR Range Images
Zhi Tian, Xiangxiang Chu, Xiaoming Wang +2
cs.CVarXiv:2205.13764v22022Branched Multi-Task Networks: Deciding What Layers To Share
Simon Vandenhende, Stamatios Georgoulis, Bert De Brabandere +1
cs.CVarXiv:1904.02920v52019Semi-Supervised Learning with Context-Conditional Generative Adversarial Networks
Remi Denton, Sam Gross, Rob Fergus
cs.CVarXiv:1611.06430v12016One MLLM, One Call: Efficient Zero-Shot Vision-and-Language Navigation via Spatial-Aware Waypoints
Shiqi Pan, Qi Zheng, Hanqin Sun +3
cs.CVcs.AIarXiv:2609.06476v12026Language and Visual Entity Relationship Graph for Agent Navigation
Yicong Hong, Cristian Rodriguez-Opazo, Yuankai Qi +2
cs.CVarXiv:2010.09304v22020Total variation regularization for fMRI-based prediction of behaviour
Vincent Michel, Alexandre Gramfort, Gaël Varoquaux +2
cs.CVq-bio.NCarXiv:1102.1101v12011Detail Preserved Point Cloud Completion via Separated Feature Aggregation
Wenxiao Zhang, Qingan Yan, Chunxia Xiao
cs.CVcs.CGarXiv:2007.02374v12020Learning Common and Specific Features for RGB-D Semantic Segmentation with Deconvolutional Networks
Jinghua Wang, Zhenhua Wang, Dacheng Tao +2
cs.CVarXiv:1608.01082v12016A Learned Representation for Scalable Vector Graphics
Raphael Gontijo Lopes, David Ha, Douglas Eck +1
cs.CVcs.LGstat.MLarXiv:1904.02632v12019Automatic Extrinsic Calibration for Lidar-Stereo Vehicle Sensor Setups
Carlos Guindel, Jorge Beltrán, David Martín +1
cs.CVcs.ROarXiv:1705.04085v32017Discriminative Localization in CNNs for Weakly-Supervised Segmentation of Pulmonary Nodules
Xinyang Feng, Jie Yang, Andrew F. Laine +1
cs.CVarXiv:1707.01086v22017Geometry Guided Adversarial Facial Expression Synthesis
Lingxiao Song, Zhihe Lu, Ran He +2
cs.CVarXiv:1712.03474v12017Detailed Human Shape Estimation from a Single Image by Hierarchical Mesh Deformation
Hao Zhu, Xinxin Zuo, Sen Wang +2
cs.CVeess.IVarXiv:1904.10506v22019The Benchmark Lottery
Mostafa Dehghani, Yi Tay, Alexey A. Gritsenko +5
cs.LGcs.AIcs.CLarXiv:2107.07002v12021Grounding Language Models to Images for Multimodal Inputs and Outputs
Jing Yu Koh, Ruslan Salakhutdinov, Daniel Fried
cs.CLcs.AIcs.CVarXiv:2301.13823v42023Adversarial Objects Against LiDAR-Based Autonomous Driving Systems
Yulong Cao, Chaowei Xiao, Dawei Yang +4
cs.CRcs.CVcs.LGarXiv:1907.05418v12019Spatial Information Guided Convolution for Real-Time RGBD Semantic Segmentation
Lin-Zhuo Chen, Zheng Lin, Ziqin Wang +2
cs.CVarXiv:2004.04534v22020Learning to Evaluate Image Captioning
Yin Cui, Guandao Yang, Andreas Veit +2
cs.CVcs.LGarXiv:1806.06422v12018Modeling Local Geometric Structure of 3D Point Clouds using Geo-CNN
Shiyi Lan, Ruichi Yu, Gang Yu +1
cs.CVarXiv:1811.07782v12018Grounding Language with Visual Affordances over Unstructured Data
Oier Mees, Jessica Borja-Diaz, Wolfram Burgard
cs.ROcs.AIcs.CLarXiv:2210.01911v32022Multiple Myeloma Lesion Segmentation on Whole-Body Diffusion-Weighted Imaging via Efficient Anatomical Anticipation and Multimodal Confirmation
Mengmeng Zhang, Shengqian Huang, Junde Zhou +12
cs.CVcs.AIarXiv:2609.06165v12026Light Field Image Super-Resolution Using Deformable Convolution
Yingqian Wang, Jungang Yang, Longguang Wang +4
eess.IVcs.CVarXiv:2007.03535v42020SALSA: A Novel Dataset for Multimodal Group Behavior Analysis
Xavier Alameda-Pineda, Jacopo Staiano, Ramanathan Subramanian +5
cs.CVarXiv:1506.06882v12015Subdivision-Based Mesh Convolution Networks
Shi-Min Hu, Zheng-Ning Liu, Meng-Hao Guo +4
cs.CVcs.GRcs.LGarXiv:2106.02285v22021Boosting Multimodal Large Language Models with Visual Tokens Withdrawal for Rapid Inference
Zhihang Lin, Mingbao Lin, Luxi Lin +1
cs.CVcs.AIarXiv:2405.05803v32024ABAW: Learning from Synthetic Data & Multi-Task Learning Challenges
Dimitrios Kollias
cs.CVarXiv:2207.01138v22022PP-YOLOv2: A Practical Object Detector
Xin Huang, Xinxin Wang, Wenyu Lv +10
cs.CVarXiv:2104.10419v12021WSOD^2: Learning Bottom-up and Top-down Objectness Distillation for Weakly-supervised Object Detection
Zhaoyang Zeng, Bei Liu, Jianlong Fu +2
cs.CVarXiv:1909.04972v12019GLF-CR: SAR-Enhanced Cloud Removal with Global-Local Fusion
Fang Xu, Yilei Shi, Patrick Ebel +4
cs.CVeess.IVarXiv:2206.02850v32022The Way to my Heart is through Contrastive Learning: Remote Photoplethysmography from Unlabelled Video
John Gideon, Simon Stent
cs.CVcs.HCarXiv:2111.09748v12021Instance-Conditioned GAN
Arantxa Casanova, Marlène Careil, Jakob Verbeek +2
cs.CVcs.LGarXiv:2109.05070v22021GALIP: Generative Adversarial CLIPs for Text-to-Image Synthesis
Ming Tao, Bing-Kun Bao, Hao Tang +1
cs.CVcs.AIarXiv:2301.12959v12023Transfer Learning from Synthetic to Real LiDAR Point Cloud for Semantic Segmentation
Aoran Xiao, Jiaxing Huang, Dayan Guan +2
cs.CVarXiv:2107.05399v22021Machine learning of hierarchical clustering to segment 2D and 3D images
Juan Nunez-Iglesias, Ryan Kennedy, Toufiq Parag +2
cs.CVcs.LGarXiv:1303.6163v32013A Machine Learning Benchmark for Facies Classification
Yazeed Alaudah, Patrycja Michalowicz, Motaz Alfarraj +1
eess.IVcs.CVphysics.geo-pharXiv:1901.07659v22019Template Adaptation for Face Verification and Identification
Nate Crosswhite, Jeffrey Byrne, Omkar M. Parkhi +3
cs.CVarXiv:1603.03958v32016Dancing Stick Figures: An Introductory Dataset for Training Video Generation Models
Jin Hyuk Cho
cs.CVarXiv:2608.29123v12026POCO: Point Convolution for Surface Reconstruction
Alexandre Boulch, Renaud Marlet
cs.CVcs.CGcs.LGarXiv:2201.01831v22022A Remote Sensing Image Dataset for Cloud Removal
Daoyu Lin, Guangluan Xu, Xiaoke Wang +3
cs.CVarXiv:1901.00600v12019Bilinear Factor Matrix Norm Minimization for Robust PCA: Algorithms and Applications
Fanhua Shang, James Cheng, Yuanyuan Liu +2
cs.LGcs.CVmath.OCarXiv:1810.05186v12018Summaries:한국어Style-Hallucinated Dual Consistency Learning for Domain Generalized Semantic Segmentation
Yuyang Zhao, Zhun Zhong, Na Zhao +2
cs.CVarXiv:2204.02548v22022An Information-Theoretic Approach to Transferability in Task Transfer Learning
Yajie Bao, Yang Li, Shao-Lun Huang +4
cs.LGcs.CVarXiv:2212.10082v12022MUTANT: A Training Paradigm for Out-of-Distribution Generalization in Visual Question Answering
Tejas Gokhale, Pratyay Banerjee, Chitta Baral +1
cs.CVcs.CLarXiv:2009.08566v22020End-to-end Trainable Deep Neural Network for Robotic Grasp Detection and Semantic Segmentation from RGB
Stefan Ainetter, Friedrich Fraundorfer
cs.CVcs.ROarXiv:2107.05287v22021Protecting Facial Privacy: Generating Adversarial Identity Masks via Style-robust Makeup Transfer
Shengshan Hu, Xiaogeng Liu, Yechao Zhang +4
cs.CVcs.CRarXiv:2203.03121v22022Semantically Tied Paired Cycle Consistency for Zero-Shot Sketch-based Image Retrieval
Anjan Dutta, Zeynep Akata
cs.CVarXiv:1903.03372v12019