Computer Vision and Pattern Recognition

Papers filed under cs.CV on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.

Search paper metadata (including unsummarized papers)

2,461 to 2,520 of 18,855

  1. Few-Example Object Detection with Model Communication

    Xuanyi Dong, Liang Zheng, Fan Ma +2

    cs.CVarXiv:1706.08249v82017
  2. Intra-Retinal Layer Segmentation of 3D Optical Coherence Tomography Using Coarse Grained Diffusion Map

    Raheleh Kafieh, Hossein Rabbani, Michael D. Abramoff +1

    cs.CVarXiv:1210.0310v22012
  3. Visual Explanations From Deep 3D Convolutional Neural Networks for Alzheimer's Disease Classification

    Chengliang Yang, Anand Rangarajan, Sanjay Ranka

    cs.CVcs.AIcs.LGarXiv:1803.02544v32018
  4. CoreDiff: Contextual Error-Modulated Generalized Diffusion Model for Low-Dose CT Denoising and Generalization

    Qi Gao, Zilong Li, Junping Zhang +2

    eess.IVcs.CVcs.LGarXiv:2304.01814v22023
  5. On the generalization of GAN image forensics

    Xinsheng Xuan, Bo Peng, Wei Wang +1

    cs.CVcs.LGstat.MLarXiv:1902.11153v22019
  6. Machine Vision for Natural Gas Methane Emissions Detection Using an Infrared Camera

    Jingfan Wang, Lyne P. Tchapmi, Arvind P. Ravikumara +5

    cs.CVcs.LGeess.IVarXiv:1904.08500v12019
  7. In-context learning enables multimodal large language models to classify cancer pathology images

    Dyke Ferber, Georg Wölflein, Isabella C. Wiest +8

    cs.CVarXiv:2403.07407v12024
  8. Deep Learning-Based Autonomous Driving Systems: A Survey of Attacks and Defenses

    Yao Deng, Tiehua Zhang, Guannan Lou +3

    cs.LGcs.CRcs.CVarXiv:2104.01789v22021
  9. Layer-Wise Gate-Controlled Prompt Truncation in a Multimodal Chest X-Ray Classifier

    Jingtao Lei, Hongji Li, Dexiang Shu

    cs.LGcs.AIcs.CVarXiv:2609.06590v12026
  10. Adversarial Attacks Beyond the Image Space

    Xiaohui Zeng, Chenxi Liu, Yu-Siang Wang +5

    cs.CVarXiv:1711.07183v62017
  11. Phonocardiographic Sensing using Deep Learning for Abnormal Heartbeat Detection

    Siddique Latif, Muhammad Usman, Rajib Rana +1

    cs.CVarXiv:1801.08322v42018
  12. Reading Decoder Trajectories: Training-Free Counterfactual Query-Trajectory Reliability for Small-Object Detection

    Zhaoning Shi, Bo Ma

    cs.CVcs.AIarXiv:2609.06581v12026
  13. LargeKernel3D: Scaling up Kernels in 3D Sparse CNNs

    Yukang Chen, Jianhui Liu, Xiangyu Zhang +2

    cs.CVcs.LGarXiv:2206.10555v22022
  14. OracleZoom: On-Policy Self-Distillation Inspired Reference-Constrained Recursive Image Super Resolution

    Shubhashis Roy Dipta, Sourajit Saha, Shaswati Saha +1

    cs.CVcs.AIcs.CLarXiv:2609.06490v12026
  15. Fully Convolutional One-Stage 3D Object Detection on LiDAR Range Images

    Zhi Tian, Xiangxiang Chu, Xiaoming Wang +2

    cs.CVarXiv:2205.13764v22022
  16. Branched Multi-Task Networks: Deciding What Layers To Share

    Simon Vandenhende, Stamatios Georgoulis, Bert De Brabandere +1

    cs.CVarXiv:1904.02920v52019
  17. Semi-Supervised Learning with Context-Conditional Generative Adversarial Networks

    Remi Denton, Sam Gross, Rob Fergus

    cs.CVarXiv:1611.06430v12016
  18. One MLLM, One Call: Efficient Zero-Shot Vision-and-Language Navigation via Spatial-Aware Waypoints

    Shiqi Pan, Qi Zheng, Hanqin Sun +3

    cs.CVcs.AIarXiv:2609.06476v12026
  19. Language and Visual Entity Relationship Graph for Agent Navigation

    Yicong Hong, Cristian Rodriguez-Opazo, Yuankai Qi +2

    cs.CVarXiv:2010.09304v22020
  20. Total variation regularization for fMRI-based prediction of behaviour

    Vincent Michel, Alexandre Gramfort, Gaël Varoquaux +2

    cs.CVq-bio.NCarXiv:1102.1101v12011
  21. Detail Preserved Point Cloud Completion via Separated Feature Aggregation

    Wenxiao Zhang, Qingan Yan, Chunxia Xiao

    cs.CVcs.CGarXiv:2007.02374v12020
  22. Learning Common and Specific Features for RGB-D Semantic Segmentation with Deconvolutional Networks

    Jinghua Wang, Zhenhua Wang, Dacheng Tao +2

    cs.CVarXiv:1608.01082v12016
  23. A Learned Representation for Scalable Vector Graphics

    Raphael Gontijo Lopes, David Ha, Douglas Eck +1

    cs.CVcs.LGstat.MLarXiv:1904.02632v12019
  24. Automatic Extrinsic Calibration for Lidar-Stereo Vehicle Sensor Setups

    Carlos Guindel, Jorge Beltrán, David Martín +1

    cs.CVcs.ROarXiv:1705.04085v32017
  25. Discriminative Localization in CNNs for Weakly-Supervised Segmentation of Pulmonary Nodules

    Xinyang Feng, Jie Yang, Andrew F. Laine +1

    cs.CVarXiv:1707.01086v22017
  26. Geometry Guided Adversarial Facial Expression Synthesis

    Lingxiao Song, Zhihe Lu, Ran He +2

    cs.CVarXiv:1712.03474v12017
  27. Detailed Human Shape Estimation from a Single Image by Hierarchical Mesh Deformation

    Hao Zhu, Xinxin Zuo, Sen Wang +2

    cs.CVeess.IVarXiv:1904.10506v22019
  28. The Benchmark Lottery

    Mostafa Dehghani, Yi Tay, Alexey A. Gritsenko +5

    cs.LGcs.AIcs.CLarXiv:2107.07002v12021
  29. Grounding Language Models to Images for Multimodal Inputs and Outputs

    Jing Yu Koh, Ruslan Salakhutdinov, Daniel Fried

    cs.CLcs.AIcs.CVarXiv:2301.13823v42023
  30. Adversarial Objects Against LiDAR-Based Autonomous Driving Systems

    Yulong Cao, Chaowei Xiao, Dawei Yang +4

    cs.CRcs.CVcs.LGarXiv:1907.05418v12019
  31. Spatial Information Guided Convolution for Real-Time RGBD Semantic Segmentation

    Lin-Zhuo Chen, Zheng Lin, Ziqin Wang +2

    cs.CVarXiv:2004.04534v22020
  32. Learning to Evaluate Image Captioning

    Yin Cui, Guandao Yang, Andreas Veit +2

    cs.CVcs.LGarXiv:1806.06422v12018
  33. Modeling Local Geometric Structure of 3D Point Clouds using Geo-CNN

    Shiyi Lan, Ruichi Yu, Gang Yu +1

    cs.CVarXiv:1811.07782v12018
  34. Grounding Language with Visual Affordances over Unstructured Data

    Oier Mees, Jessica Borja-Diaz, Wolfram Burgard

    cs.ROcs.AIcs.CLarXiv:2210.01911v32022
  35. Multiple Myeloma Lesion Segmentation on Whole-Body Diffusion-Weighted Imaging via Efficient Anatomical Anticipation and Multimodal Confirmation

    Mengmeng Zhang, Shengqian Huang, Junde Zhou +12

    cs.CVcs.AIarXiv:2609.06165v12026
  36. Light Field Image Super-Resolution Using Deformable Convolution

    Yingqian Wang, Jungang Yang, Longguang Wang +4

    eess.IVcs.CVarXiv:2007.03535v42020
  37. SALSA: A Novel Dataset for Multimodal Group Behavior Analysis

    Xavier Alameda-Pineda, Jacopo Staiano, Ramanathan Subramanian +5

    cs.CVarXiv:1506.06882v12015
  38. Subdivision-Based Mesh Convolution Networks

    Shi-Min Hu, Zheng-Ning Liu, Meng-Hao Guo +4

    cs.CVcs.GRcs.LGarXiv:2106.02285v22021
  39. Boosting Multimodal Large Language Models with Visual Tokens Withdrawal for Rapid Inference

    Zhihang Lin, Mingbao Lin, Luxi Lin +1

    cs.CVcs.AIarXiv:2405.05803v32024
  40. ABAW: Learning from Synthetic Data & Multi-Task Learning Challenges

    Dimitrios Kollias

    cs.CVarXiv:2207.01138v22022
  41. PP-YOLOv2: A Practical Object Detector

    Xin Huang, Xinxin Wang, Wenyu Lv +10

    cs.CVarXiv:2104.10419v12021
  42. WSOD^2: Learning Bottom-up and Top-down Objectness Distillation for Weakly-supervised Object Detection

    Zhaoyang Zeng, Bei Liu, Jianlong Fu +2

    cs.CVarXiv:1909.04972v12019
  43. GLF-CR: SAR-Enhanced Cloud Removal with Global-Local Fusion

    Fang Xu, Yilei Shi, Patrick Ebel +4

    cs.CVeess.IVarXiv:2206.02850v32022
  44. The Way to my Heart is through Contrastive Learning: Remote Photoplethysmography from Unlabelled Video

    John Gideon, Simon Stent

    cs.CVcs.HCarXiv:2111.09748v12021
  45. Instance-Conditioned GAN

    Arantxa Casanova, Marlène Careil, Jakob Verbeek +2

    cs.CVcs.LGarXiv:2109.05070v22021
  46. GALIP: Generative Adversarial CLIPs for Text-to-Image Synthesis

    Ming Tao, Bing-Kun Bao, Hao Tang +1

    cs.CVcs.AIarXiv:2301.12959v12023
  47. Transfer Learning from Synthetic to Real LiDAR Point Cloud for Semantic Segmentation

    Aoran Xiao, Jiaxing Huang, Dayan Guan +2

    cs.CVarXiv:2107.05399v22021
  48. Machine learning of hierarchical clustering to segment 2D and 3D images

    Juan Nunez-Iglesias, Ryan Kennedy, Toufiq Parag +2

    cs.CVcs.LGarXiv:1303.6163v32013
  49. A Machine Learning Benchmark for Facies Classification

    Yazeed Alaudah, Patrycja Michalowicz, Motaz Alfarraj +1

    eess.IVcs.CVphysics.geo-pharXiv:1901.07659v22019
  50. Template Adaptation for Face Verification and Identification

    Nate Crosswhite, Jeffrey Byrne, Omkar M. Parkhi +3

    cs.CVarXiv:1603.03958v32016
  51. Dancing Stick Figures: An Introductory Dataset for Training Video Generation Models

    Jin Hyuk Cho

    cs.CVarXiv:2608.29123v12026
  52. POCO: Point Convolution for Surface Reconstruction

    Alexandre Boulch, Renaud Marlet

    cs.CVcs.CGcs.LGarXiv:2201.01831v22022
  53. A Remote Sensing Image Dataset for Cloud Removal

    Daoyu Lin, Guangluan Xu, Xiaoke Wang +3

    cs.CVarXiv:1901.00600v12019
  54. Bilinear Factor Matrix Norm Minimization for Robust PCA: Algorithms and Applications

    Fanhua Shang, James Cheng, Yuanyuan Liu +2

    cs.LGcs.CVmath.OCarXiv:1810.05186v12018
    Summaries:한국어
  55. Style-Hallucinated Dual Consistency Learning for Domain Generalized Semantic Segmentation

    Yuyang Zhao, Zhun Zhong, Na Zhao +2

    cs.CVarXiv:2204.02548v22022
  56. An Information-Theoretic Approach to Transferability in Task Transfer Learning

    Yajie Bao, Yang Li, Shao-Lun Huang +4

    cs.LGcs.CVarXiv:2212.10082v12022
  57. MUTANT: A Training Paradigm for Out-of-Distribution Generalization in Visual Question Answering

    Tejas Gokhale, Pratyay Banerjee, Chitta Baral +1

    cs.CVcs.CLarXiv:2009.08566v22020
  58. End-to-end Trainable Deep Neural Network for Robotic Grasp Detection and Semantic Segmentation from RGB

    Stefan Ainetter, Friedrich Fraundorfer

    cs.CVcs.ROarXiv:2107.05287v22021
  59. Protecting Facial Privacy: Generating Adversarial Identity Masks via Style-robust Makeup Transfer

    Shengshan Hu, Xiaogeng Liu, Yechao Zhang +4

    cs.CVcs.CRarXiv:2203.03121v22022
  60. Semantically Tied Paired Cycle Consistency for Zero-Shot Sketch-based Image Retrieval

    Anjan Dutta, Zeynep Akata

    cs.CVarXiv:1903.03372v12019