Computer Vision and Pattern Recognition

Papers filed under cs.CV on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.

Search paper metadata (including unsummarized papers)

11,521 to 11,580 of 18,867

  1. Diversify and Match: A Domain Adaptive Representation Learning Paradigm for Object Detection

    Taekyung Kim, Minki Jeong, Seunghyeon Kim +2

    cs.CVarXiv:1905.05396v12019
  2. Comparing Kullback-Leibler Divergence and Mean Squared Error Loss in Knowledge Distillation

    Taehyeon Kim, Jaehoon Oh, NakYil Kim +2

    cs.LGcs.CVarXiv:2105.08919v12021
  3. Efficiently Identifying Task Groupings for Multi-Task Learning

    Christopher Fifty, Ehsan Amid, Zhe Zhao +3

    cs.LGcs.AIcs.CVarXiv:2109.04617v22021
  4. Mastering Atari Games with Limited Data

    Weirui Ye, Shaohuai Liu, Thanard Kurutach +2

    cs.LGcs.AIcs.CVarXiv:2111.00210v22021
  5. Semi-Supervised Medical Image Segmentation via Cross Teaching between CNN and Transformer

    Xiangde Luo, Minhao Hu, Tao Song +2

    eess.IVcs.CVarXiv:2112.04894v22021
  6. Video Captioning with Transferred Semantic Attributes

    Yingwei Pan, Ting Yao, Houqiang Li +1

    cs.CVarXiv:1611.07675v12016
  7. Super-Resolution with Deep Convolutional Sufficient Statistics

    Joan Bruna, Pablo Sprechmann, Yann LeCun

    cs.CVarXiv:1511.05666v42015
  8. Automatic Moth Detection from Trap Images for Pest Management

    Weiguang Ding, Graham Taylor

    cs.CVcs.LGcs.NEarXiv:1602.07383v12016
  9. FVC: A New Framework towards Deep Video Compression in Feature Space

    Zhihao Hu, Guo Lu, Dong Xu

    eess.IVcs.CVarXiv:2105.09600v22021
  10. Skin Lesion Classification Using Ensembles of Multi-Resolution EfficientNets with Meta Data

    Nils Gessert, Maximilian Nielsen, Mohsin Shaikh +2

    cs.CVarXiv:1910.03910v12019
  11. Deep Exemplar-based Colorization

    Mingming He, Dongdong Chen, Jing Liao +2

    cs.CVarXiv:1807.06587v22018
  12. Patch-based Probabilistic Image Quality Assessment for Face Selection and Improved Video-based Face Recognition

    Yongkang Wong, Shaokang Chen, Sandra Mau +2

    cs.CVstat.AParXiv:1304.0869v22013
  13. Learning the Best Pooling Strategy for Visual Semantic Embedding

    Jiacheng Chen, Hexiang Hu, Hao Wu +2

    cs.CVarXiv:2011.04305v52020
  14. DilateFormer: Multi-Scale Dilated Transformer for Visual Recognition

    Jiayu Jiao, Yu-Ming Tang, Kun-Yu Lin +4

    cs.CVarXiv:2302.01791v12023
  15. Unsupervised Misaligned Infrared and Visible Image Fusion via Cross-Modality Image Generation and Registration

    Di Wang, Jinyuan Liu, Xin Fan +1

    cs.CVarXiv:2205.11876v12022
  16. In Defense of Classical Image Processing: Fast Depth Completion on the CPU

    Jason Ku, Ali Harakeh, Steven L. Waslander

    cs.CVarXiv:1802.00036v12018
  17. Image Inpainting via Generative Multi-column Convolutional Neural Networks

    Yi Wang, Xin Tao, Xiaojuan Qi +2

    cs.CVarXiv:1810.08771v12018
  18. OpenVid-1M: A Large-Scale High-Quality Dataset for Text-to-video Generation

    Kepan Nan, Rui Xie, Penghao Zhou +6

    cs.CVarXiv:2407.02371v32024
  19. Fully Convolutional Cross-Scale-Flows for Image-based Defect Detection

    Marco Rudolph, Tom Wehrbein, Bodo Rosenhahn +1

    cs.CVarXiv:2110.02855v12021
  20. FastFCN: Rethinking Dilated Convolution in the Backbone for Semantic Segmentation

    Huikai Wu, Junge Zhang, Kaiqi Huang +2

    cs.CVarXiv:1903.11816v12019
  21. Occlusion Aware Unsupervised Learning of Optical Flow

    Yang Wang, Yi Yang, Zhenheng Yang +3

    cs.CVarXiv:1711.05890v22017
  22. Toward Driving Scene Understanding: A Dataset for Learning Driver Behavior and Causal Reasoning

    Vasili Ramanishka, Yi-Ting Chen, Teruhisa Misu +1

    cs.CVarXiv:1811.02307v12018
  23. A robust and efficient video representation for action recognition

    Heng Wang, Dan Oneata, Jakob Verbeek +1

    cs.CVarXiv:1504.05524v12015
  24. Rethinking Knowledge Graph Propagation for Zero-Shot Learning

    Michael Kampffmeyer, Yinbo Chen, Xiaodan Liang +3

    cs.CVarXiv:1805.11724v32018
  25. LMMs-Eval: Reality Check on the Evaluation of Large Multimodal Models

    Kaichen Zhang, Bo Li, Peiyuan Zhang +8

    cs.CLcs.CVarXiv:2407.12772v22024
  26. Back to the Feature: Learning Robust Camera Localization from Pixels to Pose

    Paul-Edouard Sarlin, Ajaykumar Unagar, Måns Larsson +8

    cs.CVarXiv:2103.09213v22021
  27. GPS-Net: Graph Property Sensing Network for Scene Graph Generation

    Xin Lin, Changxing Ding, Jinquan Zeng +1

    cs.CVarXiv:2003.12962v12020
  28. MOS: Towards Scaling Out-of-distribution Detection for Large Semantic Space

    Rui Huang, Yixuan Li

    cs.CVcs.LGarXiv:2105.01879v12021
  29. Spike-driven Transformer

    Man Yao, Jiakui Hu, Zhaokun Zhou +4

    cs.NEcs.CVarXiv:2307.01694v12023
  30. SelFlow: Self-Supervised Learning of Optical Flow

    Pengpeng Liu, Michael Lyu, Irwin King +1

    cs.CVcs.LGarXiv:1904.09117v12019
  31. Surgical Data Science -- from Concepts toward Clinical Translation

    Lena Maier-Hein, Matthias Eisenmann, Duygu Sarikaya +47

    cs.CYcs.CVcs.LGarXiv:2011.02284v22020
  32. AdaViT: Adaptive Vision Transformers for Efficient Image Recognition

    Lingchen Meng, Hengduo Li, Bor-Chun Chen +4

    cs.CVarXiv:2111.15668v12021
  33. 3D Neural Field Generation using Triplane Diffusion

    J. Ryan Shue, Eric Ryan Chan, Ryan Po +3

    cs.CVcs.AIcs.GRarXiv:2211.16677v12022
  34. BA-Net: Dense Bundle Adjustment Network

    Chengzhou Tang, Ping Tan

    cs.CVarXiv:1806.04807v32018
  35. RSGPT: A Remote Sensing Vision Language Model and Benchmark

    Yuan Hu, Jianlong Yuan, Congcong Wen +2

    cs.CVarXiv:2307.15266v12023
  36. Cross-view Asymmetric Metric Learning for Unsupervised Person Re-identification

    Hong-Xing Yu, Ancong Wu, Wei-Shi Zheng

    cs.CVarXiv:1708.08062v22017
  37. Accurate Monocular Object Detection via Color-Embedded 3D Reconstruction for Autonomous Driving

    Xinzhu Ma, Zhihui Wang, Haojie Li +3

    cs.CVarXiv:1903.11444v42019
  38. Path-SGD: Path-Normalized Optimization in Deep Neural Networks

    Behnam Neyshabur, Ruslan Salakhutdinov, Nathan Srebro

    cs.LGcs.CVcs.NEarXiv:1506.02617v12015
  39. Visual Sketchpad: Sketching as a Visual Chain of Thought for Multimodal Language Models

    Yushi Hu, Weijia Shi, Xingyu Fu +5

    cs.CVcs.CLarXiv:2406.09403v32024
  40. Capsule Networks for Brain Tumor Classification based on MRI Images and Course Tumor Boundaries

    Parnian Afshar, Konstantinos N. Plataniotis, Arash Mohammadi

    cs.CVarXiv:1811.00597v12018
  41. STMTrack: Template-free Visual Tracking with Space-time Memory Networks

    Zhihong Fu, Qingjie Liu, Zehua Fu +1

    cs.CVarXiv:2104.00324v22021
  42. Adaptive NMS: Refining Pedestrian Detection in a Crowd

    Songtao Liu, Di Huang, Yunhong Wang

    cs.CVarXiv:1904.03629v12019
  43. Dense Regression Network for Video Grounding

    Runhao Zeng, Haoming Xu, Wenbing Huang +3

    cs.CVarXiv:2004.03545v12020
  44. A Mutual Bootstrapping Model for Automated Skin Lesion Segmentation and Classification

    Yutong Xie, Jianpeng Zhang, Yong Xia +1

    cs.CVarXiv:1903.03313v42019
  45. Real-time 2D Multi-Person Pose Estimation on CPU: Lightweight OpenPose

    Daniil Osokin

    cs.CVarXiv:1811.12004v12018
  46. Surpassing Human-Level Face Verification Performance on LFW with GaussianFace

    Chaochao Lu, Xiaoou Tang

    cs.CVcs.LGstat.MLarXiv:1404.3840v32014
  47. More Control for Free! Image Synthesis with Semantic Diffusion Guidance

    Xihui Liu, Dong Huk Park, Samaneh Azadi +6

    cs.CVcs.GRarXiv:2112.05744v42021
  48. You Only Look Twice: Rapid Multi-Scale Object Detection In Satellite Imagery

    Adam Van Etten

    cs.CVarXiv:1805.09512v12018
  49. Data Uncertainty Learning in Face Recognition

    Jie Chang, Zhonghao Lan, Changmao Cheng +1

    cs.CVarXiv:2003.11339v12020
  50. HumanPlus: Humanoid Shadowing and Imitation from Humans

    Zipeng Fu, Qingqing Zhao, Qi Wu +2

    cs.ROcs.AIcs.CVarXiv:2406.10454v12024
  51. Deep Convolutional Neural Networks for Breast Cancer Histology Image Analysis

    Alexander Rakhlin, Alexey Shvets, Vladimir Iglovikov +1

    cs.CVarXiv:1802.00752v22018
  52. Deep Neural Network Concepts for Background Subtraction: A Systematic Review and Comparative Evaluation

    Thierry Bouwmans, Sajid Javed, Maryam Sultana +1

    cs.CVarXiv:1811.05255v12018
  53. Deep-Learning Inversion of Seismic Data

    Shucai Li, Bin Liu, Yuxiao Ren +4

    cs.CVcs.AIarXiv:1901.07733v22019
  54. Neural Volumes: Learning Dynamic Renderable Volumes from Images

    Stephen Lombardi, Tomas Simon, Jason Saragih +3

    cs.GRcs.CVarXiv:1906.07751v12019
  55. A Review of Single-Source Deep Unsupervised Visual Domain Adaptation

    Sicheng Zhao, Xiangyu Yue, Shanghang Zhang +8

    cs.CVcs.LGeess.IVarXiv:2009.00155v32020
  56. DreamLLM: Synergistic Multimodal Comprehension and Creation

    Runpei Dong, Chunrui Han, Yuang Peng +11

    cs.CVcs.CLcs.LGarXiv:2309.11499v22023
  57. dipIQ: Blind Image Quality Assessment by Learning-to-Rank Discriminable Image Pairs

    Kede Ma, Wentao Liu, Tongliang Liu +2

    cs.CVcs.MMarXiv:1904.06505v12019
  58. What Should Not Be Contrastive in Contrastive Learning

    Tete Xiao, Xiaolong Wang, Alexei A. Efros +1

    cs.CVarXiv:2008.05659v22020
  59. Brain-Like Object Recognition with High-Performing Shallow Recurrent ANNs

    Jonas Kubilius, Martin Schrimpf, Kohitij Kar +11

    cs.CVcs.LGcs.NEarXiv:1909.06161v22019
  60. A deep matrix factorization method for learning attribute representations

    George Trigeorgis, Konstantinos Bousmalis, Stefanos Zafeiriou +1

    cs.CVcs.LGstat.MLarXiv:1509.03248v12015