Computer Vision and Pattern Recognition

Papers filed under cs.CV on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.

Search paper metadata (including unsummarized papers)

10,681 to 10,740 of 18,866

  1. Beyond Gaussian Pyramid: Multi-skip Feature Stacking for Action Recognition

    Zhenzhong Lan, Ming Lin, Xuanchong Li +2

    cs.CVarXiv:1411.6660v42014
  2. Small Data Challenges in Big Data Era: A Survey of Recent Progress on Unsupervised and Semi-Supervised Methods

    Guo-Jun Qi, Jiebo Luo

    cs.CVarXiv:1903.11260v22019
  3. ROI-10D: Monocular Lifting of 2D Detection to 6D Pose and Metric Shape

    Fabian Manhardt, Wadim Kehl, Adrien Gaidon

    cs.CVarXiv:1812.02781v32018
  4. TextureGAN: Controlling Deep Image Synthesis with Texture Patches

    Wenqi Xian, Patsorn Sangkloy, Varun Agrawal +5

    cs.CVcs.GRarXiv:1706.02823v32017
  5. STAR: A Structure and Texture Aware Retinex Model

    Jun Xu, Yingkun Hou, Dongwei Ren +5

    cs.CVarXiv:1906.06690v52019
  6. Partial success in closing the gap between human and machine vision

    Robert Geirhos, Kantharaju Narayanappa, Benjamin Mitzkus +4

    cs.CVcs.AIcs.LGarXiv:2106.07411v22021
  7. Hierarchical Open-Vocabulary 3D Scene Graphs for Language-Grounded Robot Navigation

    Abdelrhman Werby, Chenguang Huang, Martin Büchner +2

    cs.ROcs.AIcs.CLarXiv:2403.17846v22024
  8. More is Less: A More Complicated Network with Less Inference Complexity

    Xuanyi Dong, Junshi Huang, Yi Yang +1

    cs.CVarXiv:1703.08651v22017
  9. A Bio-Inspired Multi-Exposure Fusion Framework for Low-light Image Enhancement

    Zhenqiang Ying, Ge Li, Wen Gao

    cs.CVarXiv:1711.00591v12017
  10. VisualGPT: Data-efficient Adaptation of Pretrained Language Models for Image Captioning

    Jun Chen, Han Guo, Kai Yi +2

    cs.CVcs.AIcs.CLarXiv:2102.10407v52021
  11. Self-Erasing Network for Integral Object Attention

    Qibin Hou, Peng-Tao Jiang, Yunchao Wei +1

    cs.CVarXiv:1810.09821v12018
  12. Investigating Bi-Level Optimization for Learning and Vision from a Unified Perspective: A Survey and Beyond

    Risheng Liu, Jiaxin Gao, Jin Zhang +2

    cs.LGcs.CVmath.DSarXiv:2101.11517v32021
  13. Editing Conditional Radiance Fields

    Steven Liu, Xiuming Zhang, Zhoutong Zhang +3

    cs.CVcs.GRcs.LGarXiv:2105.06466v22021
  14. Learning to Self-Train for Semi-Supervised Few-Shot Classification

    Xinzhe Li, Qianru Sun, Yaoyao Liu +4

    cs.CVcs.LGstat.MLarXiv:1906.00562v22019
  15. MVTN: Multi-View Transformation Network for 3D Shape Recognition

    Abdullah Hamdi, Silvio Giancola, Bernard Ghanem

    cs.CVcs.LGarXiv:2011.13244v32020
  16. Video Object Segmentation with Episodic Graph Memory Networks

    Xiankai Lu, Wenguan Wang, Martin Danelljan +3

    cs.CVcs.LGarXiv:2007.07020v42020
  17. CityGaussian: Real-time High-quality Large-Scale Scene Rendering with Gaussians

    Yang Liu, He Guan, Chuanchen Luo +4

    cs.CVarXiv:2404.01133v32024
  18. Yedrouj-Net: An efficient CNN for spatial steganalysis

    Mehdi Yedroudj, Frederic Comby, Marc Chaumont

    cs.CVcs.CRarXiv:1803.00407v12018
  19. Point-SLAM: Dense Neural Point Cloud-based SLAM

    Erik Sandström, Yue Li, Luc Van Gool +1

    cs.CVarXiv:2304.04278v32023
  20. Neural Architecture Search on ImageNet in Four GPU Hours: A Theoretically Inspired Perspective

    Wuyang Chen, Xinyu Gong, Zhangyang Wang

    cs.CVcs.LGarXiv:2102.11535v42021
  21. NeW CRFs: Neural Window Fully-connected CRFs for Monocular Depth Estimation

    Weihao Yuan, Xiaodong Gu, Zuozhuo Dai +2

    cs.CVarXiv:2203.01502v22022
  22. Rethinking Visual Geo-localization for Large-Scale Applications

    Gabriele Berton, Carlo Masone, Barbara Caputo

    cs.CVarXiv:2204.02287v22022
  23. ParticleNet: Jet Tagging via Particle Clouds

    Huilin Qu, Loukas Gouskos

    hep-phcs.CVhep-exarXiv:1902.08570v32019
  24. AnatomyNet: Deep Learning for Fast and Fully Automated Whole-volume Segmentation of Head and Neck Anatomy

    Wentao Zhu, Yufang Huang, Liang Zeng +6

    cs.CVcs.LGcs.NEarXiv:1808.05238v22018
  25. Reconstruction of three-dimensional porous media using generative adversarial neural networks

    Lukas Mosser, Olivier Dubrule, Martin J. Blunt

    cs.CVcond-mat.mtrl-sciphysics.flu-dynarXiv:1704.03225v12017
  26. FPNN: Field Probing Neural Networks for 3D Data

    Yangyan Li, Soeren Pirk, Hao Su +2

    cs.CVarXiv:1605.06240v32016
  27. Self-supervised Pretraining of Visual Features in the Wild

    Priya Goyal, Mathilde Caron, Benjamin Lefaudeux +8

    cs.CVcs.AIarXiv:2103.01988v22021
  28. Unconstrained Face Verification using Deep CNN Features

    Jun-Cheng Chen, Vishal M. Patel, Rama Chellappa

    cs.CVarXiv:1508.01722v22015
  29. CLIPDraw: Exploring Text-to-Drawing Synthesis through Language-Image Encoders

    Kevin Frans, L. B. Soros, Olaf Witkowski

    cs.CVarXiv:2106.14843v12021
  30. Reviving Iterative Training with Mask Guidance for Interactive Segmentation

    Konstantin Sofiiuk, Ilia A. Petrov, Anton Konushin

    cs.CVarXiv:2102.06583v12021
  31. PIoU Loss: Towards Accurate Oriented Object Detection in Complex Environments

    Zhiming Chen, Kean Chen, Weiyao Lin +4

    cs.CVarXiv:2007.09584v12020
  32. DynaSLAM II: Tightly-Coupled Multi-Object Tracking and SLAM

    Berta Bescos, Carlos Campos, Juan D. Tardós +1

    cs.ROcs.CVarXiv:2010.07820v12020
  33. GLEAN: Generative Latent Bank for Large-Factor Image Super-Resolution

    Kelvin C. K. Chan, Xintao Wang, Xiangyu Xu +2

    cs.CVarXiv:2012.00739v12020
  34. An Efficient Sampling-based Method for Online Informative Path Planning in Unknown Environments

    Lukas Schmid, Michael Pantic, Raghav Khanna +3

    cs.ROcs.CVarXiv:1909.09548v22019
  35. Uncertainty-aware Joint Salient Object and Camouflaged Object Detection

    Aixuan Li, Jing Zhang, Yunqiu Lv +3

    cs.CVarXiv:2104.02628v12021
  36. SAR image despeckling through convolutional neural networks

    G. Chierchia, D. Cozzolino, G. Poggi +1

    cs.CVarXiv:1704.00275v22017
  37. Dash: Semi-Supervised Learning with Dynamic Thresholding

    Yi Xu, Lei Shang, Jinxing Ye +5

    cs.LGcs.CVstat.MLarXiv:2109.00650v12021
  38. A Convex Relaxation Barrier to Tight Robustness Verification of Neural Networks

    Hadi Salman, Greg Yang, Huan Zhang +2

    cs.LGcs.AIcs.CRarXiv:1902.08722v52019
  39. Uncertainty Inspired Underwater Image Enhancement

    Zhenqi Fu, Wu Wang, Yue Huang +2

    cs.CVarXiv:2207.09689v12022
  40. Dual Encoding for Zero-Example Video Retrieval

    Jianfeng Dong, Xirong Li, Chaoxi Xu +4

    cs.CVarXiv:1809.06181v32018
  41. ExpandNet: A Deep Convolutional Neural Network for High Dynamic Range Expansion from Low Dynamic Range Content

    Demetris Marnerides, Thomas Bashford-Rogers, Jonathan Hatchett +1

    cs.CVcs.GRarXiv:1803.02266v22018
  42. Seesaw Loss for Long-Tailed Instance Segmentation

    Jiaqi Wang, Wenwei Zhang, Yuhang Zang +7

    cs.CVarXiv:2008.10032v42020
  43. TransMVSNet: Global Context-aware Multi-view Stereo Network with Transformers

    Yikang Ding, Wentao Yuan, Qingtian Zhu +4

    cs.CVarXiv:2111.14600v12021
  44. Hierarchical Conditional Relation Networks for Video Question Answering

    Thao Minh Le, Vuong Le, Svetha Venkatesh +1

    cs.CVarXiv:2002.10698v32020
  45. Image Manipulation Detection by Multi-View Multi-Scale Supervision

    Xinru Chen, Chengbo Dong, Jiaqi Ji +2

    cs.CVcs.AIarXiv:2104.06832v22021
  46. NeRF-Editing: Geometry Editing of Neural Radiance Fields

    Yu-Jie Yuan, Yang-Tian Sun, Yu-Kun Lai +3

    cs.GRcs.CVarXiv:2205.04978v12022
  47. Learning Visual Commonsense for Robust Scene Graph Generation

    Alireza Zareian, Zhecan Wang, Haoxuan You +1

    cs.CVcs.LGarXiv:2006.09623v22020
  48. Unsupervised Learning by Predicting Noise

    Piotr Bojanowski, Armand Joulin

    stat.MLcs.CVcs.LGarXiv:1704.05310v12017
  49. Enhancing Adversarial Example Transferability with an Intermediate Level Attack

    Qian Huang, Isay Katsman, Horace He +3

    cs.LGcs.CRcs.CVarXiv:1907.10823v32019
  50. Adversarial Feature Hallucination Networks for Few-Shot Learning

    Kai Li, Yulun Zhang, Kunpeng Li +1

    cs.CVarXiv:2003.13193v22020
  51. CycleMorph: Cycle Consistent Unsupervised Deformable Image Registration

    Boah Kim, Dong Hwan Kim, Seong Ho Park +3

    cs.CVcs.LGeess.IVarXiv:2008.05772v12020
  52. EdgeViTs: Competing Light-weight CNNs on Mobile Devices with Vision Transformers

    Junting Pan, Adrian Bulat, Fuwen Tan +5

    cs.CVarXiv:2205.03436v22022
  53. Recursive Cascaded Networks for Unsupervised Medical Image Registration

    Shengyu Zhao, Yue Dong, Eric I-Chao Chang +1

    cs.CVarXiv:1907.12353v32019
  54. Graph HyperNetworks for Neural Architecture Search

    Chris Zhang, Mengye Ren, Raquel Urtasun

    cs.LGcs.CVstat.MLarXiv:1810.05749v32018
  55. Adaptive Wing Loss for Robust Face Alignment via Heatmap Regression

    Xinyao Wang, Liefeng Bo, Li Fuxin

    cs.CVarXiv:1904.07399v32019
  56. StoryGAN: A Sequential Conditional GAN for Story Visualization

    Yitong Li, Zhe Gan, Yelong Shen +6

    cs.CVarXiv:1812.02784v22018
  57. LEEP: A New Measure to Evaluate Transferability of Learned Representations

    Cuong V. Nguyen, Tal Hassner, Matthias Seeger +1

    cs.LGcs.CVstat.MLarXiv:2002.12462v22020
  58. Exploiting Temporal Contexts with Strided Transformer for 3D Human Pose Estimation

    Wenhao Li, Hong Liu, Runwei Ding +3

    cs.CVarXiv:2103.14304v82021
  59. Visually-Aware Fashion Recommendation and Design with Generative Image Models

    Wang-Cheng Kang, Chen Fang, Zhaowen Wang +1

    cs.CVcs.AIcs.HCarXiv:1711.02231v12017
  60. MUREL: Multimodal Relational Reasoning for Visual Question Answering

    Remi Cadene, Hedi Ben-younes, Matthieu Cord +1

    cs.CVcs.AIcs.CLarXiv:1902.09487v12019