Computer Vision and Pattern Recognition
Papers filed under cs.CV on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.
Search paper metadata (including unsummarized papers)
10,801 to 10,860 of 18,981
Partial success in closing the gap between human and machine vision
Robert Geirhos, Kantharaju Narayanappa, Benjamin Mitzkus +4
cs.CVcs.AIcs.LGarXiv:2106.07411v22021Hierarchical Open-Vocabulary 3D Scene Graphs for Language-Grounded Robot Navigation
Abdelrhman Werby, Chenguang Huang, Martin Büchner +2
cs.ROcs.AIcs.CLarXiv:2403.17846v22024More is Less: A More Complicated Network with Less Inference Complexity
Xuanyi Dong, Junshi Huang, Yi Yang +1
cs.CVarXiv:1703.08651v22017A Bio-Inspired Multi-Exposure Fusion Framework for Low-light Image Enhancement
Zhenqiang Ying, Ge Li, Wen Gao
cs.CVarXiv:1711.00591v12017VisualGPT: Data-efficient Adaptation of Pretrained Language Models for Image Captioning
Jun Chen, Han Guo, Kai Yi +2
cs.CVcs.AIcs.CLarXiv:2102.10407v52021Self-Erasing Network for Integral Object Attention
Qibin Hou, Peng-Tao Jiang, Yunchao Wei +1
cs.CVarXiv:1810.09821v12018Investigating Bi-Level Optimization for Learning and Vision from a Unified Perspective: A Survey and Beyond
Risheng Liu, Jiaxin Gao, Jin Zhang +2
cs.LGcs.CVmath.DSarXiv:2101.11517v32021Editing Conditional Radiance Fields
Steven Liu, Xiuming Zhang, Zhoutong Zhang +3
cs.CVcs.GRcs.LGarXiv:2105.06466v22021Learning to Self-Train for Semi-Supervised Few-Shot Classification
Xinzhe Li, Qianru Sun, Yaoyao Liu +4
cs.CVcs.LGstat.MLarXiv:1906.00562v22019MVTN: Multi-View Transformation Network for 3D Shape Recognition
Abdullah Hamdi, Silvio Giancola, Bernard Ghanem
cs.CVcs.LGarXiv:2011.13244v32020Video Object Segmentation with Episodic Graph Memory Networks
Xiankai Lu, Wenguan Wang, Martin Danelljan +3
cs.CVcs.LGarXiv:2007.07020v42020CityGaussian: Real-time High-quality Large-Scale Scene Rendering with Gaussians
Yang Liu, He Guan, Chuanchen Luo +4
cs.CVarXiv:2404.01133v32024Yedrouj-Net: An efficient CNN for spatial steganalysis
Mehdi Yedroudj, Frederic Comby, Marc Chaumont
cs.CVcs.CRarXiv:1803.00407v12018Point-SLAM: Dense Neural Point Cloud-based SLAM
Erik Sandström, Yue Li, Luc Van Gool +1
cs.CVarXiv:2304.04278v32023Neural Architecture Search on ImageNet in Four GPU Hours: A Theoretically Inspired Perspective
Wuyang Chen, Xinyu Gong, Zhangyang Wang
cs.CVcs.LGarXiv:2102.11535v42021NeW CRFs: Neural Window Fully-connected CRFs for Monocular Depth Estimation
Weihao Yuan, Xiaodong Gu, Zuozhuo Dai +2
cs.CVarXiv:2203.01502v22022Rethinking Visual Geo-localization for Large-Scale Applications
Gabriele Berton, Carlo Masone, Barbara Caputo
cs.CVarXiv:2204.02287v22022ParticleNet: Jet Tagging via Particle Clouds
Huilin Qu, Loukas Gouskos
hep-phcs.CVhep-exarXiv:1902.08570v32019AnatomyNet: Deep Learning for Fast and Fully Automated Whole-volume Segmentation of Head and Neck Anatomy
Wentao Zhu, Yufang Huang, Liang Zeng +6
cs.CVcs.LGcs.NEarXiv:1808.05238v22018Reconstruction of three-dimensional porous media using generative adversarial neural networks
Lukas Mosser, Olivier Dubrule, Martin J. Blunt
cs.CVcond-mat.mtrl-sciphysics.flu-dynarXiv:1704.03225v12017FPNN: Field Probing Neural Networks for 3D Data
Yangyan Li, Soeren Pirk, Hao Su +2
cs.CVarXiv:1605.06240v32016Self-supervised Pretraining of Visual Features in the Wild
Priya Goyal, Mathilde Caron, Benjamin Lefaudeux +8
cs.CVcs.AIarXiv:2103.01988v22021Unconstrained Face Verification using Deep CNN Features
Jun-Cheng Chen, Vishal M. Patel, Rama Chellappa
cs.CVarXiv:1508.01722v22015CLIPDraw: Exploring Text-to-Drawing Synthesis through Language-Image Encoders
Kevin Frans, L. B. Soros, Olaf Witkowski
cs.CVarXiv:2106.14843v12021Reviving Iterative Training with Mask Guidance for Interactive Segmentation
Konstantin Sofiiuk, Ilia A. Petrov, Anton Konushin
cs.CVarXiv:2102.06583v12021PIoU Loss: Towards Accurate Oriented Object Detection in Complex Environments
Zhiming Chen, Kean Chen, Weiyao Lin +4
cs.CVarXiv:2007.09584v12020DynaSLAM II: Tightly-Coupled Multi-Object Tracking and SLAM
Berta Bescos, Carlos Campos, Juan D. Tardós +1
cs.ROcs.CVarXiv:2010.07820v12020GLEAN: Generative Latent Bank for Large-Factor Image Super-Resolution
Kelvin C. K. Chan, Xintao Wang, Xiangyu Xu +2
cs.CVarXiv:2012.00739v12020An Efficient Sampling-based Method for Online Informative Path Planning in Unknown Environments
Lukas Schmid, Michael Pantic, Raghav Khanna +3
cs.ROcs.CVarXiv:1909.09548v22019Uncertainty-aware Joint Salient Object and Camouflaged Object Detection
Aixuan Li, Jing Zhang, Yunqiu Lv +3
cs.CVarXiv:2104.02628v12021SAR image despeckling through convolutional neural networks
G. Chierchia, D. Cozzolino, G. Poggi +1
cs.CVarXiv:1704.00275v22017Dash: Semi-Supervised Learning with Dynamic Thresholding
Yi Xu, Lei Shang, Jinxing Ye +5
cs.LGcs.CVstat.MLarXiv:2109.00650v12021A Convex Relaxation Barrier to Tight Robustness Verification of Neural Networks
Hadi Salman, Greg Yang, Huan Zhang +2
cs.LGcs.AIcs.CRarXiv:1902.08722v52019Uncertainty Inspired Underwater Image Enhancement
Zhenqi Fu, Wu Wang, Yue Huang +2
cs.CVarXiv:2207.09689v12022Dual Encoding for Zero-Example Video Retrieval
Jianfeng Dong, Xirong Li, Chaoxi Xu +4
cs.CVarXiv:1809.06181v32018ExpandNet: A Deep Convolutional Neural Network for High Dynamic Range Expansion from Low Dynamic Range Content
Demetris Marnerides, Thomas Bashford-Rogers, Jonathan Hatchett +1
cs.CVcs.GRarXiv:1803.02266v22018Seesaw Loss for Long-Tailed Instance Segmentation
Jiaqi Wang, Wenwei Zhang, Yuhang Zang +7
cs.CVarXiv:2008.10032v42020TransMVSNet: Global Context-aware Multi-view Stereo Network with Transformers
Yikang Ding, Wentao Yuan, Qingtian Zhu +4
cs.CVarXiv:2111.14600v12021Hierarchical Conditional Relation Networks for Video Question Answering
Thao Minh Le, Vuong Le, Svetha Venkatesh +1
cs.CVarXiv:2002.10698v32020Image Manipulation Detection by Multi-View Multi-Scale Supervision
Xinru Chen, Chengbo Dong, Jiaqi Ji +2
cs.CVcs.AIarXiv:2104.06832v22021NeRF-Editing: Geometry Editing of Neural Radiance Fields
Yu-Jie Yuan, Yang-Tian Sun, Yu-Kun Lai +3
cs.GRcs.CVarXiv:2205.04978v12022Learning Visual Commonsense for Robust Scene Graph Generation
Alireza Zareian, Zhecan Wang, Haoxuan You +1
cs.CVcs.LGarXiv:2006.09623v22020Unsupervised Learning by Predicting Noise
Piotr Bojanowski, Armand Joulin
stat.MLcs.CVcs.LGarXiv:1704.05310v12017Enhancing Adversarial Example Transferability with an Intermediate Level Attack
Qian Huang, Isay Katsman, Horace He +3
cs.LGcs.CRcs.CVarXiv:1907.10823v32019Adversarial Feature Hallucination Networks for Few-Shot Learning
Kai Li, Yulun Zhang, Kunpeng Li +1
cs.CVarXiv:2003.13193v22020CycleMorph: Cycle Consistent Unsupervised Deformable Image Registration
Boah Kim, Dong Hwan Kim, Seong Ho Park +3
cs.CVcs.LGeess.IVarXiv:2008.05772v12020EdgeViTs: Competing Light-weight CNNs on Mobile Devices with Vision Transformers
Junting Pan, Adrian Bulat, Fuwen Tan +5
cs.CVarXiv:2205.03436v22022Recursive Cascaded Networks for Unsupervised Medical Image Registration
Shengyu Zhao, Yue Dong, Eric I-Chao Chang +1
cs.CVarXiv:1907.12353v32019Graph HyperNetworks for Neural Architecture Search
Chris Zhang, Mengye Ren, Raquel Urtasun
cs.LGcs.CVstat.MLarXiv:1810.05749v32018Adaptive Wing Loss for Robust Face Alignment via Heatmap Regression
Xinyao Wang, Liefeng Bo, Li Fuxin
cs.CVarXiv:1904.07399v32019StoryGAN: A Sequential Conditional GAN for Story Visualization
Yitong Li, Zhe Gan, Yelong Shen +6
cs.CVarXiv:1812.02784v22018LEEP: A New Measure to Evaluate Transferability of Learned Representations
Cuong V. Nguyen, Tal Hassner, Matthias Seeger +1
cs.LGcs.CVstat.MLarXiv:2002.12462v22020Exploiting Temporal Contexts with Strided Transformer for 3D Human Pose Estimation
Wenhao Li, Hong Liu, Runwei Ding +3
cs.CVarXiv:2103.14304v82021Visually-Aware Fashion Recommendation and Design with Generative Image Models
Wang-Cheng Kang, Chen Fang, Zhaowen Wang +1
cs.CVcs.AIcs.HCarXiv:1711.02231v12017MUREL: Multimodal Relational Reasoning for Visual Question Answering
Remi Cadene, Hedi Ben-younes, Matthieu Cord +1
cs.CVcs.AIcs.CLarXiv:1902.09487v12019Domain Agnostic Learning with Disentangled Representations
Xingchao Peng, Zijun Huang, Ximeng Sun +1
cs.CVcs.LGarXiv:1904.12347v12019PointOdyssey: A Large-Scale Synthetic Dataset for Long-Term Point Tracking
Yang Zheng, Adam W. Harley, Bokui Shen +2
cs.CVarXiv:2307.15055v12023A Large-Scale Study on Unsupervised Spatiotemporal Representation Learning
Christoph Feichtenhofer, Haoqi Fan, Bo Xiong +2
cs.CVcs.AIcs.LGarXiv:2104.14558v12021Are adversarial examples inevitable?
Ali Shafahi, W. Ronny Huang, Christoph Studer +2
cs.LGcs.CVstat.MLarXiv:1809.02104v32018Deep Multi-instance Networks with Sparse Label Assignment for Whole Mammogram Classification
Wentao Zhu, Qi Lou, Yeeleng Scott Vang +1
cs.CVcs.LGarXiv:1612.05968v12016