Computer Vision and Pattern Recognition
Papers filed under cs.CV on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.
Search paper metadata (including unsummarized papers)
10,921 to 10,980 of 18,801
ROAD: Reality Oriented Adaptation for Semantic Segmentation of Urban Scenes
Yuhua Chen, Wen Li, Luc Van Gool
cs.CVarXiv:1711.11556v22017Diffusion models as plug-and-play priors
Alexandros Graikos, Nikolay Malkin, Nebojsa Jojic +1
cs.LGcs.CVarXiv:2206.09012v32022Universal Adversarial Perturbations Against Semantic Image Segmentation
Jan Hendrik Metzen, Mummadi Chaithanya Kumar, Thomas Brox +1
stat.MLcs.AIcs.CVarXiv:1704.05712v32017Do Convnets Learn Correspondence?
Jonathan Long, Ning Zhang, Trevor Darrell
cs.CVcs.LGcs.NEarXiv:1411.1091v12014Auxiliary Tasks in Multi-task Learning
Lukas Liebel, Marco Körner
cs.CVcs.LGarXiv:1805.06334v22018Transferring Rich Feature Hierarchies for Robust Visual Tracking
Naiyan Wang, Siyi Li, Abhinav Gupta +1
cs.CVcs.NEarXiv:1501.04587v22015Loss Functions for Neural Networks for Image Processing
Hang Zhao, Orazio Gallo, Iuri Frosio +1
cs.CVarXiv:1511.08861v32015A-NeRF: Articulated Neural Radiance Fields for Learning Human Shape, Appearance, and Pose
Shih-Yang Su, Frank Yu, Michael Zollhoefer +1
cs.CVcs.GRarXiv:2102.06199v32021Projected GANs Converge Faster
Axel Sauer, Kashyap Chitta, Jens Müller +1
cs.CVcs.LGarXiv:2111.01007v12021TopFormer: Token Pyramid Transformer for Mobile Semantic Segmentation
Wenqiang Zhang, Zilong Huang, Guozhong Luo +5
cs.CVarXiv:2204.05525v12022Robust Subspace Learning: Robust PCA, Robust Subspace Tracking, and Robust Subspace Recovery
Namrata Vaswani, Thierry Bouwmans, Sajid Javed +1
cs.ITcs.CVstat.MEarXiv:1711.09492v42017Latent Action Pretraining from Videos
Seonghyeon Ye, Joel Jang, Byeongguk Jeon +13
cs.ROcs.CLcs.CVarXiv:2410.11758v22024ASpanFormer: Detector-Free Image Matching with Adaptive Span Transformer
Hongkai Chen, Zixin Luo, Lei Zhou +6
cs.CVarXiv:2208.14201v12022LSQ+: Improving low-bit quantization through learnable offsets and better initialization
Yash Bhalgat, Jinwon Lee, Markus Nagel +2
cs.CVcs.LGstat.MLarXiv:2004.09576v12020DeepSentiBank: Visual Sentiment Concept Classification with Deep Convolutional Neural Networks
Tao Chen, Damian Borth, Trevor Darrell +1
cs.CVcs.LGcs.MMarXiv:1410.8586v12014Robot Parkour Learning
Ziwen Zhuang, Zipeng Fu, Jianren Wang +4
cs.ROcs.AIcs.CVarXiv:2309.05665v22023Generating Natural Questions About an Image
Nasrin Mostafazadeh, Ishan Misra, Jacob Devlin +3
cs.CLcs.AIcs.CVarXiv:1603.06059v32016Single Shot Text Detector with Regional Attention
Pan He, Weilin Huang, Tong He +3
cs.CVarXiv:1709.00138v12017Learning N:M Fine-grained Structured Sparse Neural Networks From Scratch
Aojun Zhou, Yukun Ma, Junnan Zhu +5
cs.CVcs.ARarXiv:2102.04010v22021Efficient Diffusion Training via Min-SNR Weighting Strategy
Tiankai Hang, Shuyang Gu, Chen Li +5
cs.CVarXiv:2303.09556v32023End-to-end Audio-visual Speech Recognition with Conformers
Pingchuan Ma, Stavros Petridis, Maja Pantic
cs.CVeess.ASarXiv:2102.06657v12021TailorNet: Predicting Clothing in 3D as a Function of Human Pose, Shape and Garment Style
Chaitanya Patel, Zhouyingcheng Liao, Gerard Pons-Moll
cs.CVcs.GRarXiv:2003.04583v22020Infinity: Scaling Bitwise AutoRegressive Modeling for High-Resolution Image Synthesis
Jian Han, Jinlai Liu, Yi Jiang +5
cs.CVarXiv:2412.04431v22024Weakly-supervised Disentangling with Recurrent Transformations for 3D View Synthesis
Jimei Yang, Scott Reed, Ming-Hsuan Yang +1
cs.LGcs.AIcs.CVarXiv:1601.00706v12016MSR-net:Low-light Image Enhancement Using Deep Convolutional Network
Liang Shen, Zihan Yue, Fan Feng +3
cs.CVarXiv:1711.02488v12017Seeing Out of tHe bOx: End-to-End Pre-training for Vision-Language Representation Learning
Zhicheng Huang, Zhaoyang Zeng, Yupan Huang +3
cs.CVarXiv:2104.03135v22021IconQA: A New Benchmark for Abstract Diagram Understanding and Visual Language Reasoning
Pan Lu, Liang Qiu, Jiaqi Chen +6
cs.CVcs.AIcs.CLarXiv:2110.13214v42021SCTransNet: Spatial-channel Cross Transformer Network for Infrared Small Target Detection
Shuai Yuan, Hanlin Qin, Xiang Yan +2
cs.CVarXiv:2401.15583v32024Single-source Domain Expansion Network for Cross-Scene Hyperspectral Image Classification
Yuxiang Zhang, Wei Li, Weidong Sun +2
cs.CVarXiv:2209.01634v12022SteganoGAN: High Capacity Image Steganography with GANs
Kevin Alex Zhang, Alfredo Cuesta-Infante, Lei Xu +1
cs.CVcs.LGcs.MMarXiv:1901.03892v22019DeepSeeNet: A deep learning model for automated classification of patient-based age-related macular degeneration severity from color fundus photographs
Yifan Peng, Shazia Dharssi, Qingyu Chen +5
cs.CVarXiv:1811.07492v22018Quaternion Fourier Transform on Quaternion Fields and Generalizations
Eckhard Hitzer
math.RAcs.CVmath-pharXiv:1306.1023v12013Deep Multimodal Fusion by Channel Exchanging
Yikai Wang, Wenbing Huang, Fuchun Sun +3
cs.CVcs.LGarXiv:2011.05005v22020Multimodal Residual Learning for Visual QA
Jin-Hwa Kim, Sang-Woo Lee, Dong-Hyun Kwak +4
cs.CVarXiv:1606.01455v22016ZegCLIP: Towards Adapting CLIP for Zero-shot Semantic Segmentation
Ziqin Zhou, Bowen Zhang, Yinjie Lei +2
cs.CVarXiv:2212.03588v32022Focal Sparse Convolutional Networks for 3D Object Detection
Yukang Chen, Yanwei Li, Xiangyu Zhang +2
cs.CVcs.LGarXiv:2204.12463v12022MDTv2: Masked Diffusion Transformer is a Strong Image Synthesizer
Shanghua Gao, Pan Zhou, Ming-Ming Cheng +1
cs.CVarXiv:2303.14389v22023AMP-Net: Denoising based Deep Unfolding for Compressive Image Sensing
Zhonghao Zhang, Yipeng Liu, Jiani Liu +2
eess.IVcs.CVarXiv:2004.10078v22020Cluster Contrast for Unsupervised Person Re-Identification
Zuozhuo Dai, Guangyuan Wang, Weihao Yuan +3
cs.CVarXiv:2103.11568v42021Review: deep learning on 3D point clouds
Saifullahi Aminu Bello, Shangshu Yu, Cheng Wang
cs.CVarXiv:2001.06280v12020Going Deeper into First-Person Activity Recognition
Minghuang Ma, Haoqi Fan, Kris M. Kitani
cs.CVarXiv:1605.03688v12016MetaFormer Baselines for Vision
Weihao Yu, Chenyang Si, Pan Zhou +5
cs.CVcs.AIcs.LGarXiv:2210.13452v42022CFA: Coupled-hypersphere-based Feature Adaptation for Target-Oriented Anomaly Localization
Sungwook Lee, Seunghyun Lee, Byung Cheol Song
cs.CVcs.LGarXiv:2206.04325v12022Fast Hyperspectral Image Denoising and Inpainting Based on Low-Rank and Sparse Representations
Lina Zhuang, Jose M. Bioucas-Dias
eess.IVcs.CVarXiv:2103.06842v12021Learning Depth-Guided Convolutions for Monocular 3D Object Detection
Mingyu Ding, Yuqi Huo, Hongwei Yi +4
cs.CVarXiv:1912.04799v22019DeepHuman: 3D Human Reconstruction from a Single Image
Zerong Zheng, Tao Yu, Yixuan Wei +2
cs.CVarXiv:1903.06473v22019GSPN: Generative Shape Proposal Network for 3D Instance Segmentation in Point Cloud
Li Yi, Wang Zhao, He Wang +2
cs.CVarXiv:1812.03320v12018Neural Descriptor Fields: SE(3)-Equivariant Object Representations for Manipulation
Anthony Simeonov, Yilun Du, Andrea Tagliasacchi +4
cs.ROcs.AIcs.CVarXiv:2112.05124v12021Language-Grounded Indoor 3D Semantic Segmentation in the Wild
David Rozenberszki, Or Litany, Angela Dai
cs.CVarXiv:2204.07761v22022V2V4Real: A Real-world Large-scale Dataset for Vehicle-to-Vehicle Cooperative Perception
Runsheng Xu, Xin Xia, Jinlong Li +10
cs.CVarXiv:2303.07601v22023A Review and Comparative Study on Probabilistic Object Detection in Autonomous Driving
Di Feng, Ali Harakeh, Steven Waslander +1
cs.CVcs.ROarXiv:2011.10671v22020Interpreting Deep Visual Representations via Network Dissection
Bolei Zhou, David Bau, Aude Oliva +1
cs.CVarXiv:1711.05611v22017ConvNet Architecture Search for Spatiotemporal Feature Learning
Du Tran, Jamie Ray, Zheng Shou +2
cs.CVarXiv:1708.05038v12017SafetyNet: Detecting and Rejecting Adversarial Examples Robustly
Jiajun Lu, Theerasit Issaranon, David Forsyth
cs.CVcs.LGarXiv:1704.00103v22017Interpretable Learning for Self-Driving Cars by Visualizing Causal Attention
Jinkyu Kim, John Canny
cs.CVcs.LGarXiv:1703.10631v12017DR2-Net: Deep Residual Reconstruction Network for Image Compressive Sensing
Hantao Yao, Feng Dai, Dongming Zhang +4
cs.CVarXiv:1702.05743v42017Non-Local Color Image Denoising with Convolutional Neural Networks
Stamatios Lefkimmiatis
cs.CVcs.AIarXiv:1611.06757v220163D Object Proposals using Stereo Imagery for Accurate Object Class Detection
Xiaozhi Chen, Kaustav Kundu, Yukun Zhu +3
cs.CVarXiv:1608.07711v22016Structured Domain Randomization: Bridging the Reality Gap by Context-Aware Synthetic Data
Aayush Prakash, Shaad Boochoon, Mark Brophy +5
cs.CVarXiv:1810.10093v22018Object Contour Detection with a Fully Convolutional Encoder-Decoder Network
Jimei Yang, Brian Price, Scott Cohen +2
cs.CVcs.LGarXiv:1603.04530v12016