Computer Vision and Pattern Recognition
Papers filed under cs.CV on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.
Search paper metadata (including unsummarized papers)
8,161 to 8,220 of 18,855
Learning to Count Objects in Natural Images for Visual Question Answering
Yan Zhang, Jonathon Hare, Adam Prügel-Bennett
cs.CVcs.CLarXiv:1802.05766v12018Gradient Centralization: A New Optimization Technique for Deep Neural Networks
Hongwei Yong, Jianqiang Huang, Xiansheng Hua +1
cs.CVarXiv:2004.01461v22020Dimensionality Reduction on SPD Manifolds: The Emergence of Geometry-Aware Methods
Mehrtash Harandi, Mathieu Salzmann, Richard Hartley
cs.CVarXiv:1605.06182v12016Semantic Adversarial Examples
Hossein Hosseini, Radha Poovendran
cs.CVcs.AIcs.LGarXiv:1804.00499v12018Text-Adaptive Generative Adversarial Networks: Manipulating Images with Natural Language
Seonghyeon Nam, Yunji Kim, Seon Joo Kim
cs.CVarXiv:1810.11919v22018ForgeryNet: A Versatile Benchmark for Comprehensive Forgery Analysis
Yinan He, Bei Gan, Siyu Chen +6
cs.CVcs.LGarXiv:2103.05630v22021Robust Collaborative 3D Object Detection in Presence of Pose Errors
Yifan Lu, Quanhao Li, Baoan Liu +4
cs.CVcs.MAcs.ROarXiv:2211.07214v32022IFRNet: Intermediate Feature Refine Network for Efficient Frame Interpolation
Lingtong Kong, Boyuan Jiang, Donghao Luo +5
cs.CVarXiv:2205.14620v12022Fast and Robust Multi-Person 3D Pose Estimation from Multiple Views
Junting Dong, Wen Jiang, Qixing Huang +2
cs.CVarXiv:1901.04111v12019Monocular Dynamic View Synthesis: A Reality Check
Hang Gao, Ruilong Li, Shubham Tulsiani +2
cs.CVarXiv:2210.13445v12022Neural Video Compression with Feature Modulation
Jiahao Li, Bin Li, Yan Lu
cs.CVeess.IVarXiv:2402.17414v22024APT: Anchor-aligned Perturbations for Tamper Localization in Fully Regenerated Images
Suhyeon Ha, Woo Jae Kim, Joonsung Jeon +2
cs.CVarXiv:2608.30656v12026A Convolutional Neural Network Neutrino Event Classifier
A. Aurisano, A. Radovic, D. Rocco +7
hep-excs.CVarXiv:1604.01444v32016Fast Learning of Temporal Action Proposal via Dense Boundary Generator
Chuming Lin, Jian Li, Yabiao Wang +7
cs.CVarXiv:1911.04127v12019HEp-2 Cell Image Classification with Deep Convolutional Neural Networks
Zhimin Gao, Lei Wang, Luping Zhou +1
cs.CVarXiv:1504.02531v22015SOFT: Softmax-free Transformer with Linear Complexity
Jiachen Lu, Jinghan Yao, Junge Zhang +6
cs.CVcs.AIcs.LGarXiv:2110.11945v32021Blind Image Super-Resolution: A Survey and Beyond
Anran Liu, Yihao Liu, Jinjin Gu +2
cs.CVarXiv:2107.03055v12021Summaries:한국어ImageCAS-X: a dataset and benchmark for coronary artery segmentation and centerline extraction in coronary CT angiography
Kit M. Bransby, Esther Øksnebjerg, Kristoffer Kjær +7
cs.CVcs.AIarXiv:2608.30404v12026Transfer Learning from Synthetic to Real-Noise Denoising with Adaptive Instance Normalization
Yoonsik Kim, Jae Woong Soh, Gu Yong Park +1
cs.CVeess.IVarXiv:2002.11244v22020The Surprising Effectiveness of Representation Learning for Visual Imitation
Jyothish Pari, Nur Muhammad Shafiullah, Sridhar Pandian Arunachalam +1
cs.ROcs.AIcs.CVarXiv:2112.01511v22021Defensive Quantization: When Efficiency Meets Robustness
Ji Lin, Chuang Gan, Song Han
cs.LGcs.CVstat.MLarXiv:1904.08444v12019MR-JEPA: A General Purpose Video Foundation Model for Cardiac MRI
Athira J. Jacob, Puneet Sharma, Dorin Comaniciu +1
cs.CVcs.AIarXiv:2608.30975v12026SIGMA: Semantic-complete Graph Matching for Domain Adaptive Object Detection
Wuyang Li, Xinyu Liu, Yixuan Yuan
cs.CVarXiv:2203.06398v32022MARS: An Instance-aware, Modular and Realistic Simulator for Autonomous Driving
Zirui Wu, Tianyu Liu, Liyi Luo +13
cs.CVarXiv:2307.15058v12023Arbitrary Style Transfer via Multi-Adaptation Network
Yingying Deng, Fan Tang, Weiming Dong +3
cs.CVcs.AIarXiv:2005.13219v22020VidTr: Video Transformer Without Convolutions
Yanyi Zhang, Xinyu Li, Chunhui Liu +6
cs.CVarXiv:2104.11746v22021Fully Convolutional Networks with Sequential Information for Robust Crop and Weed Detection in Precision Farming
Philipp Lottes, Jens Behley, Andres Milioto +1
cs.CVarXiv:1806.03412v12018Autoregressive Mosaics: Probing 2D Spatial Reasoning in Text-Only Language Models
Ashwin Nedungadi, Stefan Oehmcke, Stefan Lüdtke
cs.AIcs.CVarXiv:2608.30751v22026VeriCam: A Verification Baseline for the Classification of Unknown Data
Lucas Wojcik, Gabriel E. Lima, Sergio M. Silva +2
cs.CVarXiv:2608.31107v12026XNOR-Net++: Improved Binary Neural Networks
Adrian Bulat, Georgios Tzimiropoulos
cs.CVcs.LGeess.IVarXiv:1909.13863v12019Real-Time Video Anomaly Detection Using YOLO Pose Estimation and CLIP-Based Semantic Scoring
Vanodhya G. Warnasooriya, Amir Hajian, Watchara Ruangsang +1
cs.CVcs.AIeess.IVarXiv:2608.31074v12026DePlot: One-shot visual language reasoning by plot-to-table translation
Fangyu Liu, Julian Martin Eisenschlos, Francesco Piccinno +7
cs.CLcs.AIcs.CVarXiv:2212.10505v22022Learning Canonical Shape Space for Category-Level 6D Object Pose and Size Estimation
Dengsheng Chen, Jun Li, Zheng Wang +1
cs.CVarXiv:2001.09322v32020SER-FIQ: Unsupervised Estimation of Face Image Quality Based on Stochastic Embedding Robustness
Philipp Terhörst, Jan Niklas Kolf, Naser Damer +2
cs.CVarXiv:2003.09373v12020End-to-End Learning of Motion Representation for Video Understanding
Lijie Fan, Wenbing Huang, Chuang Gan +3
cs.CVarXiv:1804.00413v12018Domain-invariant Stereo Matching Networks
Feihu Zhang, Xiaojuan Qi, Ruigang Yang +3
cs.CVarXiv:1911.13287v12019APQ: Joint Search for Network Architecture, Pruning and Quantization Policy
Tianzhe Wang, Kuan Wang, Han Cai +3
cs.LGcs.CVstat.MLarXiv:2006.08509v12020V2X-Seq: A Large-Scale Sequential Dataset for Vehicle-Infrastructure Cooperative Perception and Forecasting
Haibao Yu, Wenxian Yang, Hongzhi Ruan +11
cs.CVcs.AIarXiv:2305.05938v12023Guaranteed Outlier Removal for Point Cloud Registration with Correspondences
Álvaro Parra Bustos, Tat-Jun Chin
cs.CVarXiv:1711.10209v12017RIDCP: Revitalizing Real Image Dehazing via High-Quality Codebook Priors
Rui-Qi Wu, Zheng-Peng Duan, Chun-Le Guo +2
cs.CVarXiv:2304.03994v12023Train in Germany, Test in The USA: Making 3D Object Detectors Generalize
Yan Wang, Xiangyu Chen, Yurong You +5
cs.CVarXiv:2005.08139v12020Leveraging Photometric Consistency over Time for Sparsely Supervised Hand-Object Reconstruction
Yana Hasson, Bugra Tekin, Federica Bogo +3
cs.CVarXiv:2004.13449v12020Modeling Indirect Illumination for Inverse Rendering
Yuanqing Zhang, Jiaming Sun, Xingyi He +3
cs.CVarXiv:2204.06837v12022Contextual-based Image Inpainting: Infer, Match, and Translate
Yuhang Song, Chao Yang, Zhe Lin +4
cs.CVarXiv:1711.08590v52017LISynSeg: Data-Centric Label-to-Image Synthesis for Cross-Modality Whole-Heart Segmentation
Jiacheng Wang, Ivana Isgum, Ipek Oguz
cs.CVeess.IVarXiv:2608.31073v12026Online Human Action Detection using Joint Classification-Regression Recurrent Neural Networks
Yanghao Li, Cuiling Lan, Junliang Xing +3
cs.CVarXiv:1604.05633v22016Training Sparse Neural Networks
Suraj Srinivas, Akshayvarun Subramanya, R. Venkatesh Babu
cs.CVcs.LGarXiv:1611.06694v12016CedarCypress3D: an annotated UAV-LiDAR dataset of individual trees in planted cedar and cypress forests
Katsuto Shimizu, Fumiaki Kitahara, Tomohiro Nishizono +8
cs.CVarXiv:2608.30149v12026Compositional Explanations of Neurons
Jesse Mu, Jacob Andreas
cs.LGcs.AIcs.CLarXiv:2006.14032v22020CentralNet: a Multilayer Approach for Multimodal Fusion
Valentin Vielzeuf, Alexis Lechervy, Stéphane Pateux +1
cs.AIcs.CVcs.MMarXiv:1808.07275v12018Vision Models Predict Urban Scene Appraisal with Limited Neural Alignment
Kaizhen Tan, Yuantao Deng
cs.CVarXiv:2608.30964v12026DeepCaps: Going Deeper with Capsule Networks
Jathushan Rajasegaran, Vinoj Jayasundara, Sandaru Jayasekara +3
cs.CVarXiv:1904.09546v12019DreamGaussian4D: Generative 4D Gaussian Splatting
Jiawei Ren, Liang Pan, Jiaxiang Tang +4
cs.CVcs.GRarXiv:2312.17142v32023APRIL-GAN: A Zero-/Few-Shot Anomaly Classification and Segmentation Method for CVPR 2023 VAND Workshop Challenge Tracks 1&2: 1st Place on Zero-shot AD and 4th Place on Few-shot AD
Xuhai Chen, Yue Han, Jiangning Zhang
cs.CVarXiv:2305.17382v32023Adversarial Example Does Good: Preventing Painting Imitation from Diffusion Models via Adversarial Examples
Chumeng Liang, Xiaoyu Wu, Yang Hua +6
cs.CVcs.AIcs.CRarXiv:2302.04578v22023Generative Adversarial Network for Medical Images (MI-GAN)
Talha Iqbal, Hazrat Ali
cs.LGcs.CVeess.IVarXiv:1810.00551v12018Natural and Effective Obfuscation by Head Inpainting
Qianru Sun, Liqian Ma, Seong Joon Oh +3
cs.CVcs.CRcs.CYarXiv:1711.09001v52017MonoCap: Monocular Human Motion Capture using a CNN Coupled with a Geometric Prior
Xiaowei Zhou, Menglong Zhu, Georgios Pavlakos +3
cs.CVarXiv:1701.02354v22017SeaDronesSee: A Maritime Benchmark for Detecting Humans in Open Water
Leon Amadeus Varga, Benjamin Kiefer, Martin Messmer +1
cs.CVarXiv:2105.01922v22021VinDr-Mammo: A large-scale benchmark dataset for computer-aided diagnosis in full-field digital mammography
Hieu T. Nguyen, Ha Q. Nguyen, Hieu H. Pham +4
eess.IVcs.CVarXiv:2203.11205v22022