Computer Vision and Pattern Recognition

Papers filed under cs.CV on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.

Search paper metadata (including unsummarized papers)

8,161 to 8,220 of 18,855

  1. Learning to Count Objects in Natural Images for Visual Question Answering

    Yan Zhang, Jonathon Hare, Adam Prügel-Bennett

    cs.CVcs.CLarXiv:1802.05766v12018
  2. Gradient Centralization: A New Optimization Technique for Deep Neural Networks

    Hongwei Yong, Jianqiang Huang, Xiansheng Hua +1

    cs.CVarXiv:2004.01461v22020
  3. Dimensionality Reduction on SPD Manifolds: The Emergence of Geometry-Aware Methods

    Mehrtash Harandi, Mathieu Salzmann, Richard Hartley

    cs.CVarXiv:1605.06182v12016
  4. Semantic Adversarial Examples

    Hossein Hosseini, Radha Poovendran

    cs.CVcs.AIcs.LGarXiv:1804.00499v12018
  5. Text-Adaptive Generative Adversarial Networks: Manipulating Images with Natural Language

    Seonghyeon Nam, Yunji Kim, Seon Joo Kim

    cs.CVarXiv:1810.11919v22018
  6. ForgeryNet: A Versatile Benchmark for Comprehensive Forgery Analysis

    Yinan He, Bei Gan, Siyu Chen +6

    cs.CVcs.LGarXiv:2103.05630v22021
  7. Robust Collaborative 3D Object Detection in Presence of Pose Errors

    Yifan Lu, Quanhao Li, Baoan Liu +4

    cs.CVcs.MAcs.ROarXiv:2211.07214v32022
  8. IFRNet: Intermediate Feature Refine Network for Efficient Frame Interpolation

    Lingtong Kong, Boyuan Jiang, Donghao Luo +5

    cs.CVarXiv:2205.14620v12022
  9. Fast and Robust Multi-Person 3D Pose Estimation from Multiple Views

    Junting Dong, Wen Jiang, Qixing Huang +2

    cs.CVarXiv:1901.04111v12019
  10. Monocular Dynamic View Synthesis: A Reality Check

    Hang Gao, Ruilong Li, Shubham Tulsiani +2

    cs.CVarXiv:2210.13445v12022
  11. Neural Video Compression with Feature Modulation

    Jiahao Li, Bin Li, Yan Lu

    cs.CVeess.IVarXiv:2402.17414v22024
  12. APT: Anchor-aligned Perturbations for Tamper Localization in Fully Regenerated Images

    Suhyeon Ha, Woo Jae Kim, Joonsung Jeon +2

    cs.CVarXiv:2608.30656v12026
  13. A Convolutional Neural Network Neutrino Event Classifier

    A. Aurisano, A. Radovic, D. Rocco +7

    hep-excs.CVarXiv:1604.01444v32016
  14. Fast Learning of Temporal Action Proposal via Dense Boundary Generator

    Chuming Lin, Jian Li, Yabiao Wang +7

    cs.CVarXiv:1911.04127v12019
  15. HEp-2 Cell Image Classification with Deep Convolutional Neural Networks

    Zhimin Gao, Lei Wang, Luping Zhou +1

    cs.CVarXiv:1504.02531v22015
  16. SOFT: Softmax-free Transformer with Linear Complexity

    Jiachen Lu, Jinghan Yao, Junge Zhang +6

    cs.CVcs.AIcs.LGarXiv:2110.11945v32021
  17. Blind Image Super-Resolution: A Survey and Beyond

    Anran Liu, Yihao Liu, Jinjin Gu +2

    cs.CVarXiv:2107.03055v12021
    Summaries:한국어
  18. ImageCAS-X: a dataset and benchmark for coronary artery segmentation and centerline extraction in coronary CT angiography

    Kit M. Bransby, Esther Øksnebjerg, Kristoffer Kjær +7

    cs.CVcs.AIarXiv:2608.30404v12026
  19. Transfer Learning from Synthetic to Real-Noise Denoising with Adaptive Instance Normalization

    Yoonsik Kim, Jae Woong Soh, Gu Yong Park +1

    cs.CVeess.IVarXiv:2002.11244v22020
  20. The Surprising Effectiveness of Representation Learning for Visual Imitation

    Jyothish Pari, Nur Muhammad Shafiullah, Sridhar Pandian Arunachalam +1

    cs.ROcs.AIcs.CVarXiv:2112.01511v22021
  21. Defensive Quantization: When Efficiency Meets Robustness

    Ji Lin, Chuang Gan, Song Han

    cs.LGcs.CVstat.MLarXiv:1904.08444v12019
  22. MR-JEPA: A General Purpose Video Foundation Model for Cardiac MRI

    Athira J. Jacob, Puneet Sharma, Dorin Comaniciu +1

    cs.CVcs.AIarXiv:2608.30975v12026
  23. SIGMA: Semantic-complete Graph Matching for Domain Adaptive Object Detection

    Wuyang Li, Xinyu Liu, Yixuan Yuan

    cs.CVarXiv:2203.06398v32022
  24. MARS: An Instance-aware, Modular and Realistic Simulator for Autonomous Driving

    Zirui Wu, Tianyu Liu, Liyi Luo +13

    cs.CVarXiv:2307.15058v12023
  25. Arbitrary Style Transfer via Multi-Adaptation Network

    Yingying Deng, Fan Tang, Weiming Dong +3

    cs.CVcs.AIarXiv:2005.13219v22020
  26. VidTr: Video Transformer Without Convolutions

    Yanyi Zhang, Xinyu Li, Chunhui Liu +6

    cs.CVarXiv:2104.11746v22021
  27. Fully Convolutional Networks with Sequential Information for Robust Crop and Weed Detection in Precision Farming

    Philipp Lottes, Jens Behley, Andres Milioto +1

    cs.CVarXiv:1806.03412v12018
  28. Autoregressive Mosaics: Probing 2D Spatial Reasoning in Text-Only Language Models

    Ashwin Nedungadi, Stefan Oehmcke, Stefan Lüdtke

    cs.AIcs.CVarXiv:2608.30751v22026
  29. VeriCam: A Verification Baseline for the Classification of Unknown Data

    Lucas Wojcik, Gabriel E. Lima, Sergio M. Silva +2

    cs.CVarXiv:2608.31107v12026
  30. XNOR-Net++: Improved Binary Neural Networks

    Adrian Bulat, Georgios Tzimiropoulos

    cs.CVcs.LGeess.IVarXiv:1909.13863v12019
  31. Real-Time Video Anomaly Detection Using YOLO Pose Estimation and CLIP-Based Semantic Scoring

    Vanodhya G. Warnasooriya, Amir Hajian, Watchara Ruangsang +1

    cs.CVcs.AIeess.IVarXiv:2608.31074v12026
  32. DePlot: One-shot visual language reasoning by plot-to-table translation

    Fangyu Liu, Julian Martin Eisenschlos, Francesco Piccinno +7

    cs.CLcs.AIcs.CVarXiv:2212.10505v22022
  33. Learning Canonical Shape Space for Category-Level 6D Object Pose and Size Estimation

    Dengsheng Chen, Jun Li, Zheng Wang +1

    cs.CVarXiv:2001.09322v32020
  34. SER-FIQ: Unsupervised Estimation of Face Image Quality Based on Stochastic Embedding Robustness

    Philipp Terhörst, Jan Niklas Kolf, Naser Damer +2

    cs.CVarXiv:2003.09373v12020
  35. End-to-End Learning of Motion Representation for Video Understanding

    Lijie Fan, Wenbing Huang, Chuang Gan +3

    cs.CVarXiv:1804.00413v12018
  36. Domain-invariant Stereo Matching Networks

    Feihu Zhang, Xiaojuan Qi, Ruigang Yang +3

    cs.CVarXiv:1911.13287v12019
  37. APQ: Joint Search for Network Architecture, Pruning and Quantization Policy

    Tianzhe Wang, Kuan Wang, Han Cai +3

    cs.LGcs.CVstat.MLarXiv:2006.08509v12020
  38. V2X-Seq: A Large-Scale Sequential Dataset for Vehicle-Infrastructure Cooperative Perception and Forecasting

    Haibao Yu, Wenxian Yang, Hongzhi Ruan +11

    cs.CVcs.AIarXiv:2305.05938v12023
  39. Guaranteed Outlier Removal for Point Cloud Registration with Correspondences

    Álvaro Parra Bustos, Tat-Jun Chin

    cs.CVarXiv:1711.10209v12017
  40. RIDCP: Revitalizing Real Image Dehazing via High-Quality Codebook Priors

    Rui-Qi Wu, Zheng-Peng Duan, Chun-Le Guo +2

    cs.CVarXiv:2304.03994v12023
  41. Train in Germany, Test in The USA: Making 3D Object Detectors Generalize

    Yan Wang, Xiangyu Chen, Yurong You +5

    cs.CVarXiv:2005.08139v12020
  42. Leveraging Photometric Consistency over Time for Sparsely Supervised Hand-Object Reconstruction

    Yana Hasson, Bugra Tekin, Federica Bogo +3

    cs.CVarXiv:2004.13449v12020
  43. Modeling Indirect Illumination for Inverse Rendering

    Yuanqing Zhang, Jiaming Sun, Xingyi He +3

    cs.CVarXiv:2204.06837v12022
  44. Contextual-based Image Inpainting: Infer, Match, and Translate

    Yuhang Song, Chao Yang, Zhe Lin +4

    cs.CVarXiv:1711.08590v52017
  45. LISynSeg: Data-Centric Label-to-Image Synthesis for Cross-Modality Whole-Heart Segmentation

    Jiacheng Wang, Ivana Isgum, Ipek Oguz

    cs.CVeess.IVarXiv:2608.31073v12026
  46. Online Human Action Detection using Joint Classification-Regression Recurrent Neural Networks

    Yanghao Li, Cuiling Lan, Junliang Xing +3

    cs.CVarXiv:1604.05633v22016
  47. Training Sparse Neural Networks

    Suraj Srinivas, Akshayvarun Subramanya, R. Venkatesh Babu

    cs.CVcs.LGarXiv:1611.06694v12016
  48. CedarCypress3D: an annotated UAV-LiDAR dataset of individual trees in planted cedar and cypress forests

    Katsuto Shimizu, Fumiaki Kitahara, Tomohiro Nishizono +8

    cs.CVarXiv:2608.30149v12026
  49. Compositional Explanations of Neurons

    Jesse Mu, Jacob Andreas

    cs.LGcs.AIcs.CLarXiv:2006.14032v22020
  50. CentralNet: a Multilayer Approach for Multimodal Fusion

    Valentin Vielzeuf, Alexis Lechervy, Stéphane Pateux +1

    cs.AIcs.CVcs.MMarXiv:1808.07275v12018
  51. Vision Models Predict Urban Scene Appraisal with Limited Neural Alignment

    Kaizhen Tan, Yuantao Deng

    cs.CVarXiv:2608.30964v12026
  52. DeepCaps: Going Deeper with Capsule Networks

    Jathushan Rajasegaran, Vinoj Jayasundara, Sandaru Jayasekara +3

    cs.CVarXiv:1904.09546v12019
  53. DreamGaussian4D: Generative 4D Gaussian Splatting

    Jiawei Ren, Liang Pan, Jiaxiang Tang +4

    cs.CVcs.GRarXiv:2312.17142v32023
  54. APRIL-GAN: A Zero-/Few-Shot Anomaly Classification and Segmentation Method for CVPR 2023 VAND Workshop Challenge Tracks 1&2: 1st Place on Zero-shot AD and 4th Place on Few-shot AD

    Xuhai Chen, Yue Han, Jiangning Zhang

    cs.CVarXiv:2305.17382v32023
  55. Adversarial Example Does Good: Preventing Painting Imitation from Diffusion Models via Adversarial Examples

    Chumeng Liang, Xiaoyu Wu, Yang Hua +6

    cs.CVcs.AIcs.CRarXiv:2302.04578v22023
  56. Generative Adversarial Network for Medical Images (MI-GAN)

    Talha Iqbal, Hazrat Ali

    cs.LGcs.CVeess.IVarXiv:1810.00551v12018
  57. Natural and Effective Obfuscation by Head Inpainting

    Qianru Sun, Liqian Ma, Seong Joon Oh +3

    cs.CVcs.CRcs.CYarXiv:1711.09001v52017
  58. MonoCap: Monocular Human Motion Capture using a CNN Coupled with a Geometric Prior

    Xiaowei Zhou, Menglong Zhu, Georgios Pavlakos +3

    cs.CVarXiv:1701.02354v22017
  59. SeaDronesSee: A Maritime Benchmark for Detecting Humans in Open Water

    Leon Amadeus Varga, Benjamin Kiefer, Martin Messmer +1

    cs.CVarXiv:2105.01922v22021
  60. VinDr-Mammo: A large-scale benchmark dataset for computer-aided diagnosis in full-field digital mammography

    Hieu T. Nguyen, Ha Q. Nguyen, Hieu H. Pham +4

    eess.IVcs.CVarXiv:2203.11205v22022