Image and Video Processing

Papers filed under eess.IV on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.

Search paper metadata (including unsummarized papers)

961 to 1,020 of 1,302

  1. Anomaly Detection in Video via Self-Supervised and Multi-Task Learning

    Mariana-Iuliana Georgescu, Antonio Barbalau, Radu Tudor Ionescu +3

    cs.CVcs.LGeess.IVarXiv:2011.07491v32020
  2. Real-World Single Image Super-Resolution: A Brief Review

    Honggang Chen, Xiaohai He, Linbo Qing +3

    eess.IVcs.CVarXiv:2103.02368v12021
  3. Artificial Intelligence for Digital and Computational Pathology

    Andrew H. Song, Guillaume Jaume, Drew F. K. Williamson +4

    eess.IVcs.AIcs.CVarXiv:2401.06148v12023
  4. SCRDet++: Detecting Small, Cluttered and Rotated Objects via Instance-Level Feature Denoising and Rotation Loss Smoothing

    Xue Yang, Junchi Yan, Wenlong Liao +3

    cs.CVcs.LGeess.IVarXiv:2004.13316v22020
  5. Lung and Colon Cancer Histopathological Image Dataset (LC25000)

    Andrew A. Borkowski, Marilyn M. Bui, L. Brannon Thomas +3

    eess.IVcs.CVq-bio.QMarXiv:1912.12142v12019
  6. MS-Net: Multi-Site Network for Improving Prostate Segmentation with Heterogeneous MRI Data

    Quande Liu, Qi Dou, Lequan Yu +1

    eess.IVcs.CVarXiv:2002.03366v22020
  7. Federated Learning Enables Big Data for Rare Cancer Boundary Detection

    Sarthak Pati, Ujjwal Baid, Brandon Edwards +276

    cs.LGeess.IVarXiv:2204.10836v22022
  8. Loam_livox: A fast, robust, high-precision LiDAR odometry and mapping package for LiDARs of small FoV

    Jiarong Lin, Fu Zhang

    cs.ROcs.CVeess.IVarXiv:1909.06700v12019
  9. CLIP-Driven Universal Model for Organ Segmentation and Tumor Detection

    Jie Liu, Yixiao Zhang, Jie-Neng Chen +7

    eess.IVcs.CVcs.LGarXiv:2301.00785v52023
  10. Deep Learning for Visual Tracking: A Comprehensive Survey

    Seyed Mojtaba Marvasti-Zadeh, Li Cheng, Hossein Ghanei-Yakhdan +1

    cs.CVcs.LGeess.IVarXiv:1912.00535v22019
  11. Segment Anything Model for Medical Image Segmentation: Current Applications and Future Directions

    Yichi Zhang, Zhenrong Shen, Rushi Jiao

    eess.IVcs.CVarXiv:2401.03495v12024
  12. MedSegDiff-V2: Diffusion based Medical Image Segmentation with Transformer

    Junde Wu, Wei Ji, Huazhu Fu +3

    eess.IVcs.CVarXiv:2301.11798v22023
  13. Mamba-UNet: UNet-Like Pure Visual Mamba for Medical Image Segmentation

    Ziyang Wang, Jian-Qing Zheng, Yichi Zhang +2

    eess.IVcs.CVarXiv:2402.05079v22024
  14. DeepFaceLab: Integrated, flexible and extensible face-swapping framework

    Ivan Perov, Daiheng Gao, Nikolay Chervoniy +11

    cs.CVcs.LGcs.MMarXiv:2005.05535v52020
  15. A Survey on Generative Adversarial Networks: Variants, Applications, and Training

    Abdul Jabbar, Xi Li, Bourahla Omar

    cs.CVcs.LGeess.IVarXiv:2006.05132v12020
  16. Hi-Net: Hybrid-fusion Network for Multi-modal MR Image Synthesis

    Tao Zhou, Huazhu Fu, Geng Chen +2

    cs.CVeess.IVarXiv:2002.05000v12020
  17. Semantic Segmentation of Underwater Imagery: Dataset and Benchmark

    Md Jahidul Islam, Chelsey Edge, Yuyang Xiao +5

    cs.CVeess.IVarXiv:2004.01241v32020
  18. Large Deformation Diffeomorphic Image Registration with Laplacian Pyramid Networks

    Tony C. W. Mok, Albert C. S. Chung

    eess.IVcs.CVarXiv:2006.16148v22020
  19. History Repeats Itself: Human Motion Prediction via Motion Attention

    Wei Mao, Miaomiao Liu, Mathieu Salzmann

    cs.CVcs.LGeess.IVarXiv:2007.11755v12020
  20. HiFuse: Hierarchical Multi-Scale Feature Fusion Network for Medical Image Classification

    Xiangzuo Huo, Gang Sun, Shengwei Tian +5

    eess.IVcs.CVarXiv:2209.10218v12022
  21. Improving performance of CNN to predict likelihood of COVID-19 using chest X-ray images with preprocessing algorithms

    Morteza Heidari, Seyedehnafiseh Mirniaharikandehei, Abolfazl Zargari Khuzani +3

    eess.IVcs.LGarXiv:2006.12229v12020
  22. Learning Image-adaptive 3D Lookup Tables for High Performance Photo Enhancement in Real-time

    Hui Zeng, Jianrui Cai, Lida Li +2

    eess.IVcs.CVarXiv:2009.14468v12020
  23. RTM3D: Real-time Monocular 3D Detection from Object Keypoints for Autonomous Driving

    Peixuan Li, Huaici Zhao, Pengfei Liu +1

    cs.CVcs.ROeess.IVarXiv:2001.03343v12020
  24. CANet: Cross-disease Attention Network for Joint Diabetic Retinopathy and Diabetic Macular Edema Grading

    Xiaomeng Li, Xiaowei Hu, Lequan Yu +3

    eess.IVcs.CVarXiv:1911.01376v12019
  25. Sharp U-Net: Depthwise Convolutional Network for Biomedical Image Segmentation

    Hasib Zunair, A. Ben Hamza

    eess.IVcs.CVarXiv:2107.12461v12021
  26. Deep Learning Based Brain Tumor Segmentation: A Survey

    Zhihua Liu, Lei Tong, Zheheng Jiang +6

    eess.IVcs.CVarXiv:2007.09479v32020
  27. Classification of Hyperspectral and LiDAR Data Using Coupled CNNs

    Renlong Hang, Zhu Li, Pedram Ghamisi +3

    cs.CVeess.IVarXiv:2002.01144v12020
  28. Uncertainty-Aware Blind Image Quality Assessment in the Laboratory and Wild

    Weixia Zhang, Kede Ma, Guangtao Zhai +1

    cs.CVcs.LGcs.MMarXiv:2005.13983v62020
  29. Plug-and-Play Priors for Bright Field Electron Tomography and Sparse Interpolation

    Suhas Sreehari, S. V. Venkatakrishnan, Brendt Wohlberg +3

    cs.CVeess.IVarXiv:1512.07331v12015
  30. PAD-UFES-20: a skin lesion dataset composed of patient data and clinical images collected from smartphones

    Andre G. C. Pacheco, Gustavo R. Lima, Amanda S. Salomão +16

    eess.IVq-bio.QMarXiv:2007.00478v22020
  31. VerSe: A Vertebrae Labelling and Segmentation Benchmark for Multi-detector CT Images

    Anjany Sekuboyina, Malek E. Husseini, Amirhossein Bayat +66

    cs.CVeess.IVarXiv:2001.09193v62020
  32. ARCH: Animatable Reconstruction of Clothed Humans

    Zeng Huang, Yuanlu Xu, Christoph Lassner +2

    cs.GRcs.CVcs.LGarXiv:2004.04572v22020
  33. Axiom-based Grad-CAM: Towards Accurate Visualization and Explanation of CNNs

    Ruigang Fu, Qingyong Hu, Xiaohu Dong +3

    cs.CVcs.AIcs.LGarXiv:2008.02312v42020
  34. Autoencoders for Unsupervised Anomaly Segmentation in Brain MR Images: A Comparative Study

    Christoph Baur, Stefan Denner, Benedikt Wiestler +2

    eess.IVcs.CVcs.LGarXiv:2004.03271v22020
  35. Anatomy-Guided Foundation Model Adaptation with Within-Case Prototype Supervision for Standard Plane Detection in Fetal Ultrasound Blind Sweeps

    Yuzhe Zhao

    cs.CVeess.IVarXiv:2608.27051v12026
  36. CheXclusion: Fairness gaps in deep chest X-ray classifiers

    Laleh Seyyed-Kalantari, Guanxiong Liu, Matthew McDermott +2

    cs.CVcs.AIcs.LGarXiv:2003.00827v22020
  37. PhaseStain: Digital staining of label-free quantitative phase microscopy images using deep learning

    Yair Rivenson, Tairan Liu, Zhensong Wei +2

    eess.IVcs.CVphysics.med-pharXiv:1807.07701v12018
  38. Structure-Preserving Super Resolution with Gradient Guidance

    Cheng Ma, Yongming Rao, Yean Cheng +3

    eess.IVcs.CVarXiv:2003.13081v12020
  39. Deblurring via Stochastic Refinement

    Jay Whang, Mauricio Delbracio, Hossein Talebi +3

    cs.CVeess.IVarXiv:2112.02475v22021
  40. MetaIQA: Deep Meta-learning for No-Reference Image Quality Assessment

    Hancheng Zhu, Leida Li, Jinjian Wu +2

    eess.IVcs.CVarXiv:2004.05508v12020
  41. Cross-Modality Fusion Transformer for Multispectral Object Detection

    Fang Qingyun, Han Dapeng, Wang Zhaokui

    eess.IVcs.CVarXiv:2111.00273v42021
  42. Unsupervised Bidirectional Cross-Modality Adaptation via Deeply Synergistic Image and Feature Alignment for Medical Image Segmentation

    Cheng Chen, Qi Dou, Hao Chen +2

    eess.IVcs.CVarXiv:2002.02255v12020
  43. Estimating Uncertainty and Interpretability in Deep Learning for Coronavirus (COVID-19) Detection

    Biraja Ghoshal, Allan Tucker

    eess.IVcs.CVcs.LGarXiv:2003.10769v22020
  44. NeRV: Neural Representations for Videos

    Hao Chen, Bo He, Hanyu Wang +3

    cs.CVeess.IVarXiv:2110.13903v12021
  45. UniVL: A Unified Video and Language Pre-Training Model for Multimodal Understanding and Generation

    Huaishao Luo, Lei Ji, Botian Shi +6

    cs.CVcs.CLcs.LGarXiv:2002.06353v32020
  46. DIDFuse: Deep Image Decomposition for Infrared and Visible Image Fusion

    Zixiang Zhao, Shuang Xu, Chunxia Zhang +3

    eess.IVcs.CVarXiv:2003.09210v32020
  47. Mask-guided Spectral-wise Transformer for Efficient Hyperspectral Image Reconstruction

    Yuanhao Cai, Jing Lin, Xiaowan Hu +5

    eess.IVcs.CVarXiv:2111.07910v22021
  48. Expression, Affect, Action Unit Recognition: Aff-Wild2, Multi-Task Learning and ArcFace

    Dimitrios Kollias, Stefanos Zafeiriou

    cs.CVcs.HCcs.LGarXiv:1910.04855v12019
  49. Models Genesis: Generic Autodidactic Models for 3D Medical Image Analysis

    Zongwei Zhou, Vatsal Sodha, Md Mahfuzur Rahman Siddiquee +4

    eess.IVcs.CVarXiv:1908.06912v12019
  50. TransAttUnet: Multi-level Attention-guided U-Net with Transformer for Medical Image Segmentation

    Bingzhi Chen, Yishu Liu, Zheng Zhang +2

    eess.IVcs.CVarXiv:2107.05274v22021
  51. Multi-stage image denoising with the wavelet transform

    Chunwei Tian, Menghua Zheng, Wangmeng Zuo +3

    eess.IVcs.CVarXiv:2209.12394v32022
  52. Self supervised contrastive learning for digital histopathology

    Ozan Ciga, Tony Xu, Anne L. Martel

    eess.IVcs.CVarXiv:2011.13971v22020
  53. Dynamic Fusion with Intra- and Inter- Modality Attention Flow for Visual Question Answering

    Gao Peng, Zhengkai Jiang, Haoxuan You +4

    cs.CVeess.IVarXiv:1812.05252v42018
  54. Fringe pattern analysis using deep learning

    Shijie Feng, Qian Chen, Guohua Gu +5

    eess.IVarXiv:1807.02757v12018
  55. Multi-Task Temporal Shift Attention Networks for On-Device Contactless Vitals Measurement

    Xin Liu, Josh Fromm, Shwetak Patel +1

    eess.SPcs.CVeess.IVarXiv:2006.03790v22020
  56. Lung Infection Quantification of COVID-19 in CT Images with Deep Learning

    Fei Shan, Yaozong Gao, Jun Wang +6

    cs.CVeess.IVq-bio.QMarXiv:2003.04655v32020
  57. Diagnosing COVID-19 Pneumonia from X-Ray and CT Images using Deep Learning and Transfer Learning Algorithms

    Halgurd S. Maghdid, Aras T. Asaad, Kayhan Zrar Ghafoor +2

    eess.IVcs.CVcs.LGarXiv:2004.00038v12020
  58. When Radiology Report Generation Meets Knowledge Graph

    Yixiao Zhang, Xiaosong Wang, Ziyue Xu +3

    cs.CVcs.AIcs.LGarXiv:2002.08277v12020
  59. Bi-Directional ConvLSTM U-Net with Densley Connected Convolutions

    Reza Azad, Maryam Asadi-Aghbolaghi, Mahmood Fathy +1

    eess.IVcs.CVarXiv:1909.00166v12019
  60. Diffusion Models for Medical Anomaly Detection

    Julia Wolleb, Florentin Bieder, Robin Sandkühler +1

    eess.IVcs.CVarXiv:2203.04306v22022