Image and Video Processing

Papers filed under eess.IV on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.

Search paper metadata (including unsummarized papers)

1,261 to 1,301 of 1,301

  1. Covid-19: Automatic detection from X-Ray images utilizing Transfer Learning with Convolutional Neural Networks

    Ioannis D. Apostolopoulos, Tzani Bessiana

    eess.IVcs.CVcs.LGarXiv:2003.11617v12020
  2. Real-ESRGAN: Training Real-World Blind Super-Resolution with Pure Synthetic Data

    Xintao Wang, Liangbin Xie, Chao Dong +1

    eess.IVcs.CVarXiv:2107.10833v22021
  3. RandLA-Net: Efficient Semantic Segmentation of Large-Scale Point Clouds

    Qingyong Hu, Bo Yang, Linhai Xie +5

    cs.CVcs.LGeess.IVarXiv:1911.11236v32019
  4. SEAOTTER: Sensor Embedded Autoencoding with One-Time Transcode for Efficient Reconstruction

    Dan Jacobellis, Neeraja J. Yadwadkar

    eess.IVcs.CVcs.LGarXiv:2606.03940v12026
  5. PV-RCNN: Point-Voxel Feature Set Abstraction for 3D Object Detection

    Shaoshuai Shi, Chaoxu Guo, Li Jiang +4

    cs.CVcs.LGeess.IVarXiv:1912.13192v22019
  6. Deep Learning for 3D Point Clouds: A Survey

    Yulan Guo, Hanyun Wang, Qingyong Hu +3

    cs.CVcs.LGcs.ROarXiv:1912.12033v22019
  7. Variational image compression with a scale hyperprior

    Johannes Ballé, David Minnen, Saurabh Singh +2

    eess.IVcs.ITarXiv:1802.01436v22018
  8. EnlightenGAN: Deep Light Enhancement without Paired Supervision

    Yifan Jiang, Xinyu Gong, Ding Liu +6

    cs.CVeess.IVarXiv:1906.06972v22019
  9. Swin UNETR: Swin Transformers for Semantic Segmentation of Brain Tumors in MRI Images

    Ali Hatamizadeh, Vishwesh Nath, Yucheng Tang +3

    eess.IVcs.CVcs.LGarXiv:2201.01266v12022
  10. Image Super-Resolution via Iterative Refinement

    Chitwan Saharia, Jonathan Ho, William Chan +3

    eess.IVcs.CVcs.LGarXiv:2104.07636v22021
  11. UNet 3+: A Full-Scale Connected UNet for Medical Image Segmentation

    Huimin Huang, Lanfen Lin, Ruofeng Tong +6

    eess.IVcs.CVcs.LGarXiv:2004.08790v12020
  12. COVID-Net: A Tailored Deep Convolutional Neural Network Design for Detection of COVID-19 Cases from Chest X-Ray Images

    Linda Wang, Alexander Wong

    eess.IVcs.CVcs.LGarXiv:2003.09871v42020
  13. UNETR: Transformers for 3D Medical Image Segmentation

    Ali Hatamizadeh, Yucheng Tang, Vishwesh Nath +5

    eess.IVcs.CVcs.LGarXiv:2103.10504v32021
  14. Swin-Unet: Unet-like Pure Transformer for Medical Image Segmentation

    Hu Cao, Yueyue Wang, Joy Chen +4

    eess.IVcs.CVarXiv:2105.05537v12021
  15. MMDetection: Open MMLab Detection Toolbox and Benchmark

    Kai Chen, Jiaqi Wang, Jiangmiao Pang +22

    cs.CVcs.LGeess.IVarXiv:1906.07155v12019
  16. CheXpert: A Large Chest Radiograph Dataset with Uncertainty Labels and Expert Comparison

    Jeremy Irvin, Pranav Rajpurkar, Michael Ko +17

    cs.CVcs.AIcs.LGarXiv:1901.07031v12019
  17. Implicit Neural Representations with Periodic Activation Functions

    Vincent Sitzmann, Julien N. P. Martel, Alexander W. Bergman +2

    cs.CVcs.LGeess.IVarXiv:2006.09661v12020
  18. SwinIR: Image Restoration Using Swin Transformer

    Jingyun Liang, Jiezhang Cao, Guolei Sun +3

    eess.IVcs.CVarXiv:2108.10257v12021
  19. Analyzing and Improving the Image Quality of StyleGAN

    Tero Karras, Samuli Laine, Miika Aittala +3

    cs.CVcs.LGcs.NEarXiv:1912.04958v22019
  20. EfficientDet: Scalable and Efficient Object Detection

    Mingxing Tan, Ruoming Pang, Quoc V. Le

    cs.CVcs.LGeess.IVarXiv:1911.09070v72019
  21. UNet++: A Nested U-Net Architecture for Medical Image Segmentation

    Zongwei Zhou, Md Mahfuzur Rahman Siddiquee, Nima Tajbakhsh +1

    cs.CVcs.LGeess.IVarXiv:1807.10165v12018
  22. MagViT: Interpretable Multi-Magnification Transformers with Patient-Level Model Selection for Breast Histopathology

    Nabil Ashab, Soumit Kumar Kundu, Saif Mahmud Parvez +3

    eess.IVcs.CVcs.LGarXiv:2608.16959v12026
  23. YOLOv4: Optimal Speed and Accuracy of Object Detection

    Alexey Bochkovskiy, Chien-Yao Wang, Hong-Yuan Mark Liao

    cs.CVeess.IVarXiv:2004.10934v12020
  24. ENAF: A Multi-Exit Network with an Adaptive Patch Fusion for Large Image Super Resolution

    Duong M. Nguyen, Tuan Nghia Nguyen, Xuan Truong Nguyen

    cs.CVcs.AIeess.IVarXiv:2608.15349v12026
  25. Primitive Representation Learning for Unsupervised Dynamic Contrast Enhanced MRI Reconstruction

    Veronika Spieker, Wenqi Huang, Cemre Ariyurek +5

    eess.IVcs.CVcs.LGarXiv:2608.18055v22026
  26. SP$^3$: Spherical Priors for Plug-and-Play Restoration

    Sean Man, Ron Raphaeli, Matan Kleiner +1

    cs.CVeess.IVarXiv:2606.16396v12026
  27. Iterative Grasp Pose Refinement: A Deep Reinforcement Learning Approach for 2D Vision

    Amir Arsalan Nematollahi, Shayan Ahmadi, Mehdi Tale Masouleh +1

    cs.ROcs.AIcs.LGarXiv:2608.17628v12026
  28. FRAPPE: Full Input, Residual Output Autoencoding with Projection Pursuit Encoder

    Dan Jacobellis, Neeraja J. Yadwadkar

    eess.IVarXiv:2605.28992v12026
  29. Synthesizing Post-Acetazolamide Cerebral Blood Flow Maps from Baseline MRI in Moyamoya Using 3D Generative AI

    Julia Huang, Camila Gonzalez, Rydham Goyal +5

    eess.IVcs.AIcs.CVarXiv:2608.14758v12026
  30. Cross-Modal Ultrasound-MRI Learning for Fetal Brain Ventricular Volumetry and Abnormality Screening

    Yuhao Huang, Yuanji Zhang, Yuhuan Lu +3

    eess.IVcs.AIcs.CVarXiv:2608.14763v12026
  31. Emotion Across Speech and Faces: Shared Affective Mechanisms in Multimodal Foundation Models

    Xiutian Zhao, Luqi Sun, Björn Schuller +1

    cs.CLeess.ASeess.IVarXiv:2608.17102v12026
  32. Decoupling Parcellation from Classification: Systematic Benchmark of Fast Brain Segmentation Methods for Alzheimer's Disease Detection

    Jiadao Zou, Hongyu Guo, Wei Xi

    eess.IVcs.AIcs.CVarXiv:2608.16039v12026
  33. A cross-modal generative model for incomplete and degraded prostate MRI with multicentre clinical validation

    Siyuan Ma, Liang He, Mengying Zhu +14

    eess.IVcs.AIcs.CVarXiv:2608.16233v12026
  34. PathFinder: Joint Decompositions of Linked Multimodal Datasets

    Ying-Qiu Zheng, Alex Fung, Stephen M Smith +2

    cs.LGeess.IVq-bio.QMarXiv:2608.14951v12026
  35. PixCon: Clean-Positive Contrastive Learning for Foundation-Model Semi-Supervised Segmentation

    Ebenezer Tarubinga

    cs.CVcs.AIcs.LGarXiv:2607.03068v12026
  36. OPERA: Offline Policy-guided Expert Routing and Adaptation for Universal Biomedical Image Analysis

    Zihan Li, Feiyang Liu, Dandan Shan +2

    cs.CVcs.AIcs.LGarXiv:2607.25108v12026
  37. Mage-Flow: An Efficient Native-Resolution Foundation Model for Image Generation and Editing

    Xinjie Zhang, Peng Zhang, Shicheng Zheng +21

    cs.CVcs.AIcs.LGarXiv:2607.19064v22026
  38. Human-in-the-Loop Signature Bootstrapping for UAV Hyperspectral PFM-1 Mine Detection

    Sagar Lekhak, Prasanna Reddy Pulakurthi, Emmett J. Ientilucci

    cs.CVeess.IVarXiv:2607.25310v12026
  39. Secret-Stego Dissimilarity as a Design Axis: Invertible Coverless Image Steganography with Diffusion Models

    Hongxin Xu, Jianping Mei, Can Wang +1

    eess.IVcs.AIcs.CRarXiv:2608.13597v12026
  40. Multi-Task Multi-Frame Visual Piano Transcription

    Yonghyun Kim, Hoyeol Sohn, Juhan Nam +1

    cs.SDcs.AIcs.CVarXiv:2608.03419v12026
  41. TRUE-Colon: Exposing a Consistent Transfer Asymmetry in Real-Time Polyp Detection

    Sebastian Doerrich, Andreas Franz Schwab, Francesco Di Salvo +3

    eess.IVcs.CVcs.LGarXiv:2608.13711v12026