Image and Video Processing

Papers filed under eess.IV on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.

Search paper metadata (including unsummarized papers)

1,141 to 1,200 of 1,302

  1. The state of the art in kidney and kidney tumor segmentation in contrast-enhanced CT imaging: Results of the KiTS19 Challenge

    Nicholas Heller, Fabian Isensee, Klaus H. Maier-Hein +38

    eess.IVcs.CVcs.LGarXiv:1912.01054v22019
  2. ScanRefer: 3D Object Localization in RGB-D Scans using Natural Language

    Dave Zhenyu Chen, Angel X. Chang, Matthias Nießner

    cs.CVcs.CLcs.LGarXiv:1912.08830v32019
  3. Learning Enriched Features for Fast Image Restoration and Enhancement

    Syed Waqas Zamir, Aditya Arora, Salman Khan +4

    eess.IVcs.CVarXiv:2205.01649v12022
  4. Spatiotemporal Distillation via Recurrent Bottlenecks for Aortic Tracking

    Dexter Wen Jie Teo, Nairouz Shehata, Herve Lombaert

    eess.IVcs.CVcs.LGarXiv:2608.23879v12026
  5. Deep Learning for Deepfakes Creation and Detection: A Survey

    Thanh Thi Nguyen, Quoc Viet Hung Nguyen, Dung Tien Nguyen +6

    cs.CVcs.LGeess.IVarXiv:1909.11573v52019
  6. Native-Space 3D CarveMix for Multi-Site T1w Stroke Segmentation

    Dexter Wen Jie Teo, Kumaradevan Punithakumar

    eess.IVcs.CVarXiv:2608.23882v12026
  7. U-GAT-IT: Unsupervised Generative Attentional Networks with Adaptive Layer-Instance Normalization for Image-to-Image Translation

    Junho Kim, Minjae Kim, Hyeonwoo Kang +1

    cs.CVeess.IVarXiv:1907.10830v42019
  8. Explainable deep learning models in medical image analysis

    Amitojdeep Singh, Sourya Sengupta, Vasudevan Lakshminarayanan

    cs.CVcs.LGeess.IVarXiv:2005.13799v12020
  9. PULSE: Self-Supervised Photo Upsampling via Latent Space Exploration of Generative Models

    Sachit Menon, Alexandru Damian, Shijia Hu +2

    cs.CVcs.LGeess.IVarXiv:2003.03808v32020
  10. Model Effect or Label Effect? Refined Annotations and a Human-Referenced Benchmark for Pulmonary Embolism Segmentation

    Qihang Sun, Zhongxiao Liu, Bailiang Jian +6

    eess.IVcs.CVarXiv:2608.24486v12026
  11. High-Fidelity Generative Image Compression

    Fabian Mentzer, George Toderici, Michael Tschannen +1

    eess.IVcs.CVcs.LGarXiv:2006.09965v32020
  12. Detecting and Simulating Artifacts in GAN Fake Images

    Xu Zhang, Svebor Karaman, Shih-Fu Chang

    cs.CVeess.IVarXiv:1907.06515v22019
  13. CA-Net: Comprehensive Attention Convolutional Neural Networks for Explainable Medical Image Segmentation

    Ran Gu, Guotai Wang, Tao Song +6

    eess.IVcs.CVarXiv:2009.10549v22020
  14. Deep Unfolding Network for Image Super-Resolution

    Kai Zhang, Luc Van Gool, Radu Timofte

    eess.IVcs.CVarXiv:2003.10428v12020
  15. SynthSeg: Segmentation of brain MRI scans of any contrast and resolution without retraining

    Benjamin Billot, Douglas N. Greve, Oula Puonti +5

    eess.IVcs.CVarXiv:2107.09559v42021
  16. Drone-based RGB-Infrared Cross-Modality Vehicle Detection via Uncertainty-Aware Learning

    Yiming Sun, Bing Cao, Pengfei Zhu +1

    cs.CVcs.LGeess.IVarXiv:2003.02437v22020
  17. BCN20000: Dermoscopic Lesions in the Wild

    Marc Combalia, Noel C. F. Codella, Veronica Rotemberg +8

    eess.IVcs.CVarXiv:1908.02288v22019
  18. Deep Learning Techniques for Inverse Problems in Imaging

    Gregory Ongie, Ajil Jalal, Christopher A. Metzler +3

    eess.IVcs.LGstat.MLarXiv:2005.06001v12020
  19. Deep Learning in Medical Image Registration: A Survey

    Grant Haskins, Uwe Kruger, Pingkun Yan

    q-bio.QMcs.CVeess.IVarXiv:1903.02026v22019
  20. HINet: Half Instance Normalization Network for Image Restoration

    Liangyu Chen, Xin Lu, Jie Zhang +2

    eess.IVcs.CVarXiv:2105.06086v22021
  21. Make it SING: Analyzing Semantic Invariants in Classifiers

    Harel Yadid, Meir Yossef Levi, Roy Betser +1

    cs.CVeess.IVarXiv:2603.14610v22026
  22. HistoAtlas: A Pan-Cancer Morphology Atlas Linking Histomics to Molecular Programs and Clinical Outcomes

    Pierre-Antoine Bannier

    q-bio.QMcs.CVeess.IVarXiv:2603.16587v12026
  23. Generalized ODIN: Detecting Out-of-distribution Image without Learning from Out-of-distribution Data

    Yen-Chang Hsu, Yilin Shen, Hongxia Jin +1

    cs.CVcs.LGeess.IVarXiv:2002.11297v22020
  24. Unified Focal loss: Generalising Dice and cross entropy-based losses to handle class imbalanced medical image segmentation

    Michael Yeung, Evis Sala, Carola-Bibiane Schönlieb +1

    eess.IVcs.CVcs.LGarXiv:2102.04525v42021
  25. Deep learning with noisy labels: exploring techniques and remedies in medical image analysis

    Davood Karimi, Haoran Dou, Simon K. Warfield +1

    cs.CVcs.LGeess.IVarXiv:1912.02911v42019
  26. CHIMERA Challenge: Biochemical Recurrence Prediction in Prostate Cancer Patients using multimodal datasets

    Robert N. Spaans, Catherine Chia, Tongjie Wang +7

    eess.IVcs.CVarXiv:2608.21497v12026
  27. Pattern-Derived Visual Swarm Games: Multi-Scale Drone-Vision States for Interception and Sustainability Audits

    Faruk Alpay, Levent Sarioglu

    cs.ROcs.GTeess.IVarXiv:2608.23575v12026
  28. Contrastive learning of global and local features for medical image segmentation with limited annotations

    Krishna Chaitanya, Ertunc Erdil, Neerav Karani +1

    cs.CVcs.LGeess.IVarXiv:2006.10511v22020
  29. InterFaceGAN: Interpreting the Disentangled Face Representation Learned by GANs

    Yujun Shen, Ceyuan Yang, Xiaoou Tang +1

    cs.CVcs.LGeess.IVarXiv:2005.09635v22020
  30. CMX: Cross-Modal Fusion for RGB-X Semantic Segmentation with Transformers

    Jiaming Zhang, Huayao Liu, Kailun Yang +3

    cs.CVcs.ROeess.IVarXiv:2203.04838v52022
  31. PlantDoc: A Dataset for Visual Plant Disease Detection

    Davinder Singh, Naman Jain, Pranjali Jain +3

    cs.CVeess.IVarXiv:1911.10317v12019
  32. Precision-Aware Variable Bit Processing Elements for Hardware-Efficient Systolic Array Designs

    Dantu Nandini Devi, Madhav Rao

    cs.ARcs.ETcs.LGarXiv:2608.22378v12026
  33. Deep neural network models for computational histopathology: A survey

    Chetan L. Srinidhi, Ozan Ciga, Anne L. Martel

    eess.IVcs.CVcs.LGarXiv:1912.12378v22019
  34. Convolutional Neural Networks for Classification of Alzheimer's Disease: Overview and Reproducible Evaluation

    Junhao Wen, Elina Thibeau-Sutre, Mauricio Diaz-Melo +7

    cs.LGeess.IVstat.MLarXiv:1904.07773v62019
  35. MAXIM: Multi-Axis MLP for Image Processing

    Zhengzhong Tu, Hossein Talebi, Han Zhang +4

    eess.IVcs.CVarXiv:2201.02973v22022
  36. SGDC: Structurally-Guided Dynamic Convolution for Medical Image Segmentation

    Bo Shi, Wei-ping Zhu, M. N. S. Swamy

    eess.IVcs.CVarXiv:2602.23496v12026
  37. PREDATOR: Registration of 3D Point Clouds with Low Overlap

    Shengyu Huang, Zan Gojcic, Mikhail Usvyatsov +2

    cs.CVeess.IVarXiv:2011.13005v32020
  38. Colon-Bench: An Agentic Workflow for Scalable Dense Lesion Annotation in Full-Procedure Colonoscopy Videos

    Abdullah Hamdi, Changchun Yang, Xin Gao

    eess.IVcs.CVcs.HCarXiv:2603.25645v32026
  39. GET: Generative Embedding Translation for Medical Image Segmentation

    Md Maklachur Rahman, Md Hasan Al Banna, Saraf Anjum +2

    eess.IVcs.CVcs.LGarXiv:2608.22619v12026
  40. Deep Learning COVID-19 Features on CXR using Limited Training Data Sets

    Yujin Oh, Sangjoon Park, Jong Chul Ye

    eess.IVcs.CVcs.LGarXiv:2004.05758v22020
  41. Visual Transformers: Token-based Image Representation and Processing for Computer Vision

    Bichen Wu, Chenfeng Xu, Xiaoliang Dai +7

    cs.CVcs.LGeess.IVarXiv:2006.03677v42020
  42. Big Self-Supervised Models Advance Medical Image Classification

    Shekoofeh Azizi, Basil Mustafa, Fiona Ryan +9

    eess.IVcs.CVcs.LGarXiv:2101.05224v22021
  43. VATT: Transformers for Multimodal Self-Supervised Learning from Raw Video, Audio and Text

    Hassan Akbari, Liangzhe Yuan, Rui Qian +4

    cs.CVcs.AIcs.LGarXiv:2104.11178v32021
  44. Boundary loss for highly unbalanced segmentation

    Hoel Kervadec, Jihene Bouchtiba, Christian Desrosiers +3

    eess.IVcs.CVarXiv:1812.07032v42018
  45. Solving Inverse Problems in Medical Imaging with Score-Based Generative Models

    Yang Song, Liyue Shen, Lei Xing +1

    eess.IVcs.CVcs.LGarXiv:2111.08005v22021
  46. The Universal Normal Embedding

    Chen Tasker, Roy Betser, Eyal Gofer +2

    cs.CVeess.IVarXiv:2603.21786v12026
  47. U-shape Transformer for Underwater Image Enhancement

    Lintao Peng, Chunli Zhu, Liheng Bian

    cs.CVeess.IVarXiv:2111.11843v62021
  48. MANIQA: Multi-dimension Attention Network for No-Reference Image Quality Assessment

    Sidi Yang, Tianhe Wu, Shuwei Shi +5

    cs.CVeess.IVarXiv:2204.08958v22022
  49. Diffusion Models for Medical Image Analysis: A Comprehensive Survey

    Amirhossein Kazerouni, Ehsan Khodapanah Aghdam, Moein Heidari +4

    eess.IVcs.CVarXiv:2211.07804v32022
  50. OpenVision 3: A Family of Unified Visual Encoder for Both Understanding and Generation

    Letian Zhang, Sucheng Ren, Yanqing Liu +9

    eess.IVcs.AIarXiv:2601.15369v22026
  51. Capsule-Forensics: Using Capsule Networks to Detect Forged Images and Videos

    Huy H. Nguyen, Junichi Yamagishi, Isao Echizen

    cs.CVeess.IVarXiv:1810.11215v12018
  52. Deep Semantic Segmentation of Natural and Medical Images: A Review

    Saeid Asgari Taghanaki, Kumar Abhishek, Joseph Paul Cohen +2

    cs.CVcs.LGeess.IVarXiv:1910.07655v42019
  53. DoubleU-Net: A Deep Convolutional Neural Network for Medical Image Segmentation

    Debesh Jha, Michael A. Riegler, Dag Johansen +2

    eess.IVcs.CVarXiv:2006.04868v22020
  54. HigherHRNet: Scale-Aware Representation Learning for Bottom-Up Human Pose Estimation

    Bowen Cheng, Bin Xiao, Jingdong Wang +3

    cs.CVcs.LGeess.IVarXiv:1908.10357v32019
  55. Single Image Super-Resolution via a Holistic Attention Network

    Ben Niu, Weilei Wen, Wenqi Ren +6

    eess.IVcs.CVarXiv:2008.08767v12020
  56. Multi-Scale Progressive Fusion Network for Single Image Deraining

    Kui Jiang, Zhongyuan Wang, Peng Yi +5

    cs.CVcs.LGeess.IVarXiv:2003.10985v22020
  57. Spanning the Visual Analogy Space with a Weight Basis of LoRAs

    Hila Manor, Rinon Gal, Haggai Maron +2

    cs.CVcs.AIcs.GRarXiv:2602.15727v22026
  58. PadChest: A large chest x-ray image dataset with multi-label annotated reports

    Aurelia Bustos, Antonio Pertusa, Jose-Maria Salinas +1

    eess.IVcs.CVarXiv:1901.07441v22019
  59. Video Generation Models as World Models: Efficient Paradigms, Architectures and Algorithms

    Muyang He, Hanzhong Guo, Junxiong Lin +1

    eess.IVcs.CVarXiv:2603.28489v32026
  60. COVID-19 Image Data Collection: Prospective Predictions Are the Future

    Joseph Paul Cohen, Paul Morrison, Lan Dao +3

    q-bio.QMcs.CVcs.LGarXiv:2006.11988v32020