Image and Video Processing

Papers filed under eess.IV on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.

Search paper metadata (including unsummarized papers)

1 to 60 of 1,302

  1. JAFAR: Jack up Any Feature at Any Resolution

    Paul Couairon, Loick Chambon, Louis Serrano +3

    cs.CVeess.IVarXiv:2506.11136v32025
  2. DiffNR: Diffusion-Enhanced Neural Representation Optimization for Sparse-View 3D Tomographic Reconstruction

    Shiyan Su, Ruyi Zha, Danli Shi +2

    eess.IVcs.CVarXiv:2604.21518v12026
  3. Endmember-Guided Unmixing Network (EGU-Net): A General Deep Learning Framework for Self-Supervised Hyperspectral Unmixing

    Danfeng Hong, Lianru Gao, Jing Yao +4

    eess.IVcs.CVarXiv:2105.10194v12021
  4. AtlasPatch: Efficient Tissue Detection and High-throughput Patch Extraction for Computational Pathology at Scale

    Ahmed Alagha, Christopher Leclerc, Yousef Kotp +12

    eess.IVcs.CVq-bio.QMarXiv:2602.03998v22026
  5. CardioState-JEPA: Delay-Aware Cross-Modal Learning of a Shared Cardiac Representation

    Hamza Shafiq, Hung Manh Pham, Bin Zhu +3

    cs.LGeess.IVstat.MLarXiv:2608.12944v12026
  6. DeepGCNs: Making GCNs Go as Deep as CNNs

    Guohao Li, Matthias Müller, Guocheng Qian +4

    cs.CVcs.LGeess.IVarXiv:1910.06849v32019
  7. InstructIR: High-Quality Image Restoration Following Human Instructions

    Marcos V. Conde, Gregor Geigle, Radu Timofte

    cs.CVcs.LGeess.IVarXiv:2401.16468v52024
  8. Attention Guided Anomaly Localization in Images

    Shashanka Venkataramanan, Kuan-Chuan Peng, Rajat Vikram Singh +1

    cs.CVeess.IVarXiv:1911.08616v42019
  9. Multimodal Whole Slide Foundation Model for Pathology

    Tong Ding, Sophia J. Wagner, Andrew H. Song +20

    eess.IVcs.AIcs.CVarXiv:2411.19666v12024
  10. A Survey of Convolutional Neural Networks: Analysis, Applications, and Prospects

    Zewen Li, Wenjie Yang, Shouheng Peng +1

    cs.CVcs.LGeess.IVarXiv:2004.02806v12020
  11. FPGA: Fast Patch-Free Global Learning Framework for Fully End-to-End Hyperspectral Image Classification

    Zhuo Zheng, Yanfei Zhong, Ailong Ma +1

    cs.CVeess.IVarXiv:2011.05670v12020
  12. Fruit Quality and Defect Image Classification with Conditional GAN Data Augmentation

    Jordan J. Bird, Chloe M. Barnes, Luis J. Manso +2

    cs.CVcs.LGeess.IVarXiv:2104.05647v12021
  13. Data Augmentation for Skin Lesion using Self-Attention based Progressive Generative Adversarial Network

    Ibrahim Saad Ali, Mamdouh Farouk Mohamed, Yousef Bassyouni Mahdy

    eess.IVcs.CVarXiv:1910.11960v12019
  14. Knowledge as Orbit: Finite Collections as Phases of an Exactly Periodic Latent Generator

    Siddharth Pal, Viktoria Rojkova

    cs.LGcs.CVeess.IVarXiv:2609.17417v12026
  15. Semantic-Aware Neural Video Codec for Error-Resilient Low-Latency Transmission

    Matin Mortaheb, Homa Esfahanizadeh, Jinfeng Du +1

    eess.IVcs.ITcs.LGarXiv:2609.16279v12026
  16. Light Field Reconstruction Using Convolutional Network on EPI and Extended Applications

    Gaochang Wu, Yebin Liu, Lu Fang +2

    eess.IVcs.CVarXiv:2103.13043v12021
  17. LM-PCVMNet: Pediatric Cervical Vertebral Maturation Analysis with Deep Fusion of Landmarks and Metadata

    Peng Wang, Wanzhen Song, Anli Wang +3

    eess.IVcs.CVarXiv:2609.16033v12026
  18. Convolution-Free Medical Image Segmentation using Transformers

    Davood Karimi, Serge Vasylechko, Ali Gholipour

    eess.IVcs.CVarXiv:2102.13645v22021
  19. From Foundation Embeddings to Cropland Maps: Label Efficiency, Temporal Transferability and Independent Human Validation

    Mohammad Ammar Mughees, Giovanni Montefoschi, Zhongxin Chen +1

    cs.CVcs.LGeess.IVarXiv:2609.17138v12026
  20. A Comprehensive Review of Deep Learning-based Single Image Super-resolution

    Syed Muhammad Arsalan Bashir, Yi Wang, Mahrukh Khan +1

    cs.CVcs.LGeess.IVarXiv:2102.09351v32021
  21. HyperTransformer: A Textural and Spectral Feature Fusion Transformer for Pansharpening

    Wele Gedara Chaminda Bandara, Vishal M. Patel

    cs.CVeess.IVarXiv:2203.02503v32022
  22. Automated Mobile Video Objective Testing System

    Eric Petajan, Jonathan Lynam, Morey Antebi +5

    cs.NIcs.MMeess.IVarXiv:2609.09579v12026
  23. Convolutional Neural Networks for Global Human Settlements Mapping from Sentinel-2 Satellite Imagery

    Christina Corbane, Vasileios Syrris, Filip Sabo +5

    eess.IVcs.CVcs.LGarXiv:2006.03267v22020
  24. StainBridge: Stain-Aware Pairwise Registration of Serial Renal Biopsy Whole-Slide Images Across Structural and Immunohistochemical Stains

    Ellen Wei, Bohang Jiang, Yanfan Zhu +8

    eess.IVarXiv:2609.17090v12026
  25. Prototyping QoE-Aware Rate Adaptation in Cellular Networks with Commercial Applications

    Szilveszter Nádas, Lars Ernström, Dan Druta +4

    cs.NIcs.MMeess.IVarXiv:2609.09490v12026
  26. Sparse concept attribution for histomorphological hypothesis generation from whole-slide classifiers

    Tristan Lazard, Kenza Bouzid, Julius Hense +6

    q-bio.QMeess.IVarXiv:2609.02985v12026
  27. Tree-CNN: A Hierarchical Deep Convolutional Neural Network for Incremental Learning

    Deboleena Roy, Priyadarshini Panda, Kaushik Roy

    cs.CVcs.AIeess.IVarXiv:1802.05800v32018
  28. Deep Learning for Medical Anomaly Detection -- A Survey

    Tharindu Fernando, Harshala Gammulle, Simon Denman +2

    cs.LGcs.CVeess.IVarXiv:2012.02364v22020
  29. Causal Contextual Prediction for Learned Image Compression

    Zongyu Guo, Zhizheng Zhang, Runsen Feng +1

    cs.CVeess.IVarXiv:2011.09704v52020
  30. Optimization for Medical Image Segmentation: Theory and Practice when evaluating with Dice Score or Jaccard Index

    Tom Eelbode, Jeroen Bertels, Maxim Berman +4

    eess.IVcs.CVcs.LGarXiv:2010.13499v12020
  31. Same Same But DifferNet: Semi-Supervised Defect Detection with Normalizing Flows

    Marco Rudolph, Bastian Wandt, Bodo Rosenhahn

    cs.CVcs.LGeess.IVarXiv:2008.12577v12020
  32. Video Super Resolution Based on Deep Learning: A Comprehensive Survey

    Hongying Liu, Zhubo Ruan, Peng Zhao +5

    cs.CVeess.IVarXiv:2007.12928v32020
  33. Exponential Pixelating Integral transform with dual fractal features for enhanced chest X-ray abnormality detection

    Naveenraj Kamalakannan, Sri Ram Macharla, M Kanimozhi +1

    eess.IVcs.CVarXiv:2609.10988v12026
  34. Reliability-Aware Hybrid-K Ensemble Selection for Cervical Cytology Classification: Integrating Discrimination, Calibration, and Selective Prediction

    Nisreen Albzour, Sarah S. Lam

    eess.IVcs.AIcs.CVarXiv:2609.09189v12026
  35. Pyramidal Convolution: Rethinking Convolutional Neural Networks for Visual Recognition

    Ionut Cosmin Duta, Li Liu, Fan Zhu +1

    cs.CVcs.LGeess.IVarXiv:2006.11538v12020
  36. Change Guiding Network: Incorporating Change Prior to Guide Change Detection in Remote Sensing Imagery

    Chengxi Han, Chen Wu, Haonan Guo +3

    cs.CVeess.IVarXiv:2404.09179v12024
  37. On the use of deep learning for phase recovery

    Kaiqiang Wang, Li Song, Chutian Wang +8

    physics.opticscs.LGeess.IVarXiv:2308.00942v12023
  38. Exploiting Deep Generative Prior for Versatile Image Restoration and Manipulation

    Xingang Pan, Xiaohang Zhan, Bo Dai +3

    eess.IVcs.CVarXiv:2003.13659v42020
  39. Rapid AI Development Cycle for the Coronavirus (COVID-19) Pandemic: Initial Results for Automated Detection & Patient Monitoring using Deep Learning CT Image Analysis

    Ophir Gozes, Maayan Frid-Adar, Hayit Greenspan +5

    eess.IVcs.CVcs.LGarXiv:2003.05037v32020
  40. Celeb-DF: A Large-scale Challenging Dataset for DeepFake Forensics

    Yuezun Li, Xin Yang, Pu Sun +2

    cs.CRcs.CVeess.IVarXiv:1909.12962v42019
  41. Enhancement of Underwater Images with Statistical Model of Background Light and Optimization of Transmission Map

    Wei Song, Yan Wang, Dongmei Huang +2

    eess.IVcs.MMarXiv:1906.08673v12019
  42. Rethinking the Unpretentious U-net for Medical Ultrasound Image Segmentation

    Gongping Chen, Lei Li, JianXun Zhang +1

    eess.IVcs.CVcs.LGarXiv:2209.07193v42022
  43. Recurrent Video Restoration Transformer with Guided Deformable Attention

    Jingyun Liang, Yuchen Fan, Xiaoyu Xiang +7

    cs.CVeess.IVarXiv:2206.02146v32022
  44. Decoupled-and-Coupled Networks: Self-Supervised Hyperspectral Image Super-Resolution with Subpixel Fusion

    Danfeng Hong, Jing Yao, Deyu Meng +2

    eess.IVcs.CVarXiv:2205.03742v12022
  45. Towards An End-to-End Framework for Flow-Guided Video Inpainting

    Zhen Li, Cheng-Ze Lu, Jianhua Qin +2

    eess.IVcs.CVarXiv:2204.02663v22022
  46. Polyp-PVT: Polyp Segmentation with Pyramid Vision Transformers

    Bo Dong, Wenhai Wang, Deng-Ping Fan +3

    eess.IVcs.CVarXiv:2108.06932v82021
  47. MedXIAOHE: A Comprehensive Recipe for Building Medical MLLMs

    Baorong Shi, Bo Cui, Boyuan Jiang +17

    cs.CLcs.AIcs.CVarXiv:2602.12705v42026
  48. Applications of Deep Learning Techniques for Automated Multiple Sclerosis Detection Using Magnetic Resonance Imaging: A Review

    Afshin Shoeibi, Marjane Khodatars, Mahboobeh Jafari +9

    eess.IVcs.CVarXiv:2105.04881v22021
  49. Seed3D 1.0: From Images to High-Fidelity Simulation-Ready 3D Assets

    Jiashi Feng, Xiu Li, Jing Lin +25

    eess.IVcs.CVarXiv:2510.19944v12025
  50. Deep learning-based computed tomography (CT) derived body composition classifier for colorectal cancer patients

    Eve Harling, Chattarin Pumtako, Bernd Porr +2

    eess.IVcs.LGarXiv:2608.15712v12026
  51. Explainable artificial intelligence (XAI) in deep learning-based medical image analysis

    Bas H. M. van der Velden, Hugo J. Kuijf, Kenneth G. A. Gilhuijs +1

    eess.IVcs.CVarXiv:2107.10912v12021
  52. Computational Imaging Without a Computer: Seeing Through Random Diffusers at the Speed of Light

    Yi Luo, Yifan Zhao, Jingxi Li +4

    physics.opticseess.IVarXiv:2107.06586v12021
  53. Common Limitations of Image Processing Metrics: A Picture Story

    Annika Reinke, Minu D. Tizabi, Carole H. Sudre +90

    eess.IVcs.CVarXiv:2104.05642v82021
  54. Image biomarker standardisation initiative

    Alex Zwanenburg, Stefan Leger, Martin Vallières +1

    cs.CVeess.IVarXiv:1612.07003v112016
  55. Cooperative Multi-Task Semantic Communication for Joint Classification and Regression Tasks

    Ahmad Halimi Razlighi, Mohammad Siddiqur Rahman, Maximilian H. V. Tillmann +2

    eess.SPcs.LGeess.IVarXiv:2609.03977v12026
  56. Plant Growth Estimation with a Camera-Based Vegetation Index Mapping System for Agricultural Ground Vehicles

    Lukas Pindl, Michael Maier, Timo Oksanen

    eess.IVarXiv:2609.03872v12026
  57. Defending Against Physically Realizable Attacks on Image Classification

    Tong Wu, Liang Tong, Yevgeniy Vorobeychik

    cs.LGcs.AIcs.CVarXiv:1909.09552v22019
  58. Seamless Whole Slide Label-Free Virtual Staining

    Dou Hoon Kwark, Kianoush Falahkheirkhah, Ji-hun Oh +3

    eess.IVcs.CVq-bio.QMarXiv:2609.10914v12026
  59. UBone3D: Physics-Rectified Conditional Flow Matching for Anatomical 3D Shape Completion from Ultrasound

    Weiying Chen, Yuchong Gao, Siyuan Li +3

    cs.CVeess.IVarXiv:2609.11506v12026
  60. Rethinking Handwritten Character Recognition

    Ranjit Raut, Aarav Subedi, Ashim Shrestha

    cs.CVeess.IVarXiv:2609.10572v12026