Image and Video Processing

Papers filed under eess.IV on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.

Search paper metadata (including unsummarized papers)

781 to 840 of 1,302

  1. Finding Covid-19 from Chest X-rays using Deep Learning on a Small Dataset

    Lawrence O. Hall, Rahul Paul, Dmitry B. Goldgof +1

    eess.IVcs.CVcs.LGarXiv:2004.02060v42020
  2. ViT-V-Net: Vision Transformer for Unsupervised Volumetric Medical Image Registration

    Junyu Chen, Yufan He, Eric C. Frey +2

    eess.IVcs.CVarXiv:2104.06468v12021
  3. A deep learning system for differential diagnosis of skin diseases

    Yuan Liu, Ayush Jain, Clara Eng +19

    eess.IVcs.CVarXiv:1909.05382v12019
  4. Brain MRI Super Resolution Using 3D Deep Densely Connected Neural Networks

    Yuhua Chen, Yibin Xie, Zhengwei Zhou +3

    cs.CVeess.IVarXiv:1801.02728v12018
  5. Towards an Effective and Efficient Deep Learning Model for COVID-19 Patterns Detection in X-ray Images

    Eduardo Luz, Pedro Lopes Silva, Rodrigo Silva +3

    eess.IVcs.CVcs.LGarXiv:2004.05717v52020
  6. Cross-domain Correspondence Learning for Exemplar-based Image Translation

    Pan Zhang, Bo Zhang, Dong Chen +2

    cs.CVcs.GReess.IVarXiv:2004.05571v12020
  7. VisualVoice: Audio-Visual Speech Separation with Cross-Modal Consistency

    Ruohan Gao, Kristen Grauman

    cs.CVcs.SDeess.IVarXiv:2101.03149v22021
  8. PanNuke Dataset Extension, Insights and Baselines

    Jevgenij Gamper, Navid Alemi Koohbanani, Ksenija Benes +6

    eess.IVcs.CVq-bio.QMarXiv:2003.10778v72020
  9. TResNet: High Performance GPU-Dedicated Architecture

    Tal Ridnik, Hussam Lawen, Asaf Noy +3

    cs.CVcs.LGeess.IVarXiv:2003.13630v32020
  10. X2CT-GAN: Reconstructing CT from Biplanar X-Rays with Generative Adversarial Networks

    Xingde Ying, Heng Guo, Kai Ma +3

    eess.IVcs.CVarXiv:1905.06902v12019
  11. FourLLIE: Boosting Low-Light Image Enhancement by Fourier Frequency Information

    Chenxi Wang, Hongjun Wu, Zhi Jin

    cs.CVeess.IVarXiv:2308.03033v12023
  12. Missing MRI Pulse Sequence Synthesis using Multi-Modal Generative Adversarial Network

    Anmol Sharma, Ghassan Hamarneh

    eess.IVcs.AIcs.CVarXiv:1904.12200v32019
  13. A Physics-based Noise Formation Model for Extreme Low-light Raw Denoising

    Kaixuan Wei, Ying Fu, Jiaolong Yang +1

    eess.IVcs.CVarXiv:2003.12751v22020
  14. InfoNeRF: Ray Entropy Minimization for Few-Shot Neural Volume Rendering

    Mijeong Kim, Seonguk Seo, Bohyung Han

    cs.CVcs.GReess.IVarXiv:2112.15399v22021
  15. Hyperspectral Image Classification with Attention Aided CNNs

    Renlong Hang, Zhu Li, Qingshan Liu +2

    eess.IVcs.CVarXiv:2005.11977v22020
  16. Vine disease detection in UAV multispectral images with deep learning segmentation approach

    Mohamed Kerkech, Adel Hafiane, Raphael Canals

    eess.IVarXiv:1912.05281v12019
  17. Ada-TokenCom: Rate-Adaptive Token Communications via Large-Model-Driven Token Compression and Generation

    Zijun Zhang, Li Qiao, Mahdi Boloursaz Mashhadi +3

    cs.ITcs.CVeess.IVarXiv:2608.28086v12026
  18. Deep-neural-network based sinogram synthesis for sparse-view CT image reconstruction

    Hoyeon Lee, Jongha Lee, Hyeongseok Kim +2

    physics.med-phcs.CVeess.IVarXiv:1803.00694v22018
  19. Change Detection Methods for Remote Sensing in the Last Decade: A Comprehensive Review

    Guangliang Cheng, Yunmeng Huang, Xiangtai Li +4

    cs.CVcs.LGeess.IVarXiv:2305.05813v12023
  20. Multi-Temporal Recurrent Neural Networks For Progressive Non-Uniform Single Image Deblurring With Incremental Temporal Training

    Dongwon Park, Dong Un Kang, Jisoo Kim +1

    eess.IVcs.CVarXiv:1911.07410v12019
  21. More Diverse Means Better: Multimodal Deep Learning Meets Remote Sensing Imagery Classification

    Danfeng Hong, Lianru Gao, Naoto Yokoya +4

    cs.CVeess.IVarXiv:2008.05457v12020
  22. Video Face Manipulation Detection Through Ensemble of CNNs

    Nicolò Bonettini, Edoardo Daniele Cannas, Sara Mandelli +3

    cs.CVcs.MMeess.IVarXiv:2004.07676v12020
  23. PHiSeg: Capturing Uncertainty in Medical Image Segmentation

    Christian F. Baumgartner, Kerem C. Tezcan, Krishna Chaitanya +6

    eess.IVcs.LGstat.MLarXiv:1906.04045v22019
  24. DeepISP: Towards Learning an End-to-End Image Processing Pipeline

    Eli Schwartz, Raja Giryes, Alex M. Bronstein

    eess.IVcs.CVarXiv:1801.06724v22018
  25. Machine Learning Techniques for Biomedical Image Segmentation: An Overview of Technical Aspects and Introduction to State-of-Art Applications

    Hyunseok Seo, Masoud Badiei Khuzani, Varun Vasudevan +5

    eess.IVcs.CVcs.LGarXiv:1911.02521v12019
  26. DuDoNet: Dual Domain Network for CT Metal Artifact Reduction

    Wei-An Lin, Haofu Liao, Cheng Peng +5

    eess.IVcs.CVarXiv:1907.00273v12019
  27. CheXtriev: Anatomy-Centered Representation for Case-Based Retrieval of Chest Radiographs

    Naren Akash, Arihanth Tadanki, Jayanthi Sivaswamy

    eess.IVcs.AIcs.CVarXiv:2608.28137v12026
  28. Self-supervised learning methods and applications in medical imaging analysis: A survey

    Saeed Shurrab, Rehab Duwairi

    eess.IVcs.CVcs.LGarXiv:2109.08685v32021
  29. Do Medical Vision Models Reason About Anatomy? Probing the Spatial Inductive Biases of Learned Visual Representations

    Naren Akash, Neeraja Ramanan

    eess.IVcs.AIcs.CVarXiv:2608.28092v12026
  30. 3D Self-Supervised Methods for Medical Imaging

    Aiham Taleb, Winfried Loetzsch, Noel Danz +4

    cs.CVcs.LGeess.IVarXiv:2006.03829v32020
  31. Frequency Separation for Real-World Super-Resolution

    Manuel Fritsche, Shuhang Gu, Radu Timofte

    eess.IVcs.CVarXiv:1911.07850v12019
  32. CARDINAL Predicts Cardiovascular Risk From Non-contrast Cardiac CT

    Roy Gabriel, Nattakorn Kittisut, Jamshid Hassanpour +6

    eess.IVcs.AIcs.CVarXiv:2608.27690v12026
  33. On Translation Invariance in CNNs: Convolutional Layers can Exploit Absolute Spatial Location

    Osman Semih Kayhan, Jan C. van Gemert

    cs.CVcs.LGeess.IVarXiv:2003.07064v22020
  34. SynthMorph: learning contrast-invariant registration without acquired images

    Malte Hoffmann, Benjamin Billot, Douglas N. Greve +3

    eess.IVcs.CVq-bio.NCarXiv:2004.10282v42020
  35. EGE-UNet: an Efficient Group Enhanced UNet for skin lesion segmentation

    Jiacheng Ruan, Mingye Xie, Jingsheng Gao +2

    eess.IVcs.CVarXiv:2307.08473v12023
  36. Destroy Me: Automatic Artifact Generation for Histopathology Images

    Zuzanna Krawczyk-Borysiak, Adam Krawczyk, Mateusz Miller +4

    eess.IVcs.AIcs.CVarXiv:2608.27516v12026
  37. LadderNet: Multi-path networks based on U-Net for medical image segmentation

    Juntang Zhuang

    cs.CVeess.IVarXiv:1810.07810v42018
  38. EPOS: Estimating 6D Pose of Objects with Symmetries

    Tomas Hodan, Daniel Barath, Jiri Matas

    cs.CVcs.LGcs.ROarXiv:2004.00605v12020
  39. Breast cancer detection using artificial intelligence techniques: A systematic literature review

    Ali Bou Nassif, Manar Abu Talib, Qassim Nasir +2

    eess.IVcs.AIcs.CVarXiv:2203.04308v12022
  40. M-LVC: Multiple Frames Prediction for Learned Video Compression

    Jianping Lin, Dong Liu, Houqiang Li +1

    eess.IVcs.CVcs.LGarXiv:2004.10290v12020
  41. Segment Anything Model (SAM) for Digital Pathology: Assess Zero-shot Segmentation on Whole Slide Imaging

    Ruining Deng, Can Cui, Quan Liu +13

    eess.IVcs.CVarXiv:2304.04155v12023
  42. SPACE: Unsupervised Object-Oriented Scene Representation via Spatial Attention and Decomposition

    Zhixuan Lin, Yi-Fu Wu, Skand Vishwanath Peri +5

    cs.LGcs.CVeess.IVarXiv:2001.02407v32020
  43. Retinal vessel segmentation based on Fully Convolutional Neural Networks

    Américo Oliveira, Sérgio Pereira, Carlos A. Silva

    eess.IVcs.LGstat.MLarXiv:1812.07110v22018
  44. Disentangling Light Fields for Super-Resolution and Disparity Estimation

    Yingqian Wang, Longguang Wang, Gaochang Wu +4

    eess.IVcs.CVarXiv:2202.10603v52022
  45. Hybrid Spatial-Temporal Entropy Modelling for Neural Video Compression

    Jiahao Li, Bin Li, Yan Lu

    eess.IVcs.CVcs.MMarXiv:2207.05894v12022
  46. Variable Rate Deep Image Compression With a Conditional Autoencoder

    Yoojin Choi, Mostafa El-Khamy, Jungwon Lee

    eess.IVcs.CVarXiv:1909.04802v12019
  47. Zoom To Learn, Learn To Zoom

    Xuaner Cecilia Zhang, Qifeng Chen, Ren Ng +1

    cs.CVeess.IVarXiv:1905.05169v12019
  48. DeepCervix: A Deep Learning-based Framework for the Classification of Cervical Cells Using Hybrid Deep Feature Fusion Techniques

    Md Mamunur Rahaman, Chen Li, Yudong Yao +4

    eess.IVcs.CVarXiv:2102.12191v22021
  49. Scan2Cap: Context-aware Dense Captioning in RGB-D Scans

    Dave Zhenyu Chen, Ali Gholami, Matthias Nießner +1

    cs.CVcs.LGeess.IVarXiv:2012.02206v12020
  50. Single Image Deraining: From Model-Based to Data-Driven and Beyond

    Wenhan Yang, Robby T. Tan, Shiqi Wang +2

    eess.IVcs.CVarXiv:1912.07150v22019
  51. TeCNO: Surgical Phase Recognition with Multi-Stage Temporal Convolutional Networks

    Tobias Czempiel, Magdalini Paschali, Matthias Keicher +4

    eess.IVcs.CVcs.LGarXiv:2003.10751v12020
  52. Densely Residual Laplacian Super-Resolution

    Saeed Anwar, Nick Barnes

    eess.IVcs.CVarXiv:1906.12021v22019
  53. Local Texture Estimator for Implicit Representation Function

    Jaewon Lee, Kyong Hwan Jin

    cs.CVeess.IVarXiv:2111.08918v62021
  54. Image Reconstruction: From Sparsity to Data-adaptive Methods and Machine Learning

    Saiprasad Ravishankar, Jong Chul Ye, Jeffrey A. Fessler

    eess.IVcs.LGstat.MLarXiv:1904.02816v32019
  55. Nonlinear Transform Coding

    Johannes Ballé, Philip A. Chou, David Minnen +5

    cs.ITeess.IVarXiv:2007.03034v22020
  56. mmFormer: Multimodal Medical Transformer for Incomplete Multimodal Learning of Brain Tumor Segmentation

    Yao Zhang, Nanjun He, Jiawei Yang +6

    eess.IVcs.CVarXiv:2206.02425v22022
  57. PST900: RGB-Thermal Calibration, Dataset and Segmentation Network

    Shreyas S. Shivakumar, Neil Rodrigues, Alex Zhou +3

    cs.CVcs.ROeess.IVarXiv:1909.10980v12019
  58. Wireless Semantic Communications for Video Conferencing

    Peiwen Jiang, Chao-Kai Wen, Shi Jin +1

    eess.IVeess.SParXiv:2204.07790v12022
  59. Exploring Smoothness and Class-Separation for Semi-supervised Medical Image Segmentation

    Yicheng Wu, Zhonghua Wu, Qianyi Wu +2

    eess.IVcs.CVarXiv:2203.01324v32022
  60. AutoGAN: Neural Architecture Search for Generative Adversarial Networks

    Xinyu Gong, Shiyu Chang, Yifan Jiang +1

    cs.CVcs.LGeess.IVarXiv:1908.03835v12019