Image and Video Processing

Papers filed under eess.IV on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.

Search paper metadata (including unsummarized papers)

121 to 180 of 1,302

  1. CovidAID: COVID-19 Detection Using Chest X-Ray

    Arpan Mangal, Surya Kalia, Harish Rajgopal +4

    eess.IVcs.CVcs.LGarXiv:2004.09803v12020
  2. Towards Flexible Blind JPEG Artifacts Removal

    Jiaxi Jiang, Kai Zhang, Radu Timofte

    eess.IVcs.CVarXiv:2109.14573v12021
  3. Storage-Scalable Progressive Semantic Communication via Knowledge-Base Reuse

    Heng Zhu, Ye Liu, Kun Zhu +1

    cs.LGcs.NIeess.IVarXiv:2609.10112v12026
  4. Uncertainty-Informed Deep Learning Models Enable High-Confidence Predictions for Digital Histopathology

    James M Dolezal, Andrew Srisuwananukorn, Dmitry Karpeyev +13

    q-bio.QMcs.CVeess.IVarXiv:2204.04516v12022
  5. Deep-Learning-Based Image Segmentation Integrated with Optical Microscopy for Automatically Searching for Two-Dimensional Materials

    Satoru Masubuchi, Eisuke Watanabe, Yuta Seo +5

    eess.IVcond-mat.mtrl-sciarXiv:1910.12750v12019
  6. Unsupervised Domain Adaptation via Disentangled Representations: Application to Cross-Modality Liver Segmentation

    Junlin Yang, Nicha C. Dvornek, Fan Zhang +3

    eess.IVcs.CVarXiv:1907.13590v22019
  7. Bridging 2D and 3D Segmentation Networks for Computation Efficient Volumetric Medical Image Segmentation: An Empirical Study of 2.5D Solutions

    Yichi Zhang, Qingcheng Liao, Le Ding +1

    eess.IVcs.CVarXiv:2010.06163v22020
  8. Polyp-SAM: Transfer SAM for Polyp Segmentation

    Yuheng Li, Mingzhe Hu, Xiaofeng Yang

    eess.IVcs.CVarXiv:2305.00293v12023
  9. From W-Net to CDGAN: Bi-temporal Change Detection via Deep Learning Techniques

    Bin Hou, Qingjie Liu, Heng Wang +1

    cs.CVeess.IVarXiv:2003.06583v12020
  10. A Portable Brain MRI Scanner for Underserved Settings and Point-Of-Care Imaging

    Clarissa Z. Cooley, Patrick C. McDaniel, Jason P. Stockmann +9

    eess.IVphysics.med-pharXiv:2004.13183v12020
  11. X-Net: Brain Stroke Lesion Segmentation Based on Depthwise Separable Convolution and Long-range Dependencies

    Kehan Qi, Hao Yang, Cheng Li +4

    eess.IVcs.CVarXiv:1907.07000v22019
  12. Optimized U-Net for Brain Tumor Segmentation

    Michał Futrega, Alexandre Milesi, Michal Marcinkiewicz +1

    eess.IVcs.CVcs.LGarXiv:2110.03352v22021
  13. An automatic multi-tissue human fetal brain segmentation benchmark using the Fetal Tissue Annotation Dataset

    Kelly Payette, Priscille de Dumast, Hamza Kebiri +17

    eess.IVcs.CVarXiv:2010.15526v42020
  14. Multi-scale Attention Network for Single Image Super-Resolution

    Yan Wang, Yusen Li, Gang Wang +1

    eess.IVcs.CVarXiv:2209.14145v32022
  15. Multimodal Attention-based Deep Learning for Alzheimer's Disease Diagnosis

    Michal Golovanevsky, Carsten Eickhoff, Ritambhara Singh

    cs.LGcs.CVeess.IVarXiv:2206.08826v22022
  16. On the limits of cross-domain generalization in automated X-ray prediction

    Joseph Paul Cohen, Mohammad Hashir, Rupert Brooks +1

    eess.IVcs.LGq-bio.QMarXiv:2002.02497v22020
  17. Scalable Image Coding for Humans and Machines

    Hyomin Choi, Ivan V. Bajic

    eess.IVarXiv:2107.08373v22021
  18. The AVA-Kinetics Localized Human Actions Video Dataset

    Ang Li, Meghana Thotakuri, David A. Ross +3

    cs.CVcs.LGeess.IVarXiv:2005.00214v22020
  19. Generalizability of Machine Learning Models: Quantitative Evaluation of Three Methodological Pitfalls

    Farhad Maleki, Katie Ovens, Rajiv Gupta +3

    cs.LGcs.CVeess.IVarXiv:2202.01337v22022
  20. A Review on Deep Learning in Medical Image Reconstruction

    Haimiao Zhang, Bin Dong

    eess.IVcs.CVcs.LGarXiv:1906.10643v32019
  21. NestedFormer: Nested Modality-Aware Transformer for Brain Tumor Segmentation

    Zhaohu Xing, Lequan Yu, Liang Wan +2

    eess.IVcs.CVarXiv:2208.14876v12022
  22. Predicting Effective Diffusivity of Porous Media from Images by Deep Learning

    Haiyi Wu, Wen-Zhen Fang, Qinjun Kang +2

    physics.geo-pheess.IVarXiv:1912.05532v12019
  23. SAR-U-Net: squeeze-and-excitation block and atrous spatial pyramid pooling based residual U-Net for automatic liver segmentation in Computed Tomography

    Jinke Wang, Peiqing Lv, Haiying Wang +1

    eess.IVcs.CVcs.LGarXiv:2103.06419v32021
  24. Generalized Radiograph Representation Learning via Cross-supervision between Images and Free-text Radiology Reports

    Hong-Yu Zhou, Xiaoyu Chen, Yinghao Zhang +3

    eess.IVcs.CVcs.LGarXiv:2111.03452v22021
  25. Cross-Attention in Coupled Unmixing Nets for Unsupervised Hyperspectral Super-Resolution

    Jing Yao, Danfeng Hong, Jocelyn Chanussot +3

    eess.IVcs.CVarXiv:2007.05230v32020
  26. GTC: Guided Training of CTC Towards Efficient and Accurate Scene Text Recognition

    Wenyang Hu, Xiaocong Cai, Jun Hou +2

    cs.CVcs.LGeess.IVarXiv:2002.01276v12020
  27. Coronavirus (COVID-19) Classification using Deep Features Fusion and Ranking Technique

    Umut Ozkaya, Saban Ozturk, Mucahid Barstugan

    eess.IVcs.CVcs.LGarXiv:2004.03698v12020
  28. Multi-Modal Temporal Attention Models for Crop Mapping from Satellite Time Series

    Vivien Sainte Fare Garnot, Loic Landrieu, Nesrine Chehata

    cs.CVeess.IVarXiv:2112.07558v12021
  29. Deep Image Translation with an Affinity-Based Change Prior for Unsupervised Multimodal Change Detection

    Luigi Tommaso Luppino, Michael Kampffmeyer, Filippo Maria Bianchi +4

    cs.LGcs.CVeess.IVarXiv:2001.04271v22020
  30. An Uncertainty-aware Transfer Learning-based Framework for Covid-19 Diagnosis

    Afshar Shamsi Jokandan, Hamzeh Asgharnezhad, Shirin Shamsi Jokandan +5

    eess.IVcs.CVcs.LGarXiv:2007.14846v12020
  31. In-Vivo Hyperspectral Human Brain Image Database for Brain Cancer Detection

    H. Fabelo, S. Ortega, A. Szolna +25

    eess.IVcs.CVarXiv:2402.10776v12024
  32. Fast Unsupervised Brain Anomaly Detection and Segmentation with Diffusion Models

    Walter H. L. Pinaya, Mark S. Graham, Robert Gray +12

    cs.CVeess.IVq-bio.QMarXiv:2206.03461v12022
  33. Local Class-Specific and Global Image-Level Generative Adversarial Networks for Semantic-Guided Scene Generation

    Hao Tang, Dan Xu, Yan Yan +2

    cs.CVcs.LGeess.IVarXiv:1912.12215v32019
  34. The Multi-modality Cell Segmentation Challenge: Towards Universal Solutions

    Jun Ma, Ronald Xie, Shamini Ayyadhury +37

    eess.IVcs.CVcs.LGarXiv:2308.05864v22023
  35. Deep Attentive Features for Prostate Segmentation in 3D Transrectal Ultrasound

    Yi Wang, Haoran Dou, Xiaowei Hu +7

    eess.IVcs.AIcs.CVarXiv:1907.01743v22019
  36. Don't Forget The Past: Recurrent Depth Estimation from Monocular Video

    Vaishakh Patil, Wouter Van Gansbeke, Dengxin Dai +1

    cs.CVcs.LGcs.ROarXiv:2001.02613v22020
  37. Enhanced U-Net: A Feature Enhancement Network for Polyp Segmentation

    Krushi Patel, Andres M. Bur, Guanghui Wang

    eess.IVcs.CVarXiv:2105.00999v12021
  38. Generation of 3D Brain MRI Using Auto-Encoding Generative Adversarial Networks

    Gihyun Kwon, Chihye Han, Dae-shik Kim

    eess.IVcs.CVarXiv:1908.02498v12019
  39. Light Field Spatial Super-resolution via Deep Combinatorial Geometry Embedding and Structural Consistency Regularization

    Jing Jin, Junhui Hou, Jie Chen +1

    cs.CVeess.IVarXiv:2004.02215v12020
  40. Deep Learning in Photoacoustic Tomography: Current approaches and future directions

    Andreas Hauptmann, Ben Cox

    eess.IVcs.CVcs.LGarXiv:2009.07608v12020
  41. Myocardial Strain Drift Correction in Deep Learning Based Ultrasound Tracking

    Thierry Judge, Nicolas Duchateau, Andreas Østvik +6

    eess.IVcs.AIcs.CVarXiv:2609.09577v12026
  42. Spectral Superresolution of Multispectral Imagery with Joint Sparse and Low-Rank Learning

    Lianru Gao, Danfeng Hong, Jing Yao +3

    eess.IVcs.CVarXiv:2007.14006v12020
  43. Diffusion Probabilistic Model Made Slim

    Xingyi Yang, Daquan Zhou, Jiashi Feng +1

    cs.CVeess.IVarXiv:2211.17106v12022
  44. Multi-Layer Pseudo-Supervision for Histopathology Tissue Semantic Segmentation using Patch-level Classification Labels

    Chu Han, Jiatai Lin, Jinhai Mai +15

    eess.IVcs.CVq-bio.QMarXiv:2110.08048v12021
  45. SpaceNet 6: Multi-Sensor All Weather Mapping Dataset

    Jacob Shermeyer, Daniel Hogan, Jason Brown +8

    eess.IVcs.CVarXiv:2004.06500v12020
  46. A cross Transformer for image denoising

    Chunwei Tian, Menghua Zheng, Wangmeng Zuo +3

    eess.IVcs.CVcs.LGarXiv:2310.10408v12023
  47. Control Copy-Paste: Controllable Diffusion-Based Augmentation Method for Remote Sensing Few-Shot Object Detection

    Yanxing Liu, Jiancheng Pan, Bingchen Zhang

    eess.IVcs.CVarXiv:2507.21816v12025
  48. Diverse Instance Generation via Diffusion Models for Enhanced Few-Shot Object Detection in Remote Sensing Images

    Yanxing Liu, Jiancheng Pan, Jianwei Yang +3

    eess.IVcs.CVarXiv:2511.18031v12025
  49. Cellular-Connected Wireless Virtual Reality: Requirements, Challenges, and Solutions

    Fenghe Hu, Yansha Deng, Walid Saad +2

    eess.SPeess.IVarXiv:2001.06287v22020
  50. Self Pre-training with Masked Autoencoders for Medical Image Classification and Segmentation

    Lei Zhou, Huidong Liu, Joseph Bae +3

    eess.IVcs.CVcs.LGarXiv:2203.05573v22022
  51. Comparing SNNs and RNNs on Neuromorphic Vision Datasets: Similarities and Differences

    Weihua He, YuJie Wu, Lei Deng +6

    cs.CVcs.NEeess.IVarXiv:2005.02183v12020
  52. DeepJSCC-Q: Constellation Constrained Deep Joint Source-Channel Coding

    Tze-Yang Tung, David Burth Kurka, Mikolaj Jankowski +1

    eess.IVcs.LGeess.SParXiv:2206.08100v12022
  53. Iterative energy-based projection on a normal data manifold for anomaly localization

    David Dehaene, Oriel Frigo, Sébastien Combrexelle +1

    cs.CVeess.IVarXiv:2002.03734v12020
  54. Robust and interpretable blind image denoising via bias-free convolutional neural networks

    Sreyas Mohan, Zahra Kadkhodaie, Eero P. Simoncelli +1

    eess.IVcs.CVcs.LGarXiv:1906.05478v32019
  55. SEN12MS-CR-TS: A Remote Sensing Data Set for Multi-modal Multi-temporal Cloud Removal

    Patrick Ebel, Yajin Xu, Michael Schmitt +1

    cs.CVeess.IVarXiv:2201.09613v12022
  56. Seeing What You Said: Talking Face Generation Guided by a Lip Reading Expert

    Jiadong Wang, Xinyuan Qian, Malu Zhang +2

    cs.CVcs.AIeess.IVarXiv:2303.17480v12023
  57. Source-Free Domain Adaptive Fundus Image Segmentation with Denoised Pseudo-Labeling

    Cheng Chen, Quande Liu, Yueming Jin +2

    eess.IVcs.CVarXiv:2109.09735v12021
  58. Super sub-Nyquist single-pixel imaging by means of cake-cutting Hadamard basis sort

    Wen-Kai Yu

    eess.IVphysics.opticsarXiv:1903.11175v12019
  59. HED-UNet: Combined Segmentation and Edge Detection for Monitoring the Antarctic Coastline

    Konrad Heidler, Lichao Mou, Celia Baumhoer +2

    cs.CVeess.IVarXiv:2103.01849v12021
  60. Deep Hyperspectral Unmixing using Transformer Network

    Preetam Ghosh, Swalpa Kumar Roy, Bikram Koirala +2

    cs.CVeess.IVarXiv:2203.17076v12022