Image and Video Processing

Papers filed under eess.IV on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.

Search paper metadata (including unsummarized papers)

541 to 600 of 1,302

  1. TGANet: Text-guided attention for improved polyp segmentation

    Nikhil Kumar Tomar, Debesh Jha, Ulas Bagci +1

    eess.IVcs.CVcs.LGarXiv:2205.04280v12022
  2. ResRep: Lossless CNN Pruning via Decoupling Remembering and Forgetting

    Xiaohan Ding, Tianxiang Hao, Jianchao Tan +4

    cs.LGcs.CVeess.IVarXiv:2007.03260v42020
  3. Recognition of Ischaemia and Infection in Diabetic Foot Ulcers: Dataset and Techniques

    Manu Goyal, Neil Reeves, Satyan Rajbhandari +3

    eess.IVcs.CVarXiv:1908.05317v42019
  4. Inversion by Direct Iteration: An Alternative to Denoising Diffusion for Image Restoration

    Mauricio Delbracio, Peyman Milanfar

    eess.IVcs.CVcs.LGarXiv:2303.11435v52023
  5. Use HiResCAM instead of Grad-CAM for faithful explanations of convolutional neural networks

    Rachel Lea Draelos, Lawrence Carin

    eess.IVcs.CVcs.LGarXiv:2011.08891v42020
  6. Spot the conversation: speaker diarisation in the wild

    Joon Son Chung, Jaesung Huh, Arsha Nagrani +2

    cs.SDcs.CVeess.ASarXiv:2007.01216v32020
  7. SNE-RoadSeg: Incorporating Surface Normal Information into Semantic Segmentation for Accurate Freespace Detection

    Rui Fan, Hengli Wang, Peide Cai +1

    cs.CVcs.ROeess.IVarXiv:2008.11351v12020
  8. Motion-Aware Feature for Improved Video Anomaly Detection

    Yi Zhu, Shawn Newsam

    cs.CVcs.LGeess.IVarXiv:1907.10211v12019
  9. Beyond Blur: A Semantic Tri-view Pipeline for Teledermatology Gradability via Skin Micro-relief

    Robert Engel

    eess.IVcs.CVcs.HCarXiv:2609.03095v12026
  10. Towards Interpretable Semantic Segmentation via Gradient-weighted Class Activation Mapping

    Kira Vinogradova, Alexandr Dibrov, Gene Myers

    cs.CVcs.LGeess.IVarXiv:2002.11434v12020
  11. Automatic Scene Inference for 3D Object Compositing

    Kevin Karsch, Kalyan Sunkavalli, Sunil Hadap +4

    cs.GReess.IVarXiv:1912.12297v12019
  12. GazeRefine: Expert Gaze as a Test-Time Prompt for Training-Free Medical Image Segmentation

    Mohammed Oussama Benyahia, Marouane Tliba, Mohamed Amine Kerkouri +10

    eess.IVcs.AIcs.CVarXiv:2609.01310v12026
  13. Prostate Cancer Detection using Deep Convolutional Neural Networks

    Sunghwan Yoo, Isha Gujrathi, Masoom A. Haider +1

    cs.CVeess.IVq-bio.QMarXiv:1905.13145v12019
  14. Hetero-Center Loss for Cross-Modality Person Re-Identification

    Yuanxin Zhu, Zhao Yang, Li Wang +3

    cs.CVeess.IVarXiv:1910.09830v12019
  15. Noise Flow: Noise Modeling with Conditional Normalizing Flows

    Abdelrahman Abdelhamed, Marcus A. Brubaker, Michael S. Brown

    cs.CVcs.LGeess.IVarXiv:1908.08453v12019
  16. COVID-19 Infection Localization and Severity Grading from Chest X-ray Images

    Anas M. Tahir, Muhammad E. H. Chowdhury, Amith Khandakar +11

    eess.IVcs.CVarXiv:2103.07985v12021
  17. Perceptual Video Quality Assessment: A Survey

    Xiongkuo Min, Huiyu Duan, Wei Sun +2

    cs.MMcs.CVeess.IVarXiv:2402.03413v12024
  18. BVI-DVC: A Training Database for Deep Video Compression

    Di Ma, Fan Zhang, David R. Bull

    eess.IVcs.CVarXiv:2003.13552v22020
  19. Linking Points With Labels in 3D: A Review of Point Cloud Semantic Segmentation

    Yuxing Xie, Jiaojiao Tian, Xiao Xiang Zhu

    cs.CVcs.LGeess.IVarXiv:1908.08854v32019
  20. Predicting Risk of Developing Diabetic Retinopathy using Deep Learning

    Ashish Bora, Siva Balasubramanian, Boris Babenko +13

    eess.IVcs.CVarXiv:2008.04370v12020
  21. EEG-Inception: An Accurate and Robust End-to-End Neural Network for EEG-based Motor Imagery Classification

    Ce Zhang, Young-Keun Kim, Azim Eskandarian

    eess.SPcs.HCcs.LGarXiv:2101.10932v32021
  22. HarmoFL: Harmonizing Local and Global Drifts in Federated Learning on Heterogeneous Medical Images

    Meirui Jiang, Zirui Wang, Qi Dou

    eess.IVcs.AIcs.CVarXiv:2112.10775v32021
  23. Deep Gradient Projection Networks for Pan-sharpening

    Shuang Xu, Jiangshe Zhang, Zixiang Zhao +3

    cs.CVeess.IVarXiv:2103.04584v12021
  24. MonoDETR: Depth-guided Transformer for Monocular 3D Object Detection

    Renrui Zhang, Han Qiu, Tai Wang +7

    cs.CVcs.AIeess.IVarXiv:2203.13310v52022
  25. Plug-and-Play Methods for Integrating Physical and Learned Models in Computational Imaging

    Ulugbek S. Kamilov, Charles A. Bouman, Gregery T. Buzzard +1

    eess.IVarXiv:2203.17061v32022
  26. Transmission of natural scene images through a multimode fibre

    Piergiorgio Caramazza, Oisín Moran, Roderick Murray-Smith +1

    eess.IVphysics.opticsarXiv:1904.11985v12019
  27. Fourier Space Losses for Efficient Perceptual Image Super-Resolution

    Dario Fuoli, Luc Van Gool, Radu Timofte

    eess.IVcs.CVarXiv:2106.00783v12021
  28. Lightweight Interpretable RGB-Guided Hyperspectral Super-Resolution under Real Cross-resolution Misalignment

    Mohamad Jouni, Aurélien Godet, Mauro Dalla Mura

    eess.IVcs.CVarXiv:2609.01060v12026
  29. One-pass Multi-task Networks with Cross-task Guided Attention for Brain Tumor Segmentation

    Chenhong Zhou, Changxing Ding, Xinchao Wang +2

    cs.CVcs.AIcs.LGarXiv:1906.01796v22019
  30. Fine-Tuning and Training of DenseNet for Histopathology Image Representation Using TCGA Diagnostic Slides

    Abtin Riasatian, Morteza Babaie, Danial Maleki +19

    eess.IVarXiv:2101.07903v12021
  31. Pansharpening via Detail Injection Based Convolutional Neural Networks

    Lin He, Yizhou Rao, Jun Li +2

    eess.IVarXiv:1806.08898v12018
  32. FCN-Transformer Feature Fusion for Polyp Segmentation

    Edward Sanderson, Bogdan J. Matuszewski

    eess.IVcs.CVcs.LGarXiv:2208.08352v12022
  33. Compressing AI Traffic: Standardized Neural Network Coding of Visual-Token Representations in Split Vision-Language Inference

    Reza Heidari, Hamed R. Tavakoli, Juho Kannala

    cs.CVeess.IVarXiv:2609.01200v12026
  34. SemanticAdv: Generating Adversarial Examples via Attribute-conditional Image Editing

    Haonan Qiu, Chaowei Xiao, Lei Yang +3

    cs.LGcs.CRcs.CVarXiv:1906.07927v42019
  35. LoFi RADIO: A Distilled In-Domain Backbone Applied for Artifact-Severity Grading of Ultra-Low-Field Neonatal Brain MR

    Jonathan B. Martin, Yashwant Kurmi, Charlotte R. Sappo

    eess.IVcs.CVarXiv:2609.02676v12026
  36. Learning End-to-End Lossy Image Compression: A Benchmark

    Yueyu Hu, Wenhan Yang, Zhan Ma +1

    eess.IVcs.CVarXiv:2002.03711v42020
  37. Generative Joint Source-Channel Coding for Semantic Image Transmission

    Ecenaz Erdemir, Tze-Yang Tung, Pier Luigi Dragotti +1

    eess.IVcs.AIcs.LGarXiv:2211.13772v12022
  38. VFHQ: A High-Quality Dataset and Benchmark for Video Face Super-Resolution

    Liangbin Xie. Xintao Wang, Honglun Zhang, Chao Dong +1

    eess.IVcs.AIcs.CVarXiv:2205.03409v12022
  39. Zooming Slow-Mo: Fast and Accurate One-Stage Space-Time Video Super-Resolution

    Xiaoyu Xiang, Yapeng Tian, Yulun Zhang +3

    cs.CVcs.MMeess.IVarXiv:2002.11616v12020
  40. Deep Joint Source-Channel Coding for Wireless Image Transmission with Adaptive Rate Control

    Mingyu Yang, Hun-Seok Kim

    eess.SPcs.LGeess.IVarXiv:2110.04456v12021
  41. Waste detection in Pomerania: non-profit project for detecting waste in environment

    Sylwia Majchrowska, Agnieszka Mikołajczyk, Maria Ferlin +4

    cs.CVeess.IVarXiv:2105.06808v12021
  42. The impact of patient clinical information on automated skin cancer detection

    Andre G. C. Pacheco, Renato A. Krohling

    eess.IVcs.CVcs.LGarXiv:1909.12912v12019
  43. An interpretable classifier for high-resolution breast cancer screening images utilizing weakly supervised localization

    Yiqiu Shen, Nan Wu, Jason Phang +8

    cs.CVcs.LGeess.IVarXiv:2002.07613v12020
  44. Real-time 3D reconstruction from single-photon lidar data using plug-and-play point cloud denoisers

    Julián Tachella, Yoann Altmann, Nicolas Mellado +5

    eess.IVphysics.opticsarXiv:1905.06700v22019
  45. NAM: Normalization-based Attention Module

    Yichao Liu, Zongru Shao, Yueyang Teng +1

    cs.CVeess.IVarXiv:2111.12419v12021
  46. Interpretable Survival Prediction for Colorectal Cancer using Deep Learning

    Ellery Wulczyn, David F. Steiner, Melissa Moran +20

    eess.IVcs.CVarXiv:2011.08965v12020
  47. Learning Convolutional Transforms for Lossy Point Cloud Geometry Compression

    Maurice Quach, Giuseppe Valenzise, Frederic Dufaux

    cs.CVcs.LGeess.IVarXiv:1903.08548v22019
  48. Fully Convolutional Change Detection Framework with Generative Adversarial Network for Unsupervised, Weakly Supervised and Regional Supervised Change Detection

    Chen Wu, Bo Du, Liangpei Zhang

    cs.CVcs.AIeess.IVarXiv:2201.06030v12022
  49. Learned Point Cloud Geometry Compression

    Jianqiang Wang, Hao Zhu, Zhan Ma +3

    cs.CVeess.IVarXiv:1909.12037v12019
  50. Intracranial Hemorrhage Segmentation Using Deep Convolutional Model

    Murtadha D. Hssayeni, M. S., Muayad S. Croock +9

    eess.IVcs.CVarXiv:1910.08643v22019
  51. CMasher: Scientific colormaps for making accessible, informative and 'cmashing' plots

    Ellert van der Velden

    eess.IVphysics.data-anarXiv:2003.01069v12020
  52. Multi-modal Dense Video Captioning

    Vladimir Iashin, Esa Rahtu

    cs.CVcs.CLcs.LGarXiv:2003.07758v22020
  53. Motion-Attentive Transition for Zero-Shot Video Object Segmentation

    Tianfei Zhou, Shunzhou Wang, Yi Zhou +3

    cs.CVcs.LGeess.IVarXiv:2003.04253v32020
  54. Temporal Attentive Alignment for Large-Scale Video Domain Adaptation

    Min-Hung Chen, Zsolt Kira, Ghassan AlRegib +3

    cs.CVcs.LGcs.MMarXiv:1907.12743v62019
  55. Identifying Corresponding Patches in SAR and Optical Images with a Pseudo-Siamese CNN

    Lloyd H. Hughes, Michael Schmitt, Lichao Mou +2

    eess.IVcs.CVarXiv:1801.08467v12018
  56. Seeing Beyond the Lesion: Disease Recognition from Reactive CNS Tissue

    Jan Schnorrenberg, Jan Ernsting, Enrico Küllenberg +3

    eess.IVcs.CVq-bio.TOarXiv:2609.02390v12026
  57. Test-Time Adaptable Neural Networks for Robust Medical Image Segmentation

    Neerav Karani, Ertunc Erdil, Krishna Chaitanya +1

    eess.IVcs.CVcs.LGarXiv:2004.04668v42020
  58. U-Net v2: Rethinking the Skip Connections of U-Net for Medical Image Segmentation

    Yaopeng Peng, Milan Sonka, Danny Z. Chen

    eess.IVcs.CVarXiv:2311.17791v22023
  59. Diagnosis of Coronavirus Disease 2019 (COVID-19) with Structured Latent Multi-View Representation Learning

    Hengyuan Kang, Liming Xia, Fuhua Yan +8

    eess.IVcs.CVcs.LGarXiv:2005.03227v12020
  60. Data-Efficient Networks for Multi-Contrast MRI Reconstruction based on a Generalized Content/Style Prior

    Chinmay Rao, Efe Ilıcak, Matthias J. P. van Osch +5

    eess.IVcs.CVarXiv:2609.01959v12026