Image and Video Processing

Papers filed under eess.IV on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.

Search paper metadata (including unsummarized papers)

1,081 to 1,140 of 1,302

  1. NeRF in the Dark: High Dynamic Range View Synthesis from Noisy Raw Images

    Ben Mildenhall, Peter Hedman, Ricardo Martin-Brualla +2

    cs.CVcs.GReess.IVarXiv:2111.13679v12021
  2. Virtual KITTI 2

    Yohann Cabon, Naila Murray, Martin Humenberger

    cs.CVcs.ROeess.IVarXiv:2001.10773v12020
  3. SEAN: Image Synthesis with Semantic Region-Adaptive Normalization

    Peihao Zhu, Rameen Abdal, Yipeng Qin +1

    cs.CVcs.GReess.IVarXiv:1911.12861v22019
  4. Deepfakes Generation and Detection: State-of-the-art, open challenges, countermeasures, and way forward

    Momina Masood, Marriam Nawaz, Khalid Mahmood Malik +2

    cs.CRcs.LGcs.SDarXiv:2103.00484v22021
  5. DeepJSCC-f: Deep Joint Source-Channel Coding of Images with Feedback

    David Burth Kurka, Deniz Gündüz

    cs.ITcs.LGeess.IVarXiv:1911.11174v22019
  6. SynthStrip: Skull-Stripping for Any Brain Image

    Andrew Hoopes, Jocelyn S. Mora, Adrian V. Dalca +2

    eess.IVcs.CVphysics.med-pharXiv:2203.09974v22022
  7. YOLACT++: Better Real-time Instance Segmentation

    Daniel Bolya, Chong Zhou, Fanyi Xiao +1

    cs.CVcs.LGeess.IVarXiv:1912.06218v22019
  8. Multimodal Fusion Transformer for Remote Sensing Image Classification

    Swalpa Kumar Roy, Ankur Deria, Danfeng Hong +3

    cs.CVcs.LGeess.IVarXiv:2203.16952v22022
  9. Deep Affect Prediction in-the-wild: Aff-Wild Database and Challenge, Deep Architectures, and Beyond

    Dimitrios Kollias, Panagiotis Tzirakis, Mihalis A. Nicolaou +5

    cs.CVcs.AIcs.HCarXiv:1804.10938v52018
  10. 3D-CVF: Generating Joint Camera and LiDAR Features Using Cross-View Spatial Feature Fusion for 3D Object Detection

    Jin Hyeok Yoo, Yecheol Kim, Jisong Kim +1

    cs.CVcs.LGeess.IVarXiv:2004.12636v22020
  11. Cross-City Matters: A Multimodal Remote Sensing Benchmark Dataset for Cross-City Semantic Segmentation using High-Resolution Domain Adaptation Networks

    Danfeng Hong, Bing Zhang, Hao Li +7

    cs.CVeess.IVarXiv:2309.16499v22023
  12. Unsupervised Medical Image Translation with Adversarial Diffusion Models

    Muzaffer Özbey, Onat Dalmaz, Salman UH Dar +4

    eess.IVcs.CVarXiv:2207.08208v32022
  13. A Survey on Instance Segmentation: State of the art

    Abdul Mueed Hafiz, Ghulam Mohiuddin Bhat

    cs.CVcs.LGeess.IVarXiv:2007.00047v12020
  14. VinDr-CXR: An open dataset of chest X-rays with radiologist's annotations

    Ha Q. Nguyen, Khanh Lam, Linh T. Le +21

    eess.IVarXiv:2012.15029v32020
  15. UltraPIPS: Improving model perception in B-mode ultrasound with foundation models

    Tal Grutman, Tali Ilovitsh

    cs.CVeess.IVarXiv:2608.26033v12026
  16. Deep neural networks for the evaluation and design of photonic devices

    Jiaqi Jiang, Mingkun Chen, Jonathan A. Fan

    eess.IVcs.LGphysics.app-pharXiv:2007.00084v12020
  17. Dynamic Snake Convolution based on Topological Geometric Constraints for Tubular Structure Segmentation

    Yaolei Qi, Yuting He, Xiaoming Qi +2

    cs.CVeess.IVarXiv:2307.08388v22023
  18. Improving Cross-Site Whole-Heart Segmentation

    Tanish Mudaliar, Justin Li, Daniel Lin +4

    eess.IVcs.CVarXiv:2608.25109v12026
  19. Lowering the Barrier to AI-Driven Inspection: A No-Code Workflow for Automated Structural Defect Detection

    Michael Holm, Tanner McElroy, Xinghang Zhang +1

    cs.CVcs.LGeess.IVarXiv:2608.25176v12026
  20. Modality Contribution Score - A Per-Patient Framework for Quantifying the Relative Diagnostic Contribution of Structural MRI and Amyloid PET in Alzheimer's Disease

    Dawa Chyophel Lepcha, Aaliya Ali, Sophie A. Martin +3

    eess.IVcs.AIcs.CVarXiv:2608.24931v12026
  21. Score-Based Ideal Observer Approximation via Denoising Score Matching for Signal-Known-Exactly Detection Tasks

    Weimin Zhou

    eess.IVcs.AIcs.CVarXiv:2608.24768v12026
  22. Learning spatially varying regularisation parameters of low regularity for image reconstruction

    Kostas Papafitsoros, Luca Calatroni, Andreas Kofler

    eess.IVcs.CVmath.OCarXiv:2608.25127v12026
  23. Differentiable Soft Quantization: Bridging Full-Precision and Low-Bit Neural Networks

    Ruihao Gong, Xianglong Liu, Shenghu Jiang +5

    cs.CVcs.LGeess.IVarXiv:1908.05033v12019
  24. Understanding Adversarial Attacks on Deep Learning Based Medical Image Analysis Systems

    Xingjun Ma, Yuhao Niu, Lin Gu +4

    cs.CVcs.LGeess.IVarXiv:1907.10456v22019
  25. Watch your Up-Convolution: CNN Based Generative Deep Neural Networks are Failing to Reproduce Spectral Distributions

    Ricard Durall, Margret Keuper, Janis Keuper

    cs.CVeess.IVarXiv:2003.01826v12020
  26. Road Crack Detection Using Deep Convolutional Neural Network and Adaptive Thresholding

    Rui Fan, Mohammud Junaid Bocus, Yilong Zhu +5

    cs.CVcs.LGeess.IVarXiv:1904.08582v12019
  27. Whole Slide Images based Cancer Survival Prediction using Attention Guided Deep Multiple Instance Learning Networks

    Jiawen Yao, Xinliang Zhu, Jitendra Jonnagaddala +2

    eess.IVcs.CVarXiv:2009.11169v12020
  28. Reducing the Hausdorff Distance in Medical Image Segmentation with Convolutional Neural Networks

    Davood Karimi, Septimiu E. Salcudean

    eess.IVcs.LGstat.MLarXiv:1904.10030v12019
  29. U-KAN Makes Strong Backbone for Medical Image Segmentation and Generation

    Chenxin Li, Xinyu Liu, Wuyang Li +5

    eess.IVcs.CVarXiv:2406.02918v32024
  30. A Patient-Centric Dataset of Images and Metadata for Identifying Melanomas Using Clinical Context

    Veronica Rotemberg, Nicholas Kurtansky, Brigid Betz-Stablein +21

    eess.IVcs.CVcs.CYarXiv:2008.07360v12020
  31. ResViT: Residual vision transformers for multi-modal medical image synthesis

    Onat Dalmaz, Mahmut Yurt, Tolga Çukur

    eess.IVcs.CVarXiv:2106.16031v32021
  32. Reliability- and Anatomy-Consistency-Aware Multimodal Learning for Robust Fracture Classification from Bangladeshi Radiographs

    Musa Tur Farazi, K G Subarno Bithi

    eess.IVcs.AIcs.CVarXiv:2608.21482v12026
  33. Multimodal pseudo-CT synthesis for PET attenuation correction using separate modality encoding and topogram conditioning

    Rory Bell, Artemis Bouzaki, Jiaming Cao +2

    eess.IVcs.CVphysics.med-pharXiv:2608.21481v12026
  34. TorchIO: A Python library for efficient loading, preprocessing, augmentation and patch-based sampling of medical images in deep learning

    Fernando Pérez-García, Rachel Sparks, Sébastien Ourselin

    eess.IVcs.AIcs.CVarXiv:2003.04696v52020
  35. Channel-wise Autoregressive Entropy Models for Learned Image Compression

    David Minnen, Saurabh Singh

    eess.IVcs.CVcs.ITarXiv:2007.08739v12020
  36. Transfer Learning with Deep Convolutional Neural Network (CNN) for Pneumonia Detection using Chest X-ray

    Tawsifur Rahman, Muhammad E. H. Chowdhury, Amith Khandakar +5

    eess.IVcs.CVcs.LGarXiv:2004.06578v12020
  37. Automated Gleason Grading of Prostate Biopsies using Deep Learning

    Wouter Bulten, Hans Pinckaers, Hester van Boven +6

    eess.IVcs.CVarXiv:1907.07980v12019
  38. 3D Deep Learning on Medical Images: A Review

    Satya P. Singh, Lipo Wang, Sukrit Gupta +3

    q-bio.QMcs.CVcs.LGarXiv:2004.00218v42020
  39. LiteEvent-AE: Lightweight Autoencoder for Event-Based Vision on Low-Latency Energy-Constrained Edge Devices

    Riadul Islam, Joey Mule, Dhandeep Challagundla +3

    cs.CVcs.AIeess.IVarXiv:2608.21764v12026
  40. MDFI: A Multi-Domain Features Integration for Compressed Video Quality Enhancement

    Sang NguyenQuang, Hieu Bui Minh, Dang BuiDinh +1

    eess.IVcs.CVarXiv:2608.21495v12026
  41. CompressAI: a PyTorch library and evaluation platform for end-to-end compression research

    Jean Bégaint, Fabien Racapé, Simon Feltman +1

    cs.CVeess.IVarXiv:2011.03029v12020
  42. Pretreatment DCE-MRI Resolves Response Quality Within Pathologic Endpoints in Neoadjuvant Breast Cancer

    Dattatreya Kantha, Murray H. Loew

    eess.IVcs.CVcs.LGarXiv:2608.22097v12026
  43. Brain-inspired computing: We need a master plan

    Adnan Mehonic, Anthony J Kenyon

    cs.ETcs.AIeess.IVarXiv:2104.14517v12021
  44. Recent Advances in Domain Adaptation for the Classification of Remote Sensing Data

    Devis Tuia, Claudio Persello, Lorenzo Bruzzone

    cs.CVeess.IVarXiv:2104.07778v12021
  45. Multi-Attention-Network for Semantic Segmentation of Fine Resolution Remote Sensing Images

    Rui Li, Shunyi Zheng, Chenxi Duan +3

    eess.IVcs.CVarXiv:2009.02130v42020
  46. Residual Feature Distillation Network for Lightweight Image Super-Resolution

    Jie Liu, Jie Tang, Gangshan Wu

    eess.IVcs.CVarXiv:2009.11551v12020
  47. AMOS: A Large-Scale Abdominal Multi-Organ Benchmark for Versatile Medical Image Segmentation

    Yuanfeng Ji, Haotian Bai, Jie Yang +8

    eess.IVcs.CVcs.LGarXiv:2206.08023v32022
  48. clDice -- A Novel Topology-Preserving Loss Function for Tubular Structure Segmentation

    Suprosanna Shit, Johannes C. Paetzold, Anjany Sekuboyina +6

    cs.CVcs.LGeess.IVarXiv:2003.07311v72020
  49. Deep Learning for Classification of Hyperspectral Data: A Comparative Review

    Nicolas Audebert, Bertrand Saux, Sébastien Lefèvre

    cs.LGcs.CVcs.NEarXiv:1904.10674v12019
  50. Score-based diffusion models for accelerated MRI

    Hyungjin Chung, Jong Chul Ye

    eess.IVcs.AIcs.CVarXiv:2110.05243v32021
  51. ELIC: Efficient Learned Image Compression with Unevenly Grouped Space-Channel Contextual Adaptive Coding

    Dailan He, Ziming Yang, Weikun Peng +3

    cs.CVeess.IVarXiv:2203.10886v22022
  52. Deep Global Registration

    Christopher Choy, Wei Dong, Vladlen Koltun

    cs.CVcs.CGcs.LGarXiv:2004.11540v22020
  53. Low-Light Image Enhancement with Normalizing Flow

    Yufei Wang, Renjie Wan, Wenhan Yang +3

    eess.IVcs.CVarXiv:2109.05923v12021
  54. Multi-Stage Prompt-Guided Feature Modulation for Generalizable Brain Tumor Segmentation

    Mohammad Mahdi Danesh Pajouh, Sara Saeedi

    eess.IVcs.CVarXiv:2608.23745v12026
  55. A review: Deep learning for medical image segmentation using multi-modality fusion

    Tongxue Zhou, Su Ruan, Stéphane Canu

    eess.IVcs.CVcs.LGarXiv:2004.10664v22020
  56. Deep Learning in Medical Image Registration: A Review

    Yabo Fu, Yang Lei, Tonghe Wang +3

    eess.IVcs.CVcs.LGarXiv:1912.12318v12019
  57. A Survey on Active Learning and Human-in-the-Loop Deep Learning for Medical Image Analysis

    Samuel Budd, Emma C Robinson, Bernhard Kainz

    cs.LGcs.CVcs.HCarXiv:1910.02923v22019
  58. Towards a Guideline for Evaluation Metrics in Medical Image Segmentation

    Dominik Müller, Iñaki Soto-Rey, Frank Kramer

    eess.IVcs.CVcs.LGarXiv:2202.05273v12022
  59. COVID-CAPS: A Capsule Network-based Framework for Identification of COVID-19 cases from X-ray Images

    Parnian Afshar, Shahin Heidarian, Farnoosh Naderkhani +3

    cs.CVcs.LGeess.IVarXiv:2004.02696v22020
  60. TractSeg - Fast and accurate white matter tract segmentation

    Jakob Wasserthal, Peter Neher, Klaus H. Maier-Hein

    cs.CVeess.IVarXiv:1805.07103v22018