Image and Video Processing
Papers filed under eess.IV on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.
Search paper metadata (including unsummarized papers)
1,141 to 1,200 of 1,302
The state of the art in kidney and kidney tumor segmentation in contrast-enhanced CT imaging: Results of the KiTS19 Challenge
Nicholas Heller, Fabian Isensee, Klaus H. Maier-Hein +38
eess.IVcs.CVcs.LGarXiv:1912.01054v22019ScanRefer: 3D Object Localization in RGB-D Scans using Natural Language
Dave Zhenyu Chen, Angel X. Chang, Matthias Nießner
cs.CVcs.CLcs.LGarXiv:1912.08830v32019Learning Enriched Features for Fast Image Restoration and Enhancement
Syed Waqas Zamir, Aditya Arora, Salman Khan +4
eess.IVcs.CVarXiv:2205.01649v12022Spatiotemporal Distillation via Recurrent Bottlenecks for Aortic Tracking
Dexter Wen Jie Teo, Nairouz Shehata, Herve Lombaert
eess.IVcs.CVcs.LGarXiv:2608.23879v12026Deep Learning for Deepfakes Creation and Detection: A Survey
Thanh Thi Nguyen, Quoc Viet Hung Nguyen, Dung Tien Nguyen +6
cs.CVcs.LGeess.IVarXiv:1909.11573v52019Native-Space 3D CarveMix for Multi-Site T1w Stroke Segmentation
Dexter Wen Jie Teo, Kumaradevan Punithakumar
eess.IVcs.CVarXiv:2608.23882v12026U-GAT-IT: Unsupervised Generative Attentional Networks with Adaptive Layer-Instance Normalization for Image-to-Image Translation
Junho Kim, Minjae Kim, Hyeonwoo Kang +1
cs.CVeess.IVarXiv:1907.10830v42019Explainable deep learning models in medical image analysis
Amitojdeep Singh, Sourya Sengupta, Vasudevan Lakshminarayanan
cs.CVcs.LGeess.IVarXiv:2005.13799v12020PULSE: Self-Supervised Photo Upsampling via Latent Space Exploration of Generative Models
Sachit Menon, Alexandru Damian, Shijia Hu +2
cs.CVcs.LGeess.IVarXiv:2003.03808v32020Model Effect or Label Effect? Refined Annotations and a Human-Referenced Benchmark for Pulmonary Embolism Segmentation
Qihang Sun, Zhongxiao Liu, Bailiang Jian +6
eess.IVcs.CVarXiv:2608.24486v12026High-Fidelity Generative Image Compression
Fabian Mentzer, George Toderici, Michael Tschannen +1
eess.IVcs.CVcs.LGarXiv:2006.09965v32020Detecting and Simulating Artifacts in GAN Fake Images
Xu Zhang, Svebor Karaman, Shih-Fu Chang
cs.CVeess.IVarXiv:1907.06515v22019CA-Net: Comprehensive Attention Convolutional Neural Networks for Explainable Medical Image Segmentation
Ran Gu, Guotai Wang, Tao Song +6
eess.IVcs.CVarXiv:2009.10549v22020Deep Unfolding Network for Image Super-Resolution
Kai Zhang, Luc Van Gool, Radu Timofte
eess.IVcs.CVarXiv:2003.10428v12020SynthSeg: Segmentation of brain MRI scans of any contrast and resolution without retraining
Benjamin Billot, Douglas N. Greve, Oula Puonti +5
eess.IVcs.CVarXiv:2107.09559v42021Drone-based RGB-Infrared Cross-Modality Vehicle Detection via Uncertainty-Aware Learning
Yiming Sun, Bing Cao, Pengfei Zhu +1
cs.CVcs.LGeess.IVarXiv:2003.02437v22020BCN20000: Dermoscopic Lesions in the Wild
Marc Combalia, Noel C. F. Codella, Veronica Rotemberg +8
eess.IVcs.CVarXiv:1908.02288v22019Deep Learning Techniques for Inverse Problems in Imaging
Gregory Ongie, Ajil Jalal, Christopher A. Metzler +3
eess.IVcs.LGstat.MLarXiv:2005.06001v12020Deep Learning in Medical Image Registration: A Survey
Grant Haskins, Uwe Kruger, Pingkun Yan
q-bio.QMcs.CVeess.IVarXiv:1903.02026v22019HINet: Half Instance Normalization Network for Image Restoration
Liangyu Chen, Xin Lu, Jie Zhang +2
eess.IVcs.CVarXiv:2105.06086v22021Make it SING: Analyzing Semantic Invariants in Classifiers
Harel Yadid, Meir Yossef Levi, Roy Betser +1
cs.CVeess.IVarXiv:2603.14610v22026HistoAtlas: A Pan-Cancer Morphology Atlas Linking Histomics to Molecular Programs and Clinical Outcomes
Pierre-Antoine Bannier
q-bio.QMcs.CVeess.IVarXiv:2603.16587v12026Generalized ODIN: Detecting Out-of-distribution Image without Learning from Out-of-distribution Data
Yen-Chang Hsu, Yilin Shen, Hongxia Jin +1
cs.CVcs.LGeess.IVarXiv:2002.11297v22020Unified Focal loss: Generalising Dice and cross entropy-based losses to handle class imbalanced medical image segmentation
Michael Yeung, Evis Sala, Carola-Bibiane Schönlieb +1
eess.IVcs.CVcs.LGarXiv:2102.04525v42021Deep learning with noisy labels: exploring techniques and remedies in medical image analysis
Davood Karimi, Haoran Dou, Simon K. Warfield +1
cs.CVcs.LGeess.IVarXiv:1912.02911v42019CHIMERA Challenge: Biochemical Recurrence Prediction in Prostate Cancer Patients using multimodal datasets
Robert N. Spaans, Catherine Chia, Tongjie Wang +7
eess.IVcs.CVarXiv:2608.21497v12026Pattern-Derived Visual Swarm Games: Multi-Scale Drone-Vision States for Interception and Sustainability Audits
Faruk Alpay, Levent Sarioglu
cs.ROcs.GTeess.IVarXiv:2608.23575v12026Contrastive learning of global and local features for medical image segmentation with limited annotations
Krishna Chaitanya, Ertunc Erdil, Neerav Karani +1
cs.CVcs.LGeess.IVarXiv:2006.10511v22020InterFaceGAN: Interpreting the Disentangled Face Representation Learned by GANs
Yujun Shen, Ceyuan Yang, Xiaoou Tang +1
cs.CVcs.LGeess.IVarXiv:2005.09635v22020CMX: Cross-Modal Fusion for RGB-X Semantic Segmentation with Transformers
Jiaming Zhang, Huayao Liu, Kailun Yang +3
cs.CVcs.ROeess.IVarXiv:2203.04838v52022PlantDoc: A Dataset for Visual Plant Disease Detection
Davinder Singh, Naman Jain, Pranjali Jain +3
cs.CVeess.IVarXiv:1911.10317v12019Precision-Aware Variable Bit Processing Elements for Hardware-Efficient Systolic Array Designs
Dantu Nandini Devi, Madhav Rao
cs.ARcs.ETcs.LGarXiv:2608.22378v12026Deep neural network models for computational histopathology: A survey
Chetan L. Srinidhi, Ozan Ciga, Anne L. Martel
eess.IVcs.CVcs.LGarXiv:1912.12378v22019Convolutional Neural Networks for Classification of Alzheimer's Disease: Overview and Reproducible Evaluation
Junhao Wen, Elina Thibeau-Sutre, Mauricio Diaz-Melo +7
cs.LGeess.IVstat.MLarXiv:1904.07773v62019MAXIM: Multi-Axis MLP for Image Processing
Zhengzhong Tu, Hossein Talebi, Han Zhang +4
eess.IVcs.CVarXiv:2201.02973v22022SGDC: Structurally-Guided Dynamic Convolution for Medical Image Segmentation
Bo Shi, Wei-ping Zhu, M. N. S. Swamy
eess.IVcs.CVarXiv:2602.23496v12026PREDATOR: Registration of 3D Point Clouds with Low Overlap
Shengyu Huang, Zan Gojcic, Mikhail Usvyatsov +2
cs.CVeess.IVarXiv:2011.13005v32020Colon-Bench: An Agentic Workflow for Scalable Dense Lesion Annotation in Full-Procedure Colonoscopy Videos
Abdullah Hamdi, Changchun Yang, Xin Gao
eess.IVcs.CVcs.HCarXiv:2603.25645v32026GET: Generative Embedding Translation for Medical Image Segmentation
Md Maklachur Rahman, Md Hasan Al Banna, Saraf Anjum +2
eess.IVcs.CVcs.LGarXiv:2608.22619v12026Deep Learning COVID-19 Features on CXR using Limited Training Data Sets
Yujin Oh, Sangjoon Park, Jong Chul Ye
eess.IVcs.CVcs.LGarXiv:2004.05758v22020Visual Transformers: Token-based Image Representation and Processing for Computer Vision
Bichen Wu, Chenfeng Xu, Xiaoliang Dai +7
cs.CVcs.LGeess.IVarXiv:2006.03677v42020Big Self-Supervised Models Advance Medical Image Classification
Shekoofeh Azizi, Basil Mustafa, Fiona Ryan +9
eess.IVcs.CVcs.LGarXiv:2101.05224v22021VATT: Transformers for Multimodal Self-Supervised Learning from Raw Video, Audio and Text
Hassan Akbari, Liangzhe Yuan, Rui Qian +4
cs.CVcs.AIcs.LGarXiv:2104.11178v32021Boundary loss for highly unbalanced segmentation
Hoel Kervadec, Jihene Bouchtiba, Christian Desrosiers +3
eess.IVcs.CVarXiv:1812.07032v42018Solving Inverse Problems in Medical Imaging with Score-Based Generative Models
Yang Song, Liyue Shen, Lei Xing +1
eess.IVcs.CVcs.LGarXiv:2111.08005v22021The Universal Normal Embedding
Chen Tasker, Roy Betser, Eyal Gofer +2
cs.CVeess.IVarXiv:2603.21786v12026U-shape Transformer for Underwater Image Enhancement
Lintao Peng, Chunli Zhu, Liheng Bian
cs.CVeess.IVarXiv:2111.11843v62021MANIQA: Multi-dimension Attention Network for No-Reference Image Quality Assessment
Sidi Yang, Tianhe Wu, Shuwei Shi +5
cs.CVeess.IVarXiv:2204.08958v22022Diffusion Models for Medical Image Analysis: A Comprehensive Survey
Amirhossein Kazerouni, Ehsan Khodapanah Aghdam, Moein Heidari +4
eess.IVcs.CVarXiv:2211.07804v32022OpenVision 3: A Family of Unified Visual Encoder for Both Understanding and Generation
Letian Zhang, Sucheng Ren, Yanqing Liu +9
eess.IVcs.AIarXiv:2601.15369v22026Capsule-Forensics: Using Capsule Networks to Detect Forged Images and Videos
Huy H. Nguyen, Junichi Yamagishi, Isao Echizen
cs.CVeess.IVarXiv:1810.11215v12018Deep Semantic Segmentation of Natural and Medical Images: A Review
Saeid Asgari Taghanaki, Kumar Abhishek, Joseph Paul Cohen +2
cs.CVcs.LGeess.IVarXiv:1910.07655v42019DoubleU-Net: A Deep Convolutional Neural Network for Medical Image Segmentation
Debesh Jha, Michael A. Riegler, Dag Johansen +2
eess.IVcs.CVarXiv:2006.04868v22020HigherHRNet: Scale-Aware Representation Learning for Bottom-Up Human Pose Estimation
Bowen Cheng, Bin Xiao, Jingdong Wang +3
cs.CVcs.LGeess.IVarXiv:1908.10357v32019Single Image Super-Resolution via a Holistic Attention Network
Ben Niu, Weilei Wen, Wenqi Ren +6
eess.IVcs.CVarXiv:2008.08767v12020Multi-Scale Progressive Fusion Network for Single Image Deraining
Kui Jiang, Zhongyuan Wang, Peng Yi +5
cs.CVcs.LGeess.IVarXiv:2003.10985v22020Spanning the Visual Analogy Space with a Weight Basis of LoRAs
Hila Manor, Rinon Gal, Haggai Maron +2
cs.CVcs.AIcs.GRarXiv:2602.15727v22026PadChest: A large chest x-ray image dataset with multi-label annotated reports
Aurelia Bustos, Antonio Pertusa, Jose-Maria Salinas +1
eess.IVcs.CVarXiv:1901.07441v22019Video Generation Models as World Models: Efficient Paradigms, Architectures and Algorithms
Muyang He, Hanzhong Guo, Junxiong Lin +1
eess.IVcs.CVarXiv:2603.28489v32026COVID-19 Image Data Collection: Prospective Predictions Are the Future
Joseph Paul Cohen, Paul Morrison, Lan Dao +3
q-bio.QMcs.CVcs.LGarXiv:2006.11988v32020