Image and Video Processing
Papers filed under eess.IV on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.
Search paper metadata (including unsummarized papers)
301 to 360 of 1,302
Anchor Forcing: Anchor Memory and Tri-Region RoPE for Interactive Streaming Video Diffusion
Yang Yang, Tianyi Zhang, Wei Huang +6
cs.CVeess.IVarXiv:2603.13405v12026A Review of Uncertainty Estimation and its Application in Medical Imaging
Ke Zou, Zhihao Chen, Xuedong Yuan +3
eess.IVcs.CVarXiv:2302.08119v32023Bidirectional Mapping Generative Adversarial Networks for Brain MR to PET Synthesis
Shengye Hu, Baiying Lei, Yong Wang +3
eess.IVcs.LGarXiv:2008.03483v12020Gen2Act: Human Video Generation in Novel Scenarios enables Generalizable Robot Manipulation
Homanga Bharadhwaj, Debidatta Dwibedi, Abhinav Gupta +7
cs.ROcs.CVcs.LGarXiv:2409.16283v12024Automatic Lung Cancer Prediction from Chest X-ray Images Using Deep Learning Approach
Worawate Ausawalaithong, Sanparith Marukatat, Arjaree Thirach +1
eess.IVcs.CVarXiv:1808.10858v12018Coupled Convolutional Neural Network with Adaptive Response Function Learning for Unsupervised Hyperspectral Super-Resolution
Ke Zheng, Lianru Gao, Wenzhi Liao +4
eess.IVcs.CVarXiv:2007.14007v12020Scanner-Induced Domain Shifts Undermine the Robustness of Pathology Foundation Models
Erik Thiringer, Fredrik K. Gustafsson, Kajsa Ledesma Eriksson +1
eess.IVcs.CVcs.LGarXiv:2601.04163v12026Multi-Sensor Data Fusion for Cloud Removal in Global and All-Season Sentinel-2 Imagery
Patrick Ebel, Andrea Meraner, Michael Schmitt +1
eess.IVcs.CVarXiv:2009.07683v12020Invertible Denoising Network: A Light Solution for Real Noise Removal
Yang Liu, Zhenyue Qin, Saeed Anwar +4
eess.IVcs.CVarXiv:2104.10546v12021Nighttime Dehazing with a Synthetic Benchmark
Jing Zhang, Yang Cao, Zheng-Jun Zha +1
cs.CVcs.LGeess.IVarXiv:2008.03864v32020Rethinking Data Augmentation for Image Super-resolution: A Comprehensive Analysis and a New Strategy
Jaejun Yoo, Namhyuk Ahn, Kyung-Ah Sohn
eess.IVcs.CVarXiv:2004.00448v22020The reliability of a deep learning model in clinical out-of-distribution MRI data: a multicohort study
Gustav Mårtensson, Daniel Ferreira, Tobias Granberg +22
physics.med-phcs.CVcs.LGarXiv:1911.00515v12019MR Image Denoising and Super-Resolution Using Regularized Reverse Diffusion
Hyungjin Chung, Eun Sun Lee, Jong Chul Ye
eess.IVcs.AIcs.CVarXiv:2203.12621v12022Efficient Medical Image Segmentation Based on Knowledge Distillation
Dian Qin, Jiajun Bu, Zhe Liu +6
eess.IVcs.CVarXiv:2108.09987v12021RoIMix: Proposal-Fusion among Multiple Images for Underwater Object Detection
Wei-Hong Lin, Jia-Xing Zhong, Shan Liu +2
cs.CVcs.LGeess.IVarXiv:1911.03029v22019Dual Residual Attention Network for Image Denoising
Wencong Wu, Shijie Liu, Yi Zhou +2
eess.IVcs.CVarXiv:2305.04269v12023Physics-based Noise Modeling for Extreme Low-light Photography
Kaixuan Wei, Ying Fu, Yinqiang Zheng +1
eess.IVcs.CVarXiv:2108.02158v12021Video TokenCom: Textual Intent-Guided Multi-Rate Video Token Communications with UEP-Based Adaptive Source-Channel Coding
Jingxuan Men, Mahdi Boloursaz Mashhadi, Ning Wang +3
cs.ITcs.LGcs.MMarXiv:2603.02470v12026Landslide4Sense: Reference Benchmark Data and Deep Learning Models for Landslide Detection
Omid Ghorbanzadeh, Yonghao Xu, Pedram Ghamisi +2
cs.CVeess.IVarXiv:2206.00515v32022Deep Gaussian Scale Mixture Prior for Spectral Compressive Imaging
Tao Huang, Weisheng Dong, Xin Yuan +2
eess.IVcs.CVarXiv:2103.07152v22021Deep DIC: Deep Learning-Based Digital Image Correlation for End-to-End Displacement and Strain Measurement
Ru Yang, Yang Li, Danielle Zeng +1
eess.IVcond-mat.mtrl-scics.CVarXiv:2110.13720v22021No-Reference Quality Assessment for 3D Colored Point Cloud and Mesh Models
Zicheng Zhang, Wei Sun, Xiongkuo Min +3
cs.CVcs.GReess.IVarXiv:2107.02041v62021CIPS-3D: A 3D-Aware Generator of GANs Based on Conditionally-Independent Pixel Synthesis
Peng Zhou, Lingxi Xie, Bingbing Ni +1
cs.CVeess.IVarXiv:2110.09788v12021Bayesian-Optimized Superpixel-GrabCut for Traceable Optic Disc Segmentation
Shraddha Changune, Vivek Noel Soren, Gautam Das +1
cs.CVeess.IVarXiv:2608.29196v12026Continuous Dice Coefficient: a Method for Evaluating Probabilistic Segmentations
Reuben R Shamir, Yuval Duchin, Jinyoung Kim +2
cs.CVeess.IVarXiv:1906.11031v12019Unsupervised Image Translation using Adversarial Networks for Improved Plant Disease Recognition
Haseeb Nazki, Sook Yoon, Alvaro Fuentes +1
cs.CVcs.LGeess.IVarXiv:1909.11915v12019Multi-source Domain Adaptation for Semantic Segmentation
Sicheng Zhao, Bo Li, Xiangyu Yue +5
cs.CVcs.LGeess.IVarXiv:1910.12181v12019DSNet: Automatic Dermoscopic Skin Lesion Segmentation
Md. Kamrul Hasan, Lavsen Dahal, Prasad N. Samarakoon +2
eess.IVcs.CVarXiv:1907.04305v22019Multi-institutional Collaborations for Improving Deep Learning-based Magnetic Resonance Image Reconstruction Using Federated Learning
Pengfei Guo, Puyang Wang, Jinyuan Zhou +2
eess.IVcs.CVarXiv:2103.02148v42021A multi-centre polyp detection and segmentation dataset for generalisability assessment
Sharib Ali, Debesh Jha, Noha Ghatwary +12
eess.IVcs.CVcs.LGarXiv:2106.04463v32021Hierarchical Regression Network for Spectral Reconstruction from RGB Images
Yuzhi Zhao, Lai-Man Po, Qiong Yan +2
eess.IVcs.CVcs.LGarXiv:2005.04703v12020Single-photon computational 3D imaging at 45 km
Zheng-Ping Li, Xin Huang, Yuan Cao +9
eess.IVphysics.opticsarXiv:1904.10341v12019A Comprehensive Review of Computer Vision in Sports: Open Issues, Future Trends and Research Directions
Banoth Thulasya Naik, Mohammad Farukh Hashmi, Neeraj Dhanraj Bokde
cs.CVeess.IVarXiv:2203.02281v22022DeepInverse: A Python package for solving imaging inverse problems with deep learning
Julián Tachella, Matthieu Terris, Samuel Hurault +24
eess.IVarXiv:2505.20160v22025Learning from Scarce Labels: Multi-View Echocardiography for Ejection Fraction Prediction
Zhiyuan Gao, Dominic Yurk, Yaser S. Abu-Mostafa
eess.IVcs.CVcs.LGarXiv:2609.02969v12026mustGAN: Multi-Stream Generative Adversarial Networks for MR Image Synthesis
Mahmut Yurt, Salman Ul Hassan Dar, Aykut Erdem +2
eess.IVcs.CVarXiv:1909.11504v12019Patch-Based Diffusion Reconstruction for Accelerated Cardiac Cine
Xuan Lei, Philip Schniter, Juliet Varghese +1
eess.IVeess.SParXiv:2608.28927v12026Building Damage Detection in Satellite Imagery Using Convolutional Neural Networks
Joseph Z. Xu, Wenhan Lu, Zebo Li +2
cs.CVcs.LGeess.IVarXiv:1910.06444v12019Artificial Intelligence Assistance Significantly Improves Gleason Grading of Prostate Biopsies by Pathologists
Wouter Bulten, Maschenka Balkenhol, Jean-Joël Awoumou Belinga +17
eess.IVcs.CVq-bio.QMarXiv:2002.04500v12020Temperate Fish Detection and Classification: a Deep Learning based Approach
Kristian Muri Knausgård, Arne Wiklund, Tonje Knutsen Sørdalen +4
cs.CVcs.LGeess.IVarXiv:2005.07518v12020RTNet: Relation Transformer Network for Diabetic Retinopathy Multi-lesion Segmentation
Shiqi Huang, Jianan Li, Yuze Xiao +2
eess.IVcs.CVarXiv:2201.11037v12022Assessing Reliability and Challenges of Uncertainty Estimations for Medical Image Segmentation
Alain Jungo, Mauricio Reyes
eess.IVcs.CVarXiv:1907.03338v22019AlignTransformer: Hierarchical Alignment of Visual Regions and Disease Tags for Medical Report Generation
Di You, Fenglin Liu, Shen Ge +3
eess.IVcs.CVarXiv:2203.10095v12022Annotation-efficient deep learning for automatic medical image segmentation
Shanshan Wang, Cheng Li, Rongpin Wang +12
eess.IVcs.CVcs.LGarXiv:2012.04885v32020ESRGAN+ : Further Improving Enhanced Super-Resolution Generative Adversarial Network
Nathanaël Carraz Rakotonirina, Andry Rasoanaivo
eess.IVcs.LGarXiv:2001.08073v22020Cascaded deep monocular 3D human pose estimation with evolutionary training data
Shichao Li, Lei Ke, Kevin Pratama +3
cs.CVcs.LGeess.IVarXiv:2006.07778v32020A Comprehensive Review of U-Net and Its Variants: Advances and Applications in Medical Image Segmentation
Wang Jiangtao, Nur Intan Raihana Ruhaiyem, Fu Panpan
eess.IVcs.CVarXiv:2502.06895v12025LUVLi Face Alignment: Estimating Landmarks' Location, Uncertainty, and Visibility Likelihood
Abhinav Kumar, Tim K. Marks, Wenxuan Mou +6
cs.CVcs.LGeess.IVarXiv:2004.02980v12020Actor-Context-Actor Relation Network for Spatio-Temporal Action Localization
Junting Pan, Siyu Chen, Mike Zheng Shou +3
cs.CVcs.LGeess.IVarXiv:2006.07976v32020A survey on deep learning in medical image registration: new technologies, uncertainty, evaluation metrics, and beyond
Junyu Chen, Yihao Liu, Shuwen Wei +5
eess.IVcs.CVarXiv:2307.15615v42023FALCON: Fault-Tolerant Magnetic Tunnel Junction-Based In-Memory Stochastic Architecture for Reliability-Critical Edge AI Applications
Farzad Razi, Mehran Moghadam, Sercan Aygun +2
cs.ETcs.AReess.IVarXiv:2609.00701v12026Large-scale neuromorphic optoelectronic computing with a reconfigurable diffractive processing unit
Tiankuang Zhou, Xing Lin, Jiamin Wu +7
eess.IVcs.LGcs.NEarXiv:2008.11659v12020Comparing Different Deep Learning Architectures for Classification of Chest Radiographs
Keno K. Bressem, Lisa Adams, Christoph Erxleben +3
cs.LGcs.CVeess.IVarXiv:2002.08991v12020Unsupervised Learning for Real-World Super-Resolution
Andreas Lugmayr, Martin Danelljan, Radu Timofte
eess.IVcs.CVarXiv:1909.09629v12019FrePGAN: Robust Deepfake Detection Using Frequency-level Perturbations
Yonghyun Jeong, Doyeon Kim, Youngmin Ro +1
cs.CVcs.LGeess.IVarXiv:2202.03347v12022Edge-aware Guidance Fusion Network for RGB Thermal Scene Parsing
Wujie Zhou, Shaohua Dong, Caie Xu +1
cs.CVeess.IVarXiv:2112.05144v12021Contrastive Cross-site Learning with Redesigned Net for COVID-19 CT Classification
Zhao Wang, Quande Liu, Qi Dou
eess.IVcs.CVcs.LGarXiv:2009.07652v12020FaceQnet: Quality Assessment for Face Recognition based on Deep Learning
Javier Hernandez-Ortega, Javier Galbally, Julian Fierrez +2
cs.CVcs.LGeess.IVarXiv:1904.01740v22019Resolving challenges in deep learning-based analyses of histopathological images using explanation methods
Miriam Hägele, Philipp Seegerer, Sebastian Lapuschkin +5
eess.IVcs.CVq-bio.QMarXiv:1908.06943v22019Cross-domain Hyperspectral Image Classification based on Bi-directional Domain Adaptation
Yuxiang Zhang, Wei Li, Wen Jia +3
cs.CVeess.IVarXiv:2507.02268v12025