Image and Video Processing
Papers filed under eess.IV on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.
Search paper metadata (including unsummarized papers)
61 to 120 of 1,302
Robust breast cancer detection in mammography and digital breast tomosynthesis using annotation-efficient deep learning approach
William Lotter, Abdul Rahman Diab, Bryan Haslam +10
eess.IVcs.CVcs.LGarXiv:1912.11027v22019Your Model Already Knows Don't Teach It, Learn to Ask It: Soft Prompting for Few-Shot Adaptation of Vision-Language Models
Gautam Rajendrakumar Gare, Siyi Li, Hewei Wang +5
cs.CVcs.AIcs.LGarXiv:2609.11310v12026DeepQTMT: A Deep Learning Approach for Fast QTMT-based CU Partition of Intra-mode VVC
Tianyi Li, Mai Xu, Runzhi Tang +2
eess.IVcs.MMarXiv:2006.13125v32020Scale-Aware 3D Deep Learning for Robust Brain Metastasis Detection in Multimodal MRI
Sylvain Jaume, Hongming Wang, Simon K. Warfield
eess.IVcs.CVcs.LGarXiv:2609.10825v12026Modality-aware Mutual Learning for Multi-modal Medical Image Segmentation
Yao Zhang, Jiawei Yang, Jiang Tian +4
eess.IVcs.CVarXiv:2107.09842v12021Fully-automated Body Composition Analysis in Routine CT Imaging Using 3D Semantic Segmentation Convolutional Neural Networks
Sven Koitka, Lennard Kroll, Eugen Malamutmann +2
eess.IVcs.CVarXiv:2002.10776v12020InfiniPot-V: Memory-Constrained KV Cache Compression for Streaming Video Understanding
Minsoo Kim, Kyuhong Shim, Jungwook Choi +1
eess.IVcs.LGarXiv:2506.15745v22025Synthetic CT Generation from MRI using 3D Transformer-based Denoising Diffusion Model
Shaoyan Pan, Elham Abouei, Jacob Wynne +10
eess.IVcs.CVarXiv:2305.19467v12023Hybrid Convolutional and Attention Network for Hyperspectral Image Denoising
Shuai Hu, Feng Gao, Xiaowei Zhou +2
eess.IVcs.CVarXiv:2403.10067v12024Interactive Sketch & Fill: Multiclass Sketch-to-Image Translation
Arnab Ghosh, Richard Zhang, Puneet K. Dokania +4
cs.CVcs.LGeess.IVarXiv:1909.11081v22019Mind Reader: Reconstructing complex images from brain activities
Sikun Lin, Thomas Sprague, Ambuj K Singh
q-bio.NCcs.CVcs.HCarXiv:2210.01769v12022HOI Analysis: Integrating and Decomposing Human-Object Interaction
Yong-Lu Li, Xinpeng Liu, Xiaoqian Wu +2
cs.CVcs.LGeess.IVarXiv:2010.16219v22020BCI: Breast Cancer Immunohistochemical Image Generation through Pyramid Pix2pix
Shengjie Liu, Chuang Zhu, Feng Xu +3
eess.IVcs.CVarXiv:2204.11425v22022Moiré Photo Restoration Using Multiresolution Convolutional Neural Networks
Yujing Sun, Yizhou Yu, Wenping Wang
cs.CVeess.IVarXiv:1805.02996v12018A Deep Convolutional Neural Network for COVID-19 Detection Using Chest X-Rays
Pedro R. A. S. Bassi, Romis Attux
eess.IVcs.CVcs.LGarXiv:2005.01578v42020Deepfakes Detection with Automatic Face Weighting
Daniel Mas Montserrat, Hanxiang Hao, S. K. Yarlagadda +8
cs.CVeess.IVarXiv:2004.12027v22020COAST: COntrollable Arbitrary-Sampling NeTwork for Compressive Sensing
Di You, Jian Zhang, Jingfen Xie +2
cs.CVeess.IVarXiv:2107.07225v12021VCT: A Video Compression Transformer
Fabian Mentzer, George Toderici, David Minnen +4
cs.CVcs.LGeess.IVarXiv:2206.07307v22022From voxels to pixels and back: Self-supervision in natural-image reconstruction from fMRI
Roman Beliy, Guy Gaziv, Assaf Hoogi +3
eess.IVcs.LGq-bio.NCarXiv:1907.02431v12019ImageCAS: A Large-Scale Dataset and Benchmark for Coronary Artery Segmentation based on Computed Tomography Angiography Images
An Zeng, Chunbiao Wu, Meiping Huang +10
eess.IVcs.LGarXiv:2211.01607v22022Vita-CLIP: Video and text adaptive CLIP via Multimodal Prompting
Syed Talal Wasim, Muzammal Naseer, Salman Khan +2
cs.CVeess.IVarXiv:2304.03307v12023DMD: A Large-Scale Multi-Modal Driver Monitoring Dataset for Attention and Alertness Analysis
Juan Diego Ortega, Neslihan Kose, Paola Cañas +5
cs.CVcs.LGeess.IVarXiv:2008.12085v12020Virchow: A Million-Slide Digital Pathology Foundation Model
Eugene Vorontsov, Alican Bozkurt, Adam Casson +28
eess.IVcs.CVcs.LGarXiv:2309.07778v62023Automatic Crack Detection on Road Pavements Using Encoder Decoder Architecture
Zhun Fan, Chong Li, Ying Chen +4
cs.CVcs.LGeess.IVarXiv:2007.00477v12020CMU-Net: A Strong ConvMixer-based Medical Ultrasound Image Segmentation Network
Fenghe Tang, Lingtao Wang, Chunping Ning +2
eess.IVcs.CVarXiv:2210.13012v42022Radiomics in Medical Imaging: Methods, Applications, and Challenges
Fnu Neha, Deepak kumar Shukla
eess.IVcs.AIcs.LGarXiv:2602.00102v12026Adversarial Attack Vulnerability of Medical Image Analysis Systems: Unexplored Factors
Gerda Bortsova, Cristina González-Gonzalo, Suzanne C. Wetstein +9
cs.CRcs.CVeess.IVarXiv:2006.06356v32020Soft-Attention Improves Skin Cancer Classification Performance
Soumyya Kanti Datta, Seyed Mohammad Abuzar Hashemi, Sargur N. Srihari +1
eess.IVcs.CVcs.LGarXiv:2105.03358v42021Supervised Raw Video Denoising with a Benchmark Dataset on Dynamic Scenes
Huanjing Yue, Cong Cao, Lei Liao +2
eess.IVcs.CVcs.LGarXiv:2003.14013v12020Fine Perceptive GANs for Brain MR Image Super-Resolution in Wavelet Domain
Senrong You, Yong Liu, Baiying Lei +1
eess.IVcs.CVcs.LGarXiv:2011.04145v12020MR image reconstruction using deep density priors
Kerem C. Tezcan, Christian F. Baumgartner, Roger Luechinger +2
cs.CVeess.IVstat.MLarXiv:1711.11386v42017Spatio-spectral classification of hyperspectral images for brain cancer detection during surgical operations
H. Fabelo, S. Ortega, D. Ravi +19
eess.IVcs.CVarXiv:2402.07192v12024COVID-19 CT Image Synthesis with a Conditional Generative Adversarial Network
Yifan Jiang, Han Chen, Murray Loew +1
eess.IVcs.CVcs.LGarXiv:2007.14638v22020Resolution Dependent GAN Interpolation for Controllable Image Synthesis Between Domains
Justin N. M. Pinkney, Doron Adler
cs.CVeess.IVarXiv:2010.05334v32020CHIMERA Challenge Task 2 and 3: Response Subtypes Classification and Progression Survival Prediction in Bladder Cancer Patients using Multimodal Datasets
Catherine Chia, Tongjie Wang, Robert Spaans +12
eess.IVcs.CVq-bio.QMarXiv:2609.09510v12026Morphological Decoupling-Based Skeletal Classification for Clinical Assessment of Malocclusion
Zhichun Jin, Zhicheng He, Hao Xu +4
eess.IVcs.CVarXiv:2609.09801v12026A Survey on Deep Learning for Localization and Mapping: Towards the Age of Spatial Machine Intelligence
Changhao Chen, Bing Wang, Chris Xiaoxuan Lu +2
cs.CVcs.LGcs.ROarXiv:2006.12567v22020Single-shot multispectral imaging with a monochromatic camera
Sujit Kumar Sahoo, Dongliang Tang, Cuong Dang
physics.opticseess.IVeess.SParXiv:1707.09453v12017Ensemble of Deep Convolutional Neural Networks for Automatic Pavement Crack Detection and Measurement
Zhun Fan, Chong Li, Ying Chen +4
cs.CVcs.LGeess.IVarXiv:2002.03241v12020Using U-Net Network for Efficient Brain Tumor Segmentation in MRI Images
Jason Walsh, Alice Othmani, Mayank Jain +1
eess.IVcs.CVq-bio.QMarXiv:2211.01885v12022Deep Learning in Breast Cancer Imaging: A Decade of Progress and Future Directions
Luyang Luo, Xi Wang, Yi Lin +7
eess.IVcs.CVarXiv:2304.06662v42023Fusion of convolution neural network, support vector machine and Sobel filter for accurate detection of COVID-19 patients using X-ray images
Danial Sharifrazi, Roohallah Alizadehsani, Mohamad Roshanzamir +13
eess.IVcs.CVarXiv:2102.06883v12021Efficient and Degradation-Adaptive Network for Real-World Image Super-Resolution
Jie Liang, Hui Zeng, Lei Zhang
cs.CVeess.IVarXiv:2203.14216v12022Kvasir-Instrument: Diagnostic and therapeutic tool segmentation dataset in gastrointestinal endoscopy
Debesh Jha, Sharib Ali, Krister Emanuelsen +9
physics.med-phcs.CVcs.LGarXiv:2011.08065v12020Recovering Biomechanical Signals from Missing Keypoints Using Temporal Interpolation in Monocular Gait Analysis
Shubham Jariwala
cs.CVeess.IVarXiv:2609.09670v12026Bayesian imaging using Plug & Play priors: when Langevin meets Tweedie
Rémi Laumont, Valentin de Bortoli, Andrés Almansa +3
stat.MEcs.CVeess.IVarXiv:2103.04715v62021DW-GAN: A Discrete Wavelet Transform GAN for NonHomogeneous Dehazing
Minghan Fu, Huan Liu, Yankun Yu +2
eess.IVarXiv:2104.08911v22021Equivariant Imaging: Learning Beyond the Range Space
Dongdong Chen, Julián Tachella, Mike E. Davies
cs.CVeess.IVeess.SParXiv:2103.14756v22021Flow-based Kernel Prior with Application to Blind Super-Resolution
Jingyun Liang, Kai Zhang, Shuhang Gu +2
cs.CVeess.IVarXiv:2103.15977v12021Phase Imaging with Computational Specificity (PICS) for measuring dry mass changes in sub-cellular compartments
Mikhail E. Kandel, Yuchen R. He, Young Jae Lee +7
eess.IVphysics.bio-phphysics.opticsarXiv:2002.08361v22020Exploiting Epistemic Uncertainty of Anatomy Segmentation for Anomaly Detection in Retinal OCT
Philipp Seeböck, José Ignacio Orlando, Thomas Schlegl +5
eess.IVcs.CVcs.LGarXiv:1905.12806v12019Automatic Target Recognition on Synthetic Aperture Radar Imagery: A Survey
O. Kechagias-Stamatis, N. Aouf
cs.CVeess.IVeess.SParXiv:2007.02106v22020Skin disease diagnosis with deep learning: a review
Hongfeng Li, Yini Pan, Jie Zhao +1
eess.IVcs.CVcs.LGarXiv:2011.05627v22020Hyperspectral Image Classification With Context-Aware Dynamic Graph Convolutional Network
Sheng Wan, Chen Gong, Ping Zhong +3
cs.LGcs.CVeess.IVarXiv:1909.11953v12019A deep learning approach to detecting volcano deformation from satellite imagery using synthetic datasets
Nantheera Anantrasirichai, Juliet Biggs, Fabien Albino +1
cs.CVeess.IVarXiv:1905.07286v12019A General Pipeline for 3D Detection of Vehicles
Xinxin Du, Marcelo H. Ang, Sertac Karaman +1
cs.CVeess.IVstat.MLarXiv:1803.00387v12018Joint super-resolution and synthesis of 1 mm isotropic MP-RAGE volumes from clinical MRI exams with scans of different orientation, resolution and contrast
Juan Eugenio Iglesias, Benjamin Billot, Yael Balbastre +6
eess.IVcs.CVarXiv:2012.13340v12020Auxiliary Signal-Guided Knowledge Encoder-Decoder for Medical Report Generation
Mingjie Li, Fuyu Wang, Xiaojun Chang +1
cs.CVcs.CLeess.IVarXiv:2006.03744v12020FBNETGEN: Task-aware GNN-based fMRI Analysis via Functional Brain Network Generation
Xuan Kan, Hejie Cui, Joshua Lukemire +2
cs.LGcs.NEeess.IVarXiv:2205.12465v22022TINYCD: A (Not So) Deep Learning Model For Change Detection
Andrea Codegoni, Gabriele Lombardi, Alessandro Ferrari
cs.CVcs.LGeess.IVarXiv:2207.13159v22022