Image and Video Processing

Papers filed under eess.IV on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.

Search paper metadata (including unsummarized papers)

901 to 960 of 1,301

  1. A Topological Loss Function for Deep-Learning based Image Segmentation using Persistent Homology

    James R. Clough, Nicholas Byrne, Ilkay Oksuz +3

    cs.CVeess.IVarXiv:1910.01877v22019
  2. NeRP: Implicit Neural Representation Learning with Prior Embedding for Sparsely Sampled Image Reconstruction

    Liyue Shen, John Pauly, Lei Xing

    eess.IVcs.CVarXiv:2108.10991v22021
  3. A Review on Deep Learning Techniques for Video Prediction

    Sergiu Oprea, Pablo Martinez-Gonzalez, Alberto Garcia-Garcia +4

    cs.CVcs.LGeess.IVarXiv:2004.05214v22020
  4. Noisier2Noise: Learning to Denoise from Unpaired Noisy Data

    Nick Moran, Dan Schmidt, Yu Zhong +1

    eess.IVcs.CVarXiv:1910.11908v12019
  5. Memory Enhanced Global-Local Aggregation for Video Object Detection

    Yihong Chen, Yue Cao, Han Hu +1

    cs.CVcs.LGeess.IVarXiv:2003.12063v12020
  6. ClearGrasp: 3D Shape Estimation of Transparent Objects for Manipulation

    Shreeyak S. Sajjan, Matthew Moore, Mike Pan +4

    cs.CVcs.ROeess.IVarXiv:1910.02550v22019
  7. FISTA-Net: Learning A Fast Iterative Shrinkage Thresholding Network for Inverse Problems in Imaging

    Jinxi Xiang, Yonggui Dong, Yunjie Yang

    eess.IVphysics.app-pharXiv:2008.02683v32020
  8. Deep Generalized Unfolding Networks for Image Restoration

    Chong Mou, Qian Wang, Jian Zhang

    cs.CVeess.IVarXiv:2204.13348v12022
  9. Foreground-Aware Relation Network for Geospatial Object Segmentation in High Spatial Resolution Remote Sensing Imagery

    Zhuo Zheng, Yanfei Zhong, Junjue Wang +1

    cs.CVcs.LGeess.IVarXiv:2011.09766v12020
  10. Weakly Supervised Deep Learning for COVID-19 Infection Detection and Classification from CT Images

    Shaoping Hu, Yuan Gao, Zhangming Niu +9

    eess.IVcs.CVcs.LGarXiv:2004.06689v12020
  11. Deep learning for smart fish farming: applications, opportunities and challenges

    Xinting Yang, Song Zhang, Jintao Liu +3

    cs.CVcs.LGeess.IVarXiv:2004.11848v22020
  12. Deep learning-based transformation of the H&E stain into special stains

    Kevin de Haan, Yijie Zhang, Jonathan E. Zuckerman +11

    eess.IVcs.CVcs.LGarXiv:2008.08871v22020
  13. R3Det: Refined Single-Stage Detector with Feature Refinement for Rotating Object

    Xue Yang, Junchi Yan, Ziming Feng +1

    cs.CVcs.LGeess.IVarXiv:1908.05612v62019
  14. Cable Manipulation with a Tactile-Reactive Gripper

    Yu She, Shaoxiong Wang, Siyuan Dong +3

    cs.ROeess.IVeess.SYarXiv:1910.02860v32019
  15. Automatic Detection of Coronavirus Disease (COVID-19) in X-ray and CT Images: A Machine Learning-Based Approach

    Sara Hosseinzadeh Kassani, Peyman Hosseinzadeh Kassasni, Michal J. Wesolowski +2

    eess.IVcs.CVarXiv:2004.10641v12020
  16. Neural Image Compression via Non-Local Attention Optimization and Improved Context Modeling

    Tong Chen, Haojie Liu, Zhan Ma +3

    eess.IVarXiv:1910.06244v12019
  17. Deep Unsupervised Domain Adaptation: A Review of Recent Advances and Perspectives

    Xiaofeng Liu, Chaehwa Yoo, Fangxu Xing +4

    cs.CVcs.AIcs.LGarXiv:2208.07422v12022
  18. Decoupling of brain function from structure reveals regional behavioral specialization in humans

    Maria Giulia Preti, Dimitri Van De Ville

    q-bio.NCeess.IVarXiv:1905.07813v22019
  19. Cross-view Semantic Segmentation for Sensing Surroundings

    Bowen Pan, Jiankai Sun, Ho Yin Tiga Leung +2

    cs.CVeess.IVarXiv:1906.03560v32019
  20. Spatially-Attentive Patch-Hierarchical Network for Adaptive Motion Deblurring

    Maitreya Suin, Kuldeep Purohit, A. N. Rajagopalan

    cs.CVeess.IVarXiv:2004.05343v12020
  21. Neural Video Compression with Diverse Contexts

    Jiahao Li, Bin Li, Yan Lu

    eess.IVcs.CVcs.MMarXiv:2302.14402v32023
  22. On the Challenges and Perspectives of Foundation Models for Medical Image Analysis

    Shaoting Zhang, Dimitris Metaxas

    eess.IVcs.CVarXiv:2306.05705v22023
  23. U-Net Transformer: Self and Cross Attention for Medical Image Segmentation

    Olivier Petit, Nicolas Thome, Clément Rambour +1

    eess.IVcs.CVarXiv:2103.06104v22021
  24. Attention Guided Low-light Image Enhancement with a Large Scale Low-light Simulation Dataset

    Feifan Lv, Yu Li, Feng Lu

    eess.IVcs.CVarXiv:1908.00682v32019
  25. Remote Heart Rate Measurement from Highly Compressed Facial Videos: an End-to-end Deep Learning Solution with Video Enhancement

    Zitong Yu, Wei Peng, Xiaobai Li +2

    eess.IVcs.CVarXiv:1907.11921v12019
  26. Complex diffusion-weighted image estimation via matrix recovery under general noise models

    Lucilio Cordero-Grande, Daan Christiaens, Jana Hutter +2

    eess.IVstat.AParXiv:1812.05954v22018
  27. COIN: COmpression with Implicit Neural representations

    Emilien Dupont, Adam Goliński, Milad Alizadeh +2

    eess.IVcs.CVcs.LGarXiv:2103.03123v22021
  28. CellViT: Vision Transformers for Precise Cell Segmentation and Classification

    Fabian Hörst, Moritz Rempe, Lukas Heine +8

    eess.IVcs.CVcs.LGarXiv:2306.15350v22023
  29. Semi-Supervised Medical Image Segmentation via Cross Teaching between CNN and Transformer

    Xiangde Luo, Minhao Hu, Tao Song +2

    eess.IVcs.CVarXiv:2112.04894v22021
  30. FVC: A New Framework towards Deep Video Compression in Feature Space

    Zhihao Hu, Guo Lu, Dong Xu

    eess.IVcs.CVarXiv:2105.09600v22021
  31. Surgical Data Science -- from Concepts toward Clinical Translation

    Lena Maier-Hein, Matthias Eisenmann, Duygu Sarikaya +47

    cs.CYcs.CVcs.LGarXiv:2011.02284v22020
  32. A Review of Single-Source Deep Unsupervised Visual Domain Adaptation

    Sicheng Zhao, Xiangyu Yue, Shanghang Zhang +8

    cs.CVcs.LGeess.IVarXiv:2009.00155v32020
  33. Brain-Like Object Recognition with High-Performing Shallow Recurrent ANNs

    Jonas Kubilius, Martin Schrimpf, Kohitij Kar +11

    cs.CVcs.LGcs.NEarXiv:1909.06161v22019
  34. Towards Photo-Realistic Virtual Try-On by Adaptively Generating$\leftrightarrow$Preserving Image Content

    Han Yang, Ruimao Zhang, Xiaobao Guo +3

    cs.CVcs.GReess.IVarXiv:2003.05863v12020
  35. An Overview of Deep-Learning-Based Audio-Visual Speech Enhancement and Separation

    Daniel Michelsanti, Zheng-Hua Tan, Shi-Xiong Zhang +4

    eess.AScs.LGeess.IVarXiv:2008.09586v22020
  36. Association of genomic subtypes of lower-grade gliomas with shape features automatically extracted by a deep learning algorithm

    Mateusz Buda, Ashirbani Saha, Maciej A Mazurowski

    eess.IVcs.CVcs.LGarXiv:1906.03720v12019
  37. Uncertainty Guided Multi-Scale Residual Learning-using a Cycle Spinning CNN for Single Image De-Raining

    Rajeev Yasarla, Vishal M. Patel

    cs.CVcs.LGeess.IVarXiv:1906.11129v12019
  38. MALUNet: A Multi-Attention and Light-weight UNet for Skin Lesion Segmentation

    Jiacheng Ruan, Suncheng Xiang, Mingye Xie +2

    eess.IVcs.CVarXiv:2211.01784v12022
  39. 3D Photography using Context-aware Layered Depth Inpainting

    Meng-Li Shih, Shih-Yang Su, Johannes Kopf +1

    cs.CVeess.IVarXiv:2004.04727v32020
  40. AAU-net: An Adaptive Attention U-net for Breast Lesions Segmentation in Ultrasound Images

    Gongping Chen, Yu Dai, Jianxun Zhang +1

    eess.IVcs.CVcs.LGarXiv:2204.12077v32022
  41. ASD-DiagNet: A hybrid learning approach for detection of Autism Spectrum Disorder using fMRI data

    Taban Eslami, Vahid Mirjalili, Alvis Fong +2

    cs.LGeess.IVstat.MLarXiv:1904.07577v12019
  42. Aerial Imagery Pile burn detection using Deep Learning: the FLAME dataset

    Alireza Shamsoshoara, Fatemeh Afghah, Abolfazl Razi +3

    cs.CVcs.AIcs.LGarXiv:2012.14036v12020
  43. Optimizing the Dice Score and Jaccard Index for Medical Image Segmentation: Theory & Practice

    Jeroen Bertels, Tom Eelbode, Maxim Berman +4

    cs.CVcs.LGeess.IVarXiv:1911.01685v12019
  44. A U-Net Based Discriminator for Generative Adversarial Networks

    Edgar Schönfeld, Bernt Schiele, Anna Khoreva

    cs.CVcs.LGeess.IVarXiv:2002.12655v22020
  45. Reliable Tuberculosis Detection using Chest X-ray with Deep Learning, Segmentation and Visualization

    Tawsifur Rahman, Amith Khandakar, Muhammad Abdul Kadir +8

    eess.IVcs.CVarXiv:2007.14895v12020
  46. EMCAD: Efficient Multi-scale Convolutional Attention Decoding for Medical Image Segmentation

    Md Mostafijur Rahman, Mustafa Munir, Radu Marculescu

    eess.IVcs.CVarXiv:2405.06880v12024
  47. Space-Time Correspondence as a Contrastive Random Walk

    Allan Jabri, Andrew Owens, Alexei A. Efros

    cs.CVcs.LGeess.IVarXiv:2006.14613v22020
  48. Efficient and Accurate MRI Super-Resolution using a Generative Adversarial Network and 3D Multi-Level Densely Connected Network

    Yuhua Chen, Feng Shi, Anthony G. Christodoulou +3

    cs.CVeess.IVarXiv:1803.01417v32018
  49. A Survey: Deep Learning for Hyperspectral Image Classification with Few Labeled Samples

    Sen Jia, Shuguo Jiang, Zhijie Lin +3

    cs.CVcs.AIeess.IVarXiv:2112.01800v12021
  50. Building Instance Classification Using Street View Images

    Jian Kang, Marco Körner, Yuanyuan Wang +2

    cs.CVeess.IVarXiv:1802.09026v12018
  51. Integrating Spatial Configuration into Heatmap Regression Based CNNs for Landmark Localization

    Christian Payer, Darko Štern, Horst Bischof +1

    eess.IVcs.CVarXiv:1908.00748v12019
  52. CLIPstyler: Image Style Transfer with a Single Text Condition

    Gihyun Kwon, Jong Chul Ye

    cs.CVcs.CLeess.IVarXiv:2112.00374v32021
  53. Seeing What a GAN Cannot Generate

    David Bau, Jun-Yan Zhu, Jonas Wulff +4

    cs.CVcs.GRcs.LGarXiv:1910.11626v12019
  54. UGC-VQA: Benchmarking Blind Video Quality Assessment for User Generated Content

    Zhengzhong Tu, Yilin Wang, Neil Birkbeck +2

    cs.CVeess.IVarXiv:2005.14354v22020
  55. Mining Cross-Image Semantics for Weakly Supervised Semantic Segmentation

    Guolei Sun, Wenguan Wang, Jifeng Dai +1

    cs.CVcs.LGeess.IVarXiv:2007.01947v22020
  56. Deep Learning Methods for Parallel Magnetic Resonance Image Reconstruction

    Florian Knoll, Kerstin Hammernik, Chi Zhang +4

    eess.SPcs.CVcs.LGarXiv:1904.01112v12019
  57. The Devil Is in the Details: Window-based Attention for Image Compression

    Renjie Zou, Chunfeng Song, Zhaoxiang Zhang

    eess.IVcs.CVarXiv:2203.08450v12022
  58. Self-attention for raw optical Satellite Time Series Classification

    Marc Rußwurm, Marco Körner

    cs.LGeess.IVstat.MLarXiv:1910.10536v32019
  59. BBDM: Image-to-image Translation with Brownian Bridge Diffusion Models

    Bo Li, Kaitao Xue, Bin Liu +1

    cs.CVeess.IVarXiv:2205.07680v22022
  60. Anomaly Detection in Video via Self-Supervised and Multi-Task Learning

    Mariana-Iuliana Georgescu, Antonio Barbalau, Radu Tudor Ionescu +3

    cs.CVcs.LGeess.IVarXiv:2011.07491v32020