A Survey on Biomedical Image Captioning
Image captioning applied to biomedical images can assist and accelerate the diagnosis process followed by clinicians. This article is the first survey of biomedical image captioning, discussing datasets, evaluation measures, and state of the art methods. Additionally, we suggest two baselines, a weak and a stronger one; the latter outperforms all current state of the art systems on one of the datasets.
Code (2)
Tasks
Image CaptioningSurveySimilar Papers 제목 키워드 기반
Biomedical Image Reconstruction: A Survey
Biomedical image reconstruction research has been developed for more than five decades, giving rise to various techniques such as central and filtered back projection. With the rise of deep learning technology, biomedica…
Deep LearningImage ReconstructionSurveyA Comprehensive Survey of Deep Learning for Image Captioning
Generating a description of an image is called image captioning. Image captioning requires to recognize the important objects, their attributes and their relationships in an image. It also needs to generate syntactically…
Deep LearningImage CaptioningSurveyAttention-based transformer models for image captioning across languages: An in-depth survey and evaluation
Image captioning involves generating textual descriptions from input images, bridging the gap between computer vision and natural language processing. Recent advancements in transformer-based models have significantly im…
Caption GenerationImage CaptioningScene UnderstandingSurveySurveying the Landscape of Image Captioning Evaluation: A Comprehensive Taxonomy and Novel Ensemble Method
The task of image captioning has recently been gaining popularity, and with it the complex task of evaluating the quality of image captioning models. In this work, we present the first survey and taxonomy of over 70 diff…
DiversityImage CaptioningImage Captioning based on Deep Learning Methods: A Survey
Image captioning is a challenging task and attracting more and more attention in the field of Artificial Intelligence, and which can be applied to efficient image retrieval, intelligent blind guidance and human-computer …
DecoderDeep LearningImage CaptioningImage Retrieval+2