Combining Fine- and Coarse-Grained Classifiers for Diabetic Retinopathy Detection
Visual artefacts of early diabetic retinopathy in retinal fundus images are usually small in size, inconspicuous, and scattered all over retina. Detecting diabetic retinopathy requires physicians to look at the whole image and fixate on some specific regions to locate potential biomarkers of the disease. Therefore, getting inspiration from ophthalmologist, we propose to combine coarse-grained classifiers that detect discriminating features from the whole images, with a recent breed of fine-grained classifiers that discover and pay particular attention to pathologically significant regions. To evaluate the performance of this proposed ensemble, we used publicly available EyePACS and Messidor datasets. Extensive experimentation for binary, ternary and quaternary classification shows that this ensemble largely outperforms individual image classifiers as well as most of the published works in most training setups for diabetic retinopathy detection. Furthermore, the performance of fine-grained classifiers is found notably superior than coarse-grained image classifiers encouraging the development of task-oriented fine-grained classifiers modelled after specialist ophthalmologists.
Code (0)
등록된 구현이 없습니다.
Tasks
Diabetic Retinopathy DetectionSimilar Papers 제목 키워드 기반
Label Stability in Multiple Instance Learning
We address the problem of \emph{instance label stability} in multiple instance learning (MIL) classifiers. These classifiers are trained only on globally annotated images (bags), but often can provide fine-grained annota…
Medical Image AnalysisMultiple Instance LearningBidirectional Logits Tree: Pursuing Granularity Reconcilement in Fine-Grained Classification
This paper addresses the challenge of Granularity Competition in fine-grained classification tasks, which arises due to the semantic gap between multi-granularity labels. Existing approaches typically develop independent…
Lost in Space: Probing Fine-grained Spatial Understanding in Vision and Language Resamplers
An effective method for combining frozen large language models (LLM) and visual encoders involves a resampler module that creates a `visual prompt' which is provided to the LLM, along with the textual prompt. While this …
DiagnosticImage CaptioningQuestion AnsweringVisual Question AnsweringClassifying Object Manipulation Actions based on Grasp-types and Motion-Constraints
In this work, we address a challenging problem of fine-grained and coarse-grained recognition of object manipulation actions. Due to the variations in geometrical and motion constraints, there are different manipulations…
Action RecognitionObjectTemporal Action LocalizationFine-grained Category Discovery under Coarse-grained supervision with Hierarchical Weighted Self-contrastive Learning
Novel category discovery aims at adapting models trained on known categories to novel categories. Previous works only focus on the scenario where known and novel categories are of the same granularity. In this paper, we …
Contrastive Learning