paper-with-me

Papers

U-Net with Hierarchical Bottleneck Attention for Landmark Detection in Fundus Images of the Degenerated Retina

2021-07-09 · Shuyun Tang, Ziming Qi, Jacob Granley, Michael Beyeler

Fundus photography has routinely been used to document the presence and severity of retinal degenerative diseases such as age-related macular degeneration (AMD), glaucoma, and diabetic retinopathy (DR) in clinical practice, for which the fovea and optic disc (OD) are important retinal landmarks. However, the occurrence of lesions, drusen, and other retinal abnormalities during retinal degeneration severely complicates automatic landmark detection and segmentation. Here we propose HBA-U-Net: a U-Net backbone enriched with hierarchical bottleneck attention. The network consists of a novel bottleneck attention block that combines and refines self-attention, channel attention, and relative-position attention to highlight retinal abnormalities that may be important for fovea and OD segmentation in the degenerated retina. HBA-U-Net achieved state-of-the-art results on fovea detection across datasets and eye conditions (ADAM: Euclidean Distance (ED) of 25.4 pixels, REFUGE: 32.5 pixels, IDRiD: 32.1 pixels), on OD segmentation for AMD (ADAM: Dice Coefficient (DC) of 0.947), and on OD detection for DR (IDRiD: ED of 20.5 pixels). Our results suggest that HBA-U-Net may be well suited for landmark detection in the presence of a variety of retinal degenerative diseases.

📄 PDF Abstract BibTeX arXiv:2107.04721

Code (1)

bionicvisionlab/2021-HBA-U-Net tf

Tasks

Fovea DetectionOptic Disc DetectionOptic Disc SegmentationSegmentation

Methods 이 논문이 사용한 방법론

Concatenated Skip Connection A Concatenated Skip Connection is a type of skip connection that seeks to reuse features by concatenating them to new layers, allowing more information to be retained from…
Max Pooling Max Pooling is a pooling operation that calculates the maximum value for patches of a feature map, and uses it to create a downsampled (pooled) feature map. It is usually…
Convolution A convolution is a type of matrix operation, consisting of a kernel, a small matrix of weights, that slides over input data performing element-wise multiplication with the…
ReLU How Do I Communicate to Expedia? How Do I Communicate to Expedia? – Call ☎️ +1-(888) 829 (0881) or +1-805-330-4056 or +1-805-330-4056 for Live Support & Special Travel…
U-Net 설명 없음

Similar Papers 제목 키워드 기반

Localizing Anatomical Landmarks in Ocular Images using Zoom-In Attentive Networks

2022-09-25 · Xiaofeng Lei, Shaohua Li, Xinxing Xu, Huazhu Fu 외

Localizing anatomical landmarks are important tasks in medical image analysis. However, the landmarks to be localized often lack prominent visual features. Their locations are elusive and easily confused with the backgro…

Medical Image Analysisobject-detectionObject Detection

JOINED : Prior Guided Multi-task Learning for Joint Optic Disc/Cup Segmentation and Fovea Detection

2022-03-01 · Huaqing He, Li Lin, Zhiyuan Cai, Xiaoying Tang

Fundus photography has been routinely used to document the presence and severity of various retinal degenerative diseases such as age-related macula degeneration, glaucoma, and diabetic retinopathy, for which the fovea, …

Fovea DetectionMulti-Task LearningSegmentation

HSQ-VLM: A Novel Spatially-Constrained Quadrant Segmentation VLM Model for Explainability in Diabetic Retinopathy

2026-06-11 · Shivum Telang arxiv

Diabetic Retinopathy (DR) is an aggressive retinal disease and a leading cause of global blindness, yet its clinical management is currently hindered by the black-box nature of diagnostic AI. While deep learning models a…

Learning Robust Facial Landmark Detection via Hierarchical Structured Ensemble

2019-10-01 · ICCV 2019 10 · Xu Zou, Sheng Zhong, Luxin Yan, Xiangyun Zhao 외

Heatmap regression-based models have significantly advanced the progress of facial landmark detection. However, the lack of structural constraints always generates inaccurate heatmaps resulting in poor landmark detection…

Face AlignmentFacial Landmark Detection

Attention-Driven Cropping for Very High Resolution Facial Landmark Detection

2020-06-01 · CVPR 2020 6 · Prashanth Chandran, Derek Bradley, Markus Gross, Thabo Beeler

Facial landmark detection is a fundamental task for many consumer and high-end applications and is almost entirely solved by machine learning methods today. Existing datasets used to train such algorithms are primarily m…

4kFacial Landmark DetectionVocal Bursts Intensity Prediction