paper-with-me

홈 › Papers

A Deep Step Pattern Representation for Multimodal Retinal Image Registration

2019-10-01 · ICCV 2019 10 · Jimmy Addison Lee, Peng Liu, Jun Cheng, Huazhu Fu

This paper presents a novel feature-based method that is built upon a convolutional neural network (CNN) to learn the deep representation for multimodal retinal image registration. We coined the algorithm deep step patterns, in short DeepSPa. Most existing deep learning based methods require a set of manually labeled training data with known corresponding spatial transformations, which limits the size of training datasets. By contrast, our method is fully automatic and scale well to different image modalities with no human intervention. We generate feature classes from simple step patterns within patches of connecting edges formed by vascular junctions in multiple retinal imaging modalities. We leverage CNN to learn and optimize the input patches to be used for image registration. Spatial transformations are estimated based on the output possibility of the fully connected layer of CNN for a pair of images. One of the key advantages of the proposed algorithm is its robustness to non-linear intensity changes, which widely exist on retinal images due to the difference of acquisition modalities. We validate our algorithm on extensive challenging datasets comprising poor quality multimodal retinal images which are adversely affected by pathologies (diseases), speckle noise and low resolutions. The experimental results demonstrate the robustness and accuracy over state-of-the-art multimodal image registration algorithms.

📄 PDF Abstract BibTeX

Code (0)

등록된 구현이 없습니다.

Tasks

Image Registration

Similar Papers 제목 키워드 기반

A Low-Dimensional Step Pattern Analysis Algorithm With Application to Multimodal Retinal Image Registration

2015-06-01 · CVPR 2015 6 · Jimmy Addison Lee, Jun Cheng, Beng Hai Lee, Ee Ping Ong 외

Existing feature descriptor-based methods on retinal image registration are mainly based on scale-invariant feature transform (SIFT) or partial intensity invariant feature descriptor (PIIFD). While these descriptors are …

Image Registration

UrFound: Towards Universal Retinal Foundation Models via Knowledge-Guided Masked Modeling

2024-08-10 · Kai Yu, Yang Zhou, Yang Bai, Zhi Da Soh 외

Retinal foundation models aim to learn generalizable representations from diverse retinal images, facilitating label-efficient model adaptation across various ophthalmic tasks. Despite their success, current retinal foun…

Representation Learning

Predicting Stroke through Retinal Graphs and Multimodal Self-supervised Learning

2024-11-08 · Yuqing Huang, Bastian Wittmann, Olga Demler, Bjoern Menze 외

Early identification of stroke is crucial for intervention, requiring reliable models. We proposed an efficient retinal image representation together with clinical information to capture a comprehensive overview of cardi…

Contrastive LearningSelf-Supervised LearningTransfer Learning

Bridging the Gap in Ophthalmic AI: MM-Retinal-Reason Dataset and OphthaReason Model toward Dynamic Multimodal Reasoning

2025-08-22 · Ruiqi Wu, Yuang Yao, Tengfei Ma, Chenran Zhang 외 arxiv

Multimodal large language models (MLLMs) have recently demonstrated remarkable reasoning abilities with reinforcement learning paradigm. Although several multimodal reasoning models have been explored in the medical doma…

Reinforcement LearningMultimodal Reasoning

REVEAL: Multimodal Vision-Language Alignment of Retinal Morphometry and Clinical Risks for Incident AD and Dementia Prediction

2026-04-20 · Seowung Leem, Lin Gu, Chenyu You, Kuang Gong 외 arxiv

The retina provides a unique, noninvasive window into Alzheimer's disease (AD) and dementia, capturing early structural changes through morphometric features, while systemic and lifestyle risk factors reflect well-establ…

Representation LearningContrastive Learning