Multi-view Cross-Modality MR Image Translation for Vestibular Schwannoma and Cochlea Segmentation
In this work, we propose a multi-view image translation framework, which can translate contrast-enhanced T1 (ceT1) MR imaging to high-resolution T2 (hrT2) MR imaging for unsupervised vestibular schwannoma and cochlea segmentation. We adopt two image translation models in parallel that use a pixel-level consistent constraint and a patch-level contrastive constraint, respectively. Thereby, we can augment pseudo-hrT2 images reflecting different perspectives, which eventually lead to a high-performing segmentation model. Our experimental results on the CrossMoDA challenge show that the proposed method achieved enhanced performance on the vestibular schwannoma and cochlea segmentation.
Code (0)
등록된 구현이 없습니다.
Tasks
SegmentationTranslationSimilar Papers 제목 키워드 기반
Towards General Modality Translation with Contrastive and Predictive Latent Diffusion Bridge
Recent advances in generative modeling have positioned diffusion models as state-of-the-art tools for sampling from complex data distributions. While these models have shown remarkable success across single-modality doma…
Image Super-ResolutionCM-Diff: A Single Generative Network for Bidirectional Cross-Modality Translation Diffusion Model Between Infrared and Visible Images
The image translation method represents a crucial approach for mitigating information deficiencies in the infrared and visible modalities, while also facilitating the enhancement of modality-specific datasets. However, e…
TranslationTMT: Tri-Modal Translation between Speech, Image, and Text by Processing Different Modalities as Different Languages
The capability to jointly process multi-modal information is becoming an essential task. However, the limited number of paired multi-modal data and the large computational requirements in multi-modal learning hinder the …
DecoderMachine TranslationTranslationMultimodal Machine Translation through Visuals and Speech
Multimodal machine translation involves drawing information from more than one modality, based on the assumption that the additional modalities will contain useful alternative views of the input data. The most prominent …
Image CaptioningMachine TranslationMultimodal Machine Translationspeech-recognition+3TarGAN: Target-Aware Generative Adversarial Networks for Multi-modality Medical Image Translation
Paired multi-modality medical images, can provide complementary information to help physicians make more reasonable decisions than single modality medical images. But they are difficult to generate due to multiple factor…
Generative Adversarial NetworkTranslation