Learning Site-specific Styles for Multi-institutional Unsupervised Cross-modality Domain Adaptation
Unsupervised cross-modality domain adaptation is a challenging task in medical image analysis, and it becomes more challenging when source and target domain data are collected from multiple institutions. In this paper, we present our solution to tackle the multi-institutional unsupervised domain adaptation for the crossMoDA 2023 challenge. First, we perform unpaired image translation to translate the source domain images to the target domain, where we design a dynamic network to generate synthetic target domain images with controllable, site-specific styles. Afterwards, we train a segmentation model using the synthetic images and further reduce the domain gap by self-training. Our solution achieved the 1st place during both the validation and testing phases of the challenge. The code repository is publicly available at https://github.com/MedICL-VU/crossmoda2023.
Code (1)
Tasks
Domain AdaptationMedical Image AnalysisMedical Image SegmentationStyle TransferUnsupervised Domain AdaptationSimilar Papers 제목 키워드 기반
Federated Learning of Generative Image Priors for MRI Reconstruction
Multi-institutional efforts can facilitate training of deep MRI reconstruction models, albeit privacy risks arise during cross-site sharing of imaging data. Federated learning (FL) has recently been introduced to address…
Federated LearningMRI ReconstructionSpecificityMISS GAN: A Multi-IlluStrator Style Generative Adversarial Network for image to illustration translation
Unsupervised style transfer that supports diverse input styles using only one trained generator is a challenging and interesting task in computer vision. This paper proposes a Multi-IlluStrator Style Generative Adversari…
Generative Adversarial NetworkStyle TransferTranslationInferring Restaurant Styles by Mining Crowd Sourced Photos from User-Review Websites
When looking for a restaurant online, user uploaded photos often give people an immediate and tangible impression about a restaurant. Due to their informativeness, such user contributed photos are leveraged by restaurant…
InformativenessMulti-Label LearningTAGToward Federated Large Language Models in Medicine: A Parameter-Efficient Framework for Privacy-Preserving, Multi-Institutional Adaptation
Large language models (LLMs) are increasingly adapted for medical applications, but most are trained using data from a single institution because privacy and governance constraints prevent multi-institutional data sharin…
Information ExtractionFederated LearningQuestion AnsweringImproving Speech Emotion Recognition with Unsupervised Speaking Style Transfer
Humans can effortlessly modify various prosodic attributes, such as the placement of stress and the intensity of sentiment, to convey a specific emotion while maintaining consistent linguistic content. Motivated by this …
Automatic Speech RecognitionAutomatic Speech Recognition (ASR)Data AugmentationDecoder+6