paper-with-me

Papers

Unimodal Cyclic Regularization for Training Multimodal Image Registration Networks

2020-11-12 · Zhe Xu, Jiangpeng Yan, Jie Luo, William Wells, Xiu Li, Jayender Jagadeesan

The loss function of an unsupervised multimodal image registration framework has two terms, i.e., a metric for similarity measure and regularization. In the deep learning era, researchers proposed many approaches to automatically learn the similarity metric, which has been shown effective in improving registration performance. However, for the regularization term, most existing multimodal registration approaches still use a hand-crafted formula to impose artificial properties on the estimated deformation field. In this work, we propose a unimodal cyclic regularization training pipeline, which learns task-specific prior knowledge from simpler unimodal registration, to constrain the deformation field of multimodal registration. In the experiment of abdominal CT-MR registration, the proposed method yields better results over conventional regularization methods, especially for severely deformed local regions.

📄 PDF Abstract BibTeX arXiv:2011.06214

Code (0)

등록된 구현이 없습니다.

Tasks

Image Registration

Similar Papers 제목 키워드 기반

Diverse via bounded Agreement: Geometric Regularization for Multimodal Fusion

2026-01-29 · Zixuan Xia, Hao Wang, Pengcheng Weng, Yanyu Qian 외 arxiv

Multimodal fusion is often treated as an optimization-balancing problem, where training signals are adjusted to prevent one modality from dominating the others. However, balanced optimization does not fully determine the…

Representation Learning

Transformer Decoders with MultiModal Regularization for Cross-Modal Food Retrieval

2022-04-20 · Mustafa Shukor, Guillaume Couairon, Asya Grechka, Matthieu Cord

Cross-modal image-recipe retrieval has gained significant attention in recent years. Most work focuses on improving cross-modal embeddings using unimodal encoders, that allow for efficient retrieval in large-scale databa…

Cross-Modal RetrievalRetrievalTriplet

CSA: Data-efficient Mapping of Unimodal Features to Multimodal Features

2024-10-10 · Po-han Li, Sandeep P. Chinchali, Ufuk Topcu

Multimodal encoders like CLIP excel in tasks such as zero-shot image classification and cross-modal retrieval. However, they require excessive training data. We propose canonical similarity analysis (CSA), which uses two…

Cross-Modal RetrievalGPUimage-classificationImage Classification+1

A survey on Self Supervised learning approaches for improving Multimodal representation learning

2022-10-20 · Naman Goyal

Recently self supervised learning has seen explosive growth and use in variety of machine learning tasks because of its ability to avoid the cost of annotating large-scale datasets. This paper gives an overview for best …

Representation LearningSelf-Supervised LearningTranslation

MCA: Modality Composition Awareness for Robust Composed Multimodal Retrieval

2025-10-17 · Qiyu Wu, Shuyang Cui, Satoshi Hayakawa, Wei-Yao Wang 외 arxiv

Multimodal retrieval, which seeks to retrieve relevant content across modalities such as text or image, supports applications from AI search to contents production. Despite the success of separate-encoder approaches like…

Contrastive Learning