paper-with-me

Papers

Search-MIND: Training-Free Multi-Modal Medical Image Registration

2026-04-10 · Boya Wang, Ruizhe Li, Chao Chen, Xin Chen arxiv

Multi-modal image registration plays a critical role in precision medicine but faces challenges from non-linear intensity relationships and local optima. While deep learning models enable rapid inference, they often suffer from generalization collapse on unseen modalities. To address this, we propose Search-MIND, a training-free, iterative optimization framework for instance-specific registration. Our pipeline utilizes a coarse-to-fine strategy: a hierarchical coarse alignment stage followed by deformable refinement. We introduce two novel loss functions: Variance-Weighted Mutual Information (VWMI), which prioritizes informative tissue regions to shield global alignment from background noise and uniform regions, and Search-MIND (S-MIND), which broadens the convergence basin of structural descriptors by considering larger local search range. Evaluations on CARE Liver 2025 and CHAOS Challenge datasets show that Search-MIND consistently outperforms classical baselines like ANTs and foundation model-based approaches like DINO-reg, offering superior stability across diverse modalities.

📄 PDF Abstract BibTeX arXiv:2604.09743

Code (0)

등록된 구현이 없습니다.

Tasks

Medical Image Registration

Similar Papers 제목 키워드 기반

TokaMind: A Multi-Modal Transformer Foundation Model for Tokamak Plasma Dynamics

2026-02-16 · Tobia Boschi, Andrea Loreti, Nicola C. Amorisco, Rodrigo H. Ordonez-Hurtado 외 arxiv

We present TokaMind, to our knowledge the first open-source foundation model for tokamak plasma dynamics, based on a Multi-Modal Transformer (MMT) and pretrained on heterogeneous diagnostics from the publicly available M…

MindVL: Towards Efficient and Effective Training of Multimodal Large Language Models on Ascend NPUs

2025-09-15 · Feilong Chen, Yijiang Liu, Yi Huang, Hao Wang 외 arxiv

We propose MindVL, a multimodal large language model (MLLMs) trained on Ascend NPUs. The training of state-of-the-art MLLMs is often confined to a limited set of hardware platforms and relies heavily on massive, undisclo…

From Black Boxes to Transparent Minds: Evaluating and Enhancing the Theory of Mind in Multimodal Large Language Models

2025-06-17 · Xinyang Li, SiQi Liu, Bochao Zou, Jiansheng Chen 외

As large language models evolve, there is growing anticipation that they will emulate human-like Theory of Mind (ToM) to assist with routine tasks. However, existing methods for evaluating machine ToM focus primarily on …

Mind Artist: Creating Artistic Snapshots with Human Thought

2024-01-01 · CVPR 2024 1 · Jiaxuan Chen, Yu Qi, Yueming Wang, Gang Pan

We introduce Mind Artist (MindArt) a novel and efficient neural decoding architecture to snap artistic photographs from our mind in a controllable manner. Recently progress has been made in image reconstruction with …

Graph MatchingImage ReconstructionRepresentation Learning

MindWatcher: Toward Smarter Multimodal Tool-Integrated Reasoning

2025-12-29 · Jiawei Chen, Xintian Shen, Lihao Zheng, Zhenwei Shao 외 arxiv

Traditional workflow-based agents exhibit limited intelligence when addressing real-world problems requiring tool invocation. Tool-integrated reasoning (TIR) agents capable of autonomous reasoning and tool invocation are…

Object RecognitionImage Retrieval