paper-with-me

홈 › Papers

When Two Tracers Disagree: An Investigation of Multimodal Fusion for Clinical PET/CT Segmentation

2026-08-19 · Jack A. Johnson, Bartłomiej W. Papież arxiv

PSMA and FDG PET/CT visualise complementary biological information in prostate cancer. Combining both tracers could capture heterogeneous tumour phenotypes that may be missed by either alone, yet there is no consensus on effective deep learning architectures for fusing these modalities. We evaluated multimodal image-fusion strategies for automatic whole-body PET/CT lesion segmentation to estimate total tumour burden. Using the public DEEP-PSMA Challenge dataset, we trained tracer-specific 3D nnU-Net baselines and compared (i) early fusion with a single encoder and one decoder (OEOD) or two decoders (OETD), and (ii) intermediate fusion via a dual-encoder cross-attention U-Net (DECA-UNet). Tracer-specific baselines performed strongly (PSMA Dice = 0.93; FDG = 0.81). Fusion yielded mixed results: OEOD produced a combined Dice of 0.90 (on an easier, non-tracer-specific task), whilst the tracer-specific fusion models reached PSMA/FDG = 0.69/0.64 (OETD) and 0.76/0.57 (DECA-UNet). Whilst fusion often provided reasonable PSMA segmentation, FDG performance degraded and no strategy consistently exceeded the single-tracer baselines. Under the evaluated setting, tracer-specific models remain the stronger baseline; clinically useful gains from multimodal fusion will likely require architectures that better preserve tracer specific representations. Our code is available at: https://github.com/JackJ3636/DEEP_PSMA_code

📄 PDF Abstract BibTeX arXiv:2608.19063

Code (0)

등록된 구현이 없습니다.

Tasks

Lesion Segmentation

Similar Papers 제목 키워드 기반

Mixed Signals: Understanding Model Disagreement in Multimodal Empathy Detection

2025-05-20 · Maya Srikanth, Run Chen, Julia Hirschberg

Multimodal models play a key role in empathy detection, but their performance can suffer when modalities provide conflicting cues. To understand these failures, we examine cases where unimodal and multimodal predictions …

DiagnosticPosition

Deep residual inception encoder-decoder network for amyloid PET harmonization

2022-02-09 · Alzheimer's and Dementia 2022 2 · Jay Shah, Fei Gao, Baoxin Li, Valentina Ghisays 외

Multiple positron emission tomography (PET) tracers are available for amyloid imaging, posing a significant challenge to consensus interpretation and quantitative analysis. We accordingly developed and validated a deep l…

DecoderImage HarmonizationImage-to-Image Translation

MUST-PET: MUltimodal Self-supervised learning across Tracers for whole-body PET/CT-based lesion segmentation

2026-08-20 · Bashirul Azam Biswas, Amartya Bhattacharya, Biratal Raj Wagle, Matthew E. Maeder 외 arxiv

Deep learning-based whole-body PET-CT lesion segmentation can support cancer staging, treatment planning, and response assessment, but generalization is limited by scarce annotations and domain shifts. Self-supervised le…

Self-Supervised LearningLesion Segmentation

LLaDA-V: Large Language Diffusion Models with Visual Instruction Tuning

2025-05-22 · Zebin You, Shen Nie, Xiaolu Zhang, Jun Hu 외

In this work, we introduce LLaDA-V, a purely diffusion-based Multimodal Large Language Model (MLLM) that integrates visual instruction tuning with masked diffusion models, representing a departure from the autoregressive…

Language ModelingLanguage ModellingLarge Language ModelMultimodal Large Language Model

PET Tracer Separation Using Conditional Diffusion Transformer with Multi-latent Space Learning

2025-06-20 · Bin Huang, Feihong Xu, Xinchong Shi, Shan Huang 외

In clinical practice, single-radiotracer positron emission tomography (PET) is commonly used for imaging. Although multi-tracer PET imaging can provide supplementary information of radiotracers that are sensitive to phys…

Computational Efficiency