paper-with-me

홈 › Papers

Multimodal and multicontrast image fusion via deep generative models

2023-03-28 · Giovanna Maria Dimitri, Simeon Spasov, Andrea Duggento, Luca Passamonti, Pietro Li`o, Nicola Toschi

Recently, it has become progressively more evident that classic diagnostic labels are unable to reliably describe the complexity and variability of several clinical phenotypes. This is particularly true for a broad range of neuropsychiatric illnesses (e.g., depression, anxiety disorders, behavioral phenotypes). Patient heterogeneity can be better described by grouping individuals into novel categories based on empirically derived sections of intersecting continua that span across and beyond traditional categorical borders. In this context, neuroimaging data carry a wealth of spatiotemporally resolved information about each patient's brain. However, they are usually heavily collapsed a priori through procedures which are not learned as part of model training, and consequently not optimized for the downstream prediction task. This is because every individual participant usually comes with multiple whole-brain 3D imaging modalities often accompanied by a deep genotypic and phenotypic characterization, hence posing formidable computational challenges. In this paper we design a deep learning architecture based on generative models rooted in a modular approach and separable convolutional blocks to a) fuse multiple 3D neuroimaging modalities on a voxel-wise level, b) convert them into informative latent embeddings through heavy dimensionality reduction, c) maintain good generalizability and minimal information loss. As proof of concept, we test our architecture on the well characterized Human Connectome Project database demonstrating that our latent embeddings can be clustered into easily separable subject strata which, in turn, map to different phenotypical information which was not included in the embedding creation process. This may be of aid in predicting disease evolution as well as drug response, hence supporting mechanistic disease understanding and empowering clinical trials.

📄 PDF Abstract BibTeX arXiv:2303.15963

Code (0)

등록된 구현이 없습니다.

Tasks

DiagnosticDimensionality Reduction

Methods 이 논문이 사용한 방법론

Test 설명 없음

Similar Papers 제목 키워드 기반

Isotropic multichannel total variation framework for joint reconstruction of multicontrast parallel MRI

2020-06-07 · Erfan Ebrahim Esfahani

Purpose: To develop a synergistic image reconstruction framework that exploits multicontrast (MC), multicoil, and compressed sensing (CS) redundancies in magnetic resonance imaging (MRI). Approach: CS, MC acquisition, an…

compressed sensingImage Reconstruction

Query-Kontext: An Unified Multimodal Model for Image Generation and Editing

2025-09-30 · Yuxin Song, Wenkai Dong, Shizun Wang, Qi Zhang 외 arxiv

Unified Multimodal Models (UMMs) have demonstrated remarkable performance in text-to-image generation (T2I) and editing (TI2I), whether instantiated as assembled unified frameworks which couple powerful vision-language m…

Text-to-Image Generation

LMFusion: Adapting Pretrained Language Models for Multimodal Generation

2024-12-19 · Weijia Shi, Xiaochuang Han, Chunting Zhou, Weixin Liang 외

We present LMFusion, a framework for empowering pretrained text-only large language models (LLMs) with multimodal generative capabilities, enabling them to understand and generate both text and images in arbitrary sequen…

Image Generationmultimodal generation

Unified Multimodal Discrete Diffusion

2025-03-26 · Alexander Swerdlow, Mihir Prabhudesai, Siddharth Gandhi, Deepak Pathak 외

Multimodal generative models that can understand and generate across multiple modalities are dominated by autoregressive (AR) approaches, which process tokens sequentially from left to right, or top to bottom. These mode…

Image CaptioningImage GenerationQuestion AnsweringText Generation

There and Back Again: Bidirectional Diffusion Bridges for Multimodality Translation

2026-08-28 · Gabe Guo, Elon Litman, Thanawat Sornwanee, Jose Blanchet 외 arxiv

Multimodality translation (e.g., text-to-image) is a core generative AI task. However, existing approaches (1) follow generative paths that do not directly represent the source modality, limiting the flexibility of some …