paper-with-me

홈 › Papers

Focus on Focus: Focus-oriented Representation Learning and Multi-view Cross-modal Alignment for Glioma Grading

2024-08-16 · Li Pan, Yupei Zhang, Qiushi Yang, Tan Li, Xiaohan Xing, Maximus C. F. Yeung, Zhen Chen

Recently, multimodal deep learning, which integrates histopathology slides and molecular biomarkers, has achieved a promising performance in glioma grading. Despite great progress, due to the intra-modality complexity and inter-modality heterogeneity, existing studies suffer from inadequate histopathology representation learning and inefficient molecular-pathology knowledge alignment. These two issues hinder existing methods to precisely interpret diagnostic molecular-pathology features, thereby limiting their grading performance. Moreover, the real-world applicability of existing multimodal approaches is significantly restricted as molecular biomarkers are not always available during clinical deployment. To address these problems, we introduce a novel Focus on Focus (FoF) framework with paired pathology-genomic training and applicable pathology-only inference, enhancing molecular-pathology representation effectively. Specifically, we propose a Focus-oriented Representation Learning (FRL) module to encourage the model to identify regions positively or negatively related to glioma grading and guide it to focus on the diagnostic areas with a consistency constraint. To effectively link the molecular biomarkers to morphological features, we propose a Multi-view Cross-modal Alignment (MCA) module that projects histopathology representations into molecular subspaces, aligning morphological features with corresponding molecular biomarker status by supervised contrastive learning. Experiments on the TCGA GBM-LGG dataset demonstrate that our FoF framework significantly improves the glioma grading. Remarkably, our FoF achieves superior performance using only histopathology slides compared to existing multimodal methods. The source code is available at https://github.com/peterlipan/FoF.

📄 PDF Abstract BibTeX arXiv:2408.08527

Code (1)

peterlipan/fof 공식 구현 pytorch

Tasks

Contrastive Learningcross-modal alignmentDiagnosticMultimodal Deep LearningRepresentation Learning

Methods 이 논문이 사용한 방법론

Focus 설명 없음

Similar Papers 제목 키워드 기반

Hyperdimensional Probe: Decoding LLM Representations via Vector Symbolic Architectures

2025-09-29 · Marco Bronzini, Carlo Nicolini, Bruno Lepri, Jacopo Staiano 외 arxiv

Despite their capabilities, Large Language Models (LLMs) remain opaque with limited understanding of their internal representations. Current interpretability methods either focus on input-oriented feature extraction, suc…

Text Generation

OBJECT-ORIENTED REPRESENTATION OF 3D SCENES

2019-09-25 · Chang Chen, Sungjin Ahn

In this paper, we propose a generative model, called ROOTS (Representation of Object-Oriented Three-dimension Scenes), for unsupervised object-wise 3D-scene decomposition and and rendering. For 3D scene modeling, ROOTS b…

DisentanglementObject

Improving Sequence-to-Sequence Semantic Parser for Task Oriented Dialog

2020-11-01 · EMNLP (intexsempar) 2020 11 · Chaoting Xuan

Task Oriented Parsing (TOP) attempts to map utterances to compositional requests, including multiple intents and their slots. Previous work focus on a tree-based hierarchical meaning representation, and applying constitu…

Constituency Parsing

Prompt-Guided Attention Head Selection for Focus-Oriented Image Retrieval

2025-04-02 · Yuji Nozawa, Yu-Chieh Lin, Kazumoto Nakamura, Youyang Ng

The goal of this paper is to enhance pretrained Vision Transformer (ViT) models for focus-oriented image retrieval with visual prompting. In real-world image retrieval scenarios, both query and database images often exhi…

Image RetrievalRetrievalVisual Prompting

DMM: Disparity-guided Multispectral Mamba for Oriented Object Detection in Remote Sensing

2024-07-11 · Minghang Zhou, Tianyu Li, Chaofan Qiao, Dongyu Xie 외

Multispectral oriented object detection faces challenges due to both inter-modal and intra-modal discrepancies. Recent studies often rely on transformer-based models to address these issues and achieve cross-modal fusion…

Computational EfficiencyMambaobject-detectionObject Detection+1