paper-with-me

홈 › Papers

Cross-Task Attention Network: Improving Multi-Task Learning for Medical Imaging Applications

2023-09-07 · SangWook Kim, Thomas G. Purdie, Chris McIntosh

Multi-task learning (MTL) is a powerful approach in deep learning that leverages the information from multiple tasks during training to improve model performance. In medical imaging, MTL has shown great potential to solve various tasks. However, existing MTL architectures in medical imaging are limited in sharing information across tasks, reducing the potential performance improvements of MTL. In this study, we introduce a novel attention-based MTL framework to better leverage inter-task interactions for various tasks from pixel-level to image-level predictions. Specifically, we propose a Cross-Task Attention Network (CTAN) which utilizes cross-task attention mechanisms to incorporate information by interacting across tasks. We validated CTAN on four medical imaging datasets that span different domains and tasks including: radiation treatment planning prediction using planning CT images of two different target cancers (Prostate, OpenKBP); pigmented skin lesion segmentation and diagnosis using dermatoscopic images (HAM10000); and COVID-19 diagnosis and severity prediction using chest CT scans (STOIC). Our study demonstrates the effectiveness of CTAN in improving the accuracy of medical imaging tasks. Compared to standard single-task learning (STL), CTAN demonstrated a 4.67% improvement in performance and outperformed both widely used MTL baselines: hard parameter sharing (HPS) with an average performance improvement of 3.22%; and multi-task attention network (MTAN) with a relative decrease of 5.38%. These findings highlight the significance of our proposed MTL framework in solving medical imaging tasks and its potential to improve their accuracy across domains.

📄 PDF Abstract BibTeX arXiv:2309.03837

Code (0)

등록된 구현이 없습니다.

Tasks

COVID-19 DiagnosisLesion SegmentationMulti-Task Learningseverity predictionSkin Lesion Segmentation

Similar Papers 제목 키워드 기반

Multi-Modal Brain Tumor Segmentation via 3D Multi-Scale Self-attention and Cross-attention

2025-04-12 · Yonghao Huang, Leiting Chen, Chuan Zhou

Due to the success of CNN-based and Transformer-based models in various computer vision tasks, recent works study the applicability of CNN-Transformer hybrid architecture models in 3D multi-modality medical segmentation …

Brain Tumor SegmentationDecoderImage SegmentationMedical Image Segmentation+3

BERT-GT: Cross-sentence n-ary relation extraction with BERT and Graph Transformer

2021-01-11 · Po-Ting Lai, Zhiyong Lu

A biomedical relation statement is commonly expressed in multiple sentences and consists of many concepts, including gene, disease, chemical, and mutation. To automatically extract information from biomedical literature,…

BenchmarkingBinary Relation ExtractionGraph Neural NetworkRelation+2

CATNet: Cross-event Attention-based Time-aware Network for Medical Event Prediction

2022-04-29 · Sicen Liu, Xiaolong Wang, Yang Xiang, Hui Xu 외

Medical event prediction (MEP) is a fundamental task in the medical domain, which needs to predict medical events, including medications, diagnosis codes, laboratory tests, procedures, outcomes, and so on, according to h…

Time Series Analysis

Studying the Effects of Self-Attention for Medical Image Analysis

2021-09-02 · Adrit Rao, Jongchan Park, Sanghyun Woo, Joon-Young Lee 외

When the trained physician interprets medical images, they understand the clinical importance of visual features. By applying cognitive attention, they apply greater focus onto clinically relevant regions while disregard…

Medical Image Analysis

Med-2E3: A 2D-Enhanced 3D Medical Multimodal Large Language Model

2024-11-19 · Yiming Shi, Xun Zhu, Ying Hu, Chenyi Guo 외

The analysis of 3D medical images is crucial for modern healthcare, yet traditional task-specific models are becoming increasingly inadequate due to limited generalizability across diverse clinical scenarios. Multimodal …

Language ModelingLanguage ModellingLarge Language ModelMedical Image Analysis+5