Multi-task UNet: Jointly Boosting Saliency Prediction and Disease Classification on Chest X-ray Images
Human visual attention has recently shown its distinct capability in boosting machine learning models. However, studies that aim to facilitate medical tasks with human visual attention are still scarce. To support the use of visual attention, this paper describes a novel deep learning model for visual saliency prediction on chest X-ray (CXR) images. To cope with data deficiency, we exploit the multi-task learning method and tackles disease classification on CXR simultaneously. For a more robust training process, we propose a further optimized multi-task learning scheme to better handle model overfitting. Experiments show our proposed deep learning model with our new learning scheme can outperform existing methods dedicated either for saliency prediction or image classification. The code used in this paper is available at https://github.com/hz-zhu/MT-UNet.
Code (1)
Tasks
Deep Learningimage-classificationImage ClassificationMulti-Task LearningSaliency PredictionSimilar Papers 제목 키워드 기반
DiffSal: Joint Audio and Video Learning for Diffusion Saliency Prediction
Audio-visual saliency prediction can draw support from diverse modality complements, but further performance enhancement is still challenged by customized architectures as well as task-specific loss functions. In recent …
DenoisingPredictionSaliency PredictionCapSal: Leveraging Captioning to Boost Semantics for Salient Object Detection
Detecting salient objects in cluttered scenes is a big challenge. To address this problem, we argue that the model needs to learn discriminative semantic features for salient objects. To this end, we propose to leverage…
Image Captioningobject-detectionObject DetectionRGB Salient Object Detection+1SAUNet: Shape Attentive U-Net for Interpretable Medical Image Segmentation
Medical image segmentation is a difficult but important task for many clinical operations such as cardiac bi-ventricular volume estimation. More recently, there has been a shift to utilizing deep learning and fully convo…
Image SegmentationMedical Image SegmentationSegmentationSemantic SegmentationSmaAT-QMix-UNet: A Parameter-Efficient Vector-Quantized UNet for Precipitation Nowcasting
Weather forecasting supports critical socioeconomic activities and complements environmental protection, yet operational Numerical Weather Prediction (NWP) systems remain computationally intensive, thus being inefficient…
Weather ForecastingSLTUNET: A Simple Unified Model for Sign Language Translation
Despite recent successes with neural models for sign language translation (SLT), translation quality still lags behind spoken languages because of the data scarcity and modality gap between sign video and text. To addres…
Machine TranslationSign Language TranslationTranslation