paper-with-me

Papers

Multi-task UNet: Jointly Boosting Saliency Prediction and Disease Classification on Chest X-ray Images

2022-02-15 · Hongzhi Zhu, Robert Rohling, Septimiu Salcudean

Human visual attention has recently shown its distinct capability in boosting machine learning models. However, studies that aim to facilitate medical tasks with human visual attention are still scarce. To support the use of visual attention, this paper describes a novel deep learning model for visual saliency prediction on chest X-ray (CXR) images. To cope with data deficiency, we exploit the multi-task learning method and tackles disease classification on CXR simultaneously. For a more robust training process, we propose a further optimized multi-task learning scheme to better handle model overfitting. Experiments show our proposed deep learning model with our new learning scheme can outperform existing methods dedicated either for saliency prediction or image classification. The code used in this paper is available at https://github.com/hz-zhu/MT-UNet.

📄 PDF Abstract BibTeX arXiv:2202.07118

Code (1)

hz-zhu/mt-unet 공식 구현 pytorch

Tasks

Deep Learningimage-classificationImage ClassificationMulti-Task LearningSaliency Prediction

Similar Papers 제목 키워드 기반

DiffSal: Joint Audio and Video Learning for Diffusion Saliency Prediction

2024-03-02 · CVPR 2024 1 · Junwen Xiong, Peng Zhang, Tao You, Chuanyue Li 외

Audio-visual saliency prediction can draw support from diverse modality complements, but further performance enhancement is still challenged by customized architectures as well as task-specific loss functions. In recent …

DenoisingPredictionSaliency Prediction

CapSal: Leveraging Captioning to Boost Semantics for Salient Object Detection

2019-06-01 · CVPR 2019 6 · Lu Zhang, Jianming Zhang, Zhe Lin, Huchuan Lu 외

Detecting salient objects in cluttered scenes is a big challenge. To address this problem, we argue that the model needs to learn discriminative semantic features for salient objects. To this end, we propose to leverage…

Image Captioningobject-detectionObject DetectionRGB Salient Object Detection+1

SAUNet: Shape Attentive U-Net for Interpretable Medical Image Segmentation

2020-01-21 · Jesse Sun, Fatemeh Darbehani, Mark Zaidi, Bo wang

Medical image segmentation is a difficult but important task for many clinical operations such as cardiac bi-ventricular volume estimation. More recently, there has been a shift to utilizing deep learning and fully convo…

Image SegmentationMedical Image SegmentationSegmentationSemantic Segmentation

SmaAT-QMix-UNet: A Parameter-Efficient Vector-Quantized UNet for Precipitation Nowcasting

2026-03-23 · Nikolas Stavrou, Siamak Mehrkanoon arxiv

Weather forecasting supports critical socioeconomic activities and complements environmental protection, yet operational Numerical Weather Prediction (NWP) systems remain computationally intensive, thus being inefficient…

Weather Forecasting

SLTUNET: A Simple Unified Model for Sign Language Translation

2023-05-02 · International Conference on Learning Representations (ICLR) 2023 2 · Biao Zhang, Mathias Müller, Rico Sennrich

Despite recent successes with neural models for sign language translation (SLT), translation quality still lags behind spoken languages because of the data scarcity and modality gap between sign video and text. To addres…

Machine TranslationSign Language TranslationTranslation