paper-with-me

홈 › Papers

Reinforced Auto-Zoom Net: Towards Accurate and Fast Breast Cancer Segmentation in Whole-slide Images

2018-07-29 · Nanqing Dong, Michael Kampffmeyer, Xiaodan Liang, Zeya Wang, Wei Dai, Eric P. Xing

Convolutional neural networks have led to significant breakthroughs in the domain of medical image analysis. However, the task of breast cancer segmentation in whole-slide images (WSIs) is still underexplored. WSIs are large histopathological images with extremely high resolution. Constrained by the hardware and field of view, using high-magnification patches can slow down the inference process and using low-magnification patches can cause the loss of information. In this paper, we aim to achieve two seemingly conflicting goals for breast cancer segmentation: accurate and fast prediction. We propose a simple yet efficient framework Reinforced Auto-Zoom Net (RAZN) to tackle this task. Motivated by the zoom-in operation of a pathologist using a digital microscope, RAZN learns a policy network to decide whether zooming is required in a given region of interest. Because the zoom-in action is selective, RAZN is robust to unbalanced and noisy ground truth labels and can efficiently reduce overfitting. We evaluate our method on a public breast cancer dataset. RAZN outperforms both single-scale and multi-scale baseline approaches, achieving better accuracy at low inference cost.

📄 PDF Abstract BibTeX arXiv:1807.11113

Code (0)

등록된 구현이 없습니다.

Tasks

Medical Image Analysiswhole slide images

Similar Papers 제목 키워드 기반

Zoom in to where it matters: a hierarchical graph based model for mammogram analysis

2019-12-16 · Hao Du, Jiashi Feng, Mengling Feng

In clinical practice, human radiologists actually review medical images with high resolution monitors and zoom into region of interests (ROIs) for a close-up examination. Inspired by this observation, we propose a hierar…

ClassificationGeneral ClassificationGraph AttentionGraph Classification+2

Zoom-Zero: Reinforced Coarse-to-Fine Video Understanding via Temporal Zoom-in

2025-12-16 · Xiaoqian Shen, Min-Hung Chen, Yu-Chiang Frank Wang, Mohamed Elhoseiny 외 arxiv

Grounded video question answering (GVQA) aims to localize relevant temporal segments in videos and generate accurate answers to a given question; however, large video-language models (LVLMs) exhibit limited temporal awar…

Video Question AnsweringAnswer Generation

Look-Closer-Then-Diagnose: Confidence-Aware Ultrasound VQA via Active Zooming

2026-05-20 · Yue Zhou, Erxuan Wu, Yikang Sun, Hongjoo Lee 외 arxiv

Vision-Language Models (VLMs) have significantly advanced medical visual question answering, yet their performance in ultrasound remains suboptimal. In clinical practice, sonographers explicitly focus on lesion regions t…

Visual Question Answering

Automatic Calibration of a Multi-Camera System with Limited Overlapping Fields of View for 3D Surgical Scene Reconstruction

2025-01-27 · Tim Flückiger, Jonas Hein, Valery Fischer, Philipp Fürnstahl 외

The purpose of this study is to develop an automated and accurate external camera calibration method for multi-camera systems used in 3D surgical scene reconstruction (3D-SSR), eliminating the need for operator intervent…

Camera Calibration

A lightweight deep learning pipeline with DRDA-Net and MobileNet for breast cancer classification

2024-03-17 · Mahdie Ahmadi, Nader Karimi, Shadrokh Samavi

Accurate and early detection of breast cancer is essential for successful treatment. This paper introduces a novel deep-learning approach for improved breast cancer classification in histopathological images, a crucial s…

Cancer ClassificationComputational Efficiency