paper-with-me

Papers

Magnification Invariant Medical Image Analysis: A Comparison of Convolutional Networks, Vision Transformers, and Token Mixers

2023-02-22 · Pranav Jeevan, Nikhil Cherian Kurian, Amit Sethi

Convolution Neural Networks (CNNs) are widely used in medical image analysis, but their performance degrade when the magnification of testing images differ from the training images. The inability of CNNs to generalize across magnification scales can result in sub-optimal performance on external datasets. This study aims to evaluate the robustness of various deep learning architectures in the analysis of breast cancer histopathological images with varying magnification scales at training and testing stages. Here we explore and compare the performance of multiple deep learning architectures, including CNN-based ResNet and MobileNet, self-attention-based Vision Transformers and Swin Transformers, and token-mixing models, such as FNet, ConvMixer, MLP-Mixer, and WaveMix. The experiments are conducted using the BreakHis dataset, which contains breast cancer histopathological images at varying magnification levels. We show that performance of WaveMix is invariant to the magnification of training and testing data and can provide stable and good classification accuracy. These evaluations are critical in identifying deep learning architectures that can robustly handle changes in magnification scale, ensuring that scale changes across anatomical structures do not disturb the inference results.

📄 PDF Abstract BibTeX arXiv:2302.11488

Code (0)

등록된 구현이 없습니다.

Tasks

Breast Cancer Histology Image ClassificationDeep LearningImage ClassificationMedical Image Analysis

Methods 이 논문이 사용한 방법론

Batch Normalization 설명 없음
1x1 Convolution A 1 x 1 Convolution is a convolution with some special properties in that it can be used for dimensionality reduction,…
ReLU How Do I Communicate to Expedia? How Do I Communicate to Expedia? – Call ☎️ +1-(888) 829 (0881) or +1-805-330-4056 or +1-805-330-4056 for Live Support & Special Travel…
Convolution A convolution is a type of matrix operation, consisting of a kernel, a small matrix of weights, that slides over input data performing element-wise multiplication with the…
Kaiming Initialization 설명 없음
Max Pooling Max Pooling is a pooling operation that calculates the maximum value for patches of a feature map, and uses it to create a downsampled (pooled) feature map. It is usually…
Average Pooling 설명 없음
Residual Block Residual Blocks are skip-connection blocks that learn residual functions with reference to the layer inputs, instead of learning unreferenced functions. They were introduced…

Similar Papers 제목 키워드 기반

Unsupervised Magnification of Posture Deviations Across Subjects

2020-06-01 · CVPR 2020 6 · Michael Dorkenwald, Uta Buchler, Bjorn Ommer

Analyzing human posture and precisely comparing it across different subjects is essential for accurate understanding of behavior and numerous vision applications such as medical diagnostics, sports, or surveillance. Moti…

Motion Magnification

Magnification-Aware Distillation (MAD): A Self-Supervised Framework for Unified Representation Learning in Gigapixel Whole-Slide Images

2025-12-16 · Mahmut S. Gokmen, Mitchell A. Klusty, Peter T. Nelson, Allison M. Neltner 외 arxiv

Whole-slide images (WSIs) contain tissue information distributed across multiple magnification levels, yet most self-supervised methods treat these scales as independent views. This separation prevents models from learni…

Representation Learning

A Joint Spatial and Magnification Based Attention Framework for Large Scale Histopathology Classification

2021-06-19 · CVPR 2021 6 · Jingwei Zhang, Ke Ma, John Van Arnam, Rajarsi Gupta 외

Deep learning has achieved great success in process- ing large size medical images such as histopathology slides. However, conventional deep learning methods cannot han- dle the enormous image sizes; instead, they spl…

Deep Learning

Surgical Video Motion Magnification with Suppression of Instrument Artefacts

2020-09-16 · Mirek Janatka, Hani J. Marcus, Neil L. Dorward, Danail Stoyanov

Video motion magnification could directly highlight subsurface blood vessels in endoscopic video in order to prevent inadvertent damage and bleeding. Applying motion filters to the full surgical image is however sensitiv…

Motion MagnificationSSIM

STB-VMM: Swin Transformer Based Video Motion Magnification

2023-02-20 · Ricard Lado-Roigé, Marco A. Pérez

The goal of video motion magnification techniques is to magnify small motions in a video to reveal previously invisible or unseen movement. Its uses extend from bio-medical applications and deepfake detection to structur…

Motion Magnification