paper-with-me

홈 › Papers

Image Thresholding: Understanding Bias of Evaluation Metrics towards Specific Evaluation Functions

2026-05-26 · Eslam Hegazy, Mohamed Gabr arxiv

Multilevel image thresholding is widely used for segmentation in applications ranging from medical imaging to remote sensing. Classical objective functions, such as Otsu's between-class variance and Kapur's entropy, are often optimized using metaheuristic algorithms, with performance evaluated via metrics like Structural Similarity Index (SSIM) and Peak Signal-to-Noise Ratio (PSNR). These evaluations implicitly assume that SSIM and PSNR provide unbiased measures of segmentation quality. In this study, we examine this assumption by analyzing the correlation between thresholding objective functions and quality metrics across all possible thresholds for images in the BSDS500 dataset. Results show that Otsu's criterion consistently exhibits high correlation with both SSIM and PSNR, while Kapur's entropy demonstrates weaker and more variable correlation. Otsu outperforms Kapur in correlation with PSNR for all images and with SSIM for over 91%. Our findings reveal an inherent metric-objective-function bias. This work highlights the need for more neutral evaluation frameworks and motivates extending the analysis to additional thresholding criteria and domains. Source code of this paper can be found at https://w3id.org/met-dp/icpr26-95

📄 PDF Abstract BibTeX arXiv:2605.27132

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Quantitative and Qualitative Evaluation of NLM and Wavelet Methods in Image Enhancement

2024-09-22 · Cameron Khanpour

This paper presents a comprehensive analysis of image denoising techniques, primarily focusing on Non-local Means (NLM) and Daubechies Soft Wavelet Thresholding, and their efficacy across various datasets. These methods …

DenoisingImage DenoisingImage EnhancementImage Quality Assessment+1

Gender Biases in Automatic Evaluation Metrics for Image Captioning

2023-05-24 · Haoyi Qiu, Zi-Yi Dou, Tianlu Wang, Asli Celikyilmaz 외

Model-based evaluation metrics (e.g., CLIPScore and GPTScore) have demonstrated decent correlations with human judgments in various language generation tasks. However, their impact on fairness remains largely unexplored.…

FairnessImage CaptioningText Generation

A psychophysical evaluation of techniques for Mooney image generation

2024-03-18 · Lars C. Reining, Thomas S. A. Wallis

Mooney images can contribute to our understanding of the processes involved in visual perception, because they allow a dissociation between image content and image understanding. Mooney images are generated by first smoo…

Image Generation

Automated Workflow for the Detection of Vugs

2025-07-01 · M. Quamer Nasim, T. Maiti, N. Mosavat, P. V. Grech 외 arxiv

Image logs are crucial in capturing high-quality geological information about subsurface formations. Among the various geological features that can be gleaned from Formation Micro Imager log, vugs are essential for reser…

Enhancing Evaluation Methods for Infrared Small-Target Detection in Real-world Scenarios

2023-01-10 · Saed Moradi, Alireza Memarmoghadam, Payman Moallem, Mohamad Farzan Sabahi

Infrared small target detection (IRSTD) poses a significant challenge in the field of computer vision. While substantial efforts have been made over the past two decades to improve the detection capabilities of IRSTD alg…