paper-with-me

홈 › Papers

On The Direct Maximization of Quadratic Weighted Kappa

2015-09-23 · David Vaughn, Derek Justice

In recent years, quadratic weighted kappa has been growing in popularity in the machine learning community as an evaluation metric in domains where the target labels to be predicted are drawn from integer ratings, usually obtained from human experts. For example, it was the metric of choice in several recent, high profile machine learning contests hosted on Kaggle : https://www.kaggle.com/c/asap-aes , https://www.kaggle.com/c/asap-sas , https://www.kaggle.com/c/diabetic-retinopathy-detection . Yet, little is understood about the nature of this metric, its underlying mathematical properties, where it fits among other common evaluation metrics such as mean squared error (MSE) and correlation, or if it can be optimized analytically, and if so, how. Much of this is due to the cumbersome way that this metric is commonly defined. In this paper we first derive an equivalent but much simpler, and more useful, definition for quadratic weighted kappa, and then employ this alternate form to address the above issues.

📄 PDF Abstract BibTeX arXiv:1509.07107

Code (0)

등록된 구현이 없습니다.

Tasks

BIG-bench Machine LearningDiabetic Retinopathy Detection

Similar Papers 제목 키워드 기반

Deep learning-based algorithm for assessment of knee osteoarthritis severity in radiographs matches performance of radiologists

2022-07-25 · Albert Swiecicki, Nianyi Li, Jonathan O'Donnell, Nicholas Said 외

A fully-automated deep learning algorithm matched performance of radiologists in assessment of knee osteoarthritis severity in radiographs using the Kellgren-Lawrence grading system. To develop an automated deep learning…

“So You Think You’re Funny?”: Rating the Humour Quotient in Standup Comedy

2021-11-01 · EMNLP 2021 11 · Anirudh Mittal, Pranav Jeevan P, Prerak Gandhi, Diptesh Kanojia 외

Computational Humour (CH) has attracted the interest of Natural Language Processing and Computational Linguistics communities. Creating datasets for automatic measurement of humour quotient is difficult due to multiple p…

"So You Think You're Funny?": Rating the Humour Quotient in Standup Comedy

2021-10-25 · Anirudh Mittal, Pranav Jeevan, Prerak Gandhi, Diptesh Kanojia 외

Computational Humour (CH) has attracted the interest of Natural Language Processing and Computational Linguistics communities. Creating datasets for automatic measurement of humour quotient is difficult due to multiple p…

SALMA: Arabic Sense-Annotated Corpus and WSD Benchmarks

2023-10-29 · Mustafa Jarrar, Sanad Malaysha, Tymaa Hammouda, Mohammed Khalilia

SALMA, the first Arabic sense-annotated corpus, consists of ~34K tokens, which are all sense-annotated. The corpus is annotated using two different sense inventories simultaneously (Modern and Ghani). SALMA novelty lies …

Word Sense Disambiguation

Deep Learning-Based Grading of Ductal Carcinoma In Situ in Breast Histopathology Images

2020-10-07 · Suzanne C. Wetstein, Nikolas Stathonikos, Josien P. W. Pluim, Yujing J. Heng 외

Ductal carcinoma in situ (DCIS) is a non-invasive breast cancer that can progress into invasive ductal carcinoma (IDC). Studies suggest DCIS is often overtreated since a considerable part of DCIS lesions may never progre…

Deep Learning