paper-with-me

홈 › Papers

Establishing Annotation Quality in Multi-label Annotations

2022-10-01 · COLING 2022 10 · Marian Marchal, Merel Scholman, Frances Yung, Vera Demberg

In many linguistic fields requiring annotated data, multiple interpretations of a single item are possible. Multi-label annotations more accurately reflect this possibility. However, allowing for multi-label annotations also affects the chance that two coders agree with each other. Calculating inter-coder agreement for multi-label datasets is therefore not trivial. In the current contribution, we evaluate different metrics for calculating agreement on multi-label annotations: agreement on the intersection of annotated labels, an augmented version of Cohen’s Kappa, and precision, recall and F1. We propose a bootstrapping method to obtain chance agreement for each measure, which allows us to obtain an adjusted agreement coefficient that is more interpretable. We demonstrate how various measures affect estimates of agreement on simulated datasets and present a case study of discourse relation annotations. We also show how the proportion of double labels, and the entropy of the label distribution, influences the measures outlined above and how a bootstrapped adjusted agreement can make agreement measures more comparable across datasets in multi-label scenarios.

📄 PDF Abstract BibTeX

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Robust and Label-Efficient Deep Waste Detection

2025-08-26 · Hassan Abid, Khan Muhammad, Muhammad Haris Khan arxiv

Effective waste sorting is critical for sustainable recycling, yet AI research in this domain continues to lag behind commercial systems due to limited datasets and reliance on legacy object detectors. In this work, we a…

Object Detection

Dynamic Supervisor for Cross-dataset Object Detection

2022-04-01 · Ze Chen, Zhihang Fu, Jianqiang Huang, Mingyuan Tao 외

The application of cross-dataset training in object detection tasks is complicated because the inconsistency in the category range across datasets transforms fully supervised learning into semi-supervised learning. To ad…

Objectobject-detectionObject Detection

Enhancing chest X-ray datasets with privacy-preserving large language models and multi-type annotations: a data-driven approach for improved classification

2024-03-06 · Ricardo Bigolin Lanfredi, Pritam Mukherjee, Ronald Summers

In chest X-ray (CXR) image analysis, rule-based systems are usually employed to extract labels from reports for dataset releases. However, there is still room for improvement in label quality. These labelers typically ou…

Language ModelingLanguage ModellingLarge Language ModelPrivacy Preserving

Efficient Online Crowdsourcing with Complex Annotations

2024-01-25 · Reshef Meir, Viet-An Nguyen, Xu Chen, Jagdish Ramakrishnan 외

Crowdsourcing platforms use various truth discovery algorithms to aggregate annotations from multiple labelers. In an online setting, however, the main challenge is to decide whether to ask for more annotations for each …

Quality Sentinel: Estimating Label Quality and Errors in Medical Segmentation Datasets

2024-06-01 · Yixiong Chen, Zongwei Zhou, Alan Yuille

An increasing number of public datasets have shown a transformative impact on automated medical segmentation. However, these datasets are often with varying label quality, ranging from manual expert annotations to AI-gen…