paper-with-me

홈 › Papers

Towards Professional Level Crowd Annotation of Expert Domain Data

2023-01-01 · CVPR 2023 1 · Pei Wang, Nuno Vasconcelos

Image recognition on expert domains is usually fine-grained and requires expert labeling, which is costly. This limits dataset sizes and the accuracy of learning systems. To address this challenge, we consider annotating expert data with crowdsourcing. This is denoted as PrOfeSsional lEvel cRowd (POSER) annotation. A new approach, based on semi-supervised learning (SSL) and denoted as SSL with human filtering (SSL-HF) is proposed. It is a human-in-the-loop SSL method, where crowd-source workers act as filters of pseudo-labels, replacing the unreliable confidence thresholding used by state-of-the-art SSL methods. To enable annotation by non-experts, classes are specified implicitly, via positive and negative sets of examples and augmented with deliberative explanations, which highlight regions of class ambiguity. In this way, SSL-HF leverages the strong low-shot learning and confidence estimation ability of humans to create an intuitive but effective labeling experience. Experiments show that SSL-HF significantly outperforms various alternative approaches in several benchmarks.

📄 PDF Abstract BibTeX

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Stigma Annotation Scheme and Stigmatized Language Detection in Health-Care Discussions on Social Media

2020-05-01 · LREC 2020 5 · Nadiya Straton, Hyeju Jang, Raymond Ng

Much research has been done within the social sciences on the interpretation and influence of stigma on human behaviour and health, which result in out-of-group exclusion, distancing, cognitive separation, status loss, d…

Data Augmentation

Crowdsourcing Learning as Domain Adaptation: A Case Study on Named Entity Recognition

2021-05-31 · ACL 2021 5 · Xin Zhang, Guangwei Xu, Yueheng Sun, Meishan Zhang 외

Crowdsourcing is regarded as one prospective solution for effective supervised learning, aiming to build large-scale annotated training data by crowd workers. Previous studies focus on reducing the influences from the no…

Domain Adaptationnamed-entity-recognitionNamed Entity RecognitionNamed Entity Recognition (NER)+2

Exposing ambiguities in a relation-extraction gold standard with crowdsourcing

2015-05-23 · Tong Shu Li, Benjamin M. Good, Andrew I. Su

Semantic relation extraction is one of the frontiers of biomedical natural language processing research. Gold standards are key tools for advancing this research. It is challenging to generate these standards because of …

RelationRelation Extraction

DEXA: Supporting Non-Expert Annotators with Dynamic Examples from Experts

2020-05-17 · Markus Zlabinger, Marta Sabou, Sebastian Hofstätter, Mete Sertkan 외

The success of crowdsourcing based annotation of text corpora depends on ensuring that crowdworkers are sufficiently well-trained to perform the annotation task accurately. To that end, a frequent approach to train annot…

AvgSentenceSentence Similarity

Towards Extracting Medical Family History from Natural Language Interactions: A New Dataset and Baselines

2019-11-01 · IJCNLP 2019 11 · Mahmoud Azab, Stephane Dadian, Vivi Nastase, Larry An 외

We introduce a new dataset consisting of natural language interactions annotated with medical family histories, obtained during interactions with a genetic counselor and through crowdsourcing, following a questionnaire c…

Relation Extraction