paper-with-me

홈 › Papers

Modeling Human Annotation Errors to Design Bias-Aware Systems for Social Stream Processing

2019-07-16 · Rahul Pandey, Carlos Castillo, Hemant Purohit

High-quality human annotations are necessary to create effective machine learning systems for social media. Low-quality human annotations indirectly contribute to the creation of inaccurate or biased learning systems. We show that human annotation quality is dependent on the ordering of instances shown to annotators (referred as 'annotation schedule'), and can be improved by local changes in the instance ordering provided to the annotators, yielding a more accurate annotation of the data stream for efficient real-time social media analytics. We propose an error-mitigating active learning algorithm that is robust with respect to some cases of human errors when deciding an annotation schedule. We validate the human error model and evaluate the proposed algorithm against strong baselines by experimenting on classification tasks of relevant social media posts during crises. According to these experiments, considering the order in which data instances are presented to human annotators leads to both an increase in accuracy for machine learning and awareness toward some potential biases in human learning that may affect the automated classifier.

📄 PDF Abstract BibTeX arXiv:1907.07228

Code (0)

등록된 구현이 없습니다.

Tasks

Active LearningBIG-bench Machine Learning

Similar Papers 제목 키워드 기반

Modeling Annotator Preference and Stochastic Annotation Error for Medical Image Segmentation

2021-11-26 · Zehui Liao, Shishuai Hu, Yutong Xie, Yong Xia

Manual annotation of medical images is highly subjective, leading to inevitable and huge annotation biases. Deep learning models may surpass human performance on a variety of tasks, but they may also mimic or amplify the…

Image SegmentationMedical Image SegmentationSegmentationSemantic Segmentation

Surrogate Representation Inference for Text and Image Annotations

2025-09-15 · Kentaro Nakamura arxiv

As researchers increasingly rely on machine learning models and LLMs to annotate unstructured data, such as texts or images, various approaches have been proposed to correct bias in downstream statistical analysis. Howev…

CausalRM: Causal-Theoretic Reward Modeling for RLHF from Observational User Feedbacks

2026-03-19 · Hao Wang, Licheng Pan, Zhichao Chen, Chunyuan Zheng 외 arxiv

Despite the success of reinforcement learning from human feedback (RLHF) in aligning language models, current reward modeling heavily relies on experimental feedback data collected from human annotators under controlled …

Reinforcement Learning

Evaluating and Correcting Human Annotation Bias in Dynamic Micro-Expression Recognition

2026-03-05 · Feng Liu, Bingyu Nan, Xuezhong Qian, Xiaolan Fu arxiv

Existing manual labeling of micro-expressions is subject to errors in accuracy, especially in cross-cultural scenarios where deviation in labeling of key frames is more prominent. To address this issue, this paper presen…

Micro-Expression Recognition

BiasLab: Toward Explainable Political Bias Detection with Dual-Axis Annotations and Rationale Indicators

2025-05-21 · KMA Solaiman

We present BiasLab, a dataset of 300 political news articles annotated for perceived ideological bias. These articles were selected from a curated 900-document pool covering diverse political events and source biases. Ea…

ArticlesBias Detection