paper-with-me

홈 › Papers

A Pitfall of Learning from User-generated Data: In-depth Analysis of Subjective Class Problem

2020-03-24 · Kei Nemoto, Shweta Jain

Research in the supervised learning algorithms field implicitly assumes that training data is labeled by domain experts or at least semi-professional labelers accessible through crowdsourcing services like Amazon Mechanical Turk. With the advent of the Internet, data has become abundant and a large number of machine learning based systems started being trained with user-generated data, using categorical data as true labels. However, little work has been done in the area of supervised learning with user-defined labels where users are not necessarily experts and might be motivated to provide incorrect labels in order to improve their own utility from the system. In this article, we propose two types of classes in user-defined labels: subjective class and objective class - showing that the objective classes are as reliable as if they were provided by domain experts, whereas the subjective classes are subject to bias and manipulation by the user. We define this as a subjective class issue and provide a framework for detecting subjective labels in a dataset without querying oracle. Using this framework, data mining practitioners can detect a subjective class at an early stage of their projects, and avoid wasting their precious time and resources by dealing with subjective class problem with traditional machine learning techniques.

📄 PDF Abstract BibTeX arXiv:2003.10621

Code (1)

box-key/Subjective-Class-Issue 공식 구현

Tasks

BIG-bench Machine Learning

Similar Papers 제목 키워드 기반

Not All Explanations are Created Equal: Investigating the Pitfalls of Current XAI Evaluation

2025-09-27 · Joe Shymanski, Jacob Brue, Sandip Sen arxiv

Explainable Artificial Intelligence (XAI) aims to create transparency in modern AI models by offering explanations of the models to human users. There are many ways in which researchers have attempted to evaluate the qua…

Single-GPU GNN Systems: Traps and Pitfalls

2024-02-05 · Yidong Gong, Arnab Tarafder, Saima Afrin, Pradeep Kumar

The current graph neural network (GNN) systems have established a clear trend of not showing training accuracy results, and directly or indirectly relying on smaller datasets for evaluations majorly. Our in-depth analysi…

GPUGraph Neural Network

Generative Explore-Exploit: Training-free Optimization of Generative Recommender Systems using LLM Optimizers

2024-06-07 · Lütfi Kerem Senel, Besnik Fetahu, Davis Yoshida, Zhiyu Chen 외

Recommender systems are widely used to suggest engaging content, and Large Language Models (LLMs) have given rise to generative recommenders. Such systems can directly generate items, including for open-set tasks like qu…

General KnowledgeQuestion GenerationQuestion-GenerationRecommendation Systems+1

Can Users Detect Biases or Factual Errors in Generated Responses in Conversational Information-Seeking?

2024-10-28 · Weronika Łajewska, Krisztian Balog, Damiano Spina, Johanne Trippas

Information-seeking dialogues span a wide range of questions, from simple factoid to complex queries that require exploring multiple facets and viewpoints. When performing exploratory searches in unfamiliar domains, user…

DiversityMisinformationResponse Generation

Human-AI Interactions and Societal Pitfalls

2023-09-19 · Francisco Castro, Jian Gao, Sébastien Martin

When working with generative artificial intelligence (AI), users may see productivity gains, but the AI-generated content may not match their preferences exactly. To study this effect, we introduce a Bayesian framework i…