paper-with-me

Papers

Creating Training Sets via Weak Indirect Supervision

2021-10-07 · ICLR 2022 4 · Jieyu Zhang, Bohan Wang, Xiangchen Song, Yujing Wang, Yaming Yang, Jing Bai, Alexander Ratner

Creating labeled training sets has become one of the major roadblocks in machine learning. To address this, recent \emph{Weak Supervision (WS)} frameworks synthesize training labels from multiple potentially noisy supervision sources. However, existing frameworks are restricted to supervision sources that share the same output space as the target task. To extend the scope of usable sources, we formulate Weak Indirect Supervision (WIS), a new research problem for automatically synthesizing training labels based on indirect supervision sources that have different output label spaces. To overcome the challenge of mismatched output spaces, we develop a probabilistic modeling approach, PLRM, which uses user-provided label relations to model and leverage indirect supervision sources. Moreover, we provide a theoretically-principled test of the distinguishability of PLRM for unseen labels, along with a generalization bound. On both image and text classification tasks as well as an industrial advertising application, we demonstrate the advantages of PLRM by outperforming baselines by a margin of 2%-9%.

📄 PDF Abstract BibTeX arXiv:2110.03484

Code (0)

등록된 구현이 없습니다.

Tasks

text-classificationText Classification

Methods 이 논문이 사용한 방법론

Test 설명 없음

Similar Papers 제목 키워드 기반

Learning from Indirect Observations

2019-10-10 · Yivan Zhang, Nontawat Charoenphakdee, Masashi Sugiyama

Weakly-supervised learning is a paradigm for alleviating the scarcity of labeled data by leveraging lower-quality but larger-scale supervision signals. While existing work mainly focuses on utilizing a certain type of we…

Weakly-supervised Learning

Weak Supervision for Real World Graphs

2025-06-03 · Pratheeksha Nair, Reihaneh Rabbany

Node classification in real world graphs often suffers from label scarcity and noise, especially in high stakes domains like human trafficking detection and misinformation monitoring. While direct supervision is limited,…

Contrastive LearningMisinformationNode ClassificationRepresentation Learning

PABI: A Unified PAC-Bayesian Informativeness Measure for Incidental Supervision Signals

2021-01-01 · Hangfeng He, Mingyuan Zhang, Qiang Ning, Dan Roth

Real-world applications often require making use of {\em a range of incidental supervision signals}. However, we currently lack a principled way to measure the benefit an incidental training dataset can bring, and the co…

Informativenessnamed-entity-recognitionNamed Entity RecognitionNamed Entity Recognition (NER)+2

Weak Supervision for Improved Precision in Search Systems

2025-03-10 · Sriram Vasudevan

Labeled datasets are essential for modern search engines, which increasingly rely on supervised learning methods like Learning to Rank and massive amounts of data to power deep learning models. However, creating these da…

Learning-To-Rank

Detecting Online Hate Speech: Approaches Using Weak Supervision and Network Embedding Models

2020-07-24 · Michael Ridenhour, Arunkumar Bagavathi, Elaheh Raisi, Siddharth Krishnan

The ubiquity of social media has transformed online interactions among individuals. Despite positive effects, it has also allowed anti-social elements to unite in alternative social media environments (eg. Gab.com) like …

Network Embedding