paper-with-me

홈 › Papers

WebRED: Effective Pretraining And Finetuning For Relation Extraction On The Web

2021-02-18 · Robert Ormandi, Mohammad Saleh, Erin Winter, Vinay Rao

Relation extraction is used to populate knowledge bases that are important to many applications. Prior datasets used to train relation extraction models either suffer from noisy labels due to distant supervision, are limited to certain domains or are too small to train high-capacity models. This constrains downstream applications of relation extraction. We therefore introduce: WebRED (Web Relation Extraction Dataset), a strongly-supervised human annotated dataset for extracting relationships from a variety of text found on the World Wide Web, consisting of ~110K examples. We also describe the methods we used to collect ~200M examples as pre-training data for this task. We show that combining pre-training on a large weakly supervised dataset with fine-tuning on a small strongly-supervised dataset leads to better relation extraction performance. We provide baselines for this new dataset and present a case for the importance of human annotation in improving the performance of relation extraction from text found on the web.

📄 PDF Abstract BibTeX arXiv:2102.09681

Code (1)

shimorina/relation-extraction-db-wikidata

Tasks

RelationRelation Extraction

Similar Papers 제목 키워드 기반

Continual Contrastive Finetuning Improves Low-Resource Relation Extraction

2022-12-21 · Wenxuan Zhou, Sheng Zhang, Tristan Naumann, Muhao Chen 외

Relation extraction (RE), which has relied on structurally annotated corpora for model training, has been particularly challenging in low-resource scenarios and domains. Recent literature has tackled low-resource RE by s…

Contrastive LearningRelationRelation ExtractionRepresentation Learning+1

MiniConGTS: A Near Ultimate Minimalist Contrastive Grid Tagging Scheme for Aspect Sentiment Triplet Extraction

2024-06-17 · Qiao Sun, Liujia Yang, Minghao Ma, Nanyang Ye 외

Aspect Sentiment Triplet Extraction (ASTE) aims to co-extract the sentiment triplets in a given corpus. Existing approaches within the pretraining-finetuning paradigm tend to either meticulously craft complex tagging sch…

Aspect Sentiment Triplet ExtractionTriplet

Exploring Transferability for Randomized Smoothing

2023-12-14 · Kai Qiu, Huishuai Zhang, Zhirong Wu, Stephen Lin

Training foundation models on extensive datasets and then finetuning them on specific tasks has emerged as the mainstream approach in artificial intelligence. However, the model robustness, which is a critical aspect for…

Understanding Finetuning for Factual Knowledge Extraction

2024-06-20 · Gaurav Ghosal, Tatsunori Hashimoto, aditi raghunathan

In this work, we study the impact of QA fine-tuning data on downstream factuality. We show that fine-tuning on lesser-known facts that are poorly stored during pretraining yields significantly worse factuality than fine-…

MMLUQuestion Answering

AutoPrompt: Eliciting Knowledge from Language Models with Automatically Generated Prompts

2020-10-29 · EMNLP 2020 11 · Taylor Shin, Yasaman Razeghi, Robert L. Logan IV, Eric Wallace 외

The remarkable success of pretrained language models has motivated the study of what kinds of knowledge these models learn during pretraining. Reformulating tasks as fill-in-the-blanks problems (e.g., cloze tests) is a n…

Natural Language InferenceRelationRelation ExtractionSentiment Analysis