paper-with-me

홈 › Papers

Different Strokes for Different Folks: Investigating Appropriate Further Pre-training Approaches for Diverse Dialogue Tasks

2021-09-14 · EMNLP 2021 11 · Yao Qiu, Jinchao Zhang, Jie zhou

Loading models pre-trained on the large-scale corpus in the general domain and fine-tuning them on specific downstream tasks is gradually becoming a paradigm in Natural Language Processing. Previous investigations prove that introducing a further pre-training phase between pre-training and fine-tuning phases to adapt the model on the domain-specific unlabeled data can bring positive effects. However, most of these further pre-training works just keep running the conventional pre-training task, e.g., masked language model, which can be regarded as the domain adaptation to bridge the data distribution gap. After observing diverse downstream tasks, we suggest that different tasks may also need a further pre-training phase with appropriate training tasks to bridge the task formulation gap. To investigate this, we carry out a study for improving multiple task-oriented dialogue downstream tasks through designing various tasks at the further pre-training phase. The experiment shows that different downstream tasks prefer different further pre-training tasks, which have intrinsic correlation and most further pre-training tasks significantly improve certain target tasks rather than all. Our investigation indicates that it is of great importance and effectiveness to design appropriate further pre-training tasks modeling specific information that benefit downstream tasks. Besides, we present multiple constructive empirical conclusions for enhancing task-oriented dialogues.

📄 PDF Abstract BibTeX arXiv:2109.06524

Code (0)

등록된 구현이 없습니다.

Tasks

Domain AdaptationLanguage Modelling

Similar Papers 제목 키워드 기반

The Wisdom of the Few? "Supertaggers" in Collaborative Tagging Systems

2015-03-17 · Lorince Jared, Zorowitz Sam, Murdock Jaimie, Todd Peter M.

A folksonomy is ostensibly an information structure built up by the "wisdom of the crowd", but is the "crowd" really doing the work? Tagging is in fact a sharply skewed process in which a small minority of "supertagger" …

TAG

Emergent Behaviors from Folksonomy Driven Interactions

2019-12-31 · Massimiliano Dal Mas

To reflect the evolving knowledge on the Web this paper considers ontologies based on folksonomies according to a new concept structure called "Folksodriven" to represent folksonomies. This paper describes a research pro…

Automatic Data Deformation Analysis on Evolving Folksonomy Driven Environment

2016-12-30 · Massimiliano Dal Mas

The Folksodriven framework makes it possible for data scientists to define an ontology environment where searching for buried patterns that have some kind of predictive power to build predictive models more effectively. …

Kategorisasi dokumen web secara otomatis berdasarkan folksonomy menggunakan multinomial naive Bayes classifier

2016-06-24 · Irawan Hendy

Folksonomy is a non-hierarchical document categorizing system, that treats every category in a flat manner, dan every category is entered freely by anyone who submitted a document in these categories. Categorization is d…

TAG

Distributed Vector Representations of Folksong Motifs

2019-03-20 · Aitor Arronte-Alvarez, Francisco Gómez-Martin

This article presents a distributed vector representation model for learning folksong motifs. A skip-gram version of word2vec with negative sampling is used to represent high quality embeddings. Motifs from the Essen Fol…