Supervised Fine Tuning for Word Embedding with Integrated Knowledge
Learning vector representation for words is an important research field which may benefit many natural language processing tasks. Two limitations exist in nearly all available models, which are the bias caused by the context definition and the lack of knowledge utilization. They are difficult to tackle because these algorithms are essentially unsupervised learning approaches. Inspired by deep learning, the authors propose a supervised framework for learning vector representation of words to provide additional supervised fine tuning after unsupervised learning. The framework is knowledge rich approacher and compatible with any numerical vectors word representation. The authors perform both intrinsic evaluation like attributional and relational similarity prediction and extrinsic evaluations like the sentence completion and sentiment analysis. Experiments results on 6 embeddings and 4 tasks with 10 datasets show that the proposed fine tuning framework may significantly improve the quality of the vector representation of words.
Code (0)
등록된 구현이 없습니다.
Tasks
SentenceSentence CompletionSentiment AnalysisSimilar Papers 제목 키워드 기반
Unsupervised Sentence Representation Learning with Frequency-induced Adversarial Tuning and Incomplete Sentence Filtering
Pre-trained Language Model (PLM) is nowadays the mainstay of Unsupervised Sentence Representation Learning (USRL). However, PLMs are sensitive to the frequency information of words from their pre-training corpora, result…
Language ModellingRepresentation LearningSentenceSentence EmbeddingsDelta Embedding Learning
Unsupervised word embeddings have become a popular approach of word representation in NLP tasks. However there are limitations to the semantics represented by unsupervised embeddings, and inadequate fine-tuning of embedd…
Reading ComprehensionWord EmbeddingsDictionary-Assisted Supervised Contrastive Learning
Text analysis in the social sciences often involves using specialized dictionaries to reason with abstract concepts, such as perceptions about the economy or abuse on social media. These dictionaries allow researchers to…
Contrastive LearningData AugmentationFew-Shot LearningMoRTy: Unsupervised Learning of Task-specialized Word Embeddings by Autoencoding
Word embeddings have undoubtedly revolutionized NLP. However, pretrained embeddings do not always work for a specific task (or set of tasks), particularly in limited resource setups. We introduce a simple yet effective, …
Word EmbeddingsUsing Optimal Transport as Alignment Objective for fine-tuning Multilingual Contextualized Embeddings
Recent studies have proposed different methods to improve multilingual word representations in contextualized settings including techniques that align between source and target embedding spaces. For contextualized embedd…
Cross-Lingual TransferWord Alignment