Applying Multi-Sense Embeddings for German Verbs to Determine Semantic Relatedness and to Detect Non-Literal Language
Up to date, the majority of computational models still determines the semantic relatedness between words (or larger linguistic units) on the type level. In this paper, we compare and extend multi-sense embeddings, in order to model and utilise word senses on the token level. We focus on the challenging class of complex verbs, and evaluate the model variants on various semantic tasks: semantic classification; predicting compositionality; and detecting non-literal language usage. While there is no overall best model, all models significantly outperform a word2vec single-sense skip baseline, thus demonstrating the need to distinguish between word senses in a distributional semantic model.
Code (0)
등록된 구현이 없습니다.
Tasks
Semantic Textual SimilarityWord EmbeddingsWord Sense DisambiguationSimilar Papers 제목 키워드 기반
Graph-based Clustering of Synonym Senses for German Particle Verbs
Factoring Ambiguity out of the Prediction of Compositionality for German Multi-Word Expressions
Ambiguity represents an obstacle for distributional semantic models(DSMs), which typically subsume the contexts of all word senses within one vector. While individual vector space approaches have been concerned with sens…
ClusteringMachine TranslationCross-lingual Visual Verb Sense Disambiguation
Recent work has shown that visual context improves cross-lingual sense disambiguation for nouns. We extend this line of work to the more challenging task of cross-lingual verb sense disambiguation, introducing the MultiS…
Machine TranslationTranslationUnsupervised Visual Sense Disambiguation for Verbs using Multimodal Embeddings
We introduce a new task, visual sense disambiguation for verbs: given an image and a verb, assign the correct sense of the verb, i.e., the one that describes the action depicted in the image. Just as textual word sense d…
Image DescriptionImage RetrievalRetrievalWord Sense DisambiguationLexical Substitution Dataset for German
This article describes a lexical substitution dataset for German. The whole dataset contains 2,040 sentences from the German Wikipedia, with one target word in each sentence. There are 51 target nouns, 51 adjectives, and…
LEMMAQuestion AnsweringSentenceText Simplification+1