paper-with-me

Papers

A Corpus of Literal and Idiomatic Uses of German Infinitive-Verb Compounds

2016-05-01 · LREC 2016 5 · Andrea Horbach, Andrea Hensler, Sabine Krome, Jakob Prange, Werner Scholze-Stubenrecht, Diana Steffen, Stefan Thater, Christian Wellner, Manfred Pinkal

We present an annotation study on a representative dataset of literal and idiomatic uses of German infinitive-verb compounds in newspaper and journal texts. Infinitive-verb compounds form a challenge for writers of German, because spelling regulations are different for literal and idiomatic uses. Through the participation of expert lexicographers we were able to obtain a high-quality corpus resource which offers itself as a testbed for automatic idiomaticity detection and coarse-grained word-sense disambiguation. We trained a classifier on the corpus which was able to distinguish literal and idiomatic uses with an accuracy of 85 {\%}.

📄 PDF Abstract BibTeX

Code (0)

등록된 구현이 없습니다.

Tasks

Word Sense Disambiguation

Similar Papers 제목 키워드 기반

One Size Fits All? A simple LSTM for non-literal token and construction-level classification

2018-08-01 · COLING 2018 8 · Erik-L{\^a}n Do Dinh, Steffen Eger, Iryna Gurevych

In this paper, we tackle four different tasks of non-literal language classification: token and construction level metaphor detection, classification of idiomatic use of infinitive-verb compounds, and classification of n…

AllClassificationGeneral ClassificationMulti-Task Learning

EPIE Dataset: A Corpus For Possible Idiomatic Expressions

2020-06-16 · Prateek Saxena, Soma Paul

Idiomatic expressions have always been a bottleneck for language comprehension and natural language understanding, specifically for tasks like Machine Translation(MT). MT systems predominantly produce literal translation…

Machine TranslationNatural Language Understanding

Literal or idiomatic? Identifying the reading of single occurrences of German multiword expressions using word embeddings

2017-04-01 · EACL 2017 4 · Rafael Ehren

Non-compositional multiword expressions (MWEs) still pose serious issues for a variety of natural language processing tasks and their ubiquity makes it impossible to get around methods which automatically identify these …

Machine TranslationSemantic SimilaritySemantic Textual SimilarityWord Embeddings

Casting a Wide Net: Robust Extraction of Potentially Idiomatic Expressions

2019-11-20 · Hessel Haagsma, Malvina Nissim, Johan Bos

Idiomatic expressions like `out of the woods' and `up the ante' present a range of difficulties for natural language processing applications. We present work on the annotation and extraction of what we term potentially i…

Understanding Idiomatic Variation

2017-04-01 · WS 2017 4 · Kristina Geeraert, R. Harald Baayen, John Newman

This study investigates the processing of idiomatic variants through an eye-tracking experiment. Four types of idiom variants were included, in addition to the canonical form and the literal meaning. Results suggest that…

Semantic Textual Similarity