Light Verb Constructions in the SzegedParalellFX English--Hungarian Parallel Corpus
In this paper, we describe the first English-Hungarian parallel corpus annotated for light verb constructions, which contains 14,261 sentence alignment units. Annotation principles and statistical data on the corpus are also provided, and English and Hungarian data are contrasted. On the basis of corpus data, a database containing pairs of English-Hungarian light verb constructions has been created as well. The corpus and the database can contribute to the automatic detection of light verb constructions and it is also shown how they can enhance performance in several fields of NLP (e.g. parsing, information extraction/retrieval and machine translation).
Code (0)
등록된 구현이 없습니다.
Tasks
Machine TranslationRetrievalSentenceTranslationSimilar Papers 제목 키워드 기반
Identifying English and Hungarian Light Verb Constructions: A Contrastive Approach
4FX: Light Verb Constructions in a Multilingual Parallel Corpus
In this paper, we describe 4FX, a quadrilingual (English-Spanish-German-Hungarian) parallel corpus annotated for light verb constructions. We present the annotation process, and report statistical data on the frequency o…
Machine TranslationDependency Parsing for Identifying Hungarian Light Verb Constructions
Automatic Error Detection concerning the Definite and Indefinite Conjugation in the HunLearner Corpus
In this paper we present the results of automatic error detection, concerning the definite and indefinite conjugation in the extended version of the HunLearner corpus, the learners corpus of the Hungarian language. We p…