paper-with-me

Papers

Some Notes on p(e)re-Reduplication in Bulgarian and Ukrainian: A Corpus-based Study

2022-09-01 · CLIB 2022 9 · Ivan Derzhanski, Olena Siruk

We present a comparative study of p(e)re-reduplication in Bulgarian and Ukrainian, based on material from a parallel corpus of bilingual texts. We analyse all occurrences found in the corpus of close sequences and conjunctions of two cognate words, the second of which features the intensive and recursive prefix pre- (Bulgarian) or pere- (Ukrainian). We find that in Bulgarian this construction occurs more frequently with finite verb forms, and in Ukrainian with participles and nouns. There is also a correlation with the mode of action denoted by the prefix: in its intensive meaning it turns up more often in Bulgarian, in its recursive meaning in the two languages equally, and in Ukrainian there are more occasions where it cannot be identified as either intensive or recursive. Finally, in both languages instances of p(e)re-reduplication are most common, by a wide marge, in texts with Ukrainian originals.

📄 PDF Abstract BibTeX

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

A Bilingual Lexicosemantic Network of Bread Based on a Parallel Corpus

2020-09-01 · CLIB 2020 9 · Ivan Derzhanski, Olena Siruk

We present an experiment in using a corpus of Bulgarian and Ukrainian parallel texts for the automatised construction of a bilingual lexicosemantic network representing the semantic field of BREAD. We discuss the extract…

Cross-lingual Named Entity Corpus for Slavic Languages

2024-03-30 · Jakub Piskorski, Michał Marcińczuk, Roman Yangarber

This paper presents a corpus manually annotated with named entities for six Slavic languages - Bulgarian, Czech, Polish, Slovenian, Russian, and Ukrainian. This work is the result of a series of shared tasks, conducted i…

LEMMALemmatization

MULTEXT-East

2020-03-31 · Tomaž Erjavec

MULTEXT-East language resources, a multilingual dataset for language engineering research, focused on the morphosyntactic level of linguistic description. The MULTEXT-East dataset includes the EAGLES-based morphosyntacti…

Sentence

The Political Speech Corpus of Bulgarian

2012-05-01 · LREC 2012 5 · Petya Osenova, Kiril Simov

The paper introduces the Political Speech Corpus of Bulgarian. First, its current state has been discussed with respect to its size, coverage, genre specification and related online services. Then, the focus goes to the …

LemmatizationMorphological AnalysisSentiment Analysis

Evidential strategies and grammatical marking in clauses governed by verba dicendi in Bulgarian

2022-09-01 · CLIB 2022 9 · Ekaterina Tarpomanova, Krasimira Aleksova

Thе study explores the interaction between the participants in the communication process with respect to their knowledge about the situation presented in the utterance when transforming direct into indirect speech using …