paper-with-me

홈 › Papers

Annotating Cognates and Etymological Origin in Turkic Languages

2015-01-13 · Benjamin S. Mericli, Michael Bloodgood

Turkic languages exhibit extensive and diverse etymological relationships among lexical items. These relationships make the Turkic languages promising for exploring automated translation lexicon induction by leveraging cognate and other etymological information. However, due to the extent and diversity of the types of relationships between words, it is not clear how to annotate such information. In this paper, we present a methodology for annotating cognates and etymological origin in Turkic languages. Our method strives to balance the amount of research effort the annotator expends with the utility of the annotations for supporting research on improving automated translation lexicon induction.

📄 PDF Abstract BibTeX arXiv:1501.03191

Code (0)

등록된 구현이 없습니다.

Tasks

DiversityTranslation

Similar Papers 제목 키워드 기반

Caveats of Measuring Semantic Change of Cognates and Borrowings using Multilingual Word Embeddings

2022-05-01 · LChange (ACL) 2022 5 · Clémentine Fourrier, Syrielle Montariol

Cognates and borrowings carry different aspects of etymological evolution. In this work, we study semantic change of such items using multilingual word embeddings, both static and contextualised. We underline caveats ide…

Multilingual Word EmbeddingsWord Embeddings

Automatic Identification and Production of Related Words for Historical Linguistics

2019-12-01 · CL 2019 12 · Alina Maria Ciobanu, Liviu P. Dinu

Language change across space and time is one of the main concerns in historical linguistics. In this article, we develop tools to assist researchers and domain experts in the study of language evolution.First, we introdu…

Etymological Wordnet: Tracing The History of Words

2014-05-01 · LREC 2014 5 · Gerard de Melo

Research on the history of words has led to remarkable insights about language and also about the history of human civilization more generally. This paper presents the Etymological Wordnet, the first database that aims a…

Using support vector machines and state-of-the-art algorithms for phonetic alignment to identify cognates in multi-lingual wordlists

2017-04-01 · EACL 2017 4 · Gerhard J{\"a}ger, Johann-Mattis List, Pavel Sofroniev

Most current approaches in phylogenetic linguistics require as input multilingual word lists partitioned into sets of etymologically related words (cognates). Cognate identification is so far done manually by experts, wh…

Building a Dataset of Multilingual Cognates for the Romanian Lexicon

2014-05-01 · LREC 2014 5 · Liviu Dinu, Alina Maria Ciobanu

Identifying cognates is an interesting task with applications in numerous research areas, such as historical and comparative linguistics, language acquisition, cross-lingual information retrieval, readability and machine…

Cross-Lingual Information RetrievalInformation RetrievalLanguage AcquisitionMachine Translation+3