LexStat: Automatic Detection of Cognates in Multilingual Wordlists
Code (1)
Similar Papers 제목 키워드 기반
Using Sequence Similarity Networks to Identify Partial Cognates in Multilingual Wordlists
Fast and unsupervised methods for multilingual cognate clustering
In this paper we explore the use of unsupervised methods for detecting cognates in multilingual word lists. We use online EM to train sound segment similarity weights for computing similarity between two words. We tested…
ClusteringUsing support vector machines and state-of-the-art algorithms for phonetic alignment to identify cognates in multi-lingual wordlists
Most current approaches in phylogenetic linguistics require as input multilingual word lists partitioned into sets of etymologically related words (cognates). Cognate identification is so far done manually by experts, wh…
Self-Supervised Borrowing Detection on Multilingual Wordlists
This paper presents a fully self-supervised approach to borrowing detection in multilingual wordlists. The method combines two sources of information: PMI similarities based on a global correspondence model and a lightwe…
Automated Cognate Detection as a Supervised Link Prediction Task with Cognate Transformer
Identification of cognates across related languages is one of the primary problems in historical linguistics. Automated cognate identification is helpful for several downstream tasks including identifying sound correspon…
Link Prediction