Automatic cognate identification with gap-weighted string subsequences.
Code (0)
등록된 구현이 없습니다.
Tasks
Information RetrievalRetrievalWord SimilaritySimilar Papers 제목 키워드 기반
Gap-weighted subsequences for automatic cognate identification and phylogenetic inference
In this paper, we describe the problem of cognate identification and its relation to phylogenetic inference. We introduce subsequence based features for discriminating cognates from non-cognates. We show that subsequence…
Siamese convolutional networks based on phonetic features for cognate identification
In this paper, we explore the use of convolutional networks (ConvNets) for the purpose of cognate identification. We compare our architecture with binary classifiers based on string similarity measures on different langu…
Automatic Identification and Production of Related Words for Historical Linguistics
Language change across space and time is one of the main concerns in historical linguistics. In this article, we develop tools to assist researchers and domain experts in the study of language evolution.First, we introdu…
A Classification-Based Approach to Cognate Detection Combining Orthographic and Semantic Similarity Information
This paper presents proof-of-concept experiments for combining orthographic and semantic information to distinguish cognates from non-cognates. To this end, a context-independent gold standard is developed by manually la…
Binary ClassificationFormGeneral ClassificationSemantic Similarity+2Tracking Semantic Change in Cognate Sets for English and Romance Languages
Semantic divergence in related languages is a key concern of historical linguistics. We cross-linguistically investigate the semantic divergence of cognate pairs in English and Romance languages, by means of word embeddi…
Word Embeddings