Rethinking Phonotactic Complexity
In this work, we propose the use of phone-level language models to estimate phonotactic complexity{---}measured in bits per phoneme{---}which makes cross-linguistic comparison straightforward. We compare the entropy across languages using this simple measure, gaining insight on how complex different language{'}s phonotactics are. Finally, we show a very strong negative correlation between phonotactic complexity and the average length of words{---}Spearman rho=-0.744{---}when analysing a collection of 106 languages with 1016 basic concepts each.
Code (0)
등록된 구현이 없습니다.
Similar Papers 제목 키워드 기반
Phonotactic Complexity across Dialects
Received wisdom in linguistic typology holds that if the structure of a language becomes more complex in one dimension, it will simplify in another, building on the assumption that all languages are equally complex (Jose…
Language ModelingLanguage ModellingCorrelation Does Not Imply Compensation: Complexity and Irregularity in the Lexicon
It has been claimed that within a language, morphologically irregular words are more likely to be phonotactically simple and morphologically regular words are more likely to be phonotactically complex. This inverse corre…
Phonotactic Complexity and its Trade-offs
We present methods for calculating a measure of phonotactic complexity---bits per phoneme---that permits a straightforward cross-linguistic comparison. When given a word, represented as a sequence of phonemic segments su…
Learning nonlocal phonotactics in Strictly Piecewise phonotactic model
Using LSTMs to Assess the Obligatoriness of Phonological Distinctive Features for Phonotactic Learning
To ascertain the importance of phonetic information in the form of phonological distinctive features for the purpose of segment-level phonotactic acquisition, we compare the performance of two recurrent neural network mo…