Word Familiarity Rate Estimation Using a Bayesian Linear Mixed Model
This paper presents research on word familiarity rate estimation using the {}Word List by Semantic Principles{'}. We collected rating information on 96,557 words in the {}Word List by Semantic Principles{'} via Yahoo! crowdsourcing. We asked 3,392 subject participants to use their introspection to rate the familiarity of words based on the five perspectives of {}KNOW{'}, {}WRITE{'}, {}READ{'}, {}SPEAK{'}, and {}LISTEN{'}, and each word was rated by at least 16 subject participants. We used Bayesian linear mixed models to estimate the word familiarity rates. We also explored the ratings with the semantic labels used in the {}Word List by Semantic Principles{'}.
Code (0)
등록된 구현이 없습니다.
Similar Papers 제목 키워드 기반
Word Familiarity and Frequency
Word frequency is assumed to correlate with word familiarity, but the strength of this correlation has not been thoroughly investigated. In this paper, we report on our analysis of the correlation between a word familiar…
A Practical Introduction to Bayesian Estimation of Causal Effects: Parametric and Nonparametric Approaches
Substantial advances in Bayesian methods for causal inference have been developed in recent years. We provide an introduction to Bayesian inference for causal effects for practicing statisticians who have some familiarit…
Bayesian InferenceCausal InferenceStructure leads and dominates comprehension in naturalistic reading
The hierarchical account and statistical or sequential account have long been framed as rival theories in explaining online comprehension. A lot of evidence has shown that both hierarchical and non-hierarchical factors c…
What makes a word hard to learn? Modeling L1 influence on English vocabulary difficulty
What makes a word difficult to learn, and how does the difficulty depend on the learner's native language? We computationally model vocabulary difficulty for English learners whose first language is Spanish, German, or C…
Beyond Film Subtitles: Is YouTube the Best Approximation of Spoken Vocabulary?
Word frequency is a key variable in psycholinguistics, useful for modeling human familiarity with words even in the era of large language models (LLMs). Frequency in film subtitles has proved to be a particularly good ap…
Lexical Complexity PredictionWord Embeddings