paper-with-me

홈 › Papers

Toward the Evaluation of Written Proficiency on a Collaborative Social Network for Learning Languages: Yask

2019-03-23 · Fabio N. Silva, Sergio Jimenez, George Dueñas

Yask is an online social collaborative network for practicing languages in a framework that includes requests, answers, and votes. Since measuring linguistic competence using current approaches is difficult, expensive and in many cases imprecise, we present a new alternative approach based on social networks. Our method, called Proficiency Rank, extends the well-known Page Rank algorithm to measure the reputation of users in a collaborative social graph. First, we extended Page Rank so that it not only considers positive links (votes) but also negative links. Second, in addition to using explicit links, we also incorporate other 4 types of signals implicit in the social graph. These extensions allow Proficiency Rank to produce proficiency rankings for almost all users in the data set used, where only a minority contributes by answering, while the majority contributes only by voting. This overcomes the intrinsic limitation of Page Rank of only being able to rank the nodes that have incoming links. Our experimental validation showed that the reputation/importance of the users in Yask is significantly correlated with their language proficiency. In contrast, their written production was poorly correlated with the vocabulary profiles of the Common European Framework of Reference. In addition, we found that negative signals (votes) are considerably more informative than positive ones. We concluded that the use of this technology is a promising tool for measuring second language proficiency, even for relatively small groups of people.

📄 PDF Abstract BibTeX arXiv:1903.09846

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Challenges and Considerations with Code-Mixed NLP for Multilingual Societies

2021-06-15 · Vivek Srivastava, Mayank Singh

Multilingualism refers to the high degree of proficiency in two or more languages in the written and oral communication modes. It often results in language mixing, a.k.a. code-mixing, when a multilingual speaker switches…

ManagementMultilingual NLP

Coursebook Texts as a Helping Hand for Classifying Linguistic Complexity in Language Learners' Writings

2016-12-01 · WS 2016 12 · Ildik{\'o} Pil{\'a}n, David Alfter, Elena Volodina

We bring together knowledge from two different types of language learning data, texts learners read and texts they write, to improve linguistic complexity classification in the latter. Linguistic complexity in the foreig…

ClassificationDomain AdaptationGeneral Classification

The MERLIN corpus: Learner language and the CEFR

2014-05-01 · LREC 2014 5 · Adriane Boyd, Jirka Hana, Lionel Nicolas, Detmar Meurers 외

The MERLIN corpus is a written learner corpus for Czech, German,and Italian that has been designed to illustrate the Common European Framework of Reference for Languages (CEFR) with authentic learner data. The corpus con…

Language AcquisitionLanguage IdentificationNative Language Identification

Modeling Proficiency with Implicit User Representations

2021-10-15 · Kim Breitwieser, Allison Lahnala, Charles Welch, Lucie Flek 외

We introduce the problem of proficiency modeling: Given a user's posts on a social media platform, the task is to identify the subset of posts or topics for which the user has some level of proficiency. This enables the …

WMT24++: Expanding the Language Coverage of WMT24 to 55 Languages & Dialects

2025-02-18 · Daniel Deutsch, Eleftheria Briakou, Isaac Caswell, Mara Finkelstein 외

As large language models (LLM) become more and more capable in languages other than English, it is important to collect benchmark datasets in order to evaluate their multilingual performance, including on tasks like mach…

Machine Translation