Identifications of Speaker Ethnicity in South-East England: Multicultural London English as a Divisible Perceptual Variety
This study uses crowdsourcing through LanguageARC to collect data on levels of accuracy in the identification of speakers{'} ethnicities. Ten participants (5 US; 5 South-East England) classified lexically identical speech stimuli from a corpus of 227 speakers aged 18-33yrs from South-East England into the main {``}ethnic{''} groups in Britain: White British, Black British and Asian British. Firstly, the data reveals that there is no significant geographic proximity effect on performance between US and British participants. Secondly, results contribute to recent work suggesting that despite the varying heritages of young, ethnic minority speakers in London, they speak an innovative and emerging variety: Multicultural London English (MLE) (e.g. Cheshire et al., 2011). Countering this, participants found perceptual linguistic differences between speakers of all 3 ethnicities (80.7{\%} accuracy). The highest rate of accuracy (96{\%}) was when identifying the ethnicity of Black British speakers from London whose speech seems to form a distinct, perceptual category. Participants also perform substantially better than chance at identifying Black British and Asian British speakers who are not from London (80{\%} and 60{\%} respectively). This suggests that MLE is not a single, homogeneous variety but instead, there are perceptual linguistic differences by ethnicity which transcend the borders of London.
Code (0)
등록된 구현이 없습니다.
Similar Papers 제목 키워드 기반
Crowdsourced Participants’ Accuracy at Identifying the Social Class of Speakers from South East England
Five participants, each located in distinct locations (USA, Canada, South Africa, Scotland and (South East) England), identified the self-determined social class of a corpus of 227 speakers (born 1986–2001; from South Ea…
Breast Cancer Data Analytics With Missing Values: A study on Ethnic, Age and Income Groups
An analysis of breast cancer incidences in women and the relationship between ethnicity and survival rate has been an ongoing study with recorded incidences of missing values in the secondary data. In this paper, we stud…
Missing ValuesOpen-source Multi-speaker Corpora of the English Accents in the British Isles
This paper presents a dataset of transcribed high-quality audio of English sentences recorded by volunteers speaking with different accents of the British Isles. The dataset is intended for linguistic analysis as well as…
Estimating the Potential Impact of Combined Race and Ethnicity Reporting on Long-Term Earnings Statistics
We use place of birth information from the Social Security Administration linked to earnings data from the Longitudinal Employer-Household Dynamics Program and detailed race and ethnicity data from the 2010 Census to stu…
SEADialogues: A Multilingual Culturally Grounded Multi-turn Dialogue Dataset on Southeast Asian Languages
Although numerous datasets have been developed to support dialogue systems, most existing chit-chat datasets overlook the cultural nuances inherent in natural human conversations. To address this gap, we introduce SEADia…