Strategies and Challenges for Crowdsourcing Regional Dialect Perception Data for Swiss German and Swiss French
Code (0)
등록된 구현이 없습니다.
Similar Papers 제목 키워드 기반
Vashantor: A Large-scale Multilingual Benchmark Dataset for Automated Translation of Bangla Regional Dialects to Bangla Language
The Bangla linguistic variety is a fascinating mix of regional dialects that adds to the cultural diversity of the Bangla-speaking community. Despite extensive study into translating Bangla to English, English to Bangla,…
Machine TranslationTranslationCrowdsourcing Dialect Characterization through Twitter
We perform a large-scale analysis of language diatopic variation using geotagged microblogging datasets. By collecting all Twitter messages written in Spanish over more than two years, we build a corpus from which a care…
Large Language Models Discriminate Against Speakers of German Dialects
Dialects represent a significant component of human culture and are found across all regions of the world. In Germany, more than 40% of the population speaks a regional dialect (Adler and Hansen, 2022). However, despite …
Decision MakingBanglaDialecto: An End-to-End AI-Powered Regional Speech Standardization
This study focuses on recognizing Bangladeshi dialects and converting diverse Bengali accents into standardized formal Bengali speech. Dialects, often referred to as regional languages, are distinctive variations of a la…
Machine Translationspeech-recognitionSpeech RecognitionIdiom Understanding as a Tool to Measure the Dialect Gap
The tasks of idiom understanding and dialect understanding are both well-established benchmarks in natural language processing. In this paper, we propose combining them, and using regional idioms as a test of dialect und…