paper-with-me

Papers

Comparing Methods for Measuring Dialect Similarity in Norwegian

2020-05-01 · LREC 2020 5 · Janne Johannessen, Andre K{\aa}sen, Kristin Hagen, Anders N{\o}klestad, Joel Priestley

The present article presents four experiments with two different methods for measuring dialect similarity in Norwegian: the Levenshtein method and the neural long short term memory (LSTM) autoencoder network, a machine learning algorithm. The visual output in the form of dialect maps is then compared with canonical maps found in the dialect literature. All of this enables us to say that one does not need fine-grained transcriptions of speech to replicate classical classification patterns.

📄 PDF Abstract BibTeX

Code (0)

등록된 구현이 없습니다.

Tasks

BIG-bench Machine Learning

Methods 이 논문이 사용한 방법론

Solana Customer Service Number +1-833-534-1729 설명 없음

Similar Papers 제목 키워드 기반

The Norwegian Dialect Corpus Treebank

2022-06-01 · LREC 2022 6 · Andre Kåsen, Kristin Hagen, Anders Nøklestad, Joel Priestly 외

This paper presents the NDC Treebank of spoken Norwegian dialects in the Bokmål variety of Norwegian. It consists of dialect recordings made between 2006 and 2012 which have been digitised, segmented, transcribed and sub…

Similarities between Arabic Dialects: Investigating Geographical Proximity

2021-05-10 · Abdulkareem Alsudais, Wafa Alotaibi, Faye Alomary

The automatic classification of Arabic dialects is an ongoing research challenge, which has been explored in recent work that defines dialects based on increasingly limited geographic areas like cities and provinces. Thi…

Add Noise, Tasks, or Layers? MaiNLP at the VarDial 2025 Shared Task on Norwegian Dialectal Slot and Intent Detection

2025-01-07 · Verena Blaschke, Felicia Körner, Barbara Plank

Slot and intent detection (SID) is a classic natural language understanding task. Despite this, research has only more recently begun focusing on SID for dialectal and colloquial varieties. Many approaches for low-resour…

Intent DetectionNatural Language Understanding

NorDial: A Preliminary Corpus of Written Norwegian Dialect Use

2021-04-11 · NoDaLiDa 2021 5 · Jeremy Barnes, Petter Mæhlum, Samia Touileb

Norway has a large amount of dialectal variation, as well as a general tolerance to its use in the public sphere. There are, however, few available resources to study this variation and its change over time and in more i…

The Norwegian Parliamentary Speech Corpus

2022-01-26 · LREC 2022 6 · Per Erik Solberg, Pablo Ortiz

The Norwegian Parliamentary Speech Corpus (NPSC) is a speech dataset with recordings of meetings from Stortinget, the Norwegian parliament. It is the first, publicly available dataset containing unscripted, Norwegian spe…

Automatic Speech RecognitionAutomatic Speech Recognition (ASR)speech-recognitionSpeech Recognition