paper-with-me

Papers

Negation in Norwegian: an annotated dataset

2021-05-01 · NoDaLiDa 2021 5 · Petter Mæhlum, Jeremy Barnes, Robin Kurtz, Lilja Øvrelid, Erik Velldal

This paper introduces NorecNeg – the first annotated dataset of negation for Norwegian. Negation cues and their in-sentence scopes have been annotated across more than 11K sentences spanning more than 400 documents for a subset of the Norwegian Review Corpus (NoReC). In addition to providing in-depth discussion of the annotation guidelines, we also present a first set of benchmark results based on a graph-parsing approach.

📄 PDF Abstract BibTeX

Code (1)

ltgoslo/norec_neg 공식 구현 pytorch

Tasks

NegationSentence

Similar Papers 제목 키워드 기반

NARC – Norwegian Anaphora Resolution Corpus

2022-10-01 · COLING (CRAC) 2022 10 · Petter Mæhlum, Dag Haug, Tollef Jørgensen, Andre Kåsen 외

We present the Norwegian Anaphora Resolution Corpus (NARC), the first publicly available corpus annotated with anaphoric relations between noun phrases for Norwegian. The paper describes the annotated data for 326 docume…

Relation

The Norwegian Dialect Corpus Treebank

2022-06-01 · LREC 2022 6 · Andre Kåsen, Kristin Hagen, Anders Nøklestad, Joel Priestly 외

This paper presents the NDC Treebank of spoken Norwegian dialects in the Bokmål variety of Norwegian. It consists of dialect recordings made between 2006 and 2012 which have been digitised, segmented, transcribed and sub…

NorDiaChange: Diachronic Semantic Change Dataset for Norwegian

2022-01-13 · LREC 2022 6 · Andrey Kutuzov, Samia Touileb, Petter Mæhlum, Tita Ranveig Enstad 외

We describe NorDiaChange: the first diachronic semantic change dataset for Norwegian. NorDiaChange comprises two novel subsets, covering about 80 Norwegian nouns manually annotated with graded semantic change over time. …

The Norwegian Parliamentary Speech Corpus

2022-01-26 · LREC 2022 6 · Per Erik Solberg, Pablo Ortiz

The Norwegian Parliamentary Speech Corpus (NPSC) is a speech dataset with recordings of meetings from Stortinget, the Norwegian parliament. It is the first, publicly available dataset containing unscripted, Norwegian spe…

Automatic Speech RecognitionAutomatic Speech Recognition (ASR)speech-recognitionSpeech Recognition

Annotating Norwegian Language Varieties on Twitter for Part-of-Speech

2022-10-12 · VarDial (COLING) 2022 10 · Petter Mæhlum, Andre Kåsen, Samia Touileb, Jeremy Barnes

Norwegian Twitter data poses an interesting challenge for Natural Language Processing (NLP) tasks. These texts are difficult for models trained on standardized text in one of the two Norwegian written forms (Bokm{\aa}l a…

POS