paper-with-me

홈 › Papers

NorQuAD: Norwegian Question Answering Dataset

2023-05-03 · Sardana Ivanova, Fredrik Aas Andreassen, Matias Jentoft, Sondre Wold, Lilja Øvrelid

In this paper we present NorQuAD: the first Norwegian question answering dataset for machine reading comprehension. The dataset consists of 4,752 manually created question-answer pairs. We here detail the data collection procedure and present statistics of the dataset. We also benchmark several multilingual and Norwegian monolingual language models on the dataset and compare them against human performance. The dataset will be made freely available.

📄 PDF Abstract BibTeX arXiv:2305.01957

Code (1)

ltgoslo/norquad 공식 구현 pytorch

Tasks

Machine Reading ComprehensionQuestion AnsweringReading Comprehension

Similar Papers 제목 키워드 기반

A Collection of Question Answering Datasets for Norwegian

2025-01-19 · Vladislav Mikhailov, Petter Mæhlum, Victoria Ovedie Chruickshank Langø, Erik Velldal 외

This paper introduces a new suite of question answering datasets for Norwegian; NorOpenBookQA, NorCommonSenseQA, NorTruthfulQA, and NRK-Quiz-QA. The data covers a wide range of skills and knowledge domains, including wor…

Question AnsweringWorld Knowledge

BRENT: Bidirectional Retrieval Enhanced Norwegian Transformer

2023-04-19 · Lucas Georges Gabriel Charpentier, Sondre Wold, David Samuel, Egil Rønningstad

Retrieval-based language models are increasingly employed in question-answering tasks. These models search in a corpus of documents for relevant information instead of having all factual knowledge stored in its parameter…

Dependency ParsingExtractive Question-AnsweringLanguage ModelingLanguage Modelling+6

If I Could Turn Back Time: Temporal Reframing as a Historical Reasoning Task for LLMs

2025-11-06 · Lars Bungum, Charles Yijia Huang, Abeer Kashar arxiv

In this study, we experiment with the ability of LLMs to do temporal reasoning. Using a Norwegian book from 1940 containing trivia questions, we prompt the LLMs to answer the questions as if it were 1940. We also pose th…

Benchmarking Abstractive Summarisation: A Dataset of Human-authored Summaries of Norwegian News Articles

2025-01-13 · Samia Touileb, Vladislav Mikhailov, Marie Kroka, Lilja Øvrelid 외

We introduce a dataset of high-quality human-authored summaries of news articles in Norwegian. The dataset is intended for benchmarking the abstractive summarisation capabilities of generative language models. Each docum…

ArticlesBenchmarking

Whispering in Norwegian: Navigating Orthographic and Dialectic Challenges

2024-02-02 · Per E Kummervold, Javier de la Rosa, Freddy Wetjen, Rolv-Arild Braaten 외

This article introduces NB-Whisper, an adaptation of OpenAI's Whisper, specifically fine-tuned for Norwegian language Automatic Speech Recognition (ASR). We highlight its key contributions and summarise the results achie…

Automatic Speech RecognitionAutomatic Speech Recognition (ASR)speech-recognitionSpeech Recognition