paper-with-me

홈 › Papers

ArchivalQA: A Large-scale Benchmark Dataset for Open Domain Question Answering over Archival News Collections

2021-11-16 · ACL ARR November 2021 11 · Anonymous

In the last few years, open-domain question answering (ODQA) has advanced rapidly due to the development of deep learning techniques and the availability of large-scale QA datasets. However, the current datasets are essentially designed for synchronic document collections (e.g., Wikipedia). Temporal news collections such as long-term news archives spanning several decades, are rarely used in training the models despite they are quite valuable for our society. To foster the research in the field of ODQA on such historical collections, we present ArchivalQA, a large question answering dataset consisting of 532,444 question-answer pairs which is designed for temporal news QA. We divide our dataset into four subparts based on the question difficulty levels and the containment of temporal expressions, which we believe are useful for training and testing ODQA systems characterized by different strengths and abilities. The novel QA dataset-constructing framework that we introduce can be also applied to create datasets over other types of collections.

📄 PDF Abstract BibTeX

Code (0)

등록된 구현이 없습니다.

Tasks

Open-Domain Question AnsweringQuestion Answering

Similar Papers 제목 키워드 기반

ArchivalQA: A Large-scale Benchmark Dataset for Open Domain Question Answering over Historical News Collections

2021-09-08 · Jiexin Wang, Adam Jatowt, Masatoshi Yoshikawa

In the last few years, open-domain question answering (ODQA) has advanced rapidly due to the development of deep learning techniques and the availability of large-scale QA datasets. However, the current datasets are esse…

Open-Domain Question AnsweringQuestion Answering

TempRetriever: Fusion-based Temporal Dense Passage Retrieval for Time-Sensitive Questions

2025-02-28 · Abdelrahman Abdallah, Bhawna Piryani, Jonas Wallat, Avishek Anand 외

Temporal awareness is crucial in many information retrieval tasks, particularly in scenarios where the relevance of documents depends on their alignment with the query's temporal context. Traditional approaches such as B…

Information RetrievalPassage RetrievalQuestion AnsweringRetrieval+2

ASRank: Zero-Shot Re-Ranking with Answer Scent for Document Retrieval

2025-01-25 · Abdelrahman Abdallah, Jamshid Mozafari, Bhawna Piryani, Adam Jatowt

Retrieval-Augmented Generation (RAG) models have drawn considerable attention in modern open-domain question answering. The effectiveness of RAG depends on the quality of the top retrieved documents. However, conventiona…

Language ModelingLanguage ModellingLarge Language ModelOpen-Domain Question Answering+6

CMT: A Memory Compression Method for Continual Knowledge Learning of Large Language Models

2024-12-10 · Dongfang Li, Zetian Sun, Xinshuo Hu, Baotian Hu 외

Large Language Models (LLMs) need to adapt to the continuous changes in data, tasks, and user preferences. Due to their massive size and the high costs associated with training, LLMs are not suitable for frequent retrain…

Continual Learning

OVT-B: A New Large-Scale Benchmark for Open-Vocabulary Multi-Object Tracking

2024-10-23 · Haiji Liang, Ruize Han

Open-vocabulary object perception has become an important topic in artificial intelligence, which aims to identify objects with novel classes that have not been seen during training. Under this setting, open-vocabulary o…

Multi-Object TrackingObjectobject-detectionObject Detection+3