Extract, Integrate, Compete: Towards Verification Style Reading Comprehension
In this paper, we present a new verification style reading comprehension dataset named VGaokao from Chinese Language tests of Gaokao. Different from existing efforts, the new dataset is originally designed for native speakers' evaluation, thus requiring more advanced language understanding skills. To address the challenges in VGaokao, we propose a novel Extract-Integrate-Compete approach, which iteratively selects complementary evidence with a novel query updating mechanism and adaptively distills supportive evidence, followed by a pairwise competition to push models to learn the subtle difference among similar text pieces. Experiments show that our methods outperform various baselines on VGaokao with retrieved complementary evidence, while having the merits of efficiency and explainability. Our dataset and code are released for further research.
Code (1)
Tasks
Reading ComprehensionSimilar Papers 제목 키워드 기반
KRISTEVA: Close Reading as a Novel Task for Benchmarking Interpretive Reasoning
Each year, tens of millions of essays are written and graded in college-level English courses. Students are asked to analyze literary and cultural texts through a process known as close reading, in which they gather text…
BenchmarkingMMLUMultiple-choiceStyle-News: Incorporating Stylized News Generation and Adversarial Verification for Neural Fake News Detection
With the improvements in generative models, the issues of producing hallucinations in various domains (e.g., law, writing) have been brought to people's attention due to concerns about misinformation. In this paper, we f…
Fake News DetectionMisinformationNews GenerationTowards Cross-speaker Reading Style Transfer on Audiobook Dataset
Cross-speaker style transfer aims to extract the speech style of the given reference speech, which can be reproduced in the timbre of arbitrary target speakers. Existing methods on this topic have explored utilizing utte…
Style TransferTowards Comprehensive Stage-wise Benchmarking of Large Language Models in Fact-Checking
Large Language Models (LLMs) are increasingly deployed in real-world fact-checking systems, yet existing evaluations focus predominantly on claim verification and overlook the broader fact-checking workflow, including cl…
A Deep Content-Based Model for Persian Rumor Verification
During the development of social media, there has been a transformation in social communication. Despite their positive applications in social interactions and news spread, it also provides an ideal platform for spreadin…
Rumour DetectionWord Embeddings