paper-with-me

SCROLLS

Standardized CompaRison Over Long Language Sequences

홈페이지 · 논문 42편

SCROLLS (Standardized CompaRison Over Long Language Sequences) is an NLP benchmark consisting of a suite of tasks that require reasoning over long texts. SCROLLS contains summarization, question answering, and natural language inference tasks, covering multiple domains, including literature, science, business, and entertainment. The dataset is made available in a unified text-to-text format and host a live leaderboard to facilitate research on model architecture and pretraining methods. The SCROLLS benchmark contains the datasets [GovReport](govreport), SummScreenFD, [QMSum](qmsum), [QASPER](qasper), [NarrativeQA](NarrativeQA), QuALITY and ContractNLI.

Texts English

벤치마크

Long-range modeling on SCROLLS 결과 13개