GLUE
General Language Understanding Evaluation benchmark
홈페이지 · 논문 3,197편
General Language Understanding Evaluation (GLUE) benchmark is a collection of nine natural language understanding tasks, including single-sentence tasks CoLA and SST-2, similarity and paraphrasing tasks MRPC, STS-B and QQP, and natural language inference tasks MNLI, QNLI, RTE and WNLI. Source: Align, Mask and Select: A Simple Method for Incorporating Commonsense Knowledge into Language Representation Models Image Source: https://gluebenchmark.com/
Texts English벤치마크
Classification on SST-2
결과 6개
Classification on RTE
결과 2개