Only Connect Walls Dataset Task 1 (Grouping)
1개 벤치마크 · 논문 10편 · 이 태스크의 논문 보기 →
Benchmarks
OCW
Most implemented
RoBERTa: A Robustly Optimized BERT Pretraining Approach
Deep contextualized word representations
DistilBERT, a distilled version of BERT: smaller, faster, cheaper and lighter
GPT-4 Technical Report
MPNet: Masked and Permuted Pre-training for Language Understanding
GloVe: Global Vectors for Word Representation
Papers
Large Language Models are Fixated by Red Herrings: Exploring Creative Problem Solving and Einstellung Effect using the Only Connect Wall Dataset
The quest for human imitative AI has been an enduring topic in AI research since its inception. The technical evolution and emerging capabilities of the latest cohort of large language models (LLMs) have reinvigorated th…
Only Connect Walls Dataset Task 1 (Grouping)Only Connect Walls Dataset Task 2 (Connections)GPT-4 Technical Report
We report the development of GPT-4, a large-scale, multimodal model which can accept image and text inputs and produce text outputs. While less capable than humans in many real-world scenarios, GPT-4 exhibits human-level…
answerability predictionArithmetic ReasoningBug fixingCode Generation+18Text Embeddings by Weakly-Supervised Contrastive Pre-training
This paper presents E5, a family of state-of-the-art text embeddings that transfer well to a wide range of tasks. The model is trained in a contrastive manner with weak supervision signals from our curated large-scale te…
MTEB BenchmarkOnly Connect Walls Dataset Task 1 (Grouping)RetrievalMPNet: Masked and Permuted Pre-training for Language Understanding
BERT adopts masked language modeling (MLM) for pre-training and is one of the most successful pre-training models. Since BERT neglects dependency among predicted tokens, XLNet introduces permuted language modeling (PLM) …
Language ModelingLanguage ModellingMasked Language ModelingOnly Connect Walls Dataset Task 1 (Grouping)+2Pre-Training of Deep Bidirectional Protein Sequence Representations with Structural Information
Bridging the exponentially growing gap between the numbers of unlabeled and labeled protein sequences, several studies adopted semi-supervised learning for protein sequence modeling. In these studies, models were pre-tra…
Language ModelingLanguage ModellingMasked Language ModelingOnly Connect Walls Dataset Task 1 (Grouping)DistilBERT, a distilled version of BERT: smaller, faster, cheaper and lighter
As Transfer Learning from large-scale pre-trained models becomes more prevalent in Natural Language Processing (NLP), operating these large models in on-the-edge and/or under constrained computational training or inferen…
Hate Speech DetectionKnowledge DistillationLanguage ModelingLanguage Modelling+7