paper-with-me

홈 › Papers

FewshotQA: A simple framework for few-shot learning of question answering tasks using pre-trained text-to-text models

2021-09-04 · EMNLP 2021 11 · Rakesh Chada, Pradeep Natarajan

The task of learning from only a few examples (called a few-shot setting) is of key importance and relevance to a real-world setting. For question answering (QA), the current state-of-the-art pre-trained models typically need fine-tuning on tens of thousands of examples to obtain good results. Their performance degrades significantly in a few-shot setting (< 100 examples). To address this, we propose a simple fine-tuning framework that leverages pre-trained text-to-text models and is directly aligned with their pre-training framework. Specifically, we construct the input as a concatenation of the question, a mask token representing the answer span and a context. Given this input, the model is fine-tuned using the same objective as that of its pre-training objective. Through experimental studies on various few-shot configurations, we show that this formulation leads to significant gains on multiple QA benchmarks (an absolute gain of 34.2 F1 points on average when there are only 16 training examples). The gains extend further when used with larger models (Eg:- 72.3 F1 on SQuAD using BART-large with only 32 examples) and translate well to a multilingual setting . On the multilingual TydiQA benchmark, our model outperforms the XLM-Roberta-large by an absolute margin of upto 40 F1 points and an average of 33 F1 points in a few-shot setting (<= 64 training examples). We conduct detailed ablation studies to analyze factors contributing to these gains.

📄 PDF Abstract BibTeX arXiv:2109.01951

Code (1)

naver-ai/simseek pytorch

Tasks

Few-Shot LearningQuestion Answering

Similar Papers 제목 키워드 기반

Towards Zero-Shot and Few-Shot Table Question Answering using GPT-3

2022-10-31 · Pragya Srivastava, Tanuja Ganu, Saikat Guha

We present very early results on using GPT-3 to perform question answering on tabular data. We find that stock pre-trained GPT-3 is able to zero-shot learn the table structure from a serialized JSON array-of-arrays repre…

Prompt EngineeringQuestion Answering

Momentum Contrastive Pre-training for Question Answering

2022-12-12 · Minda Hu, Muzhi Li, Yasheng Wang, Irwin King

Existing pre-training methods for extractive Question Answering (QA) generate cloze-like queries different from natural questions in syntax structure, which could overfit pre-trained models to simple keyword matching. In…

BenchmarkingContrastive LearningExtractive Question-AnsweringNatural Questions+1

HPE:Answering Complex Questions over Text by Hybrid Question Parsing and Execution

2023-05-12 · Ye Liu, Semih Yavuz, Rui Meng, Dragomir Radev 외

The dominant paradigm of textual question answering systems is based on end-to-end neural networks, which excels at answering natural language questions but falls short on complex ones. This stands in contrast to the bro…

Knowledge GraphsQuestion AnsweringSemantic Parsing

Question-Instructed Visual Descriptions for Zero-Shot Video Question Answering

2024-02-16 · David Romero, Thamar Solorio

We present Q-ViD, a simple approach for video question answering (video QA), that unlike prior methods, which are based on complex architectures, computationally expensive pipelines or use closed models like GPTs, Q-ViD …

Language ModelingLanguage ModellingLarge Language ModelMultiple-choice+3

Answer-Me: Multi-Task Open-Vocabulary Visual Question Answering

2022-05-02 · AJ Piergiovanni, Wei Li, Weicheng Kuo, Mohammad Saffar 외

We present Answer-Me, a task-aware multi-task framework which unifies a variety of question answering tasks, such as, visual question answering, visual entailment, visual reasoning. In contrast to previous works using co…

DecoderImage CaptioningQuestion AnsweringVisual Entailment+4