paper-with-me

홈 › Papers

More Bang for Your Buck: Natural Perturbation for Robust Question Answering

2020-04-09 · EMNLP 2020 11 · Daniel Khashabi, Tushar Khot, Ashish Sabharwal

While recent models have achieved human-level scores on many NLP datasets, we observe that they are considerably sensitive to small changes in input. As an alternative to the standard approach of addressing this issue by constructing training sets of completely new examples, we propose doing so via minimal perturbation of examples. Specifically, our approach involves first collecting a set of seed examples and then applying human-driven natural perturbations (as opposed to rule-based machine perturbations), which often change the gold label as well. Local perturbations have the advantage of being relatively easier (and hence cheaper) to create than writing out completely new examples. To evaluate the impact of this phenomenon, we consider a recent question-answering dataset (BoolQ) and study the benefit of our approach as a function of the perturbation cost ratio, the relative cost of perturbing an existing question vs. creating a new one from scratch. We find that when natural perturbations are moderately cheaper to create, it is more effective to train models using them: such models exhibit higher robustness and better generalization, while retaining performance on the original BoolQ dataset.

📄 PDF Abstract BibTeX arXiv:2004.04849

Code (0)

등록된 구현이 없습니다.

Tasks

Question Answering

Similar Papers 제목 키워드 기반

Perturbing Attention Gives You More Bang for the Buck: Subtle Imaging Perturbations That Efficiently Fool Customized Diffusion Models

2024-04-23 · CVPR 2024 1 · Jingyao Xu, Yuetong Lu, Yandong Li, Siyang Lu 외

Diffusion models (DMs) embark a new era of generative modeling and offer more opportunities for efficient generating high-quality and realistic data samples. However, their widespread use has also brought forth new chall…

Bang for the Buck: Vector Search on Cloud CPUs

2025-05-12 · Leonardo Kuffo, Peter Boncz

Vector databases have emerged as a new type of systems that support efficient querying of high-dimensional vectors. Many of these offer their database as a service in the cloud. However, the variety of available CPUs and…

CPUQuantization

Explaining Classification Models Built on High-Dimensional Sparse Data

2016-07-21 · Julie Moeyersoms, Brian d'Alessandro, Foster Provost, David Martens

Predictive modeling applications increasingly use data representing people's behavior, opinions, and interactions. Fine-grained behavior data often has different structure from traditional data, being very high-dimension…

ClassificationGeneral ClassificationVocal Bursts Intensity Prediction

BanglaNLG and BanglaT5: Benchmarks and Resources for Evaluating Low-Resource Natural Language Generation in Bangla

2022-05-23 · Abhik Bhattacharjee, Tahmid Hasan, Wasi Uddin Ahmad, Rifat Shahriyar

This work presents BanglaNLG, a comprehensive benchmark for evaluating natural language generation (NLG) models in Bangla, a widely spoken yet low-resource language. We aggregate six challenging conditional text generati…

Conditional Text GenerationDialogue GenerationLanguage ModelingLanguage Modelling+1

More Bang For Your Buck: Quorum-Sensing Capabilities Improve the Efficacy of Suicidal Altruism

2014-06-02 · Anya Elaine Johnson, Eli Strauss, Rodney Pickett, Christoph Adami 외

Within the context of evolution, an altruistic act that benefits the receiving individual at the expense of the acting individual is a puzzling phenomenon. An extreme form of altruism can be found in colicinogenic E. col…