paper-with-me

홈 › Papers

A Natural Bias for Language Generation Models

2022-12-19 · Clara Meister, Wojciech Stokowiec, Tiago Pimentel, Lei Yu, Laura Rimell, Adhiguna Kuncoro

After just a few hundred training updates, a standard probabilistic model for language generation has likely not yet learnt many semantic or syntactic rules of natural language, making it difficult to estimate the probability distribution over next tokens. Yet around this point, these models have identified a simple, loss-minimising behaviour: to output the unigram distribution of the target training corpus. The use of such a heuristic raises the question: Can we initialise our models with this behaviour and save precious compute resources and model capacity? Here we show that we can effectively endow standard neural language generation models with a separate module that reflects unigram frequency statistics as prior knowledge, simply by initialising the bias term in a model's final linear layer with the log-unigram distribution. We use neural machine translation as a test bed for this simple technique and observe that it: (i) improves learning efficiency; (ii) achieves better overall performance; and perhaps most importantly (iii) appears to disentangle strong frequency effects by encouraging the model to specialise in non-frequency-related aspects of language.

📄 PDF Abstract BibTeX arXiv:2212.09686

Code (0)

등록된 구현이 없습니다.

Tasks

Machine TranslationText Generation

Methods 이 논문이 사용한 방법론

Linear Layer A Linear Layer is a projection $\mathbf{XW + b}$.

Similar Papers 제목 키워드 기반

Defining and Evaluating Fair Natural Language Generation

2020-07-28 · WS 2020 7 · Catherine Yeo, Alyssa Chen

Our work focuses on the biases that emerge in the natural language generation (NLG) task of sentence completion. In this paper, we introduce a framework of fairness for NLG followed by an evaluation of gender biases in t…

FairnessSentenceSentence CompletionText Generation

Quantifying Bias from Decoding Techniques in Natural Language Generation

2022-10-01 · COLING 2022 10 · Mayukh Das, Wolf Tilo Balke

Natural language generation (NLG) models can propagate social bias towards particular demography. Though several studies investigated bias from data and model, NLG task distinctively uses stochastic decoder that can posi…

DecoderText Generation

Role of Bias Terms in Dot-Product Attention

2023-02-16 · Mahdi Namazifar, Devamanyu Hazarika, Dilek Hakkani-Tur

Dot-product attention is a core module in the present generation of neural network models, particularly transformers, and is being leveraged across numerous areas such as natural language processing and computer vision. …

Language ModelingLanguage ModellingNatural Language UnderstandingText Generation

Viable Threat on News Reading: Generating Biased News Using Natural Language Models

2020-10-05 · EMNLP (NLP+CSS) 2020 11 · Saurabh Gupta, Huy H. Nguyen, Junichi Yamagishi, Isao Echizen

Recent advancements in natural language generation has raised serious concerns. High-performance language models are widely used for language generation tasks because they are able to produce fluent and meaningful senten…

ArticlesText Generation

Societal Biases in Language Generation: Progress and Challenges

2021-05-10 · ACL 2021 5 · Emily Sheng, Kai-Wei Chang, Premkumar Natarajan, Nanyun Peng

Technology for language generation has advanced rapidly, spurred by advancements in pre-training large models on massive amounts of data and the need for intelligent agents to communicate in a natural manner. While techn…

FairnessText Generation