A novel repetition normalized adversarial reward for headline generation
While reinforcement learning can effectively improve language generation models, it often suffers from generating incoherent and repetitive phrases \cite{paulus2017deep}. In this paper, we propose a novel repetition normalized adversarial reward to mitigate these problems. Our repetition penalized reward can greatly reduce the repetition rate and adversarial training mitigates generating incoherent phrases. Our model significantly outperforms the baseline model on ROUGE-1\,(+3.24), ROUGE-L\,(+2.25), and a decreased repetition-rate (-4.98\%).
Code (0)
등록된 구현이 없습니다.
Tasks
Headline Generationreinforcement-learningReinforcement LearningReinforcement Learning (RL)Text GenerationSimilar Papers 제목 키워드 기반
Clickbait? Sensational Headline Generation with Auto-tuned Reinforcement Learning
Sensational headlines are headlines that capture people's attention and generate reader interest. Conventional abstractive headline generation methods, unlike human writers, do not optimize for maximal reader attention. …
Headline Generationreinforcement-learningReinforcement LearningReinforcement Learning (RL)Attractive or Faithful? Popularity-Reinforced Learning for Inspired Headline Generation
With the rapid proliferation of online media sources and published news, headlines have become increasingly important for attracting readers to news articles, since users may be overwhelmed with the massive information. …
ArticlesHeadline GenerationReinforcement LearningReinforcement Learning (RL)+1Stimulating Creativity with FunLines: A Case Study of Humor Generation in Headlines
Building datasets of creative text, such as humor, is quite challenging. We introduce FunLines, a competitive game where players edit news headlines to make them funny, and where they rate the funniness of headlines edit…
Multiple News Headlines Generation using Page Metadata
Multiple headlines of a newspaper article have an important role to express the content of the article accurately and concisely. A headline depends on the content and intent of their article. While a single headline expr…
Headline GenerationImplicitly normalized forecaster with clipping for linear and non-linear heavy-tailed multi-armed bandits
The Implicitly Normalized Forecaster (INF) algorithm is considered to be an optimal solution for adversarial multi-armed bandit (MAB) problems. However, most of the existing complexity results for INF rely on restrictive…
Multi-Armed Bandits