paper-with-me

홈 › Papers

Discrete Variational Attention Models for Language Generation

2020-04-21 · Xianghong Fang, Haoli Bai, Zenglin Xu, Michael Lyu, Irwin King

Variational autoencoders have been widely applied for natural language generation, however, there are two long-standing problems: information under-representation and posterior collapse. The former arises from the fact that only the last hidden state from the encoder is transformed to the latent space, which is insufficient to summarize data. The latter comes as a result of the imbalanced scale between the reconstruction loss and the KL divergence in the objective function. To tackle these issues, in this paper we propose the discrete variational attention model with categorical distribution over the attention mechanism owing to the discrete nature in languages. Our approach is combined with an auto-regressive prior to capture the sequential dependency from observations, which can enhance the latent space for language generation. Moreover, thanks to the property of discreteness, the training of our proposed approach does not suffer from posterior collapse. Furthermore, we carefully analyze the superiority of discrete latent space over the continuous space with the common Gaussian distribution. Extensive experiments on language generation demonstrate superior advantages of our proposed approach in comparison with the state-of-the-art counterparts.

📄 PDF Abstract BibTeX arXiv:2004.09764

Code (0)

등록된 구현이 없습니다.

Tasks

Language ModellingText Generation

Methods 이 논문이 사용한 방법론

Tanh Activation 설명 없음
USD Coin Customer Service Number +1-833-534-1729 설명 없음
Sigmoid Activation 설명 없음
LSTM An LSTM is a type of recurrent neural network that addresses the vanishing gradient problem in vanilla…

Similar Papers 제목 키워드 기반

Discrete Auto-regressive Variational Attention Models for Text Modeling

2021-06-16 · Xianghong Fang, Haoli Bai, Jian Li, Zenglin Xu 외

Variational autoencoders (VAEs) have been widely applied for text modeling. In practice, however, they are troubled by two challenges: information underrepresentation and posterior collapse. The former arises as only the…

Language ModelingLanguage Modelling

Conditional Variational Autoencoder for Neural Machine Translation

2018-12-11 · Artidoro Pagnoni, Kevin Liu, Shangyan Li

We explore the performance of latent variable models for conditional text generation in the context of neural machine translation (NMT). Similar to Zhang et al., we augment the encoder-decoder NMT paradigm by introducing…

Conditional Text GenerationDecoderMachine TranslationNMT+2

Direct Simultaneous Speech-to-Speech Translation with Variational Monotonic Multihead Attention

2021-10-15 · Xutai Ma, Hongyu Gong, Danni Liu, Ann Lee 외

We present a direct simultaneous speech-to-speech translation (Simul-S2ST) model, Furthermore, the generation of translation is independent from intermediate text representations. Our approach leverages recent progress o…

Simultaneous Speech-to-Speech TranslationSpeech SynthesisSpeech-to-Speech TranslationTranslation

Improving Semantic Control in Discrete Latent Spaces with Transformer Quantized Variational Autoencoders

2024-02-01 · Yingji Zhang, Danilo S. Carvalho, Marco Valentino, Ian Pratt-Hartmann 외

Achieving precise semantic control over the latent spaces of Variational AutoEncoders (VAEs) holds significant value for downstream tasks in NLP as the underlying generative mechanisms could be better localised, explaine…

DiscoDVT: Generating Long Text with Discourse-Aware Discrete Variational Transformer

2021-10-12 · EMNLP 2021 11 · Haozhe Ji, Minlie Huang

Despite the recent advances in applying pre-trained language models to generate high-quality texts, generating long passages that maintain long-range coherence is yet challenging for these models. In this paper, we propo…

Story GenerationText Generation