Language coverage and generalization in RNN-based continuous sentence embeddings for interacting agents
Continuous sentence embeddings using recurrent neural networks (RNNs), where variable-length sentences are encoded into fixed-dimensional vectors, are often the main building blocks of architectures applied to language tasks such as dialogue generation. While it is known that those embeddings are able to learn some structures of language (e.g. grammar) in a purely data-driven manner, there is very little work on the objective evaluation of their ability to cover the whole language space and to generalize to sentences outside the language bias of the training data. Using a manually designed context-free grammar (CFG) to generate a large-scale dataset of sentences related to the content of realistic 3D indoor scenes, we evaluate the language coverage and generalization abilities of the most common continuous sentence embeddings based on RNNs. We also propose a new embedding method based on arithmetic coding, AriEL, that is not data-driven and that efficiently encodes in continuous space any sentence from the CFG. We find that RNN-based embeddings underfit the training data and cover only a small subset of the language defined by the CFG. They also fail to learn the underlying CFG and generalize to unbiased sentences from that same CFG. We found that AriEL provides an insightful baseline.
Code (0)
등록된 구현이 없습니다.
Tasks
Dialogue GenerationSentenceSentence EmbeddingsSimilar Papers 제목 키워드 기반
D2CSE: Difference-aware Deep continuous prompts for Contrastive Sentence Embeddings
This paper describes Difference-aware Deep continuous prompt for Contrastive Sentence Embeddings (D2CSE) that learns sentence embeddings. Compared to state-of-the-art approaches, D2CSE computes sentence vectors that are …
Contrastive LearningRetrievalSemantic Textual SimilaritySentence+2Unified Visual-Semantic Embeddings: Bridging Vision and Language With Structured Meaning Representations
We propose the Unified Visual-Semantic Embeddings (Unified VSE) for learning a joint space of visual representation and textual semantics. The model unifies the embeddings of concepts at different levels: objects, attrib…
Contrastive LearningCross-Modal RetrievalRetrievalSentenceRegressing Word and Sentence Embeddings for Regularization of Neural Machine Translation
In recent years, neural machine translation (NMT) has become the dominant approach in automated translation. However, like many other deep learning approaches, NMT suffers from overfitting when the amount of training dat…
ClusteringMachine TranslationNMTSentence+2COSTRA 1.0: A Dataset of Complex Sentence Transformations
We present COSTRA 1.0, a dataset of complex sentence transformations. The dataset is intended for the study of sentence-level embeddings beyond simple word alternations or standard paraphrasing. This first version of the…
SentenceSentence EmbeddingSentence-EmbeddingSentence EmbeddingsIn Search for Linear Relations in Sentence Embedding Spaces
We present an introductory investigation into continuous-space vector representations of sentences. We acquire pairs of very similar sentences differing only by a small alterations (such as change of a noun, adding an ad…
Natural Language InferenceSentenceSentence EmbeddingSentence-Embedding