Corpus-based Referring Expressions Generation
In Natural Language Generation, the task of attribute selection (AS) consists of determining the appropriate attribute-value pairs (or semantic properties) that represent the contents of a referring expression. Existing work on AS includes a wide range of algorithmic solutions to the problem, but the recent availability of corpora annotated with referring expressions data suggests that corpus-based AS strategies become possible as well. In this work we tentatively discuss a number of AS strategies using both semantic and surface information obtained from a corpus of this kind. Relying on semantic information, we attempt to learn both global and individual AS strategies that could be applied to a standard AS algorithm in order to generate descriptions found in the corpus. As an alternative, and perhaps less traditional approach, we also use surface information to build statistical language models of the referring expressions that are most likely to occur in the corpus, and let the model probabilities guide attribute selection.
Code (0)
등록된 구현이 없습니다.
Tasks
AttributeReferring ExpressionText GenerationSimilar Papers 제목 키워드 기반
G-TUNA: a corpus of referring expressions in German, including duration information
Corpora of referring expressions elicited from human participants in a controlled environment are an important resource for research on automatic referring expression generation. We here present G-TUNA, a new corpus of r…
Referring ExpressionReferring expression generationText GenerationLessons from Computational Modelling of Reference Production in Mandarin and English
Referring expression generation (REG) algorithms offer computational models of the production of referring expressions. In earlier work, a corpus of referring expressions (REs) in Mandarin was introduced. In the present …
Referring ExpressionReferring expression generationInvestigating the content and form of referring expressions in Mandarin: introducing the Mtuna corpus
East Asian languages are thought to handle reference differently from languages such as English, particularly in terms of the marking of definiteness and number. We present the first Data-Text corpus for Referring Expres…
FormText GenerationPentoRef: A Corpus of Spoken References in Task-oriented Dialogues
PentoRef is a corpus of task-oriented dialogues collected in systematically manipulated settings. The corpus is multilingual, with English and German sections, and overall comprises more than 20000 utterances. The dialog…
Referring to what you know and do not know: Making Referring Expression Generation Models Generalize To Unseen Entities
Data-to-text Natural Language Generation (NLG) is the computational process of generating natural language in the form of text or voice from non-linguistic data. A core micro-planning task within NLG is referring express…
DecoderReferring ExpressionReferring expression generationText Generation