paper-with-me

홈 › Papers

Characterizing Variation in Crowd-Sourced Data for Training Neural Language Generators to Produce Stylistically Varied Outputs

2018-09-14 · WS 2018 11 · Juraj Juraska, Marilyn Walker

One of the biggest challenges of end-to-end language generation from meaning representations in dialogue systems is making the outputs more natural and varied. Here we take a large corpus of 50K crowd-sourced utterances in the restaurant domain and develop text analysis methods that systematically characterize types of sentences in the training data. We then automatically label the training data to allow us to conduct two kinds of experiments with a neural generator. First, we test the effect of training the system with different stylistic partitions and quantify the effect of smaller, but more stylistically controlled training data. Second, we propose a method of labeling the style variants during training, and show that we can modify the style of the generated utterances using our stylistic labels. We contrast and compare these methods that can be used with any existing large corpus, showing how they vary in terms of semantic quality and stylistic control.

📄 PDF Abstract BibTeX arXiv:1809.05288

Code (0)

등록된 구현이 없습니다.

Tasks

Text Generation

Similar Papers 제목 키워드 기반

Characterizing 5G User Throughput via Uncertainty Modeling and Crowdsourced Measurements

2025-10-10 · Javier Albert-Smet, Zoraida Frias, Luis Mendo, Sergio Melones 외 arxiv

Characterizing application-layer user throughput in next-generation networks is increasingly challenging as the higher capacity of the 5G Radio Access Network (RAN) shifts connectivity bottlenecks towards deeper parts of…

Full Characterization of Adaptively Strong Majority Voting in Crowdsourcing

2021-11-11 · Margarita Boyarskaya, Panos Ipeirotis

In crowdsourcing, quality control is commonly achieved by having workers examine items and vote on their correctness. To minimize the impact of unreliable worker responses, a $\delta$-margin voting process is utilized, w…

Semi-crowdsourced Clustering with Deep Generative Models

2018-10-29 · NeurIPS 2018 12 · Yucen Luo, Tian Tian, Jiaxin Shi, Jun Zhu 외

We consider the semi-supervised clustering problem where crowdsourcing provides noisy information about the pairwise comparisons on a small subset of data, i.e., whether a sample pair is in the same cluster. We propose a…

ClusteringVariational Inference

Parsimonious Mixed-Effects HodgeRank for Crowdsourced Preference Aggregation

2016-07-12 · Qianqian Xu, Jiechao Xiong, Xiaochun Cao, Yuan YAO

In crowdsourced preference aggregation, it is often assumed that all the annotators are subject to a common preference or utility function which generates their comparison behaviors in experiments. However, in reality an…

Characterizing Datasets for Social Visual Question Answering, and the New TinySocial Dataset

2020-10-08 · Zhanwen Chen, Shiyao Li, Roxanne Rashedi, Xiaoman Zi 외

Modern social intelligence includes the ability to watch videos and answer questions about social and theory-of-mind-related content, e.g., for a scene in Harry Potter, "Is the father really upset about the boys flying t…

Question AnsweringVisual Question AnsweringVisual Question Answering (VQA)