Data Collection for Dialogue System: A Startup Perspective
Industrial dialogue systems such as Apple Siri and Google Now rely on large scale diverse and robust training data to enable their sophisticated conversation capability. Crowdsourcing provides a scalable and inexpensive way of data collection but collecting high quality data efficiently requires thoughtful orchestration of the crowdsourcing jobs. Prior study of this topic have focused on tasks only in the academia settings with limited scope or only provide intrinsic dataset analysis, lacking indication on how it affects the trained model performance. In this paper, we present a study of crowdsourcing methods for a user intent classification task in our deployed dialogue system. Our task requires classification of 47 possible user intents and contains many intent pairs with subtle differences. We consider different crowdsourcing job types and job prompts and analyze quantitatively the quality of the collected data and the downstream model performance on a test set of real user queries from production logs. Our observation provides insights into designing efficient crowdsourcing jobs and provide recommendations for future dialogue system data collection process.
Code (0)
등록된 구현이 없습니다.
Tasks
General Classificationintent-classificationIntent ClassificationText ClassificationSimilar Papers 제목 키워드 기반
Compliance Costs of AI Technology Commercialization: A Field Deployment Perspective
While Artificial Intelligence (AI) technologies are progressing fast, compliance costs have become a huge financial burden for AI startups, which are already constrained on research & development budgets. This situation …
Solving the Data Sparsity Problem in Predicting the Success of the Startups with Machine Learning Methods
Predicting the success of startup companies is of great importance for both startup companies and investors. It is difficult due to the lack of available data and appropriate general methods. With data platforms like Cru…
BIG-bench Machine LearningBeyond Isolated Investor: Predicting Startup Success via Roleplay-Based Collective Agents
Due to the high value and high failure rates of startups, predicting their success is a critical challenge. Existing approaches typically model startup success from a single decision-maker's perspective, overlooking the …
Bipartite-play Dialogue Collection for Practical Automatic Evaluation of Dialogue Systems
Automation of dialogue system evaluation is a driving force for the efficient development of dialogue systems. This paper introduces the bipartite-play method, a dialogue collection method for automating dialogue system …
Business-cycles and Cash-on-Market: Pre-money Startup Valuation in the Macroeconomic Environment
How do business-cycles impact startup-valuations? While several studies explore VC startupecosystems and pre-money valuations, relatively-few delve deeper into the role of macro-level economic factors in influencing thos…