paper-with-me

홈 › Papers

GCDF1: A Goal- and Context- Driven F-Score for Evaluating User Models

2021-11-01 · EANCS 2021 11 · Alexandru Coca, Bo-Hsiang Tseng, Bill Byrne

The evaluation of dialogue systems in interaction with simulated users has been proposed to improve turn-level, corpus-based metrics which can only evaluate test cases encountered in a corpus and cannot measure system’s ability to sustain multi-turn interactions. Recently, little emphasis was put on automatically assessing the quality of the user model itself, so unless correlations with human studies are measured, the reliability of user model based evaluation is unknown. We propose GCDF1, a simple but effective measure of the quality of semantic-level conversations between a goal-driven user agent and a system agent. In contrast with previous approaches we measure the F-score at dialogue level and consider user and system behaviours to improve recall and precision estimation. We facilitate scores interpretation by providing a rich hierarchical structure with information about conversational patterns present in the test data and tools to efficiently query the conversations generated. We apply our framework to assess the performance and weaknesses of a Convlab2 user model.

📄 PDF Abstract BibTeX

Code (1)

alexcoca/gcdf1 공식 구현

Tasks

Dialogue EvaluationTask-Oriented Dialogue Systems

Similar Papers 제목 키워드 기반

Explore Better Relative Position Embeddings from Encoding Perspective for Transformer Models

2021-11-01 · EMNLP 2021 11 · Anlin Qu, Jianwei Niu, Shasha Mo

Relative position embedding (RPE) is a successful method to explicitly and efficaciously encode position information into Transformer models. In this paper, we investigate the potential problems in Shaw-RPE and XL-RPE, w…

Position

Fast and Safe Trajectory Optimization for Mobile Manipulators With Neural Configuration Space Distance Field

2026-01-26 · Yulin Li, Zhiyuan Song, Yiming Li, Zhicheng Song 외 arxiv

Mobile manipulators promise agile, long-horizon behavior by coordinating base and arm motion, yet whole-body trajectory optimization in cluttered, confined spaces remains difficult due to high-dimensional nonconvexity an…

On the implicit minimization of alternative loss functions when training deep networks

2019-09-25 · Alexandre Lemire Paquin, Brahim Chaib-Draa, Philippe Giguère

Understanding the implicit bias of optimization algorithms is important in order to improve generalization of neural networks. One approach to try to exploit such understanding would be to then make the bias explicit in …

Inductive Bias

IndigoVX: Where Human Intelligence Meets AI for Optimal Decision Making

2023-07-21 · Kais Dukes

This paper defines a new approach for augmenting human intelligence with AI for optimal goal solving. Our proposed AI, Indigo, is an acronym for Informed Numerical Decision-making through Iterative Goal-Oriented optimiza…

Decision Making

AirDialogue: An Environment for Goal-Oriented Dialogue Research

2018-10-01 · EMNLP 2018 10 · Wei Wei, Quoc Le, Andrew Dai, Jia Li

Recent progress in dialogue generation has inspired a number of studies on dialogue systems that are capable of accomplishing tasks through natural language interactions. A promising direction among these studies is the …

Dialogue GenerationReinforcement LearningText Generation