paper-with-me

홈 › Papers

One "Ruler" for All Languages: Multi-Lingual Dialogue Evaluation with Adversarial Multi-Task Learning

2018-05-08 · Xiaowei Tong, Zhenxin Fu, Mingyue Shang, Dongyan Zhao, Rui Yan

Automatic evaluating the performance of Open-domain dialogue system is a challenging problem. Recent work in neural network-based metrics has shown promising opportunities for automatic dialogue evaluation. However, existing methods mainly focus on monolingual evaluation, in which the trained metric is not flexible enough to transfer across different languages. To address this issue, we propose an adversarial multi-task neural metric (ADVMT) for multi-lingual dialogue evaluation, with shared feature extraction across languages. We evaluate the proposed model in two different languages. Experiments show that the adversarial multi-task neural metric achieves a high correlation with human annotation, which yields better performance than monolingual ones and various existing metrics.

📄 PDF Abstract BibTeX arXiv:1805.02914

Code (0)

등록된 구현이 없습니다.

Tasks

AllDialogue EvaluationMulti-Task Learning

Similar Papers 제목 키워드 기반

One ruler to measure them all: Benchmarking multilingual long-context language models

2025-03-03 · Yekyung Kim, Jenna Russell, Marzena Karpinska, Mohit Iyyer

We present ONERULER, a multilingual benchmark designed to evaluate long-context language models across 26 languages. ONERULER adapts the English-only RULER benchmark (Hsieh et al., 2024) by including seven synthetic task…

8kAllBenchmarking

xDial-Eval: A Multilingual Open-Domain Dialogue Evaluation Benchmark

2023-10-13 · Chen Zhang, Luis Fernando D'Haro, Chengguang Tang, Ke Shi 외

Recent advancements in reference-free learned metrics for open-domain dialogue evaluation have been driven by the progress in pre-trained language models and the availability of dialogue data with high-quality human anno…

Dialogue EvaluationMachine Translation

Prompt Learning to Mitigate Catastrophic Forgetting in Cross-lingual Transfer for Open-domain Dialogue Generation

2023-05-12 · Lei Liu, Jimmy Xiangji Huang

Dialogue systems for non-English languages have long been under-explored. In this paper, we take the first step to investigate few-shot cross-lingual transfer learning (FS-XLT) and multitask learning (MTL) in the context…

Cross-Lingual TransferDialogue GenerationLanguage ModelingLanguage Modelling+2

XPersona: Evaluating Multilingual Personalized Chatbot

2020-03-17 · EMNLP (NLP4ConvAI) 2021 11 · Zhaojiang Lin, Zihan Liu, Genta Indra Winata, Samuel Cahyawijaya 외

Personalized dialogue systems are an essential step toward better human-machine interaction. Existing personalized dialogue agents rely on properly designed conversational datasets, which are mostly monolingual (e.g., En…

ChatbotTranslation

IndicTalk: A Large-Scale Persona-Based Multilingual Conversational Corpus for Indic Languages

2026-07-25 · Sahil Deepak Gawande, Mayank Singh arxiv

Large Language Models (LLMs) have transformed conversational AI, yet high-quality multilingual code-mixed dialogue resources remain scarce, particularly for Indic languages where speakers naturally alternate between Engl…

Dialogue Generation