paper-with-me

홈 › Papers

MultiMUC: Multilingual Template Filling on MUC-4

2024-01-29 · William Gantt, Shabnam Behzad, Hannah Youngeun An, Yunmo Chen, Aaron Steven White, Benjamin Van Durme, Mahsa Yarmohammadi

We introduce MultiMUC, the first multilingual parallel corpus for template filling, comprising translations of the classic MUC-4 template filling benchmark into five languages: Arabic, Chinese, Farsi, Korean, and Russian. We obtain automatic translations from a strong multilingual machine translation system and manually project the original English annotations into each target language. For all languages, we also provide human translations for sentences in the dev and test splits that contain annotated template arguments. Finally, we present baselines on MultiMUC both with state-of-the-art template filling models and with ChatGPT.

📄 PDF Abstract BibTeX arXiv:2401.16209

Code (1)

wgantt/multimuc 공식 구현

Tasks

Machine TranslationTranslation

Similar Papers 제목 키워드 기반

Macaron: Controlled, Human-Written Benchmark for Multilingual and Multicultural Reasoning via Template-Filling

2026-02-11 · Alaa Elsetohy, Sama Hadhoud, Haryo Akbarianto Wibowo, Chenxi Whitehouse 외 arxiv

Multilingual benchmarks rarely test reasoning over culturally grounded premises: translated datasets keep English-centric scenarios, while culture-first datasets often lack control over the reasoning required. We propose…

GlobalWoZ: Globalizing MultiWoZ to Develop Multilingual Task-Oriented Dialogue Systems

2021-11-16 · ACL ARR November 2021 11 · Anonymous

Over the last few years, there has been a move towards data curation for multilingual task-oriented dialogue (ToD) systems that can serve people speaking different languages. However, existing multilingual ToD datasets e…

Task-Oriented Dialogue Systems

GlobalWoZ: Globalizing MultiWoZ to Develop Multilingual Task-Oriented Dialogue Systems

2021-10-14 · ACL 2022 5 · Bosheng Ding, Junjie Hu, Lidong Bing, Sharifah Mahani Aljunied 외

Much recent progress in task-oriented dialogue (ToD) systems has been driven by available annotation data across multiple domains for training. Over the last few years, there has been a move towards data curation for mul…

Task-Oriented Dialogue Systems

A Dataset for Cross-Domain Reasoning via Template Filling

2022-01-16 · ACL ARR January 2022 1 · Anonymous

While several benchmarks exist for reasoning tasks, reasoning across domains is an under-explored area in NLP. Towards this, we present a dataset and a prompt-template-filling approach to enable sequence to sequence mode…

Decoder

Predicting the Performance of Multilingual NLP Models

2021-10-17 · Anirudh Srinivasan, Sunayana Sitaram, Tanuja Ganu, Sandipan Dandapat 외

Recent advancements in NLP have given us models like mBERT and XLMR that can serve over 100 languages. The languages that these models are evaluated on, however, are very few in number, and it is unlikely that evaluation…

Multilingual NLP