paper-with-me

홈 › Papers

MIA 2022 Shared Task: Evaluating Cross-lingual Open-Retrieval Question Answering for 16 Diverse Languages

2022-07-02 · NAACL (MIA) 2022 7 · Akari Asai, Shayne Longpre, Jungo Kasai, Chia-Hsuan Lee, Rui Zhang, Junjie Hu, Ikuya Yamada, Jonathan H. Clark, Eunsol Choi

We present the results of the Workshop on Multilingual Information Access (MIA) 2022 Shared Task, evaluating cross-lingual open-retrieval question answering (QA) systems in 16 typologically diverse languages. In this task, we adapted two large-scale cross-lingual open-retrieval QA datasets in 14 typologically diverse languages, and newly annotated open-retrieval QA data in 2 underrepresented languages: Tagalog and Tamil. Four teams submitted their systems. The best system leveraging iteratively mined diverse negative examples and larger pretrained models achieves 32.2 F1, outperforming our baseline by 4.5 points. The second best system uses entity-aware contextualized representations for document retrieval, and achieves significant improvements in Tamil (20.8 F1), whereas most of the other systems yield nearly zero scores.

📄 PDF Abstract BibTeX arXiv:2207.00758

Code (0)

등록된 구현이 없습니다.

Tasks

Question AnsweringRetrieval

Similar Papers 제목 키워드 기반

One "Ruler" for All Languages: Multi-Lingual Dialogue Evaluation with Adversarial Multi-Task Learning

2018-05-08 · Xiaowei Tong, Zhenxin Fu, Mingyue Shang, Dongyan Zhao 외

Automatic evaluating the performance of Open-domain dialogue system is a challenging problem. Recent work in neural network-based metrics has shown promising opportunities for automatic dialogue evaluation. However, exis…

AllDialogue EvaluationMulti-Task Learning

The CogALex Shared Task on Monolingual and Multilingual Identification of Semantic Relations

2020-12-01 · COLING (CogALex) 2020 12 · Rong Xiang, Emmanuele Chersoni, Luca Iacoponi, Enrico Santus

The shared task of the CogALex-VI workshop focuses on the monolingual and multilingual identification of semantic relations. We provided training and validation data for the following languages: English, German and Chine…

Relation

Evaluating Cross-Lingual Unlearning in Multilingual Language Models

2026-01-10 · Tyler Lizzo, Larry Heck arxiv

We present the first comprehensive evaluation of cross-lingual unlearning in multilingual LLMs. Using translated TOFU benchmarks in seven language/script variants, we test major unlearning algorithms and show that most f…

Massively Multilingual Word Embeddings

2016-02-05 · Waleed Ammar, George Mulcaire, Yulia Tsvetkov, Guillaume Lample 외

We introduce new methods for estimating and evaluating embeddings of words in more than fifty languages in a single shared embedding space. Our estimation methods, multiCluster and multiCCA, use dictionaries and monoling…

Multilingual Word EmbeddingsText CategorizationWord Embeddings

MT4CrossOIE: Multi-stage Tuning for Cross-lingual Open Information Extraction

2023-08-12 · Tongliang Li, Zixiang Wang, Linzheng Chai, Jian Yang 외

Cross-lingual open information extraction aims to extract structured information from raw text across multiple languages. Previous work uses a shared cross-lingual pre-trained model to handle the different languages but …

Cross-Lingual TransferLanguage ModellingLarge Language ModelOpen Information Extraction