paper-with-me

Papers

Naver Labs Europe (SPLADE) @ TREC NeuCLIR 2022

2023-03-10 · Carlos Lassance, Stéphane Clinchant

This paper describes our participation in the 2022 TREC NeuCLIR challenge. We submitted runs to two out of the three languages (Farsi and Russian), with a focus on first-stage rankers and comparing mono-lingual strategies to Adhoc ones. For monolingual runs, we start from pretraining models on the target language using MLM+FLOPS and then finetuning using the MSMARCO translated to the language either with ColBERT or SPLADE as the retrieval model. While for the Adhoc task, we test both query translation (to the target language) and back-translation of the documents (to English). Initial result analysis shows that the monolingual strategy is strong, but that for the moment Adhoc achieved the best results, with back-translating documents being better than translating queries.

📄 PDF Abstract BibTeX arXiv:2303.11171

Code (0)

등록된 구현이 없습니다.

Tasks

RetrievalTranslation

Methods 이 논문이 사용한 방법론

Test 설명 없음

Similar Papers 제목 키워드 기반

Naver Labs Europe (SPLADE) @ TREC Deep Learning 2022

2023-02-24 · Carlos Lassance, Stéphane Clinchant

This paper describes our participation to the 2022 TREC Deep Learning challenge. We submitted runs to all four tasks, with a focus on the full retrieval passage task. The strategy is almost the same as 2021, with first s…

Deep LearningRetrieval

Naver Labs Europe’s Participation in the Robustness, Chat, and Biomedical Tasks at WMT 2020

2020-11-01 · WMT (EMNLP) 2020 11 · Alexandre Berard, Ioan Calapodescu, Vassilina Nikoulina, Jerin Philip

This paper describes Naver Labs Europe’s participation in the Robustness, Chat, and Biomedical Translation tasks at WMT 2020. We propose a bidirectional German-English model that is multi-domain, robust to noise, and whi…

Language ModelingLanguage ModellingTranslation

NAVER LABS Europe's Multilingual Speech Translation Systems for the IWSLT 2023 Low-Resource Track

2023-06-13 · Edward Gow-Smith, Alexandre Berard, Marcely Zanon Boito, Ioan Calapodescu

This paper presents NAVER LABS Europe's systems for Tamasheq-French and Quechua-Spanish speech translation in the IWSLT 2023 Low-Resource track. Our work attempts to maximize translation quality in low-resource settings …

Translation

Overview of the TREC 2022 NeuCLIR Track

2023-04-24 · Dawn Lawrie, Sean MacAvaney, James Mayfield, Paul McNamee 외

This is the first year of the TREC Neural CLIR (NeuCLIR) track, which aims to study the impact of neural approaches to cross-language information retrieval. The main task in this year's track was ad hoc ranked retrieval …

Information RetrievalRetrieval

NeuralMind-UNICAMP at 2022 TREC NeuCLIR: Large Boring Rerankers for Cross-lingual Retrieval

2023-03-28 · Vitor Jeronymo, Roberto Lotufo, Rodrigo Nogueira

This paper reports on a study of cross-lingual information retrieval (CLIR) using the mT5-XXL reranker on the NeuCLIR track of TREC 2022. Perhaps the biggest contribution of this study is the finding that despite the mT5…

Cross-Lingual Information RetrievalInformation RetrievalRetrieval