paper-with-me

Papers

BUCC Shared Task: Cross-Language Document Similarity

2015-07-01 · WS 2015 7 · Serge Sharoff, Pierre Zweigenbaum, Reinhard Rapp
📄 PDF Abstract BibTeX

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

AUT Document Alignment Framework for BUCC Workshop Shared Task

2015-07-01 · WS 2015 7 · Atefeh Zafarian, Amir Pouya Agha Sadeghi, Fatemeh Azadi, Sonia Ghiasifard 외
Information RetrievalMachine Translation

BUCC2020: Bilingual Dictionary Induction using Cross-lingual Embedding

2020-05-01 · LREC 2020 5 · Sanjanasri JP, Vijay Krishna Menon, Soman KP

This paper presents a deep learning system for the BUCC 2020 shared task: Bilingual dictionary induction from comparable corpora. We have submitted two runs for this shared Task, German (de) and English (en) language pai…

Deep LearningWord Embeddings

Overview of the Second BUCC Shared Task: Spotting Parallel Sentences in Comparable Corpora

2017-08-01 · WS 2017 8 · Pierre Zweigenbaum, Serge Sharoff, Reinhard Rapp

This paper presents the BUCC 2017 shared task on parallel sentence extraction from comparable corpora. It recalls the design of the datasets, presents their final construction and statistics and the methods used to evalu…

Machine TranslationSentence

BUCC 2017 Shared Task: a First Attempt Toward a Deep Learning Framework for Identifying Parallel Sentences in Comparable Corpora

2017-08-01 · WS 2017 8 · Francis Gr{\'e}goire, Philippe Langlais

This paper describes our participation in BUCC 2017 shared task: identifying parallel sentences in comparable corpora. Our goal is to leverage continuous vector representations and distributional semantics with a minimal…

Feature EngineeringLanguage ModelingLanguage ModellingMachine Translation+2

TALN/LS2N Participation at the BUCC Shared Task: Bilingual Dictionary Induction from Comparable Corpora

2020-05-01 · LREC 2020 5 · Martin Laville, Amir Hazem, Emmanuel Morin

This paper describes the TALN/LS2N system participation at the Building and Using Comparable Corpora (BUCC) shared task. We first introduce three strategies: (i) a word embedding approach based on fastText embeddings; (i…