BUCC Shared Task: Cross-Language Document Similarity
Code (0)
등록된 구현이 없습니다.
Similar Papers 제목 키워드 기반
AUT Document Alignment Framework for BUCC Workshop Shared Task
BUCC2020: Bilingual Dictionary Induction using Cross-lingual Embedding
This paper presents a deep learning system for the BUCC 2020 shared task: Bilingual dictionary induction from comparable corpora. We have submitted two runs for this shared Task, German (de) and English (en) language pai…
Deep LearningWord EmbeddingsOverview of the Second BUCC Shared Task: Spotting Parallel Sentences in Comparable Corpora
This paper presents the BUCC 2017 shared task on parallel sentence extraction from comparable corpora. It recalls the design of the datasets, presents their final construction and statistics and the methods used to evalu…
Machine TranslationSentenceBUCC 2017 Shared Task: a First Attempt Toward a Deep Learning Framework for Identifying Parallel Sentences in Comparable Corpora
This paper describes our participation in BUCC 2017 shared task: identifying parallel sentences in comparable corpora. Our goal is to leverage continuous vector representations and distributional semantics with a minimal…
Feature EngineeringLanguage ModelingLanguage ModellingMachine Translation+2TALN/LS2N Participation at the BUCC Shared Task: Bilingual Dictionary Induction from Comparable Corpora
This paper describes the TALN/LS2N system participation at the Building and Using Comparable Corpora (BUCC) shared task. We first introduce three strategies: (i) a word embedding approach based on fastText embeddings; (i…