paper-with-me

홈 › Papers

Production of Phrase Tables in 11 European Languages using an Improved Sub-sentential Aligner

2014-05-01 · LREC 2014 5 · Juan Luo, Yves Lepage

This paper is a partial report of an on-going Kakenhi project which aims to improve sub-sentential alignment and release multilingual syntactic patterns for statistical and example-based machine translation. Here we focus on improving a sub-sentential aligner which is an instance of the association approach. Phrase table is not only an essential component in the machine translation systems but also an important resource for research and usage in other domains. As part of this project, all phrase tables produced in the experiments will also be made freely available.

📄 PDF Abstract BibTeX

Code (0)

등록된 구현이 없습니다.

Tasks

Machine TranslationTranslation

Similar Papers 제목 키워드 기반

Exploiting Similarities among Languages for Machine Translation

2013-09-17 · Tomas Mikolov, Quoc V. Le, Ilya Sutskever

Dictionaries and phrase tables are the basis of modern statistical machine translation systems. This paper develops a method that can automate the process of generating and extending dictionaries and phrase tables. Our m…

Machine TranslationTranslation

EENLP: Cross-lingual Eastern European NLP Index

2021-08-05 · LREC 2022 6 · Alexey Tikhonov, Alex Malkhasov, Andrey Manoshin, George Dima 외

Motivated by the sparsity of NLP resources for Eastern European languages, we present a broad index of existing Eastern European language resources (90+ datasets and 45+ models) published as a github repository open for …

Cross-Lingual TransferNatural Language InferenceTransfer Learning

Cross-Lingual Transfer Learning for Phrase Break Prediction with Multilingual Language Model

2023-06-05 · Hoyeon Lee, Hyun-Wook Yoon, Jong-Hwan Kim, Jae-Min Kim

Phrase break prediction is a crucial task for improving the prosody naturalness of a text-to-speech (TTS) system. However, most proposed phrase break prediction models are monolingual, trained exclusively on a large amou…

Cross-Lingual TransferLanguage ModelingLanguage ModellingPrediction+3

Paraphrase Detection on Noisy Subtitles in Six Languages

2018-09-21 · WS 2018 11 · Eetu Sjöblom, Mathias Creutz, Mikko Aulamo

We perform automatic paraphrase detection on subtitle data from the Opusparcus corpus comprising six European languages: German, English, Finnish, French, Russian, and Swedish. We train two types of supervised sentence e…

SentenceSentence EmbeddingSentence-Embedding

EUROPA: A Legal Multilingual Keyphrase Generation Dataset

2024-03-01 · Olivier Salaün, Frédéric Piedboeuf, Guillaume Le Berre, David Alfonso Hermelo 외

Keyphrase generation has primarily been explored within the context of academic research articles, with a particular focus on scientific domains and the English language. In this work, we present EUROPA, a dataset for mu…

ArticlesKeyphrase Generation