paper-with-me

홈 › Papers

EMBEDDIA project: Cross-Lingual Embeddings for Less- Represented Languages in European News Media

2022-06-01 · EAMT 2022 6 · Senja Pollak, Andraž Pelicon

EMBEDDIA project developed a range of resources and methods for less-resourced EU languages, focusing on applications for media industry, including keyword extraction, comment moderation and article generation.

📄 PDF Abstract BibTeX

Code (0)

등록된 구현이 없습니다.

Tasks

Keyword Extraction

Similar Papers 제목 키워드 기반

EMBEDDIA at SemEval-2022 Task 8: Investigating Sentence, Image, and Knowledge Graph Representations for Multilingual News Article Similarity

2022-07-01 · SemEval (NAACL) 2022 7 · Elaine Zosa, Emanuela Boros, Boshko Koloski, Lidia Pivovarova

In this paper, we present the participation of the EMBEDDIA team in the SemEval-2022 Task 8 (Multilingual News Article Similarity). We cover several techniques and propose different methods for finding the multilingual n…

ArticlesSemantic SimilaritySemantic Textual SimilaritySentence

Interesting cross-border news discovery using cross-lingual article linking and document similarity

2021-04-01 · EACL (Hackashop) 2021 4 · Boshko Koloski, Elaine Zosa, Timen Stepišnik-Perdih, Blaž Škrlj 외

Team Name: team-8 Embeddia Tool: Cross-Lingual Document Retrieval Zosa et al. Dataset: Estonian and Latvian news datasets abstract: Contemporary news media face increasing amounts of available data that can be of use whe…

ArticlesRetrieval

EMBEDDIA Tools, Datasets and Challenges: Resources and Hackathon Contributions

2021-04-01 · EACL (Hackashop) 2021 4 · Senja Pollak, Marko Robnik-Šikonja, Matthew Purver, Michele Boggia 외

This paper presents tools and data sources collected and released by the EMBEDDIA project, supported by the European Union’s Horizon 2020 research and innovation program. The collected resources were offered to participa…

Best Practices for Learning Domain-Specific Cross-Lingual Embeddings

2019-07-06 · WS 2019 8 · Lena Shakurova, Beata Nyari, Chao Li, Mihai Rotaru

Cross-lingual embeddings aim to represent words in multiple languages in a shared vector space by capturing semantic similarities across languages. They are a crucial component for scaling tasks to multiple languages by …

Transfer Learning

Machine Translation Reference-less Evaluation using YiSi-2 with Bilingual Mappings of Massive Multilingual Language Model

2020-11-01 · WMT (EMNLP) 2020 11 · Chi-kiu Lo, Samuel Larkin

We present a study on using YiSi-2 with massive multilingual pretrained language models for machine translation (MT) reference-less evaluation. Aiming at finding better semantic representation for semantic MT evaluation,…

Language ModelingLanguage ModellingMachine TranslationSemantic Similarity+3