paper-with-me

홈 › Papers

EMBEDDIA at SemEval-2022 Task 8: Investigating Sentence, Image, and Knowledge Graph Representations for Multilingual News Article Similarity

2022-07-01 · SemEval (NAACL) 2022 7 · Elaine Zosa, Emanuela Boros, Boshko Koloski, Lidia Pivovarova

In this paper, we present the participation of the EMBEDDIA team in the SemEval-2022 Task 8 (Multilingual News Article Similarity). We cover several techniques and propose different methods for finding the multilingual news article similarity by exploring the dataset in its entirety. We take advantage of the textual content of the articles, the provided metadata (e.g., titles, keywords, topics), the translated articles, the images (those that were available), and knowledge graph-based representations for entities and relations present in the articles. We, then, compute the semantic similarity between the different features and predict through regression the similarity scores. Our findings show that, while our proposed methods obtained promising results, exploiting the semantic textual similarity with sentence representations is unbeatable. Finally, in the official SemEval-2022 Task 8, we ranked fifth in the overall team ranking cross-lingual results, and second in the English-only results.

📄 PDF Abstract BibTeX

Code (1)

EMBEDDIA/semeval-2022-MNS 공식 구현 pytorch

Tasks

ArticlesSemantic SimilaritySemantic Textual SimilaritySentence

Similar Papers 제목 키워드 기반

Embeddia at SemEval-2019 Task 6: Detecting Hate with Neural Network and Transfer Learning Approaches

2019-06-01 · SEMEVAL 2019 6 · Andra{\v{z}} Pelicon, Matej Martinc, Petra Kralj Novak

SemEval 2019 Task 6 was OffensEval: Identifying and Categorizing Offensive Language in Social Media. The task was further divided into three sub-tasks: offensive language identification, automatic categorization of offen…

Language IdentificationTransfer Learning

HITMI&T at SemEval-2022 Task 4: Investigating Task-Adaptive Pretraining And Attention Mechanism On PCL Detection

2022-07-01 · SemEval (NAACL) 2022 7 · Zihang Liu, Yancheng He, Feiqing Zhuang, Bing Xu

This paper describes the system for the Semeval-2022 Task4 ”Patronizing and Condescending Language Detection”.An entity engages in Patronizing and Condescending Language(PCL) when its language use shows a superior attitu…

Multi Label Text ClassificationMulti-Label Text ClassificationSentencetext-classification+1

EMBEDDIA project: Cross-Lingual Embeddings for Less- Represented Languages in European News Media

2022-06-01 · EAMT 2022 6 · Senja Pollak, Andraž Pelicon

EMBEDDIA project developed a range of resources and methods for less-resourced EU languages, focusing on applications for media industry, including keyword extraction, comment moderation and article generation.

Keyword Extraction

EMBEDDIA Tools, Datasets and Challenges: Resources and Hackathon Contributions

2021-04-01 · EACL (Hackashop) 2021 4 · Senja Pollak, Marko Robnik-Šikonja, Matthew Purver, Michele Boggia 외

This paper presents tools and data sources collected and released by the EMBEDDIA project, supported by the European Union’s Horizon 2020 research and innovation program. The collected resources were offered to participa…

1Cademy at Semeval-2022 Task 1: Investigating the Effectiveness of Multilingual, Multitask, and Language-Agnostic Tricks for the Reverse Dictionary Task

2022-06-08 · SemEval (NAACL) 2022 7 · Zhiyong Wang, Ge Zhang, Nineli Lashkarashvili

This paper describes our system for the SemEval2022 task of matching dictionary glosses to word embeddings. We focus on the Reverse Dictionary Track of the competition, which maps multilingual glosses to reconstructed ve…

Reverse DictionaryWord Embeddings