paper-with-me

홈 › Papers

ArbEngVec : Arabic-English Cross-Lingual Word Embedding Model

2019-08-01 · WS 2019 8 · Raki Lachraf, El Moatez Billah Nagoudi, Youcef Ayachi, Ahmed Abdelali, Didier Schwab

Word Embeddings (WE) are getting increasingly popular and widely applied in many Natural Language Processing (NLP) applications due to their effectiveness in capturing semantic properties of words; Machine Translation (MT), Information Retrieval (IR) and Information Extraction (IE) are among such areas. In this paper, we propose an open source ArbEngVec which provides several Arabic-English cross-lingual word embedding models. To train our bilingual models, we use a large dataset with more than 93 million pairs of Arabic-English parallel sentences. In addition, we perform both extrinsic and intrinsic evaluations for the different word embedding model variants. The extrinsic evaluation assesses the performance of models on the cross-language Semantic Textual Similarity (STS), while the intrinsic evaluation is based on the Word Translation (WT) task.

📄 PDF Abstract BibTeX

Code (0)

등록된 구현이 없습니다.

Tasks

Information RetrievalMachine TranslationRetrievalSemantic Textual SimilaritySTSTranslationWord EmbeddingsWord Translation

Similar Papers 제목 키워드 기반

The Cross-Lingual Arabic Information REtrieval (CLAIRE) System

2021-07-29 · Zhizhong Chen, Carsten Eickhoff

Despite advances in neural machine translation, cross-lingual retrieval tasks in which queries and documents live in different natural language spaces remain challenging. Although neural translation models may provide an…

Information RetrievalMachine TranslationRetrievalTranslation

Code-switching Language Modeling With Bilingual Word Embeddings: A Case Study for Egyptian Arabic-English

2019-09-24 · Injy Hamed, Moritz Zhu, Mohamed Elmahdy, Slim Abdennadher 외

Code-switching (CS) is a widespread phenomenon among bilingual and multilingual societies. The lack of CS resources hinders the performance of many NLP tasks. In this work, we explore the potential use of bilingual word …

Language ModelingLanguage ModellingWord Embeddings

FII\_CROSS at SemEval-2021 Task 2: Multilingual and Cross-lingual Word-in-Context Disambiguation

2021-08-01 · SEMEVAL 2021 · Ciprian Bodnar, Andrada Tapuc, Cosmin Pintilie, Daniela Gifu 외

This paper presents a word-in-context disambiguation system. The task focuses on capturing the polysemous nature of words in a multilingual and cross-lingual setting, without considering a strict inventory of word meanin…

Task 2

Visual Grounding of Inter-lingual Word-Embeddings

2022-09-08 · Wafaa Mohammed, Hassan Shahmohammadi, Hendrik P. A. Lensch, R. Harald Baayen

Visual grounding of Language aims at enriching textual representations of language with multiple sources of visual knowledge such as images and videos. Although visual grounding is an area of intense research, inter-ling…

Visual GroundingWord EmbeddingsWord Similarity

Extracting Synonyms from Bilingual Dictionaries

2020-12-01 · EACL (GWC) 2021 1 · Mustafa Jarrar, Eman Karajah, Muhammad Khalifa, Khaled Shaalan

We present our progress in developing a novel algorithm to extract synonyms from bilingual dictionaries. Identification and usage of synonyms play a significant role in improving the performance of information access app…

Translation