paper-with-me

Papers

Automatic Verification and Augmentation of Multilingual Lexicons

2016-12-01 · WS 2016 12 · Maryam Aminian, Mohamed Al-Badrashiny, Mona Diab

We present an approach for automatic verification and augmentation of multilingual lexica. We exploit existing parallel and monolingual corpora to extract multilingual correspondents via tri-angulation. We demonstrate the efficacy of our approach on two publicly available resources: Tharwa, a three-way lexicon comprising Dialectal Arabic, Modern Standard Arabic and English lemmas among other information (Diab et al., 2014); and BabelNet, a multilingual thesaurus comprising over 276 languages including Arabic variant entries (Navigli and Ponzetto, 2012). Our automated approach yields an F1-score of 71.71{\%} in generating correct multilingual correspondents against gold Tharwa, and 54.46{\%} against gold BabelNet without any human intervention.

📄 PDF Abstract BibTeX

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Language Lexicons for Hindi-English Multilingual Text Processing

2021-06-29 · Mohd Zeeshan Ansari, Tanvir Ahmad, Noaima Bari

Language Identification in textual documents is the process of automatically detecting the language contained in a document based on its content. The present Language Identification techniques presume that a document con…

Language Identification

Lexical Coverage Evaluation of Large-scale Multilingual Semantic Lexicons for Twelve Languages

2016-05-01 · LREC 2016 5 · Scott Piao, Paul Rayson, Dawn Archer, Francesca Bianchi 외

The last two decades have seen the development of various semantic lexical resources such as WordNet (Miller, 1995) and the USAS semantic lexicon (Rayson et al., 2004), which have played an important role in the areas of…

TED-MDB Lexicons: Tr-EnConnLex, Pt-EnConnLex

2020-11-01 · EMNLP (CODI) 2020 11 · Murathan Kurfali, Sibel Ozer, Deniz Zeyrek, Amália Mendes

In this work, we present two new bilingual discourse connective lexicons, namely, for Turkish-English and European Portuguese-English created automatically using the existing discourse relation-aligned TED-MDB corpus. In…

Relation

Learning Sentiment Lexicons in Spanish

2012-05-01 · LREC 2012 5 · Ver{\'o}nica P{\'e}rez-Rosas, Carmen Banea, Rada Mihalcea

In this paper we present a framework to derive sentiment lexicons in a target language by using manually or automatically annotated data available in an electronic resource rich language, such as English. We show that br…

Opinion MiningQuestion AnsweringSentiment AnalysisSpeech Synthesis+2

A Multilingual BPE Embedding Space for Universal Sentiment Lexicon Induction

2019-07-01 · ACL 2019 7 · Mengjie Zhao, Hinrich Sch{\"u}tze

We present a new method for sentiment lexicon induction that is designed to be applicable to the entire range of typological diversity of the world{'}s languages. We evaluate our method on Parallel Bible Corpus+ (PBC+), …

DiversityDomain Adaptation