paper-with-me

홈 › Papers

The Best of Both Worlds: Lexical Resources To Improve Low-Resource Part-of-Speech Tagging

2018-11-21 · Barbara Plank, Sigrid Klerke, Zeljko Agic

In natural language processing, the deep learning revolution has shifted the focus from conventional hand-crafted symbolic representations to dense inputs, which are adequate representations learned automatically from corpora. However, particularly when working with low-resource languages, small amounts of symbolic lexical resources such as user-generated lexicons are often available even when gold-standard corpora are not. Such additional linguistic information is though often neglected, and recent neural approaches to cross-lingual tagging typically rely only on word and subword embeddings. While these representations are effective, our recent work has shown clear benefits of combining the best of both worlds: integrating conventional lexical information improves neural cross-lingual part-of-speech (PoS) tagging. However, little is known on how complementary such additional information is, and to what extent improvements depend on the coverage and quality of these external resources. This paper seeks to fill this gap by providing the first thorough analysis on the contributions of lexical resources for cross-lingual PoS tagging in neural times.

📄 PDF Abstract BibTeX arXiv:1811.08757

Code (0)

등록된 구현이 없습니다.

Tasks

Cross-Lingual POS TaggingPart-Of-Speech TaggingPOSPOS Tagging

Similar Papers 제목 키워드 기반

Best of Both Worlds: Making Word Sense Embeddings Interpretable

2016-05-01 · LREC 2016 5 · Alex Panchenko, er

Word sense embeddings represent a word sense as a low-dimensional numeric vector. While this representation is potentially useful for NLP applications, its interpretability is inherently limited. We propose a simple tech…

Lexicalization of Probabilistic Linear Context-free Rewriting Systems

2020-07-01 · WS 2020 7 · Richard M{\"o}rbitz, Thomas Ruprecht

In the field of constituent parsing, probabilistic grammar formalisms have been studied to model the syntactic structure of natural language. More recently, approaches utilizing neural models gained lots of traction in t…

Lexical Resources to Enrich English Malayalam Machine Translation

2016-05-01 · LREC 2016 5 · Sreelekha. S, Pushpak Bhattacharyya

In this paper we present our work on the usage of lexical resources for the Machine Translation English and Malayalam. We describe a comparative performance between different Statistical Machine Translation (SMT) systems…

Machine TranslationTranslation

XWalk: Random Walk Based Candidate Retrieval for Product Search

2023-07-22 · Jon Eskreis-Winkler, Yubin Kim, Andrew Stanton

In e-commerce, head queries account for the vast majority of gross merchandise sales and improvements to head queries are highly impactful to the business. While most supervised approaches to search perform better in hea…

Retrieval

Lexical Resources for Hindi Marathi MT

2017-03-04 · Sreelekha. S, Pushpak Bhattacharyya

In this paper we describe some ways to utilize various lexical resources to improve the quality of statistical machine translation system. We have augmented the training corpus with various lexical resources such as Indo…

Machine TranslationTranslation