Lexical Resources for Low-Resource PoS Tagging in Neural Times
More and more evidence is appearing that integrating symbolic lexical knowledge into neural models aids learning. This contrasts the widely-held belief that neural networks largely learn their own feature representations. For example, recent work has shows benefits of integrating lexicons to aid cross-lingual part-of-speech (PoS). However, little is known on how complementary such additional information is, and to what extent improvements depend on the coverage and quality of these external resources. This paper seeks to fill this gap by providing a thorough analysis on the contributions of lexical resources for cross-lingual PoS tagging in neural times.
Code (0)
등록된 구현이 없습니다.
Tasks
Cross-Lingual POS TaggingPOSPOS TaggingSimilar Papers 제목 키워드 기반
The Best of Both Worlds: Lexical Resources To Improve Low-Resource Part-of-Speech Tagging
In natural language processing, the deep learning revolution has shifted the focus from conventional hand-crafted symbolic representations to dense inputs, which are adequate representations learned automatically from co…
Cross-Lingual POS TaggingPart-Of-Speech TaggingPOSPOS TaggingLeveraging Lexical Resources and Constraint Grammar for Rule-Based Part-of-Speech Tagging in Welsh
Evaluating the Impact of External Lexical Resources into a CRF-based Multiword Segmenter and Part-of-Speech Tagger
This paper evaluates the impact of external lexical resources into a CRF-based joint Multiword Segmenter and Part-of-Speech Tagger. We especially show different ways of integrating lexicon-based features in the tagging m…
Named Entity Recognition (NER)Part-Of-Speech TaggingSemi-Automatic Data Annotation, POS Tagging and Mildly Context-Sensitive Disambiguation: the eXtended Revised AraMorph (XRAM)
An extended, revised form of Tim Buckwalter's Arabic lexical and morphological resource AraMorph, eXtended Revised AraMorph (henceforth XRAM), is presented which addresses a number of weaknesses and inconsistencies of th…
POSPOS TaggingUbiquitous Usage of a Broad Coverage French Corpus: Processing the Est Republicain corpus
In this paper, we introduce a set of resources that we have derived from the EST R{\'E}PUBLICAIN CORPUS, a large, freely-available collection of regional newspaper articles in French, totaling 150 million words. Our reso…
ArticlesDependency ParsingLemmatizationPart-Of-Speech Tagging