Sub-label dependencies for Neural Morphological Tagging -- The Joint Submission of University of Colorado and University of Helsinki for VarDial 2018
This paper presents the submission of the UH{\&}CU team (Joint University of Colorado and University of Helsinki team) for the VarDial 2018 shared task on morphosyntactic tagging of Croatian, Slovenian and Serbian tweets. Our system is a bidirectional LSTM tagger which emits tags as character sequences using an LSTM generator in order to be able to handle unknown tags and combinations of several tags for one token which occur in the shared task data sets. To the best of our knowledge, using an LSTM generator is a novel approach. The system delivers sizable improvements of more than 6{\%}-points over a baseline trigram tagger. Overall, the performance of our system is quite even for all three languages.
Code (0)
등록된 구현이 없습니다.
Tasks
Morphological TaggingWord EmbeddingsMethods 이 논문이 사용한 방법론
Similar Papers 제목 키워드 기반
IBM Research at the CoNLL 2018 Shared Task on Multilingual Parsing
This paper presents the IBM Research AI submission to the CoNLL 2018 Shared Task on Parsing Universal Dependencies. Our system implements a new joint transition-based parser, based on the Stack-LSTM framework and the Arc…
ARCDependency ParsingMorphological TaggingPart-Of-Speech Tagging+2Multi-Team: A Multi-attention, Multi-decoder Approach to Morphological Analysis.
This paper describes our submission to SIGMORPHON 2019 Task 2: Morphological analysis and lemmatization in context. Our model is a multi-task sequence to sequence neural network, which jointly learns morphological taggin…
DecoderLEMMALemmatizationMorphological Analysis+2Learning the Structure of Variable-Order CRFs: a finite-state perspective
The computational complexity of linear-chain Conditional Random Fields (CRFs) makes it difficult to deal with very large label sets and long range dependencies. Such situations are not rare and arise when dealing with mo…
Chunkingfeature selectionNamed Entity Recognition (NER)Part-Of-Speech TaggingAn ELECTRA Model for Latin Token Tagging Tasks
This report describes the KU Leuven / Brepols-CTLO submission to EvaLatin 2022. We present the results of our current small Latin ELECTRA model, which will be expanded to a larger model in the future. For the lemmatizati…
LEMMALemmatizationmodelMorphological Tagging+2Towards JointUD: Part-of-speech Tagging and Lemmatization using Recurrent Neural Networks
This paper describes our submission to CoNLL 2018 UD Shared Task. We have extended an LSTM-based neural network designed for sequence tagging to additionally generate character-level sequences. The network was jointly tr…
Dependency ParsingLemmatizationPart-Of-Speech TaggingSentence+1