Training Parsers on Incompatible Treebanks
Code (0)
등록된 구현이 없습니다.
Similar Papers 제목 키워드 기반
One model, two languages: training bilingual parsers with harmonized treebanks
We introduce an approach to train lexicalized parsers using bilingual corpora obtained by merging harmonized treebanks of different languages, producing parsers that can analyze sentences in either of the learned languag…
Vocal Bursts Valence PredictionCross-Domain Generalization of Neural Constituency Parsers
Neural parsers obtain state-of-the-art results on benchmark treebanks for constituency parsing -- but to what degree do they generalize to other domains? We present three results about the generalization of neural parser…
Constituency ParsingDomain GeneralizationCross-lingual Inflection as a Data Augmentation Method for Parsing
We propose a morphology-based method for low-resource (LR) dependency parsing. We train a morphological inflector for target LR languages, and apply it to related rich-resource (RR) treebanks to create cross-lingual (x-i…
Data AugmentationDependency ParsingScalable Cross-lingual Treebank Synthesis for Improved Production Dependency Parsers
We present scalable Universal Dependency (UD) treebank synthesis techniques that exploit advances in language representation modeling which leverage vast amounts of unlabeled general-purpose multilingual text. We introdu…
Data AugmentationTraining a Swedish Constituency Parser on Six Incompatible Treebanks
We investigate a transition-based parser that uses Eukalyptus, a function-tagged constituent treebank for Swedish which includes discontinuous constituents. In addition, we show that the accuracy of this parser can be im…