paper-with-me

Papers

Universal Joint Morph-Syntactic Processing: The Open University of Israel's Submission to The CoNLL 2017 Shared Task

2017-08-01 · CONLL 2017 8 · Amir More, Reut Tsarfaty

We present the Open University{'}s submission to the CoNLL 2017 Shared Task on multilingual parsing from raw text to Universal Dependencies. The core of our system is a joint morphological disambiguator and syntactic parser which accepts morphologically analyzed surface tokens as input and returns morphologically disambiguated dependency trees as output. Our parser requires a lattice as input, so we generate morphological analyses of surface tokens using a data-driven morphological analyzer that derives its lexicon from the UD training corpora, and we rely on UDPipe for sentence segmentation and surface-level tokenization. We report our official macro-average LAS is 56.56. Although our model is not as performant as many others, it does not make use of neural networks, therefore we do not rely on word embeddings or any other data source other than the corpora themselves. In addition, we show the utility of a lexicon-backed morphological analyzer for the MRL Modern Hebrew. We use our results on Modern Hebrew to argue that the UD community should define a UD-compatible standard for access to lexical resources, which we argue is crucial for MRLs and low resource languages in particular.

📄 PDF Abstract BibTeX

Code (0)

등록된 구현이 없습니다.

Tasks

MORPHSentenceSentence segmentationWord Embeddings

Similar Papers 제목 키워드 기반

Enhancing Korean Dependency Parsing with Morphosyntactic Features

2025-03-26 · Jungyeul Park, Yige Chen, Kyuwon Kim, Kyungtae Lim 외

This paper introduces UniDive for Korean, an integrated framework that bridges Universal Dependencies (UD) and Universal Morphology (UniMorph) to enhance the representation and processing of Korean {morphosyntax}. Korean…

DecoderDependency Parsing

Universal Morpho-Syntactic Parsing and the Contribution of Lexica: Analyzing the ONLP Lab Submission to the CoNLL 2018 Shared Task

2018-10-01 · CONLL 2018 10 · Amit Seker, Amir More, Reut Tsarfaty

We present the contribution of the ONLP lab at the Open University of Israel to the UD shared task on multilingual parsing from raw text to Universal Dependencies. Our contribution is based on a transition-based parser c…

Morphosyntactic Analysis for CHILDES

2024-07-17 · Houjun Liu, Brian MacWhinney

Language development researchers are interested in comparing the process of language learning across languages. Unfortunately, it has been difficult to construct a consistent quantitative framework for such comparisons. …

Automatic Speech Recognitionspeech-recognitionSpeech Recognition

Universal Dependencies v2: An Evergrowing Multilingual Treebank Collection

2020-04-22 · LREC 2020 5 · Joakim Nivre, Marie-Catherine de Marneffe, Filip Ginter, Jan Hajič 외

Universal Dependencies is an open community effort to create cross-linguistically consistent treebank annotation for many languages within a dependency-based lexicalist framework. The annotation consists in a linguistica…

Towards the Conversion of National Corpus of Polish to Universal Dependencies

2020-05-01 · LREC 2020 5 · Alina Wr{\'o}blewska

The research presented in this paper aims at enriching the manually morphosyntactically annotated part of National Corpus of Polish (NKJP1M) with a syntactic layer, i.e. dependency trees of sentences, and at converting b…