paper-with-me

Papers

Preparing Korean Data for the Shared Task on Parsing Morphologically Rich Languages

2013-09-06 · Jinho D. Choi

This document gives a brief description of Korean data prepared for the SPMRL 2013 shared task. A total of 27,363 sentences with 350,090 tokens are used for the shared task. All constituent trees are collected from the KAIST Treebank and transformed to the Penn Treebank style. All dependency trees are converted from the transformed constituent trees using heuristics and labeling rules de- signed specifically for the KAIST Treebank. In addition to the gold-standard morphological analysis provided by the KAIST Treebank, two sets of automatic morphological analysis are provided for the shared task, one is generated by the HanNanum morphological analyzer, and the other is generated by the Sejong morphological analyzer.

📄 PDF Abstract BibTeX arXiv:1309.1649

Code (0)

등록된 구현이 없습니다.

Tasks

Morphological Analysis

Similar Papers 제목 키워드 기반

Representing and Parsing Korean Constituency Structure at Different Levels of Granularity

2026-08-27 · Jungyeul Park, KyungTae Lim, Zihao Huang, Eunkyul Leah Jo 외 arxiv

Korean constituency parsing raises a representational challenge because the terminal units of a phrase-structure tree do not straightforwardly correspond to simple surface words. Korean eojeols are morphologically comple…

Constituency Parsing

Constituency Structure over Eojeol in Korean Treebanks

2025-12-27 · Jungyeul Park, Chulwoo Park arxiv

The design of Korean constituency treebanks raises a central representational question concerning the choice of terminal units. Although Korean words are morphologically complex, treating morphemes as constituency termin…

Constituency Parsing

Yet Another Format of Universal Dependencies for Korean

2022-09-20 · COLING 2022 10 · Yige Chen, Eunkyul Leah Jo, Yundong Yao, Kyungtae Lim 외

In this study, we propose a morpheme-based scheme for Korean dependency parsing and adopt the proposed scheme to Universal Dependencies. We present the linguistic rationale that illustrates the motivation and the necessi…

Dependency Parsing

Analysis of the Penn Korean Universal Dependency Treebank (PKT-UD): Manual Revision to Build Robust Parsing Model in Korean

2020-05-26 · WS 2020 7 · Tae Hwan Oh, Ji Yoon Han, Hyonsu Choe, Seokwon Park 외

In this paper, we first open on important issues regarding the Penn Korean Universal Treebank (PKT-UD) and address these issues by revising the entire corpus manually with the aim of producing cleaner UD annotations that…

Enhancing Korean Dependency Parsing with Morphosyntactic Features

2025-03-26 · Jungyeul Park, Yige Chen, Kyuwon Kim, Kyungtae Lim 외

This paper introduces UniDive for Korean, an integrated framework that bridges Universal Dependencies (UD) and Universal Morphology (UniMorph) to enhance the representation and processing of Korean {morphosyntax}. Korean…

DecoderDependency Parsing