Genres in the Prague Discourse Treebank
We present the project of classification of Prague Discourse Treebank documents (Czech journalistic texts) for their genres. Our main interest lies in opening the possibility to observe how text coherence is realized in different types (in the genre sense) of language data and, in the future, in exploring the ways of using genres as a feature for multi-sentence-level language technologies. In the paper, we first describe the motivation and the concept of the genre annotation, and briefly introduce the Prague Discourse Treebank. Then, we elaborate on the process of manual annotation of genres in the treebank, from the annotators{'} manual work to post-annotation checks and to the inter-annotator agreement measurements. The annotated genres are subsequently analyzed together with discourse relations (already annotated in the treebank) ― we present distributions of the annotated genres and results of studying distinctions of distributions of discourse relations across the individual genres.
Code (0)
등록된 구현이 없습니다.
Tasks
SentenceSentiment AnalysisSimilar Papers 제목 키워드 기반
Introducing the Prague Discourse Treebank 1.0
Discourse Relations in the Prague Dependency Treebank 3.0
Use of Coreference in Automatic Searching for Multiword Discourse Markers in the Prague Dependency Treebank
Prague Dependency Treebank - Consolidated 1.0
We present a richly annotated and genre-diversified language resource, the Prague Dependency Treebank-Consolidated 1.0 (PDT-C 1.0), the purpose of which is - as it always been the case for the family of the Prague Depend…
ArticlesDiversityTranslationPrague Dependency Treebank -- Consolidated 1.0
We present a richly annotated and genre-diversified language resource, the Prague Dependency Treebank-Consolidated 1.0 (PDT-C 1.0), the purpose of which is - as it always been the case for the family of the Prague Depend…
ArticlesDiversityTranslation