Unifying Discourse Resources with Dependency Framework
For text-level discourse analysis, there are various discourse schemes but relatively few labeled data, because discourse research is still immature and it is labor-intensive to annotate the inner logic of a text. In this paper, we attempt to unify multiple Chinese discourse corpora under different annotation schemes with discourse dependency framework by designing semi-automatic methods to convert them into dependency structures. We also implement several benchmark dependency parsers and research on how they can leverage the unified data to improve performance.
Code (1)
Similar Papers 제목 키워드 기반
From Interoperable Annotations towards Interoperable Resources: A Multilingual Approach to the Analysis of Discourse
In the present paper, we analyse variation of discourse phenomena in two typologically different languages, i.e. in German and Czech. The novelty of our approach lies in the nature of the resources we are using. Advantag…
Machine TranslationSemantic SimilaritySemantic Textual SimilarityTranslationA Novel Dependency Framework for Enhancing Discourse Data Analysis
The development of different theories of discourse structure has led to the establishment of discourse corpora based on these theories. However, the existence of discourse corpora established on different theoretical bas…
Towards Unification of Discourse Annotation Frameworks
Discourse information is difficult to represent and annotate. Among the major frameworks for annotating discourse information, RST, PDTB and SDRT are widely discussed and used, each having its own theoretical foundation …
Multi-Task LearningA Dependency Perspective on RST Discourse Parsing and Evaluation
Computational text-level discourse analysis mostly happens within Rhetorical Structure Theory (RST), whose structures have classically been presented as constituency trees, and relies on data from the RST Discourse Treeb…
Constituency ParsingDependency ParsingDiscourse ParsingSciDTB: Discourse Dependency TreeBank for Scientific Abstracts
Annotation corpus for discourse relations benefits NLP tasks such as machine translation and question answering. In this paper, we present SciDTB, a domain-specific discourse treebank annotated on scientific articles. Di…
ArticlesMachine TranslationQuestion AnsweringTranslation