From Interoperable Annotations towards Interoperable Resources: A Multilingual Approach to the Analysis of Discourse
In the present paper, we analyse variation of discourse phenomena in two typologically different languages, i.e. in German and Czech. The novelty of our approach lies in the nature of the resources we are using. Advantage is taken of existing resources, which are, however, annotated on the basis of two different frameworks. We use an interoperable scheme unifying discourse phenomena in both frameworks into more abstract categories and considering only those phenomena that have a direct match in German and Czech. The discourse properties we focus on are relations of identity, semantic similarity, ellipsis and discourse relations. Our study shows that the application of interoperable schemes allows an exploitation of discourse-related phenomena analysed in different projects and on the basis of different frameworks. As corpus compilation and annotation is a time-consuming task, positive results of this experiment open up new paths for contrastive linguistics, translation studies and NLP, including machine translation.
Code (0)
등록된 구현이 없습니다.
Tasks
Machine TranslationSemantic SimilaritySemantic Textual SimilarityTranslationSimilar Papers 제목 키워드 기반
Towards Creating Interoperable Resources for Conceptual Annotation of Multilingual Domain Corpora
In this paper we focus on creation of interoperable annotation resources that make up a significant proportion of an on-going project on the development of conceptually annotated multilingual corpora for the domain of te…
Machine TranslationTranslationA Multilingual Predicate Matrix
This paper presents the Predicate Matrix 1.3, a lexical resource resulting from the integration of multiple sources of predicate information including FrameNet, VerbNet, PropBank and WordNet. This new version of the Pred…
Extending an interoperable platform to facilitate the creation of multilingual and multimodal NLP applications
The Multilingual Corpus of Survey Questionnaires Query Interface
The dawn of the digital age led to increasing demands for digital research resources, which shall be quickly processed and handled by computers. Due to the amount of data created by this digitization process, the design …
ManagementSurveyTranslationLinking the LASLA Corpus in the LiLa Knowledge Base of Interoperable Linguistic Resources for Latin
This paper describes the process of interlinking the 130 Classical Latin texts provided by an annotated corpus developed at the LASLA laboratory with the LiLa Knowledge Base, which makes linguistic resources for Latin in…