paper-with-me

Papers

Standardisation and Interoperation of Morphosyntactic and Syntactic Annotation Tools for Spanish and their Annotations

2014-05-01 · LREC 2014 5 · Antonio Pareja-Lora, Guillermo C{\'a}rcamo-Escorza, Alicia Ballesteros-Calvo

Linguistic annotation tools and linguistic annotations are scarcely syntactically and/or semantically interoperable. Their low interoperability usually results from the number of factors taken into account in their development and design. These include (i) the type of phenomena annotated (either morphosyntactic, syntactic, semantic, etc.); (ii) how these phenomena are annotated (e.g., the particular guidelines and/or schema used to encode the annotations); and (iii) the languages (Java, C++, etc.) and technologies (as standalone programs, as APIs, as web services, etc.) used to develop them. This low level of interoperability makes it difficult to reuse both the linguistic annotation tools and their annotations in new scenarios, e.g., in natural language processing (NLP) pipelines. In spite of this, developing new linguistic tools from scratch is quite a high time-consuming task that also entails a very high cost. Therefore, cost-effective ways to systematically reuse linguistic tools and annotations must be found urgently. A traditional way to overcome reuse and/or interoperability problems is standardisation. In this paper, we present a web service version of FreeLing that provides standard-compliant morpho-syntactic and syntactic annotations for Spanish, according to several ISO linguistic annotation standards and standard drafts.

📄 PDF Abstract BibTeX

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

TalkTag: Fine-Grained Morphosyntactic Error Annotation for Transcribed Speech

2026-06-01 · Shamira Venturini, Oliver Hennhöfer, Steffen Kinkel, Jannik Strötgen arxiv

Fine-grained morphosyntactic error annotation is important in clinical and developmental language research, yet it is labour-intensive, expert-dependent, and difficult to scale. We present TalkTag, an LLM-based lightweig…

Annotation of Clinical Narratives in Bulgarian language

2017-09-01 · RANLP 2017 9 · Ivajlo Radev, Kiril Simov, Galia Angelova, Svetla Boytcheva

In this paper we describe annotation process of clinical texts with morphosyntactic and semantic information. The corpus contains 1,300 discharge letters in Bulgarian language for patients with Endocrinology and Metaboli…

ChunkingDependency ParsingInformation Retrieval

What does Neural Bring? Analysing Improvements in Morphosyntactic Annotation and Lemmatisation of Slovenian, Croatian and Serbian

2019-08-01 · WS 2019 8 · Nikola Ljube{\v{s}}i{\'c}, Kaja Dobrovoljc

We present experiments on Slovenian, Croatian and Serbian morphosyntactic annotation and lemmatisation between the former state-of-the-art for these three languages and one of the best performing systems at the CoNLL 201…

Word Embeddings

Combining Ontologies and Neural Networks for Analyzing Historical Language Varieties. A Case Study in Middle Low German

2016-05-01 · LREC 2016 5 · Maria Sukhareva, Christian Chiarcos

In this paper, we describe experiments on the morphosyntactic annotation of historical language varieties for the example of Middle Low German (MLG), the official language of the German Hanse during the Middle Ages and a…

POS

Fine-grained Morphosyntactic Analysis and Generation Tools for More Than One Thousand Languages

2020-05-01 · LREC 2020 5 · Garrett Nicolai, Dylan Lewis, Arya D. McCarthy, Aaron Mueller 외

Exploiting the broad translation of the Bible into the world{'}s languages, we train and distribute morphosyntactic tools for approximately one thousand languages, vastly outstripping previous distributions of tools devo…

RerankingTranslation