paper-with-me

Papers

Representing Etymology in the LiLa Knowledge Base of Linguistic Resources for Latin

2020-05-01 · LREC 2020 5 · Francesco Mambrini, Marco Passarotti

In this paper we describe the process of inclusion of etymological information in a knowledge base of interoperable Latin linguistic resources developed in the context of the LiLa: Linking Latin project. Interoperability is obtained by applying the Linked Open Data principles. Particularly, an extensive collection of Latin lemmas is used to link the (distributed) resources. For the etymology, we rely on the Ontolex-lemon ontology and the lemonEty extension to model the information, while the source data are taken from a recent etymological dictionary of Latin. As a result, the collection of lemmas LiLa is built around now includes 1,465 Proto-Italic and 1,393 Proto-Indo-European reconstructed forms that are used to explain the history of 1,400 Latin words. We discuss the motivation, methodology and modeling strategies of the work, as well as its possible applications and potential future developments.

📄 PDF Abstract BibTeX

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Linking the LASLA Corpus in the LiLa Knowledge Base of Interoperable Linguistic Resources for Latin

2022-06-01 · LDL (ACL) 2022 6 · Margherita Fantoli, Marco Passarotti, Francesco Mambrini, Giovanni Moretti 외

This paper describes the process of interlinking the 130 Classical Latin texts provided by an annotated corpus developed at the LASLA laboratory with the LiLa Knowledge Base, which makes linguistic resources for Latin in…

The Treatment of Word Formation in the LiLa Knowledge Base of Linguistic Resources for Latin

2019-09-01 · WS 2019 9 · Eleonora Litta, Marco Passarotti, Francesco Mambrini

Linked Open Treebanks. Interlinking Syntactically Annotated Corpora in the LiLa Knowledge Base of Linguistic Resources for Latin

2019-08-01 · WS 2019 8 · Francesco Mambrini, Marco Passarotti

The Index Thomisticus Treebank as Linked Data in the LiLa Knowledge Base

2022-06-01 · LREC 2022 6 · Francesco Mambrini, Marco Passarotti, Giovanni Moretti, Matteo Pellegrini

Although the Universal Dependencies initiative today allows for cross-linguistically consistent annotation of morphology and syntax in treebanks for several languages, syntactically annotated corpora are not yet interope…

DALILA: The Dialectal Arabic Linguistic Learning Assistant

2016-05-01 · LREC 2016 5 · Salam Khalifa, Houda Bouamor, Nizar Habash

Dialectal Arabic (DA) poses serious challenges for Natural Language Processing (NLP). The number and sophistication of tools and datasets in DA are very limited in comparison to Modern Standard Arabic (MSA) and other lan…