Thomas Aquinas in the T\"uNDRA: Integrating the Index Thomisticus Treebank into CLARIN-D
This paper describes the integration of the Index Thomisticus Treebank (IT-TB) into the web-based treebank search and visualization application TueNDRA (Tuebingen aNnotated Data Retrieval {\&} Analysis). TueNDRA was originally designed to provide access via the Internet to constituency treebanks and to tools for searching and visualizing them, as well as tabulating statistics about their contents. TueNDRA has now been extended to also provide full support for dependency treebanks with non-projective dependencies, in order to integrate the IT-TB and future treebanks with similar properties. These treebanks are queried using an adapted form of the TIGERSearch query language, which can search both hierarchical and sequential information in treebanks in a single query. As a web application, making the IT-TB accessible via TueNDRA makes the treebank and the tools to use of it available to a large community without having to distribute software and show users how to install it.
Code (0)
등록된 구현이 없습니다.
Tasks
RetrievalSimilar Papers 제목 키워드 기반
Challenges in Converting the Index Thomisticus Treebank into Universal Dependencies
This paper describes the changes applied to the original process used to convert the \textit{Index Thomisticus} Treebank, a corpus including texts in Medieval Latin by Thomas Aquinas, into the annotation style of Univers…
Dependency ParsingPOSPOS TaggingDifferentia compositionem facit. A Slower-Paced and Reliable Parser for Latin
The Index Thomisticus Treebank is the largest available treebank for Latin; it contains Medieval Latin texts by Thomas Aquinas. After experimenting on its data with a number of dependency parsers based on different super…
DiversityThe Index Thomisticus Treebank as Linked Data in the LiLa Knowledge Base
Although the Universal Dependencies initiative today allows for cross-linguistically consistent annotation of morphology and syntax in treebanks for several languages, syntactically annotated corpora are not yet interope…
THIVLVC: Retrieval Augmented Dependency Parsing for Latin
We describe THIVLVC, a two-stage system for the EvaLatin 2026 Dependency Parsing task. Given a Latin sentence, we retrieve structurally similar entries from the CIRCSE treebank using sentence length and POS n-gram simila…
Dependency ParsingUsing Artificial Intelligence to Unlock Crowdfunding Success for Small Businesses
While small businesses are increasingly turning to online crowdfunding platforms for essential funding, over 40% of these campaigns may fail to raise any money, especially those from low socio-economic areas. We utilize …
Language ModelingLanguage ModellingLarge Language Model