paper-with-me

Papers

The Treebank of Vedic Sanskrit

2020-05-01 · LREC 2020 5 · Oliver Hellwig, Salvatore Scarlata, Elia Ackermann, Paul Widmer

This paper introduces the first treebank of Vedic Sanskrit, a morphologically rich ancient Indian language that is of central importance for linguistic and historical research. The selection of the more than 3,700 sentences contained in this treebank reflects the development of metrical and prose texts over a period of 600 years. We discuss how these sentences are annotated in the Universal Dependencies scheme and which syntactic constructions required special attention. In addition, we describe a syntactic labeler based on neural networks that supports the initial annotation of the treebank, and whose evaluation can be helpful for setting up a full syntactic parser of Vedic Sanskrit.

📄 PDF Abstract BibTeX

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Vedavani: A Benchmark Corpus for ASR on Vedic Sanskrit Poetry

2025-05-30 · Sujeet Kumar, Pretam Ray, Abhinay Beerukuri, Shrey Kamoji 외

Sanskrit, an ancient language with a rich linguistic heritage, presents unique challenges for automatic speech recognition (ASR) due to its phonemic complexity and the phonetic transformations that occur at word juncture…

Automatic Speech RecognitionAutomatic Speech Recognition (ASR)speech-recognitionSpeech Recognition

Accent Placement Models for Rigvedic Sanskrit Text

2025-11-28 · Akhil Rajeev P, Annarao Kulkarni arxiv

The Rigveda, among the oldest Indian texts in Vedic Sanskrit, employs a distinctive pitch-accent system : udātta, anudātta, svarita whose marks encode melodic and interpretive cues but are often absent from modern e-text…

parameter-efficient fine-tuning

Annotating “Absolute” Preverbs in the Homeric and Vedic Treebanks

2022-06-01 · LT4HALA (LREC) 2022 6 · Luca Brigada Villa, Erica Biagetti, Chiara Zanchi

Indo-European preverbs are uninflected morphemes attaching to verbs and modifying their meaning. In Early Vedic and Homeric Greek, these morphemes held ambiguous morphosyntactic status raising issues for syntactic annota…

Position

Detecting Diachronic Syntactic Developments in Presence of Bias Terms

2022-06-01 · LT4HALA (LREC) 2022 6 · Oliver Hellwig, Sven Sellmer

Corpus-based studies of diachronic syntactic changes are typically guided by the results of previous qualitative research. When such results are missing or, as is the case for Vedic Sanskrit, are restricted to small part…

One Model is All You Need: ByT5-Sanskrit, a Unified Model for Sanskrit NLP Tasks

2024-09-20 · Sebastian Nehrdich, Oliver Hellwig, Kurt Keutzer

Morphologically rich languages are notoriously challenging to process for downstream NLP applications. This paper presents a new pretrained language model, ByT5-Sanskrit, designed for NLP applications involving the morph…

AllDependency ParsingInformation RetrievalLanguage Modelling+4