paper-with-me

Papers

For the Purpose of Curry: A UD Treebank for Ashokan Prakrit

2021-11-24 · UDW (SyntaxFest) 2021 12 · Adam Farris, Aryaman Arora

We present the first linguistically annotated treebank of Ashokan Prakrit, an early Middle Indo-Aryan dialect continuum attested through Emperor Ashoka Maurya's 3rd century BCE rock and pillar edicts. For annotation, we used the multilingual Universal Dependencies (UD) formalism, following recent UD work on Sanskrit and other Indo-Aryan languages. We touch on some interesting linguistic features that posed issues in annotation: regnal names and other nominal compounds, "proto-ergative" participial constructions, and possible grammaticalizations evidenced by sandhi (phonological assimilation across morpheme boundaries). Eventually, we plan for a complete annotation of all attested Ashokan texts, towards the larger goals of improving UD coverage of different diachronic stages of Indo-Aryan and studying language change in Indo-Aryan using computational methods.

📄 PDF Abstract BibTeX arXiv:2111.12783

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

English-to-Prakrit Machine Translation via Multilingual Transfer Learning

2026-06-04 · Om Choksi, Smit Kareliya, Shrikant Malviya, Pruthwik Mishra arxiv

We study English-to-Prakrit machine translation in a low-resource setting where the target language is unsupported by IndicTrans2. We adapt the multilingual model by mapping Prakrit to the Hindi language tag (hin_Deva) w…

Machine TranslationTransfer Learning

Optical Character Recognition using Convolutional Neural Networks for Ashokan Brahmi Inscriptions

2024-12-29 · Yash Agrawal, Srinidhi Balasubramanian, Rahul Meena, Rohail Alam 외

This research paper delves into the development of an Optical Character Recognition (OCR) system for the recognition of Ashokan Brahmi characters using Convolutional Neural Networks. It utilizes a comprehensive dataset o…

Data AugmentationImage SegmentationOptical Character RecognitionOptical Character Recognition (OCR)+2

Prakriti200: A Questionnaire-Based Dataset of 200 Ayurvedic Prakriti Assessments

2025-10-05 · Aryan Kumar Singh, Janvi Singh arxiv

This dataset provides responses to a standardized, bilingual (English-Hindi) Prakriti Assessment Questionnaire designed to evaluate the physical, physiological, and psychological characteristics of individuals according …

Development of a General-Purpose Categorial Grammar Treebank

2020-05-01 · LREC 2020 5 · Yusuke Kubota, Koji Mineshima, Noritsugu Hayashi, Shinya Okano

This paper introduces ABC Treebank, a general-purpose categorial grammar (CG) treebank for Japanese. It is {`}general-purpose{'} in the sense that it is not tailored to a specific variant of CG, but rather aims to offer …

Semantic Parsing

Enhancing Ayurvedic Diagnosis using Multinomial Naive Bayes and K-modes Clustering: An Investigation into Prakriti Types and Dosha Overlapping

2023-10-04 · Pranav Bidve, Shalini Mishra, Annapurna J

The identification of Prakriti types for the human body is a long-lost medical practice in finding the harmony between the nature of human beings and their behaviour. There are 3 fundamental Prakriti types of individuals…

Clusteringfeature selection