SandhiKosh: A Benchmark Corpus for Evaluating Sanskrit Sandhi Tools
Code (0)
등록된 구현이 없습니다.
Similar Papers 제목 키워드 기반
CharSS: Character-Level Transformer Model for Sanskrit Word Segmentation
Subword tokens in Indian languages inherently carry meaning, and isolating them can enhance NLP tasks, making sub-word segmentation a crucial process. Segmenting Sanskrit and other Indian languages into subtokens is not …
An Ontology for Comprehensive Tutoring of Euphonic Conjunctions of Sanskrit Grammar
Euphonic conjunctions (sandhis) form a very important aspect of Sanskrit morphology and phonology. The traditional and modern methods of studying about euphonic conjunctions in Sanskrit follow different methodologies. Th…
Computational Algorithms Based on the Paninian System to Process Euphonic Conjunctions for Word Searches
Searching for words in Sanskrit E-text is a problem that is accompanied by complexities introduced by features of Sanskrit such as euphonic conjunctions or sandhis. A word could occur in an E-text in a transformed form o…
Neural Compound-Word (Sandhi) Generation and Splitting in Sanskrit Language
This paper describes neural network based approaches to the process of the formation and splitting of word-compounding, respectively known as the Sandhi and Vichchhed, in Sanskrit language. Sandhi is an important idea es…
Morphological AnalysisSanskrit Sandhi Splitting using seq2(seq)^2
In Sanskrit, small words (morphemes) are combined to form compound words through a process known as Sandhi. Sandhi splitting is the process of splitting a given compound word into its constituent morphemes. Although rule…
Chinese Word SegmentationDecoder