paper-with-me

Papers

Estimating related words computationally using language model from the Mahabharata - an Indian epic

2023-05-09 · Vrunda Gadesha, Keyur D Joshi, Shefali Naik

'Mahabharata' is the most popular among many Indian pieces of literature referred to in many domains for completely different purposes. This text itself is having various dimension and aspects which is useful for the human being in their personal life and professional life. This Indian Epic is originally written in the Sanskrit Language. Now in the era of Natural Language Processing, Artificial Intelligence, Machine Learning, and Human-Computer interaction this text can be processed according to the domain requirement. It is interesting to process this text and get useful insights from Mahabharata. The limitation of the humans while analyzing Mahabharata is that they always have a sentiment aspect towards the story narrated by the author. Apart from that, the human cannot memorize statistical or computational details, like which two words are frequently coming in one sentence? What is the average length of the sentences across the whole literature? Which word is the most popular word across the text, what are the lemmas of the words used across the sentences? Thus, in this paper, we propose an NLP pipeline to get some statistical and computational insights along with the most relevant word searching method from the largest epic 'Mahabharata'. We stacked the different text-processing approaches to articulate the best results which can be further used in the various domain where Mahabharata needs to be referred.

📄 PDF Abstract BibTeX arXiv:2305.05420

Code (0)

등록된 구현이 없습니다.

Tasks

Language ModelingLanguage ModellingSentence

Similar Papers 제목 키워드 기반

A Computational Analysis of Mahabharata

2016-12-01 · WS 2016 12 · Debarati Das, Bhaskarjyoti Das, Kavi Mahesh
Emotion Recognition

A Deep Dive into Identification of Characters from Mahabharata

2017-12-01 · WS 2017 12 · Apurba Paul, Dipankar Das

Measuring cross-language intelligibility between Romance languages with computational tools

2026-02-07 · Liviu P Dinu, Ana Sabina Uban, Bogdan Iordache, Anca Dinu 외 arxiv

We present an analysis of mutual intelligibility in related languages applied for languages in the Romance family. We introduce a novel computational metric for estimating intelligibility based on lexical similarity usin…

Semantic Similarity

Estimating senses with sets of lexically related words for Polish word sense disambiguation

2019-07-01 · GWC 2019 7 · Szymon Rutkowski, Piotr Rychlik, Agnieszka Mykowiecka

We propose a new algorithm for word sense disambiguation, exploiting data from a WordNet with many types of lexical relations, such as plWordNet for Polish. In this method, sense probabilities in context are approximated…

Language ModelingLanguage ModellingWord Sense Disambiguation

Streaming word similarity mining on the cheap

2018-10-01 · EMNLP 2018 10 · Olof G{\"o}rnerup, Daniel Gillblad

Accurately and efficiently estimating word similarities from text is fundamental in natural language processing. In this paper, we propose a fast and lightweight method for estimating similarities from streams by explici…

Document ClassificationWord AlignmentWord EmbeddingsWord Similarity