paper-with-me

Papers

mahaNLP: A Marathi Natural Language Processing Library

2023-11-05 · Vidula Magdum, Omkar Dhekane, Sharayu Hiwarkhedkar, Saloni Mittal, Raviraj Joshi

We present mahaNLP, an open-source natural language processing (NLP) library specifically built for the Marathi language. It aims to enhance the support for the low-resource Indian language Marathi in the field of NLP. It is an easy-to-use, extensible, and modular toolkit for Marathi text analysis built on state-of-the-art MahaBERT-based transformer models. Our work holds significant importance as other existing Indic NLP libraries provide basic Marathi processing support and rely on older models with restricted performance. Our toolkit stands out by offering a comprehensive array of NLP tasks, encompassing both fundamental preprocessing tasks and advanced NLP tasks like sentiment analysis, NER, hate speech detection, and sentence completion. This paper focuses on an overview of the mahaNLP framework, its features, and its usage. This work is a part of the L3Cube MahaNLP initiative, more information about it can be found at https://github.com/l3cube-pune/MarathiNLP .

📄 PDF Abstract BibTeX arXiv:2311.02579

Code (1)

l3cube-pune/MarathiNLP 공식 구현

Tasks

Hate Speech DetectionNERSentenceSentence CompletionSentiment Analysis

Methods 이 논문이 사용한 방법론

Library 설명 없음

Similar Papers 제목 키워드 기반

L3Cube-MahaNLP: Marathi Natural Language Processing Datasets, Models, and Library

2022-05-29 · Raviraj Joshi

Despite being the third most popular language in India, the Marathi language lacks useful NLP resources. Moreover, popular NLP libraries do not have support for the Marathi language. With L3Cube-MahaNLP, we aim to build …

Hate Speech DetectionLanguage ModelingLanguage Modellingnamed-entity-recognition+3

Curating Stopwords in Marathi: A TF-IDF Approach for Improved Text Analysis and Information Retrieval

2024-06-16 · Rohan Chavan, Gaurav Patil, Vishal Madle, Raviraj Joshi

Stopwords are commonly used words in a language that are often considered to be of little value in determining the meaning or significance of a document. These words occur frequently in most texts and don't provide much …

Information RetrievalRetrievalSentiment Analysistext-classification+1

Experimental Evaluation of Deep Learning models for Marathi Text Classification

2021-01-13 · Atharva Kulkarni, Meet Mandhane, Manali Likhitkar, Gayatri Kshirsagar 외

The Marathi language is one of the prominent languages used in India. It is predominantly spoken by the people of Maharashtra. Over the past decade, the usage of language on online platforms has tremendously increased. H…

ClassificationDeep LearningGeneral Classificationtext-classification+2

A Review of the Marathi Natural Language Processing

2024-12-20 · Asang Dani, Shailesh R Sathe

Marathi is one of the most widely used languages in the world. One might expect that the latest advances in NLP research in languages like English reach such a large community. However, NLP advancements in English didn't…

Diversity

Long Range Named Entity Recognition for Marathi Documents

2024-10-11 · Pranita Deshmukh, Nikita Kulkarni, Sanhita Kulkarni, Kareena Manghani 외

The demand for sophisticated natural language processing (NLP) methods, particularly Named Entity Recognition (NER), has increased due to the exponential growth of Marathi-language digital content. In particular, NER is …

named-entity-recognitionNamed Entity RecognitionNamed Entity Recognition (NER)NER