Survey of Pseudonymization, Abstractive Summarization & Spell Checker for Hindi and Marathi
India's vast linguistic diversity presents unique challenges and opportunities for technological advancement, especially in the realm of Natural Language Processing (NLP). While there has been significant progress in NLP applications for widely spoken languages, the regional languages of India, such as Marathi and Hindi, remain underserved. Research in the field of NLP for Indian regional languages is at a formative stage and holds immense significance. The paper aims to build a platform which enables the user to use various features like text anonymization, abstractive text summarization and spell checking in English, Hindi and Marathi language. The aim of these tools is to serve enterprise and consumer clients who predominantly use Indian Regional Languages.
Code (0)
등록된 구현이 없습니다.
Tasks
Abstractive Text SummarizationDiversityText AnonymizationText SummarizationSimilar Papers 제목 키워드 기반
Neural spell-checker: Beyond words with synthetic data generation
Spell-checkers are valuable tools that enhance communication by identifying misspelled words in written texts. Recent improvements in deep learning, and in particular in large language models, have opened new opportuniti…
Language ModelingLanguage ModellingSynthetic Data GenerationKidSpell: A Child-Oriented, Rule-Based, Phonetic Spellchecker
For help with their spelling errors, children often turn to spellcheckers integrated in software applications like word processors and search engines. However, existing spellcheckers are usually tuned to the needs of tra…
A Language Model for Spell Checking of Educational Texts in Kurdish (Sorani)
Spell checkers are an integrated feature of most software applications handling text inputs. When we write an email or compile a report on a desktop or a smartphone editor, a spell checker could be activated that assists…
Language ModelingLanguage ModellingA Survey on Neural Abstractive Summarization Methods and Factual Consistency of Summarization
Automatic summarization is the process of shortening a set of textual data computationally, to create a subset (a summary) that represents the most important pieces of information in the original text. Existing summariza…
Abstractive Text SummarizationKannada Spell Checker with Sandhi Splitter
Spelling errors are introduced in text either during typing, or when the user does not know the correct phoneme or grapheme. If a language contains complex words like sandhi where two or more morphemes join based on some…
Machine Translation