paper-with-me

홈 › Papers

Improving Neural Machine Translation for Sanskrit-English

2020-12-01 · ICON 2020 12 · Ravneet Punia, Aditya Sharma, Sarthak Pruthi, Minni Jain

Sanskrit is one of the oldest languages of the Asian Subcontinent that fell out of common usage around 600 B.C. In this paper, we attempt to translate Sanskrit to English using Neural Machine Translation approaches based on Reinforcement Learning and Transfer learning that were never tried and tested on Sanskrit. Along with the paper, we also release monolingual Sanskrit and parallel aligned Sanskrit-English corpora for the research community. Our methodologies outperform the previous approaches applied to Sanskrit by various re- searchers and will further help the linguistic community to accelerate the costly and time consuming manual translation process.

📄 PDF Abstract BibTeX

Code (0)

등록된 구현이 없습니다.

Tasks

Machine Translationreinforcement-learningReinforcement Learning (RL)Transfer LearningTranslation

Similar Papers 제목 키워드 기반

An Augmented Translation Technique for low Resource language pair: Sanskrit to Hindi translation

2020-06-09 · Rashi Kumar, Piyush Jha, Vineet Sahula

Neural Machine Translation (NMT) is an ongoing technique for Machine Translation (MT) using enormous artificial neural network. It has exhibited promising outcomes and has shown incredible potential in solving challengin…

Dimensionality ReductionMachine TranslationNMTTranslation

Sāmayik: A Benchmark and Dataset for English-Sanskrit Translation

2023-05-23 · Ayush Maheshwari, Ashim Gupta, Amrith Krishna, Atul Kumar Singh 외

We release S\={a}mayik, a dataset of around 53,000 parallel English-Sanskrit sentences, written in contemporary prose. Sanskrit is a classical language still in sustenance and has a rich documented heritage. However, due…

Machine TranslationTranslation

Anveshana: A New Benchmark Dataset for Cross-Lingual Information Retrieval On English Queries and Sanskrit Documents

2025-05-26 · Manoj Balaji Jagadeeshan, Prince Raj, Pawan Goyal

The study presents a comprehensive benchmark for retrieving Sanskrit documents using English queries, focusing on the chapters of the Srimadbhagavatam. It employs a tripartite approach: Direct Retrieval (DR), Translation…

Cross-Lingual Information RetrievalInformation RetrievalRAGRetrieval+1

Itihasa: A large-scale corpus for Sanskrit to English translation

2021-06-06 · ACL (WAT) 2021 8 · Rahul Aralikatte, Miryam de Lhoneux, Anoop Kunchukuttan, Anders Søgaard

This work introduces Itihasa, a large-scale translation dataset containing 93,000 pairs of Sanskrit shlokas and their English translations. The shlokas are extracted from two Indian epics viz., The Ramayana and The Mahab…

Machine TranslationTranslation

Mitrasamgraha: A Comprehensive Classical Sanskrit Machine Translation Dataset

2026-01-12 · Sebastian Nehrdich, David Allport, Sven Sellmer, Jivnesh Sandhan 외 arxiv

While machine translation is regarded as a "solved problem" for many high-resource languages, close analysis quickly reveals that this is not the case for content that shows challenges such as poetic language, philosophi…

Machine Translation