IndoNet: A Multilingual Lexical Knowledge Network for Indian Languages
Code (0)
등록된 구현이 없습니다.
Similar Papers 제목 키워드 기반
Efficient Multilingual Text Classification for Indian Languages
India is one of the richest language hubs on the earth and is very diverse and multilingual. But apart from a few Indian languages, most of them are still considered to be resource poor. Since most of the NLP techniques …
ClassificationMultilingual text classificationtext-classificationText Classification+1IndoUKC: A Concept-Centered Indian Multilingual Lexical Resource
We introduce the IndoUKC, a new multilingual lexical database comprised of eighteen Indian languages, with a focus on formally capturing words and word meanings specific to Indian languages and cultures. The IndoUKC reus…
DiversityThat'll Do Fine!: A Coarse Lexical Resource for English-Hindi MT, Using Polylingual Topic Models
Parallel corpora are often injected with bilingual lexical resources for improved Indian language machine translation (MT). In absence of such lexical resources, multilingual topic models have been used to create coarse …
Machine TranslationTopic ModelsTranslationInvestigating Lexical Sharing in Multilingual Machine Translation for Indian Languages
Multilingual language models have shown impressive cross-lingual transfer ability across a diverse set of languages and tasks. To improve the cross-lingual ability of these models, some strategies include transliteration…
Cross-Lingual TransferMachine TranslationTranslationTransliterationMultilingual Multi-Domain NMT for Indian Languages
India is known as the land of many tongues and dialects. Neural machine translation (NMT) is the current state-of-the-art approach for machine translation (MT) but performs better only with large datasets which Indian la…
Machine TranslationNMTTranslation