paper-with-me

Papers

Data representation methods and use of mined corpora for Indian language transliteration

2015-07-01 · WS 2015 7 · Anoop Kunchukuttan, Pushpak Bhattacharyya
📄 PDF Abstract BibTeX

Code (0)

등록된 구현이 없습니다.

Tasks

Information RetrievalMachine TranslationTransliteration

Similar Papers 제목 키워드 기반

A Large-scale Evaluation of Neural Machine Transliteration for Indic Languages

2021-04-01 · EACL 2021 2 · Anoop Kunchukuttan, Siddharth Jain, Rahul Kejriwal

We take up the task of large-scale evaluation of neural machine transliteration between English and Indic languages, with a focus on multilingual transliteration to utilize orthographic similarity between Indian language…

TranslationTransliteration

A Multilingual Parallel Corpora Collection Effort for Indian Languages

2020-07-15 · LREC 2020 5 · Shashank Siripragada, Jerin Philip, Vinay P. Namboodiri, C. V. Jawahar

We present sentence aligned parallel corpora across 10 Indian Languages - Hindi, Telugu, Tamil, Malayalam, Gujarati, Urdu, Bengali, Oriya, Marathi, Punjabi, and English - many of which are categorized as low resource. Th…

Machine TranslationRetrievalSentenceTranslation

Cross-Corpora Language Recognition: A Preliminary Investigation with Indian Languages

2021-05-10 · Spandan Dey, Goutam Saha, Md Sahidullah

In this paper, we conduct one of the very first studies for cross-corpora performance evaluation in the spoken language identification (LID) problem. Cross-corpora evaluation was not explored much in LID research, especi…

Language IdentificationSpoken language identification

Advantages of Domain Knowledge Injection for Legal Document Summarization: A Case Study on Summarizing Indian Court Judgments in English and Hindi

2026-02-07 · Debtanu Datta, Rajdeep Mukherjee, Adrijit Goswami, Saptarshi Ghosh arxiv

Summarizing Indian legal court judgments is a complex task not only due to the intricate language and unstructured nature of the legal texts, but also since a large section of the Indian population does not understand th…

Document Summarization

An Overview of Indian Spoken Language Recognition from Machine Learning Perspective

2022-11-30 · Spandan Dey, Md Sahidullah, Goutam Saha

Automatic spoken language identification (LID) is a very important research field in the era of multilingual voice-command-based human-computer interaction (HCI). A front-end LID module helps to improve the performance o…

Language IdentificationSpoken language identification