paper-with-me

홈 › Papers

IndoUKC: A Concept-Centered Indian Multilingual Lexical Resource

2022-06-01 · LREC 2022 6 · Nandu Chandran Nair, Rajendran S. Velayuthan, Yamini Chandrashekar, Gábor Bella, Fausto Giunchiglia

We introduce the IndoUKC, a new multilingual lexical database comprised of eighteen Indian languages, with a focus on formally capturing words and word meanings specific to Indian languages and cultures. The IndoUKC reuses content from the existing IndoWordNet resource while providing a new model for the cross-lingual mapping of lexical meanings that allows for a richer, diversity-aware representation. Accordingly, beyond a thorough syntactic and semantic cleaning, the IndoWordNet lexical content has been thoroughly remodeled in order to allow a more precise expression of language-specific meaning. The resulting database is made available both for browsing through a graphical web interface and for download through the LiveLanguage data catalogue.

📄 PDF Abstract BibTeX

Code (0)

등록된 구현이 없습니다.

Tasks

Diversity

Similar Papers 제목 키워드 기반

Efficient Multilingual Text Classification for Indian Languages

2021-09-01 · RANLP 2021 9 · Salil Aggarwal, Sourav Kumar, Radhika Mamidi

India is one of the richest language hubs on the earth and is very diverse and multilingual. But apart from a few Indian languages, most of them are still considered to be resource poor. Since most of the NLP techniques …

ClassificationMultilingual text classificationtext-classificationText Classification+1

IndoNet: A Multilingual Lexical Knowledge Network for Indian Languages

2013-08-01 · ACL 2013 8 · Brijesh Bhatt, Lahari Poddar, Pushpak Bhattacharyya

Fine-tuning Pre-trained Named Entity Recognition Models For Indian Languages

2024-05-08 · Sankalp Bahad, Pruthwik Mishra, Karunesh Arora, Rakesh Chandra Balabantaray 외

Named Entity Recognition (NER) is a useful component in Natural Language Processing (NLP) applications. It is used in various tasks such as Machine Translation, Summarization, Information Retrieval, and Question-Answerin…

Information RetrievalMachine TranslationMultilingual Named Entity Recognitionnamed-entity-recognition+5

That'll Do Fine!: A Coarse Lexical Resource for English-Hindi MT, Using Polylingual Topic Models

2016-05-01 · LREC 2016 5 · Diptesh Kanojia, Aditya Joshi, Pushpak Bhattacharyya, Mark James Carman

Parallel corpora are often injected with bilingual lexical resources for improved Indian language machine translation (MT). In absence of such lexical resources, multilingual topic models have been used to create coarse …

Machine TranslationTopic ModelsTranslation

Generalising Multilingual Concept-to-Text NLG with Language Agnostic Delexicalisation

2021-05-07 · ACL 2021 5 · Giulio Zhou, Gerasimos Lampouras

Concept-to-text Natural Language Generation is the task of expressing an input meaning representation in natural language. Previous approaches in this task have been able to generalise to rare or unseen instances by rely…

Text Generation