paper-with-me

홈 › Papers

EXTRACTING PARALLEL PHRASES FROM COMPARABLE ENGLISH AND PUNJABI CORPORA USING AN INTEGRATED APPROACH

2020-12-01 · ICON 2020 12 · Manpreet Singh Lehal, Vishal Goyal

Machine translation from English to Indian languages is always a difficult task due to the unavailability of a good quality corpus and morphological richness in the Indian languages. For a system to produce better translations, the size of the corpus should be huge. We have employed three similarity and distance measures for the research and developed a software to extract parallel data from comparable corpora automatically with high precision using minimal resources. The software works upon four algorithms. The three algorithms have been used for finding Cosine Similarity, Euclidean Distance Similarity and Jaccard Similarity. The fourth algorithm is to integrate the outputs of the three algorithms in order to improve the efficiency of the system.

📄 PDF Abstract BibTeX

Code (0)

등록된 구현이 없습니다.

Tasks

Machine TranslationTranslation

Similar Papers 제목 키워드 기반

Statistical Analysis of Multilingual Text Corpus and Development of Language Models

2014-05-01 · LREC 2014 5 · Shyam Sundar Agrawal, {Abhimanue}, shweta bansal, Minakshi Mahajan

This paper presents two studies, first a statistical analysis for three languages i.e. Hindi, Punjabi and Nepali and the other, development of language models for three Indian languages i.e. Indian English, Punjabi and N…

Language IdentificationLanguage ModellingSpeech Language Identification

Punjabi to English Bidirectional NMT System

2020-12-01 · ICON 2020 12 · Kamal Deep, Ajit Kumar, Vishal Goyal

Machine Translation is ongoing research for last few decades. Today, Corpus-based Machine Translation systems are very popular. Statistical Machine Translation and Neural Machine Translation are based on the parallel cor…

Machine TranslationNMTTranslation

Multimodal Comparable Corpora as Resources for Extracting Parallel Data: Parallel Phrases Extraction

2013-10-01 · IJCNLP 2013 10 · Haithem Afli, Lo{\"\i}c Barrault, Holger Schwenk
Information RetrievalLanguage ModellingMachine TranslationSpeech Recognition

Traduction automatique \`a partir de corpus comparables: extraction de phrases parall\`eles \`a partir de donn\'ees comparables multimodales (Automatic Translation from Comparable corpora : extracting parallel sentences from multimodal comparable corpora) [in French]

2012-06-01 · JEPTALNRECITAL 2012 6 · Haithem Afli, Lo{\"\i}c Barrault, Holger Schwenk
Information RetrievalMachine TranslationSpeech Recognition

Development of Hybrid Algorithm for Automatic Extraction of Multiword Expressions from Monolingual and Parallel Corpus of English and Punjabi

2020-12-01 · ICON 2020 12 · Kapil Dev Goyal, Vishal Goyal

Identification and extraction of Multiword Expressions (MWEs) is very hard and challenging task in various Natural Language processing applications like Information Retrieval (IR), Information Extraction (IE), Question-A…

Information RetrievalMachine TranslationQuestion AnsweringRetrieval+4