paper-with-me

홈 › Papers

JU_NLP at HinglishEval: Quality Evaluation of the Low-Resource Code-Mixed Hinglish Text

2022-06-16 · Prantik Guha, Rudra Dhar, Dipankar Das

In this paper we describe a system submitted to the INLG 2022 Generation Challenge (GenChal) on Quality Evaluation of the Low-Resource Synthetically Generated Code-Mixed Hinglish Text. We implement a Bi-LSTM-based neural network model to predict the Average rating score and Disagreement score of the synthetic Hinglish dataset. In our models, we used word embeddings for English and Hindi data, and one hot encodings for Hinglish data. We achieved a F1 score of 0.11, and mean squared error of 6.0 in the average rating score prediction task. In the task of Disagreement score prediction, we achieve a F1 score of 0.18, and mean squared error of 5.0.

📄 PDF Abstract BibTeX arXiv:2206.08053

Code (0)

등록된 구현이 없습니다.

Tasks

PredictionWord Embeddings

Similar Papers 제목 키워드 기반

PreCogIIITH at HinglishEval : Leveraging Code-Mixing Metrics & Language Model Embeddings To Estimate Code-Mix Quality

2022-06-16 · Prashant Kodali, Tanmay Sachan, Akshay Goindani, Anmol Goel 외

Code-Mixing is a phenomenon of mixing two or more languages in a speech event and is prevalent in multilingual societies. Given the low-resource nature of Code-Mixing, machine generation of code-mixed text is a prevalent…

Data AugmentationLanguage ModelingLanguage Modelling

niksss at HinglishEval: Language-agnostic BERT-based Contextual Embeddings with Catboost for Quality Evaluation of the Low-Resource Synthetically Generated Code-Mixed Hinglish Text

2022-06-17 · Nikhil Singh

This paper describes the system description for the HinglishEval challenge at INLG 2022. The goal of this task was to investigate the factors influencing the quality of the code-mixed text generation system. The task was…

PredictionSentenceText GenerationWord Embeddings

BITS Pilani at HinglishEval: Quality Evaluation for Code-Mixed Hinglish Text Using Transformers

2022-06-17 · Shaz Furniturewala, Vijay Kumari, Amulya Ratna Dash, Hriday Kedia 외

Code-Mixed text data consists of sentences having words or phrases from more than one language. Most multi-lingual communities worldwide communicate using multiple languages, with English usually one of them. Hinglish is…

Quality Evaluation of the Low-Resource Synthetically Generated Code-Mixed Hinglish Text

2021-08-04 · INLG (ACL) 2021 8 · Vivek Srivastava, Mayank Singh

In this shared task, we seek the participating teams to investigate the factors influencing the quality of the code-mixed text generation systems. We synthetically generate code-mixed Hinglish sentences using two distinc…

PredictionText Generation

Supervised and Unsupervised Evaluation of Synthetic Code-Switching

2022-10-01 · COLING (WNUT) 2022 10 · Evgeny Orlov, Ekaterina Artemova

Code-switching (CS) is a phenomenon of mixing words and phrases from multiple languages within a single sentence or conversation. The ever-growing amount of CS communication among multilingual speakers in social media ha…

Sentence