paper-with-me

홈 › Papers

ChatGPT Beyond English: Towards a Comprehensive Evaluation of Large Language Models in Multilingual Learning

2023-04-12 · Viet Dac Lai, Nghia Trung Ngo, Amir Pouran Ben Veyseh, Hieu Man, Franck Dernoncourt, Trung Bui, Thien Huu Nguyen

Over the last few years, large language models (LLMs) have emerged as the most important breakthroughs in natural language processing (NLP) that fundamentally transform research and developments in the field. ChatGPT represents one of the most exciting LLM systems developed recently to showcase impressive skills for language generation and highly attract public attention. Among various exciting applications discovered for ChatGPT in English, the model can process and generate texts for multiple languages due to its multilingual training data. Given the broad adoption of ChatGPT for English in different problems and areas, a natural question is whether ChatGPT can also be applied effectively for other languages or it is necessary to develop more language-specific technologies. The answer to this question requires a thorough evaluation of ChatGPT over multiple tasks with diverse languages and large datasets (i.e., beyond reported anecdotes), which is still missing or limited in current research. Our work aims to fill this gap for the evaluation of ChatGPT and similar LLMs to provide more comprehensive information for multilingual NLP applications. While this work will be an ongoing effort to include additional experiments in the future, our current paper evaluates ChatGPT on 7 different tasks, covering 37 diverse languages with high, medium, low, and extremely low resources. We also focus on the zero-shot learning setting for ChatGPT to improve reproducibility and better simulate the interactions of general users. Compared to the performance of previous models, our extensive experimental results demonstrate a worse performance of ChatGPT for different NLP tasks and languages, calling for further research to develop better models and understanding for multilingual learning.

📄 PDF Abstract BibTeX arXiv:2304.05613

Code (0)

등록된 구현이 없습니다.

Tasks

Multilingual NLPText GenerationZero-Shot Learning

Similar Papers 제목 키워드 기반

Is ChatGPT the Future of Causal Text Mining? A Comprehensive Evaluation and Analysis

2024-02-22 · Takehiro Takayanagi, Masahiro Suzuki, Ryotaro Kobayashi, Hiroki Sakaji 외

Causality is fundamental in human cognition and has drawn attention in diverse research fields. With growing volumes of textual data, discerning causalities within text data is crucial, and causal text mining plays a piv…

Domain AdaptationIn-Context Learning

GPTAraEval: A Comprehensive Evaluation of ChatGPT on Arabic NLP

2023-05-24 · Md Tawkat Islam Khondaker, Abdul Waheed, El Moatez Billah Nagoudi, Muhammad Abdul-Mageed

ChatGPT's emergence heralds a transformative phase in NLP, particularly demonstrated through its excellent performance on many English benchmarks. However, the model's efficacy across diverse linguistic contexts remains …

Natural Language Understanding

How well ChatGPT understand Malaysian English? An Evaluation on Named Entity Recognition and Relation Extraction

2023-11-20 · Mohan Raj Chanthran, Lay-Ki Soon, Huey Fang Ong, Bhawani Selvaretnam

Recently, ChatGPT has attracted a lot of interest from both researchers and the general public. While the performance of ChatGPT in named entity recognition and relation extraction from Standard English texts is satisfac…

ArticlesLarge Language Modelnamed-entity-recognitionNamed Entity Recognition+2

Is ChatGPT a Highly Fluent Grammatical Error Correction System? A Comprehensive Evaluation

2023-04-04 · Tao Fang, Shu Yang, Kaixin Lan, Derek F. Wong 외

ChatGPT, a large-scale language model based on the advanced GPT-3.5 architecture, has shown remarkable potential in various Natural Language Processing (NLP) tasks. However, there is currently a dearth of comprehensive s…

Grammatical Error CorrectionIn-Context LearningLanguage ModelingLanguage Modelling+1

LLaMA Beyond English: An Empirical Study on Language Capability Transfer

2024-01-02 · Jun Zhao, Zhihao Zhang, Luhui Gao, Qi Zhang 외

In recent times, substantial advancements have been witnessed in large language models (LLMs), exemplified by ChatGPT, showcasing remarkable proficiency across a range of complex tasks. However, many mainstream LLMs (e.g…

GPUInformativenessMMLUText Generation