paper-with-me

홈 › Papers

Tagengo: A Multilingual Chat Dataset

2024-05-21 · Peter Devine

Open source large language models (LLMs) have shown great improvements in recent times. However, many of these models are focused solely on popular spoken languages. We present a high quality dataset of more than 70k prompt-response pairs in 74 languages which consist of human generated prompts and synthetic responses. We use this dataset to train a state-of-the-art open source English LLM to chat multilingually. We evaluate our model on MT-Bench chat benchmarks in 6 languages, finding that our multilingual model outperforms previous state-of-the-art open source LLMs across each language. We further find that training on more multilingual data is beneficial to the performance in a chosen target language (Japanese) compared to simply training on only data in that language. These results indicate the necessity of training on large amounts of high quality multilingual data to make a more accessible LLM.

📄 PDF Abstract BibTeX arXiv:2405.12612

Code (1)

Peter-Devine/multilingual_mt_bench 공식 구현 pytorch

Similar Papers 제목 키워드 기반

Multilingual KokoroChat: A Multi-LLM Ensemble Translation Method for Creating a Multilingual Counseling Dialogue Dataset

2026-03-24 · Ryoma Suzuki, Zhiyang Qi, Michimasa Inaba arxiv

To address the critical scarcity of high-quality, publicly available counseling dialogue datasets, we created Multilingual KokoroChat by translating KokoroChat, a large-scale manually authored Japanese counseling corpus,…

ChatGPT Beyond English: Towards a Comprehensive Evaluation of Large Language Models in Multilingual Learning

2023-04-12 · Viet Dac Lai, Nghia Trung Ngo, Amir Pouran Ben Veyseh, Hieu Man 외

Over the last few years, large language models (LLMs) have emerged as the most important breakthroughs in natural language processing (NLP) that fundamentally transform research and developments in the field. ChatGPT rep…

Multilingual NLPText GenerationZero-Shot Learning

MEDAL: A Framework for Benchmarking LLMs as Multilingual Open-Domain Chatbots and Dialogue Evaluators

2025-05-28 · John Mendonça, Alon Lavie, Isabel Trancoso

As the capabilities of chatbots and their underlying LLMs continue to dramatically improve, evaluating their performance has increasingly become a major blocker to their further development. A major challenge is the avai…

BenchmarkingChatbotDialogue Evaluation

ENRICH4ALL: A First Luxembourgish BERT Model for a Multilingual Chatbot

2022-06-01 · SIGUL (LREC) 2022 6 · Dimitra Anastasiou

Machine Translation (MT)-empowered chatbots are not established yet, however, we see an amazing future breaking language barriers and enabling conversation in multiple languages without time-consuming language model buil…

ChatbotLanguage ModelingLanguage ModellingMachine Translation+2

MoZIP: A Multilingual Benchmark to Evaluate Large Language Models in Intellectual Property

2024-02-26 · Shiwen Ni, Minghuan Tan, Yuelin Bai, Fuqiang Niu 외

Large language models (LLMs) have demonstrated impressive performance in various natural language processing (NLP) tasks. However, there is limited understanding of how well LLMs perform in specific domains (e.g, the int…

Language ModelingLanguage ModellingLarge Language ModelMultiple-choice+1