paper-with-me

홈 › Papers

VinaLLaMA: LLaMA-based Vietnamese Foundation Model

2023-12-18 · Quan Nguyen, Huy Pham, Dung Dao

In this technical report, we present VinaLLaMA, an open-weight, state-of-the-art (SOTA) Large Language Model for the Vietnamese language, built upon LLaMA-2 with an additional 800 billion trained tokens. VinaLLaMA not only demonstrates fluency in Vietnamese but also exhibits a profound understanding of Vietnamese culture, making it a truly indigenous model. VinaLLaMA-7B-chat, trained on 1 million high-quality synthetic samples, achieves SOTA results on key benchmarks, including VLSP, VMLU, and Vicuna Benchmark Vietnamese, marking a significant advancement in the Vietnamese AI landscape and offering a versatile resource for various applications.

📄 PDF Abstract BibTeX arXiv:2312.11011

Code (0)

등록된 구현이 없습니다.

Tasks

Language ModelingLanguage ModellingLarge Language Modelmodel

Similar Papers 제목 키워드 기반

VIVID: A Culturally Grounded Benchmark Exposing the Figurative Language Gap in Vietnamese NLP

2026-08-04 · Tu Tran Do, Nhat Ngoc Nguyen, Khanh-Tung Tran, Hoang D. Nguyen 외 arxiv

We present VIVID (Vietnamese Idioms for Validation and Interpretation Depth), the first systematic benchmark for evaluating culturally grounded figurative language understanding in Vietnamese. VIVID comprises 1,636 idiom…

A Large-Scale Benchmark for Vietnamese Sentence Paraphrases

2025-02-11 · Sang Quang Nguyen, Kiet Van Nguyen

This paper presents ViSP, a high-quality Vietnamese dataset for sentence paraphrasing, consisting of 1.2M original-paraphrase pairs collected from various domains. The dataset was constructed using a hybrid approach that…

Paraphrase GenerationSentence

VietJobs: A Vietnamese Job Advertisement Dataset

2026-03-05 · Hieu Pham Dinh, Hung Nguyen Huy, Mo El-Haj arxiv

VietJobs is the first large-scale, publicly available corpus of Vietnamese job advertisements, comprising 48,092 postings and over 15 million words collected from all 34 provinces and municipalities across Vietnam. The d…

Evaluating the Symbol Binding Ability of Large Language Models for Multiple-Choice Questions in Vietnamese General Education

2023-10-18 · Duc-Vu Nguyen, Quoc-Nam Nguyen

In this paper, we evaluate the ability of large language models (LLMs) to perform multiple choice symbol binding (MCSB) for multiple choice question answering (MCQA) tasks in zero-shot, one-shot, and few-shot settings. W…

Multiple-choiceMultiple Choice Question Answering (MCQA)Question Answering

ViLLM-Eval: A Comprehensive Evaluation Suite for Vietnamese Large Language Models

2024-04-17 · Trong-Hieu Nguyen, Anh-Cuong Le, Viet-Cuong Nguyen

The rapid advancement of large language models (LLMs) necessitates the development of new benchmarks to accurately assess their capabilities. To address this need for Vietnamese, this work aims to introduce ViLLM-Eval, t…

Language ModelingLanguage ModellingLarge Language ModelMultiple-choice