paper-with-me

Papers

Measuring Taiwanese Mandarin Language Understanding

2024-03-29 · Po-Heng Chen, Sijia Cheng, Wei-Lin Chen, Yen-Ting Lin, Yun-Nung Chen

The evaluation of large language models (LLMs) has drawn substantial attention in the field recently. This work focuses on evaluating LLMs in a Chinese context, specifically, for Traditional Chinese which has been largely underrepresented in existing benchmarks. We present TMLU, a holistic evaluation suit tailored for assessing the advanced knowledge and reasoning capability in LLMs, under the context of Taiwanese Mandarin. TMLU consists of an array of 37 subjects across social science, STEM, humanities, Taiwan-specific content, and others, ranging from middle school to professional levels. In addition, we curate chain-of-thought-like few-shot explanations for each subject to facilitate the evaluation of complex reasoning skills. To establish a comprehensive baseline, we conduct extensive experiments and analysis on 24 advanced LLMs. The results suggest that Chinese open-weight models demonstrate inferior performance comparing to multilingual proprietary ones, and open-weight models tailored for Taiwanese Mandarin lag behind the Simplified-Chinese counterparts. The findings indicate great headrooms for improvement, and emphasize the goal of TMLU to foster the development of localized Taiwanese-Mandarin LLMs. We release the benchmark and evaluation scripts for the community to promote future research.

📄 PDF Abstract BibTeX arXiv:2403.20180

Code (2)

miulab/taiwan-llama
miulab/taiwan-llm

Similar Papers 제목 키워드 기반

Enhancing Taiwanese Hokkien Dual Translation by Exploring and Standardizing of Four Writing Systems

2024-03-18 · Bo-Han Lu, Yi-Hsuan Lin, En-Shiun Annie Lee, Richard Tzong-Han Tsai

Machine translation focuses mainly on high-resource languages (HRLs), while low-resource languages (LRLs) like Taiwanese Hokkien are relatively under-explored. The study aims to address this gap by developing a dual tran…

Machine TranslationTranslation

A Topic-aware Comparable Corpus of Chinese Variations

2024-11-17 · Da-Chen Lian, Shu-Kai Hsieh

This study aims to fill the gap by constructing a topic-aware comparable corpus of Mainland Chinese Mandarin and Taiwanese Mandarin from the social media in Mainland China and Taiwan, respectively. Using Dcard for Taiwan…

Building a Taiwanese Mandarin Spoken Language Model: A First Attempt

2024-11-11 · Chih-Kai Yang, Yu-Kuan Fu, Chen-An Li, Yi-Cheng Lin 외

This technical report presents our initial attempt to build a spoken large language model (LLM) for Taiwanese Mandarin, specifically tailored to enable real-time, speech-to-speech interaction in multi-turn conversations.…

DecoderLanguage ModelingLanguage ModellingLarge Language Model

Evaluating Self-supervised Speech Models on a Taiwanese Hokkien Corpus

2023-12-06 · Yi-Hui Chou, Kalvin Chang, Meng-Ju Wu, Winston Ou 외

Taiwanese Hokkien is declining in use and status due to a language shift towards Mandarin in Taiwan. This is partly why it is a low resource language in NLP and speech research today. To ensure that the state of the art …

Self-Supervised Learning

BlueMagpie-TTS: A Token-Efficient Tokenizer, Language Model, and TTS for Taiwanese-Accent Code-Switching Speech

2026-07-07 · Ho Lam Chung, Bo-Xuan Zheng, Cheng-Chieh Huang, Cheng-Han Chang 외 arxiv

Off-the-shelf TTS systems are poorly adapted to Taiwanese Mandarin. Their accent defaults to other Mandarin variants, their tokenizers over-segment common Taiwanese text, and their pronunciation degrades at code-switchin…