paper-with-me

홈 › Papers

Seamless Language Expansion: Enhancing Multilingual Mastery in Self-Supervised Models

2024-06-20 · Jing Xu, Minglin Wu, Xixin Wu, Helen Meng

Self-supervised (SSL) models have shown great performance in various downstream tasks. However, they are typically developed for limited languages, and may encounter new languages in real-world. Developing a SSL model for each new language is costly. Thus, it is vital to figure out how to efficiently adapt existed SSL models to a new language without impairing its original abilities. We propose adaptation methods which integrate LoRA to existed SSL models to extend new language. We also develop preservation strategies which include data combination and re-clustering to retain abilities on existed languages. Applied to mHuBERT, we investigate their effectiveness on speech re-synthesis task. Experiments show that our adaptation methods enable mHuBERT to be applied to a new language (Mandarin) with MOS value increased about 1.6 and the relative value of WER reduced up to 61.72%. Also, our preservation strategies ensure that the performance on both existed and new languages remains intact.

📄 PDF Abstract BibTeX arXiv:2406.14092

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Multilingual Pixel Representations for Translation and Effective Cross-lingual Transfer

2023-05-23 · Elizabeth Salesky, Neha Verma, Philipp Koehn, Matt Post

We introduce and demonstrate how to effectively train multilingual machine translation models with pixel representations. We experiment with two different data settings with a variety of language and script coverage, dem…

Cross-Lingual TransferMachine TranslationTranslation

MoE-LPR: Multilingual Extension of Large Language Models through Mixture-of-Experts with Language Priors Routing

2024-08-21 · Hao Zhou, Zhijun Wang, ShuJian Huang, Xin Huang 외

Large Language Models (LLMs) are often English-centric due to the disproportionate distribution of languages in their pre-training data. Enhancing non-English language capabilities through post-pretraining often results …

Mixture-of-Experts

Gamayun's Path to Multilingual Mastery: Cost-Efficient Training of a 1.5B-Parameter LLM

2025-12-25 · Alexander Podolskiy, Semen Molokov, Timofey Gerasin, Maksim Titov 외 arxiv

We present Gamayun, a 1.5B-parameter multilingual language model trained entirely from scratch on 2.5T tokens. Designed for efficiency and deployment in resource-constrained environments, Gamayun addresses the lack of re…

LLaMAX: Scaling Linguistic Horizons of LLM by Enhancing Translation Capabilities Beyond 100 Languages

2024-07-08 · Yinquan Lu, Wenhao Zhu, Lei LI, Yu Qiao 외

Large Language Models (LLMs) demonstrate remarkable translation capabilities in high-resource language tasks, yet their performance in low-resource languages is hindered by insufficient multilingual data during pre-train…

Data AugmentationTranslation

Bridging the Gap: Dynamic Learning Strategies for Improving Multilingual Performance in LLMs

2024-05-28 · Somnath Kumar, Vaibhav Balloli, Mercy Ranjit, Kabir Ahuja 외

Large language models (LLMs) are at the forefront of transforming numerous domains globally. However, their inclusivity and effectiveness remain limited for non-Latin scripts and low-resource languages. This paper tackle…

Question AnsweringRAGRetrieval-augmented Generation