paper-with-me

홈 › Papers

LANGALIGN: Enhancing Non-English Language Models via Cross-Lingual Embedding Alignment

2025-03-24 · Jong Myoung Kim, Young-Jun Lee, Ho-Jin Choi, SangKeun Jung

While Large Language Models have gained attention, many service developers still rely on embedding-based models due to practical constraints. In such cases, the quality of fine-tuning data directly impacts performance, and English datasets are often used as seed data for training non-English models. In this study, we propose LANGALIGN, which enhances target language processing by aligning English embedding vectors with those of the target language at the interface between the language model and the task header. Experiments on Korean, Japanese, and Chinese demonstrate that LANGALIGN significantly improves performance across all three languages. Additionally, we show that LANGALIGN can be applied in reverse to convert target language data into a format that an English-based model can process.

📄 PDF Abstract BibTeX arXiv:2503.18603

Code (0)

등록된 구현이 없습니다.

Tasks

Language ModelingLanguage Modelling

Methods 이 논문이 사용한 방법론

Golden Queue Managers 설명 없음

Similar Papers 제목 키워드 기반

MPR-GUI: Benchmarking and Enhancing Multilingual Perception and Reasoning in GUI Agents

2025-11-30 · Ruihan Chen, Qiming Li, Xiaocheng Feng, Weihong Zhong 외 arxiv

Large Vision-Language Models (LVLMs) have shown strong potential as multilingual Graphical User Interface (GUI) agents, as evidenced by existing GUI benchmarks. However, these benchmarks exhibit two primary limitations: …

Language Versatilists vs. Specialists: An Empirical Revisiting on Multilingual Transfer Ability

2023-06-11 · Jiacheng Ye, Xijia Tao, Lingpeng Kong

Multilingual transfer ability, which reflects how well the models fine-tuned on one source language can be applied to other languages, has been well studied in multilingual pre-trained models (e.g., BLOOM). However, such…

CUTE: A Multilingual Dataset for Enhancing Cross-Lingual Knowledge Transfer in Low-Resource Languages

2025-09-21 · Wenhao Zhuang, Yuan Sun arxiv

Large Language Models (LLMs) demonstrate exceptional zero-shot capabilities in various NLP tasks, significantly enhancing user experience and efficiency. However, this advantage is primarily limited to resource-rich lang…

Cross-Lingual TransferMachine Translation

Multilingual Retrieval-Augmented Generation for Knowledge-Intensive Task

2025-04-04 · Leonardo Ranaldi, Barry Haddow, Alexandra Birch

Retrieval-augmented generation (RAG) has become a cornerstone of contemporary NLP, enhancing large language models (LLMs) by allowing them to access richer factual contexts through in-context retrieval. While effective i…

Open-Domain Question AnsweringQuestion AnsweringRAGRetrieval+1

Leveraging Multilingual Training for Authorship Representation: Enhancing Generalization across Languages and Domains

2025-09-20 · Junghwan Kim, Haotian Zhang, David Jurgens arxiv

Authorship representation (AR) learning, which models an author's unique writing style, has demonstrated strong performance in authorship attribution tasks. However, prior research has primarily focused on monolingual se…

Domain GeneralizationContrastive Learning