paper-with-me

Papers

Language Mixing in Reasoning Language Models: Patterns, Impact, and Internal Causes

2025-05-20 · Mingyang Wang, Lukas Lange, Heike Adel, Yunpu Ma, Jannik Strötgen, Hinrich Schütze

Reasoning language models (RLMs) excel at complex tasks by leveraging a chain-of-thought process to generate structured intermediate steps. However, language mixing, i.e., reasoning steps containing tokens from languages other than the prompt, has been observed in their outputs and shown to affect performance, though its impact remains debated. We present the first systematic study of language mixing in RLMs, examining its patterns, impact, and internal causes across 15 languages, 7 task difficulty levels, and 18 subject areas, and show how all three factors influence language mixing. Moreover, we demonstrate that the choice of reasoning language significantly affects performance: forcing models to reason in Latin or Han scripts via constrained decoding notably improves accuracy. Finally, we show that the script composition of reasoning traces closely aligns with that of the model's internal representations, indicating that language mixing reflects latent processing preferences in RLMs. Our findings provide actionable insights for optimizing multilingual reasoning and open new directions for controlling reasoning languages to build more interpretable and adaptable RLMs.

📄 PDF Abstract BibTeX arXiv:2505.14815

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

The Impact of Language Mixing on Bilingual LLM Reasoning

2025-07-21 · Yihao Li, Jiayi Xin, Miranda Muqing Miao, Qi Long 외 arxiv

Proficient multilingual speakers often intentionally switch languages in the middle of a conversation. Similarly, recent reasoning-focused bilingual large language models (LLMs) with strong capabilities in both languages…

Reinforcement Learning

Language Patterns and Behaviour of the Peer Supporters in Multilingual Healthcare Conversational Forums

2022-06-01 · LREC 2022 6 · Ishani Mondal, Kalika Bali, Mohit Jain, Monojit Choudhury 외

In this work, we conduct a quantitative linguistic analysis of the language usage patterns of multilingual peer supporters in two health-focused WhatsApp groups in Kenya comprising of youth living with HIV. Even though t…

Language Agnostic Code-Mixing Data Augmentation by Predicting Linguistic Patterns

2022-11-14 · Shuyue Stella Li, Kenton Murray

In this work, we focus on intrasentential code-mixing and propose several different Synthetic Code-Mixing (SCM) data augmentation methods that outperform the baseline on downstream sentiment analysis tasks across various…

Data AugmentationSentiment Analysis

Building Multilingual Bridges: Data Mixing as the Pillar of Generalization for In-Language Reasoning

2026-09-09 · Mehrnaz Mofakhami, Ananya Sahu, Alejandro R. Salamanca, Daniel D'souza 외 arxiv

Reasoning language models have made substantial advances on a variety of complex tasks, yet their capabilities remain overwhelmingly English-centric: models primarily reason in English regardless of the language they are…

Instruction Following

PreCogIIITH at HinglishEval : Leveraging Code-Mixing Metrics & Language Model Embeddings To Estimate Code-Mix Quality

2022-06-16 · Prashant Kodali, Tanmay Sachan, Akshay Goindani, Anmol Goel 외

Code-Mixing is a phenomenon of mixing two or more languages in a speech event and is prevalent in multilingual societies. Given the low-resource nature of Code-Mixing, machine generation of code-mixed text is a prevalent…

Data AugmentationLanguage ModelingLanguage Modelling