paper-with-me

Papers

Benchmarking Machine Translation on Chinese Social Media Texts

2026-01-30 · Kaiyan Zhao, Zheyong Xie, Zhongtao Miao, Xinze Lyu, Yao Hu, Shaosheng Cao arxiv

The prevalence of rapidly evolving slang, neologisms, and highly stylized expressions in informal user-generated text, particularly on Chinese social media, poses significant challenges for Machine Translation (MT) benchmarking. Specifically, we identify two primary obstacles: (1) data scarcity, as high-quality parallel data requires bilingual annotators familiar with platform-specific slang, and stylistic cues in both languages; and (2) metric limitations, where traditional evaluators like COMET often fail to capture stylistic fidelity and nonstandard expressions. To bridge these gaps, we introduce CSM-MTBench, a benchmark covering five Chinese-foreign language directions and consisting of two expert-curated subsets: Fun Posts, featuring context-rich, slang- and neologism-heavy content, and Social Snippets, emphasizing concise, emotion- and style- driven expressions. Furthermore, we propose tailored evaluation approaches for each subset: measuring the translation success rate of slang and neologisms in Fun Posts, while assessing tone and style preservation in Social Snippets via a hybrid of embedding-based metrics and LLM-as-a-judge. Experiments on over 20 models reveal substantial variation in how current MT systems handle semantic fidelity and informal, social-media-specific stylistic cues. CSM-MTBench thus serves as a rigorous testbed for advancing MT systems capable of mastering real-world Chinese social media texts.

📄 PDF Abstract BibTeX arXiv:2601.22931

Code (0)

등록된 구현이 없습니다.

Tasks

Machine Translation

Similar Papers 제목 키워드 기반

Evaluating LLMs on Chinese Idiom Translation

2025-08-14 · Cai Yang, Yao Dou, David Heineman, Xiaofeng Wu 외 arxiv

Idioms, whose figurative meanings usually differ from their literal interpretations, are common in everyday language, especially in Chinese, where they often contain historical references and follow specific structural p…

Machine Translation

Improved Multilingual Language Model Pretraining for Social Media Text via Translation Pair Prediction

2021-10-20 · WNUT (ACL) 2021 11 · Shubhanshu Mishra, Aria Haghighi

We evaluate a simple approach to improving zero-shot multilingual transfer of mBERT on social media corpus by adding a pretraining task called translation pair prediction (TPP), which predicts whether a pair of cross-lin…

BenchmarkingLanguage ModelingLanguage ModellingNER+7

Addressing the Vulnerability of NMT in Input Perturbations

2021-04-20 · NAACL 2021 4 · Weiwen Xu, Ai Ti Aw, Yang Ding, Kui Wu 외

Neural Machine Translation (NMT) has achieved significant breakthrough in performance but is known to suffer vulnerability to input perturbations. As real input noise is difficult to predict during training, robustness i…

fr-enMachine TranslationNMTTranslation

Large Language Models for Classical Chinese Poetry Translation: Benchmarking, Evaluating, and Improving

2024-08-19 · Andong Chen, Lianzhang Lou, Kehai Chen, Xuefeng Bai 외

Different from the traditional translation tasks, classical Chinese poetry translation requires both adequacy and fluency in translating culturally and historically significant content and linguistic poetic elegance. Lar…

BenchmarkingMachine TranslationTranslation

Combine CRF and MMSEG to Boost Chinese Word Segmentation in Social Media

2015-10-24 · Yao Yushi, Huang Zheng

In this paper, we propose a joint algorithm for the word segmentation on Chinese social media. Previous work mainly focus on word segmentation for plain Chinese text, in order to develop a Chinese social media processing…

Chinese Word SegmentationSegmentation