paper-with-me

홈 › Papers

Aligning Translation-Specific Understanding to General Understanding in Large Language Models

2024-01-10 · Yichong Huang, Baohang Li, Xiaocheng Feng, Chengpeng Fu, Wenshuai Huo, Ting Liu, Bing Qin

Large Language models (LLMs) have exhibited remarkable abilities in understanding complex texts, offering a promising path towards human-like translation performance. However, this study reveals the misalignment between the translation-specific understanding and the general understanding inside LLMs. This understanding misalignment leads to LLMs mistakenly or literally translating some complicated concepts that they accurately comprehend in the general scenarios (e.g., QA). To align the translation-specific understanding to the general one, we propose a novel translation process, DUAT (Difficult words Understanding Aligned Translation), explicitly incorporating the general understanding on the complicated content incurring inconsistent understanding to guide the translation. Specifically, DUAT performs cross-lingual interpretation for the difficult-to-translate words and enhances the translation with the generated interpretations. Furthermore, we reframe the external tools to improve DUAT in detecting difficult words and generating helpful interpretations. We conduct experiments on the self-constructed benchmark Challenge-WMT, consisting of samples that are prone to mistranslation. Human evaluation results on high-resource and low-resource language pairs indicate that DUAT significantly facilitates the understanding alignment, which improves the translation quality (up to +3.85 COMET) and reduces the literality of the translation by -25% to -51%.

📄 PDF Abstract BibTeX arXiv:2401.05072

Code (1)

orangeinsouth/challengewmt 공식 구현

Tasks

Machine TranslationTranslation

Methods 이 논문이 사용한 방법론

ALIGN In the ALIGN method, visual and language representations are jointly trained from noisy image alt-text data. The image and text encoders are learned via contrastive loss…

Similar Papers 제목 키워드 기반

LLM-based Translation Inference with Iterative Bilingual Understanding

2024-10-16 · Andong Chen, Kehai Chen, Yang Xiang, Xuefeng Bai 외

The remarkable understanding and generation capabilities of large language models (LLMs) have greatly improved translation performance. However, incorrect understanding of the sentence to be translated can degrade transl…

SentenceTranslation

Evaluating Menu OCR and Translation: A Benchmark for Aligning Human and Automated Evaluations in Large Vision-Language Models

2025-04-16 · Zhanglin Wu, Tengfei Song, Ning Xie, Mengli Zhu 외

The rapid advancement of large vision-language models (LVLMs) has significantly propelled applications in document understanding, particularly in optical character recognition (OCR) and multilingual translation. However,…

document understandingLayout DesignOptical Character RecognitionOptical Character Recognition (OCR)+1

PART: Progressive Alignment Representation Training for Multilingual Speech-To-Text with LLMs

2025-09-24 · Pei Zhang, Andong Chen, Xi Chen, Baosong Yang 외 arxiv

Large language models (LLMs) have expanded from text to speech, giving rise to Speech Large Models (SLMs) that support recognition, translation, and synthesis. A key challenge is aligning speech and text representations,…

Text to Speech

BayLing: Bridging Cross-lingual Alignment and Instruction Following through Interactive Translation for Large Language Models

2023-06-19 · Shaolei Zhang, Qingkai Fang, Zhuocheng Zhang, Zhengrui Ma 외

Large language models (LLMs) have demonstrated remarkable prowess in language understanding and generation. Advancing from foundation LLMs to instructionfollowing LLMs, instruction tuning plays a vital role in aligning L…

Instruction FollowingText GenerationTranslation

Understanding and Improving Hidden Representations for Neural Machine Translation

2019-06-01 · NAACL 2019 6 · Guanlin Li, Lemao Liu, Xintong Li, Conghui Zhu 외

Multilayer architectures are currently the gold standard for large-scale neural machine translation. Existing works have explored some methods for understanding the hidden representations, however, they have not sought t…

Machine TranslationTranslation