paper-with-me

Papers

Automatic Translation Alignment Pipeline for Multilingual Digital Editions of Literary Works

2024-10-17 · Maria Levchenko

This paper investigates the application of translation alignment algorithms in the creation of a Multilingual Digital Edition (MDE) of Alessandro Manzoni's Italian novel "I promessi sposi" ("The Betrothed"), with translations in eight languages (English, Spanish, French, German, Dutch, Polish, Russian and Chinese) from the 19th and 20th centuries. We identify key requirements for the MDE to improve both the reader experience and support for translation studies. Our research highlights the limitations of current state-of-the-art algorithms when applied to the translation of literary texts and outlines an automated pipeline for MDE creation. This pipeline transforms raw texts into web-based, side-by-side representations of original and translated texts with different rendering options. In addition, we propose new metrics for evaluating the alignment of literary translations and suggest visualization techniques for future analysis.

📄 PDF Abstract BibTeX arXiv:2410.13255

Code (0)

등록된 구현이 없습니다.

Tasks

Translation

Similar Papers 제목 키워드 기반

PMMT: Preference Alignment in Multilingual Machine Translation via LLM Distillation

2024-10-15 · Shuqiao Sun, Yutong Yao, Peiwen Wu, Feijun Jiang 외

Translation is important for cross-language communication, and many efforts have been made to improve its accuracy. However, less investment is conducted in aligning translations with human preferences, such as translati…

Machine TranslationTranslation

NeoBabel: A Multilingual Open Tower for Visual Generation

2025-07-08 · Mohammad Mahdi Derakhshani, Dheeraj Varghese, Marzieh Fadaee, Cees G. M. Snoek

Text-to-image generation advancements have been predominantly English-centric, creating barriers for non-English speakers and perpetuating digital inequities. While existing systems rely on translation pipelines, these i…

Image GenerationText to Image GenerationText-to-Image Generation

What Drives Cross-lingual Ranking? Retrieval Approaches with Multilingual Language Models

2025-11-24 · Roksana Goworek, Olivia Macmillan-Scott, Eda B. Özyiğit arxiv

Cross-lingual information retrieval (CLIR) enables access to multilingual knowledge but remains challenging due to disparities in resources, scripts, and weak cross-lingual semantic alignment in embedding models. Existin…

Information RetrievalContrastive Learning

BYOL: Bring Your Own Language Into LLMs

2026-01-15 · Syed Waqas Zamir, Wassim Hamidouche, Boulbaba Ben Amor, Luana Marotti 외 arxiv

Large Language Models (LLMs) exhibit strong multilingual capabilities, yet remain fundamentally constrained by the severe imbalance in global language resources. While over 7,000 languages are spoken worldwide, only a sm…

Continual PretrainingMachine TranslationText Generation

Let MT simplify and speed up your Alignment for TM creation

2020-11-01 · EAMT 2020 11 · Judith Klein, Giorgio Bernardinello

Large quantities of multilingual legal documents are waiting to be regularly aligned and used for future translations. For reasons of time, effort and cost, manual alignment is not an option. Automatically aligned segmen…