paper-with-me

Papers

ComplexityMT: Benchmarking the Interaction Between Text Complexity and Machine Translation

2026-06-03 · Joseph Marvin Imperial, Junhong Liang, Belal Shoer, Abdullah Barayan, Rodrigo Wilkens, Omar Mussa, Dawn Knight, Eugénio Ribeiro, Ekaterina Kochmar, Sowmya Vajjala, Fernando Alva-Manchego, Harish Tayyar Madabushi arxiv

When a text is translated, does the translation retain the complexity of the original? We introduce ComplexityMT, a new challenge for assessing how text complexity and machine translation interact with and influence each other, using the Common European Framework of Reference for Languages (CEFR) levels as the measure of text complexity. Across six languages, including Arabic, Dutch, English, French, Hindi, and Russian, we evaluate three open-weight models, one closed model, and a commercial machine translation system on two tasks: i) correlation of CEFR with translation difficulty, and ii) shifts in CEFR levels of the source texts. Our experiments show that higher CEFR levels make texts more difficult to translate, and that machine translation shifts the CEFR level of the target text compared to the original source, for most languages. These findings provide new insights for researchers and practitioners working on multilingual pedagogical content generation and machine translation difficulty estimation.

📄 PDF Abstract BibTeX arXiv:2606.05421

Code (0)

등록된 구현이 없습니다.

Tasks

Machine Translation

Similar Papers 제목 키워드 기반

Compact Trilinear Interaction for Visual Question Answering

2019-09-26 · ICCV 2019 10 · Tuong Do, Thanh-Toan Do, Huy Tran, Erman Tjiputra 외

In Visual Question Answering (VQA), answers have a great correlation with question meaning and visual contents. Thus, to selectively utilize image, question and answer information, we propose a novel trilinear interactio…

BenchmarkingKnowledge DistillationQuestion AnsweringVisual Question Answering+1

UI2App: Benchmarking Visual Interaction Inference in Executable Web Application Generation

2026-07-07 · Grace Man Chen, Litao Guo, Yifan Wu, Yiyu Chen 외 arxiv

Large language models (LLMs) have demonstrated growing competence in web page generation. However, existing text-driven approaches rely on complex prompts that impose substantial demands on users and offer limited expres…

Beyond Prompts: Dynamic Conversational Benchmarking of Large Language Models

2024-09-30 · David Castillo-Bolado, Joseph Davidson, Finlay Gray, Marek Rosa

We introduce a dynamic benchmarking system for conversational agents that evaluates their performance through a single, simulated, and lengthy user$\leftrightarrow$agent interaction. The interaction is a conversation bet…

BenchmarkingContinual Learning

Ansatz-free Hamiltonian learning with Heisenberg-limited scaling

2025-02-17 · Hong-Ye Hu, Muzhou Ma, Weiyuan Gong, Qi Ye 외

Learning the unknown interactions that govern a quantum system is crucial for quantum information processing, device benchmarking, and quantum sensing. The problem, known as Hamiltonian learning, is well understood under…

Benchmarking

Benchmarking Large Multimodal Models against Common Corruptions

2024-01-22 · Jiawei Zhang, Tianyu Pang, Chao Du, Yi Ren 외

This technical report aims to fill a deficiency in the assessment of large multimodal models (LMMs) by specifically examining the self-consistency of their outputs when subjected to common corruptions. We investigate the…

BenchmarkingImage to textSpeech-to-Texttext-to-speech+1