paper-with-me

홈 › Papers

A Benchmark Dataset for Multi-Level Complexity-Controllable Machine Translation

2022-06-01 · LREC 2022 6 · Kazuki Tani, Ryoya Yuasa, Kazuki Takikawa, Akihiro Tamura, Tomoyuki Kajiwara, Takashi Ninomiya, Tsuneo Kato

This paper presents a new benchmark test dataset for multi-level complexity-controllable machine translation (MLCC-MT), which is MT controlling the complexity of the output at more than two levels. In previous research, MLCC-MT models have been evaluated on a test dataset automatically constructed from the Newsela corpus, which is a document-level comparable corpus with document-level complexity. The existing test dataset has the following three problems: (i) A source language sentence and its target language sentence are not necessarily an exact translation pair because they are automatically detected. (ii) A target language sentence and its simplified target language sentence are not necessarily exactly parallel because they are automatically aligned. (iii) A sentence-level complexity is not necessarily appropriate because it is transferred from an article-level complexity attached to the Newsela corpus. Therefore, we create a benchmark test dataset for Japanese-to-English MLCC-MT from the Newsela corpus by introducing an automatic filtering of data with inappropriate sentence-level complexity, manual check for parallel target language sentences with different complexity levels, and manual translation. Moreover, we implement two MLCC-NMT frameworks with a Transformer architecture and report their performance on our test dataset as baselines for future research. Our test dataset and codes are released.

📄 PDF Abstract BibTeX

Code (1)

k-t4n1/a-benchmarkdataset-for-complexitycontrollablenmt 공식 구현

Tasks

Machine TranslationNMTSentenceTranslation

Similar Papers 제목 키워드 기반

COMPILING: A Benchmark Dataset for Chinese Complexity Controllable Definition Generation

2022-09-29 · CCL 2022 10 · Jiaxin Yuan, Cunliang Kong, Chenhui Xie, Liner Yang 외

The definition generation task aims to generate a word's definition within a specific context automatically. However, owing to the lack of datasets for different complexities, the definitions produced by models tend to k…

Simple or Complex? Complexity-Controllable Question Generation with Soft Templates and Deep Mixture of Experts Model

2021-10-13 · Findings (EMNLP) 2021 11 · Sheng Bi, Xiya Cheng, Yuan-Fang Li, Lizhen Qu 외

The ability to generate natural-language questions with controlled complexity levels is highly desirable as it further expands the applicability of question generation. In this paper, we propose an end-to-end neural comp…

Mixture-of-ExpertsQuestion GenerationQuestion-GenerationQuestion Similarity

What Limits Virtual Agent Application? OmniBench: A Scalable Multi-Dimensional Benchmark for Essential Virtual Agent Capabilities

2025-06-10 · Wendong Bu, Yang Wu, Qifan Yu, Minghe Gao 외

As multimodal large language models (MLLMs) advance, MLLM-based virtual agents have demonstrated remarkable performance. However, existing benchmarks face significant limitations, including uncontrollable task complexity…

SDDMO-Bench: A Benchmark Suite for Streaming Data-Driven Dynamic Multi-Objective Optimization

2026-08-01 · Wenjie Xiao, Hui Bai, Junhao Chen arxiv

Streaming data-driven dynamic multi-objective optimization requires algorithms to track time-varying Pareto fronts using only sequential observations under concept drift. However, systematic evaluation remains difficult …

OrchDAG: Complex Tool Orchestration in Multi-Turn Interactions with Plan DAGs

2025-10-28 · Yifu Lu, Shengjie Liu, Li Dong arxiv

Agentic tool use has gained traction with the rise of agentic tool calling, yet most existing work overlooks the complexity of multi-turn tool interactions. We introduce OrchDAG, a synthetic data generation pipeline that…

Synthetic Data Generation