paper-with-me

홈 › Papers

Fine-Grained Error Analysis on English-to-Japanese Machine Translation in the Medical Domain

2020-11-01 · EAMT 2020 11 · Takeshi Hayakawa, Yuki Arase

We performed a detailed error analysis in domain-specific neural machine translation (NMT) for the English and Japanese language pair with fine-grained manual annotation. Despite its importance for advancing NMT technologies, research on the performance of domain-specific NMT and non-European languages has been limited. In this study, we designed an error typology based on the error types that were typically generated by NMT systems and might cause significant impact in technical translations: “Addition,” “Omission,” “Mistranslation,” “Grammar,” and “Terminology.” The error annotation was targeted to the medical domain and was performed by experienced professional translators specialized in medicine under careful quality control. The annotation detected 4,912 errors on 2,480 sentences, and the frequency and distribution of errors were analyzed. We found that the major errors in NMT were “Mistranslation” and “Terminology” rather than “Addition” and “Omission,” which have been reported as typical problems of NMT. Interestingly, more errors occurred in documents for professionals compared with those for the general public. The results of our annotation work will be published as a parallel corpus with error labels, which are expected to contribute to developing better NMT models, automatic evaluation metrics, and quality estimation models.

📄 PDF Abstract BibTeX

Code (0)

등록된 구현이 없습니다.

Tasks

Machine TranslationNMTTranslation

Similar Papers 제목 키워드 기반

An Empirical Study on Fine-Grained Named Entity Recognition

2018-08-01 · COLING 2018 8 · Khai Mai, Thai-Hoang Pham, Minh Trung Nguyen, Tuan Duc Nguyen 외

Named entity recognition (NER) has attracted a substantial amount of research. Recently, several neural network-based models have been proposed and achieved high performance. However, there is little research on fine-gra…

Chatbotnamed-entity-recognitionNamed Entity RecognitionNamed Entity Recognition (NER)+2

FEANEL: A Benchmark for Fine-Grained Error Analysis in K-12 English Writing

2025-11-28 · Jingheng Ye, Shen Wang, Jiaqi Chen, Hebin Wang 외 arxiv

Large Language Models (LLMs) have transformed artificial intelligence, offering profound opportunities for educational applications. However, their ability to provide fine-grained educational feedback for K-12 English wr…

MedRECT: A Medical Reasoning Benchmark for Error Correction in Clinical Texts

2025-11-01 · Naoto Iwase, Hiroki Okuyama, Junichiro Iwasawa arxiv

Large language models (LLMs) show increasing promise in medical applications, but their ability to detect and correct errors in clinical texts -- a prerequisite for safe deployment -- remains under-evaluated, particularl…

Building a Japanese Document-Level Relation Extraction Dataset Assisted by Cross-Lingual Transfer

2024-04-25 · Youmi Ma, An Wang, Naoaki Okazaki

Document-level Relation Extraction (DocRE) is the task of extracting all semantic relationships from a document. While studies have been conducted on English DocRE, limited attention has been given to DocRE in non-Englis…

AttributeCross-Lingual TransferDocument-level Relation ExtractionRelation+1

Developing a Guideline for the Labovian-Structural Analysis of Oral Narratives in Japanese

2026-03-31 · Amane Watahiki, Tomoki Doi, Akari Kikuchi, Hiroshi Ohata 외 arxiv

Narrative analysis is a cornerstone of qualitative research. One leading approach is the Labovian model, but its application is labor-intensive, requiring a holistic, recursive interpretive process that moves back and fo…