paper-with-me

홈 › Papers

Reducing Disambiguation Biases in NMT by Leveraging Explicit Word Sense Information

2022-07-01 · NAACL 2022 7 · Niccolò Campolungo, Tommaso Pasini, Denis Emelin, Roberto Navigli

Recent studies have shed some light on a common pitfall of Neural Machine Translation (NMT) models, stemming from their struggle to disambiguate polysemous words without lapsing into their most frequently occurring senses in the training corpus.In this paper, we first provide a novel approach for automatically creating high-precision sense-annotated parallel corpora, and then put forward a specifically tailored fine-tuning strategy for exploiting these sense annotations during training without introducing any additional requirement at inference time.The use of explicit senses proved to be beneficial to reduce the disambiguation bias of a baseline NMT model, while, at the same time, leading our system to attain higher BLEU scores than its vanilla counterpart in 3 language pairs.

📄 PDF Abstract BibTeX

Code (0)

등록된 구현이 없습니다.

Tasks

Machine TranslationNMTTranslation

Similar Papers 제목 키워드 기반

Quantum Visual Word Sense Disambiguation: Unraveling Ambiguities Through Quantum Inference Model

2025-12-31 · Wenbo Qiao, Peng Zhang, Qinghua Hu arxiv

Visual word sense disambiguation focuses on polysemous words, where candidate images can be easily confused. Traditional methods use classical probability to calculate the likelihood of an image matching each gloss of th…

Word Sense DisambiguationQuantum Machine LearningImage Matching

Leveraging Word-Formation Knowledge for Chinese Word Sense Disambiguation

2021-11-01 · Findings (EMNLP) 2021 11 · Hua Zheng, Lei LI, Damai Dai, Deli Chen 외

In parataxis languages like Chinese, word meanings are constructed using specific word-formations, which can help to disambiguate word senses. However, such knowledge is rarely explored in previous word sense disambiguat…

Word Sense Disambiguation

Detecting Word Sense Disambiguation Biases in Machine Translation for Model-Agnostic Adversarial Attacks

2020-11-03 · EMNLP 2020 11 · Denis Emelin, Ivan Titov, Rico Sennrich

Word sense disambiguation is a well-known source of translation errors in NMT. We posit that some of the incorrect disambiguation choices are due to models' over-reliance on dataset artifacts found in training data, spec…

Adversarial AttackMachine TranslationNMTTranslation+1

ConSeC: Word Sense Disambiguation as Continuous Sense Comprehension

2021-11-01 · EMNLP 2021 11 · Edoardo Barba, Luigi Procopio, Roberto Navigli

Supervised systems have nowadays become the standard recipe for Word Sense Disambiguation (WSD), with Transformer-based language models as their primary ingredient. However, while these systems have certainly attained un…

Word Sense Disambiguation

One Classifier for All Ambiguous Words: Overcoming Data Sparsity by Utilizing Sense Correlations Across Words

2020-05-01 · LREC 2020 5 · Prafulla Kumar Choubey, Ruihong Huang

Most supervised word sense disambiguation (WSD) systems build word-specific classifiers by leveraging labeled data. However, when using word-specific classifiers, the sparseness of annotations leads to inferior sense dis…

AllWord EmbeddingsWord Sense Disambiguation