paper-with-me

홈 › Papers

Aspects of Terminological and Named Entity Knowledge within Rule-Based Machine Translation Models for Under-Resourced Neural Machine Translation Scenarios

2020-09-28 · Daniel Torregrosa, Nivranshu Pasricha, Maraim Masoud, Bharathi Raja Chakravarthi, Juan Alonso, Noe Casas, Mihael Arcan

Rule-based machine translation is a machine translation paradigm where linguistic knowledge is encoded by an expert in the form of rules that translate text from source to target language. While this approach grants extensive control over the output of the system, the cost of formalising the needed linguistic knowledge is much higher than training a corpus-based system, where a machine learning approach is used to automatically learn to translate from examples. In this paper, we describe different approaches to leverage the information contained in rule-based machine translation systems to improve a corpus-based one, namely, a neural machine translation model, with a focus on a low-resource scenario. Three different kinds of information were used: morphological information, named entities and terminology. In addition to evaluating the general performance of the system, we systematically analysed the performance of the proposed approaches when dealing with the targeted phenomena. Our results suggest that the proposed models have limited ability to learn from external information, and most approaches do not significantly alter the results of the automatic evaluation, but our preliminary qualitative evaluation shows that in certain cases the hypothesis generated by our system exhibit favourable behaviour such as keeping the use of passive voice.

📄 PDF Abstract BibTeX arXiv:2009.13398

Code (0)

등록된 구현이 없습니다.

Tasks

Machine TranslationTranslation

Similar Papers 제목 키워드 기반

A Domain-Specific Curated Benchmark for Entity and Document-Level Relation Extraction

2026-02-04 · Marco Martinelli, Stefano Marchesin, Vanessa Bonato, Giorgio Maria Di Nunzio 외 arxiv

Information Extraction (IE), encompassing Named Entity Recognition (NER), Named Entity Linking (NEL), and Relation Extraction (RE), is critical for transforming the rapidly growing volume of scientific publications into …

Document-level Relation ExtractionInformation ExtractionEntity Linking

How to Understand Named Entities: Using Common Sense for News Captioning

2024-03-11 · Ning Xu, Yanhui Wang, Tingting Zhang, Hongshuo Tian 외

News captioning aims to describe an image with its news article body as input. It greatly relies on a set of detected named entities, including real-world people, organizations, and places. This paper exploits commonsens…

Common Sense Reasoning

Target-Oriented Fine-tuning for Zero-Resource Named Entity Recognition

2021-07-22 · Findings (ACL) 2021 8 · Ying Zhang, Fandong Meng, Yufeng Chen, Jinan Xu 외

Zero-resource named entity recognition (NER) severely suffers from data scarcity in a specific domain or language. Most studies on zero-resource NER transfer knowledge from various data by fine-tuning on different auxili…

named-entity-recognitionNamed Entity RecognitionNamed Entity Recognition (NER)NER+1

Social Media Analysis based on Semanticity of Streaming and Batch Data

2018-01-03 · Barathi Ganesh HB

Languages shared by people differ in different regions based on their accents, pronunciation and word usages. In this era sharing of language takes place mainly through social media and blogs. Every second swing of such …

Author Profilingnamed-entity-recognitionNamed Entity RecognitionNamed Entity Recognition (NER)

Defying Wikidata: Validation of Terminological Relations in the Web of Data

2020-05-01 · LREC 2020 5 · Patricia Mart{\'\i}n-Chozas, Sina Ahmadi, Elena Montiel-Ponsoda

In this paper we present an approach to validate terminological data retrieved from open encyclopaedic knowledge bases. This need arises from the enrichment of automatically extracted terms with information from existing…