paper-with-me

홈 › Papers

Automatic TM Cleaning through MT and POS Tagging: Autodesk's Submission to the NLP4TM 2016 Shared Task

2016-05-19 · Alena Zwahlen, Olivier Carnal, Samuel Läubli

We describe a machine learning based method to identify incorrect entries in translation memories. It extends previous work by Barbu (2015) through incorporating recall-based machine translation and part-of-speech-tagging features. Our system ranked first in the Binary Classification (II) task for two out of three language pairs: English-Italian and English-Spanish.

📄 PDF Abstract BibTeX arXiv:1605.05906

Code (0)

등록된 구현이 없습니다.

Tasks

BIG-bench Machine LearningBinary ClassificationGeneral ClassificationMachine TranslationPart-Of-Speech TaggingPOSPOS TaggingTranslation

Similar Papers 제목 키워드 기반

Building an Ensemble LLM Semantic Tagger for UN Security Council Resolutions

2026-03-06 · Hussein Ghaly arxiv

This paper introduces a new methodology for using LLM-based systems for accurate and efficient semantic tagging of UN Security Council resolutions. The main goal is to leverage LLM performance variability to build ensemb…

Accurate semantic textual similarity for cleaning noisy parallel corpora using semantic machine translation evaluation metric: The NRC supervised submissions to the Parallel Corpus Filtering task

2018-10-01 · WS 2018 10 · Chi-kiu Lo, Michel Simard, Darlene Stewart, Samuel Larkin 외

We present our semantic textual similarity approach in filtering a noisy web crawled parallel corpus using YiSi{---}a novel semantic machine translation evaluation metric. The systems mainly based on this supervised appr…

Machine TranslationSemantic Textual SimilarityTranslation

JHUBC's Submission to LT4HALA EvaLatin 2020

2020-05-01 · LREC 2020 5 · Winston Wu, Garrett Nicolai

We describe the JHUBC submission to the EvaLatin Shared task on lemmatization and part-of-speech tagging for Latin. We modify a hard-attentional character-based encoder-decoder to produce lemmas and POS tags with separat…

DecoderLemmatizationPart-Of-Speech TaggingPOS

Text Preprocessing and its Implications in a Digital Humanities Project

2021-09-01 · RANLP 2021 9 · Maria Kunilovskaya, Alistair Plum

This paper focuses on data cleaning as part of a preprocessing procedure applied to text data retrieved from the web. Although the importance of this early stage in a project using NLP methods is often highlighted by res…

Part-Of-Speech Taggingtext annotationtext-classificationText Classification

Toward Pan-Slavic NLP: Some Experiments with Language Adaptation

2017-04-01 · WS 2017 4 · Serge Sharoff

There is great variation in the amount of NLP resources available for Slavonic languages. For example, the Universal Dependency treebank (Nivre et al., 2016) has about 2 MW of training resources for Czech, more than 1 MW…

Domain AdaptationLanguage ModelingLanguage ModellingMachine Translation+6