paper-with-me

홈 › Papers

Searching for COMETINHO: The Little Metric That Could

2022-06-01 · EAMT 2022 6 · Ricardo Rei, Ana C Farinha, José G.C. de Souza, Pedro G. Ramos, André F.T. Martins, Luisa Coheur, Alon Lavie

In recent years, several neural fine-tuned machine translation evaluation metrics such as COMET and BLEURT have been proposed. These metrics achieve much higher correlations with human judgments than lexical overlap metrics at the cost of computational efficiency and simplicity, limiting their applications to scenarios in which one has to score thousands of translation hypothesis (e.g. scoring multiple systems or Minimum Bayes Risk decoding). In this paper, we explore optimization techniques, pruning, and knowledge distillation to create more compact and faster COMET versions. Our results show that just by optimizing the code through the use of caching and length batching we can reduce inference time between 39% and 65% when scoring multiple systems. Also, we show that pruning COMET can lead to a 21% model reduction without affecting the model’s accuracy beyond 0.01 Kendall tau correlation. Furthermore, we present DISTIL-COMET a lightweight distilled version that is 80% smaller and 2.128x faster while attaining a performance close to the original model and above strong baselines such as BERTSCORE and PRISM.

📄 PDF Abstract BibTeX

Code (0)

등록된 구현이 없습니다.

Tasks

Computational EfficiencyKnowledge DistillationMachine TranslationTranslation

Methods 이 논문이 사용한 방법론

Pruning 설명 없음

Similar Papers 제목 키워드 기반

Are References Really Needed? Unbabel-IST 2021 Submission for the Metrics Shared Task

2021-11-01 · WMT (EMNLP) 2021 11 · Ricardo Rei, Ana C Farinha, Chrysoula Zerva, Daan van Stigt 외

In this paper, we present the joint contribution of Unbabel and IST to the WMT 2021 Metrics Shared Task. With this year’s focus on Multidimensional Quality Metric (MQM) as the ground-truth human assessment, our aim was t…

CPU

Computationally Efficient Learning of Statistical Manifolds

2021-02-22 · Fan Cheng, Anastasios Panagiotelis, Rob J Hyndman

Analyzing high-dimensional data with manifold learning algorithms often requires searching for the nearest neighbors of all observations. This presents a computational bottleneck in statistical manifold learning when obs…

Computational Efficiency

A Fingerprint Indexing Method Based on Minutia Descriptor and Clustering

2018-11-21 · Gwang-Il Ri, Chol-Gyun Ri, Su-Rim Ji

In this paper we propose a novel fingerprint indexing approach for speeding up in the fingerprint recognition system. What kind of features are used for indexing and how to employ the extracted features for searching are…

Clustering

Learning Style Similarity for Searching Infographics

2015-05-05 · Babak Saleh, Mira Dontcheva, Aaron Hertzmann, Zhicheng Liu

Infographics are complex graphic designs integrating text, images, charts and sketches. Despite the increasing popularity of infographics and the rapid growth of online design portfolios, little research investigates how…

Image RetrievalRetrieval

Multi-agent Searching System for Medical Information

2022-03-23 · Mariya Evtimova-Gardair

In the paper is proposed a model of multi-agent security system for searching a medical information in Internet. The advantages when using mobile agent are described, so that to perform searching in Internet. Nowadays, m…