paper-with-me

Papers

KoBE: Knowledge-Based Machine Translation Evaluation

2020-09-23 · Findings of the Association for Computational Linguistics 2020 · Zorik Gekhman, Roee Aharoni, Genady Beryozkin, Markus Freitag, Wolfgang Macherey

We propose a simple and effective method for machine translation evaluation which does not require reference translations. Our approach is based on (1) grounding the entity mentions found in each source sentence and candidate translation against a large-scale multilingual knowledge base, and (2) measuring the recall of the grounded entities found in the candidate vs. those found in the source. Our approach achieves the highest correlation with human judgements on 9 out of the 18 language pairs from the WMT19 benchmark for evaluation without references, which is the largest number of wins for a single evaluation method on this task. On 4 language pairs, we also achieve higher correlation with human judgements than BLEU. To foster further research, we release a dataset containing 1.8 million grounded entity mentions across 18 language pairs from the WMT19 metrics track data.

📄 PDF Abstract BibTeX arXiv:2009.11027

Code (1)

zorikg/KoBE 공식 구현

Tasks

Machine TranslationSentenceTranslation

Similar Papers 제목 키워드 기반

IrokoBench: A New Benchmark for African Languages in the Age of Large Language Models

2024-06-05 · David Ifeoluwa Adelani, Jessica Ojo, Israel Abebe Azime, Jian Yun Zhuang 외

Despite the widespread adoption of Large language models (LLMs), their remarkable capabilities remain limited to a few high-resource languages. Additionally, many low-resource languages (\eg African languages) are often …

Mathematical ReasoningNatural Language InferenceQuestion Answeringtext-classification+1

KOBEST: Korean Balanced Evaluation of Significant Tasks

2022-04-09 · COLING 2022 10 · Dohyeong Kim, Myeongjun Jang, Deuk Sin Kwon, Eric Davis

A well-formulated benchmark plays a critical role in spurring advancements in the natural language processing (NLP) field, as it allows objective and precise evaluation of diverse models. As modern language models (LMs) …

UI-KOBE: Knowledge-Oriented Behavior Exploration for Lightweight Graph-Guided GUI Agents

2026-05-28 · Yuxiang Chai, Han Xiao, Xinyu Fu, Jinpeng Chen 외 arxiv

Recent advances in mobile GUI agents have shown strong potential for automating mobile tasks, but most effective systems still depend on large vision-language models for screenshot understanding and long-horizon planning…

Towards Knowledge-Based Personalized Product Description Generation in E-commerce

2019-03-29 · Qibin Chen, Junyang Lin, Yichang Zhang, Hongxia Yang 외

Quality product descriptions are critical for providing competitive customer experience in an e-commerce platform. An accurate and attractive description not only helps customers make an informed decision but also improv…

Text Generation

Meteor++: Incorporating Copy Knowledge into Machine Translation Evaluation

2018-10-01 · WS 2018 10 · Yinuo Guo, Chong Ruan, Junfeng Hu

In machine translation evaluation, a good candidate translation can be regarded as a paraphrase of the reference. We notice that some words are always copied during paraphrasing, which we call \textbf{copy knowledge}. Co…

Machine TranslationSentenceText GenerationTranslation