paper-with-me

홈 › Papers

Re-evaluating Automatic Metrics for Image Captioning

2016-12-22 · EACL 2017 4 · Mert Kilickaya, Aykut Erdem, Nazli Ikizler-Cinbis, Erkut Erdem

The task of generating natural language descriptions from images has received a lot of attention in recent years. Consequently, it is becoming increasingly important to evaluate such image captioning approaches in an automatic manner. In this paper, we provide an in-depth evaluation of the existing image captioning metrics through a series of carefully designed experiments. Moreover, we explore the utilization of the recently proposed Word Mover's Distance (WMD) document metric for the purpose of image captioning. Our findings outline the differences and/or similarities between metrics and their relative robustness by means of extensive correlation, accuracy and distraction based evaluations. Our results also demonstrate that WMD provides strong advantages over other metrics.

📄 PDF Abstract BibTeX arXiv:1612.07600

Code (0)

등록된 구현이 없습니다.

Tasks

Image Captioning

Similar Papers 제목 키워드 기반

StyleM: Stylized Metrics for Image Captioning Built with Contrastive N-grams

2022-01-04 · Chengxi Li, Brent Harrison

In this paper, we build two automatic evaluation metrics for evaluating the association between a machine-generated caption and a ground truth stylized caption: OnlyStyle and StyleCIDEr.

Image Captioning

REO-Relevance, Extraness, Omission: A Fine-grained Evaluation for Image Captioning

2019-09-05 · IJCNLP 2019 11 · Ming Jiang, Junjie Hu, Qiuyuan Huang, Lei Zhang 외

Popular metrics used for evaluating image captioning systems, such as BLEU and CIDEr, provide a single score to gauge the system's overall effectiveness. This score is often not informative enough to indicate what specif…

Image Captioning

Contrastive Semantic Similarity Learning for Image Captioning Evaluation with Intrinsic Auto-encoder

2021-06-29 · Chao Zeng, Tiesong Zhao, Sam Kwong

Automatically evaluating the quality of image captions can be very challenging since human language is quite flexible that there can be various expressions for the same meaning. Most of the current captioning metrics rel…

Image CaptioningRepresentation LearningSemantic SimilaritySemantic Textual Similarity+1

VELA: An LLM-Hybrid-as-a-Judge Approach for Evaluating Long Image Captions

2025-09-30 · Kazuki Matsuda, Yuiga Wada, Shinnosuke Hirano, Seitaro Otsuki 외 arxiv

In this study, we focus on the automatic evaluation of long and detailed image captions generated by multimodal Large Language Models (MLLMs). Most existing automatic evaluation metrics for image captioning are primarily…

Image Captioning

Text-to-Audio Grounding Based Novel Metric for Evaluating Audio Caption Similarity

2022-10-03 · Swapnil Bhosale, Rupayan Chakraborty, Sunil Kumar Kopparapu

Automatic Audio Captioning (AAC) refers to the task of translating an audio sample into a natural language (NL) text that describes the audio events, source of the events and their relationships. Unlike NL text generatio…

Audio captioningImage CaptioningTAGText Generation