paper-with-me

Human Judgment Classification

1개 벤치마크 · 논문 2편 · 이 태스크의 논문 보기 →

Benchmarks

Pascal-50S

결과 3개

Most implemented

Papers

Mutual Information Divergence: A Unified Metric for Multimodal Generative Models

2022-05-25 · Jin-Hwa Kim, Yunji Kim, Jiyoung Lee, Kang Min Yoo 외

Text-to-image generation and image captioning are recently emerged as a new experimental paradigm to assess machine intelligence. They predict continuous quantity accompanied by their sampling techniques in the generatio…

Hallucination Pair-wise Detection (1-ref)Hallucination Pair-wise Detection (4-ref)Human Judgment ClassificationHuman Judgment Correlation+5

CLIPScore: A Reference-free Evaluation Metric for Image Captioning

2021-04-18 · EMNLP 2021 11 · Jack Hessel, Ari Holtzman, Maxwell Forbes, Ronan Le Bras 외

Image captioning has conventionally relied on reference-based automatic evaluations, where machine captions are compared against captions written by humans. This is in contrast to the reference-free manner in which human…

Hallucination Pair-wise Detection (1-ref)Hallucination Pair-wise Detection (4-ref)Human Judgment ClassificationHuman Judgment Correlation+1