Human Judgment Classification
1개 벤치마크 · 논문 2편 · 이 태스크의 논문 보기 →
Benchmarks
Pascal-50S
결과 3개
Most implemented
CLIPScore: A Reference-free Evaluation Metric for Image Captioning
2021-04-18 · 구현 3개
Papers
Mutual Information Divergence: A Unified Metric for Multimodal Generative Models
2022-05-25
· Jin-Hwa Kim, Yunji Kim, Jiyoung Lee, Kang Min Yoo 외
Text-to-image generation and image captioning are recently emerged as a new experimental paradigm to assess machine intelligence. They predict continuous quantity accompanied by their sampling techniques in the generatio…
Hallucination Pair-wise Detection (1-ref)Hallucination Pair-wise Detection (4-ref)Human Judgment ClassificationHuman Judgment Correlation+5CLIPScore: A Reference-free Evaluation Metric for Image Captioning
2021-04-18 · EMNLP 2021 11
· Jack Hessel, Ari Holtzman, Maxwell Forbes, Ronan Le Bras 외
Image captioning has conventionally relied on reference-based automatic evaluations, where machine captions are compared against captions written by humans. This is in contrast to the reference-free manner in which human…
Hallucination Pair-wise Detection (1-ref)Hallucination Pair-wise Detection (4-ref)Human Judgment ClassificationHuman Judgment Correlation+1