Video-based Generative Performance Benchmarking (Temporal Understanding) 벤치마크
Video-based Generative Performance Benchmarking (Temporal Understanding) on VideoInstruct
gpt-score
- 2023-04-28 — LLaMA Adapter: gpt-score 1.98
- 2023-07-31 — MovieChat: gpt-score 2.24
- 2023-09-27 — BT-Adapter: gpt-score 2.34
- 2023-11-14 — Chat-UniVi: gpt-score 2.39
- 2023-11-28 — VideoChat2: gpt-score 2.66
- 2024-03-30 — ST-LLM: gpt-score 2.93
- 2024-11-04 — PPLLaVA-7B: gpt-score 3.21