paper-with-me

Papers

EditBoard: Towards a Comprehensive Evaluation Benchmark for Text-Based Video Editing Models

2024-09-15 · Yupeng Chen, Penglin Chen, XiaoYu Zhang, Yixian Huang, Qian Xie

The rapid development of diffusion models has significantly advanced AI-generated content (AIGC), particularly in Text-to-Image (T2I) and Text-to-Video (T2V) generation. Text-based video editing, leveraging these generative capabilities, has emerged as a promising field, enabling precise modifications to videos based on text prompts. Despite the proliferation of innovative video editing models, there is a conspicuous lack of comprehensive evaluation benchmarks that holistically assess these models' performance across various dimensions. Existing evaluations are limited and inconsistent, typically summarizing overall performance with a single score, which obscures models' effectiveness on individual editing tasks. To address this gap, we propose EditBoard, the first comprehensive evaluation benchmark for text-based video editing models. EditBoard encompasses nine automatic metrics across four dimensions, evaluating models on four task categories and introducing three new metrics to assess fidelity. This task-oriented benchmark facilitates objective evaluation by detailing model performance and providing insights into each model's strengths and weaknesses. By open-sourcing EditBoard, we aim to standardize evaluation and advance the development of robust video editing models.

📄 PDF Abstract BibTeX arXiv:2409.09668

Code (1)

samchen2003/editboard 공식 구현

Tasks

Video Editing

Methods 이 논문이 사용한 방법론

Diffusion Diffusion models generate samples by gradually removing noise from a signal, and their training objective can be expressed as a reweighted variational lower-bound…

Similar Papers 제목 키워드 기반

MLVU: Benchmarking Multi-task Long Video Understanding

2024-06-06 · CVPR 2025 1 · Junjie Zhou, Yan Shu, Bo Zhao, Boya Wu 외

The evaluation of Long Video Understanding (LVU) performance poses an important but challenging research problem. Despite previous efforts, the existing video understanding benchmarks are severely constrained by several …

BenchmarkingVideo Understanding

VABench: A Comprehensive Benchmark for Audio-Video Generation

2025-12-10 · Daili Hua, Xizhi Wang, Bohan Zeng, Xinyi Huang 외 arxiv

Recent advances in video generation have been remarkable, enabling models to produce visually compelling videos with synchronized audio. While existing video generation benchmarks provide comprehensive metrics for visual…

Video Generation

LOVE: Benchmarking and Evaluating Text-to-Video Generation and Video-to-Text Interpretation

2025-05-17 · Jiarui Wang, Huiyu Duan, Ziheng Jia, Yu Zhao 외

Recent advancements in large multimodal models (LMMs) have driven substantial progress in both text-to-video (T2V) generation and video-to-text (V2T) interpretation tasks. However, current AI-generated videos (AIGVs) sti…

BenchmarkingQuestion AnsweringText-to-Video GenerationVideo Alignment+1

AIGCBench: Comprehensive Evaluation of Image-to-Video Content Generated by AI

2024-01-03 · Fanda Fan, Chunjie Luo, Wanling Gao, Jianfeng Zhan

The burgeoning field of Artificial Intelligence Generated Content (AIGC) is witnessing rapid advancements, particularly in video generation. This paper introduces AIGCBench, a pioneering comprehensive and scalable benchm…

Video AlignmentVideo Generation

VBench++: Comprehensive and Versatile Benchmark Suite for Video Generative Models

2024-11-20 · Ziqi Huang, Fan Zhang, Xiaojie Xu, Yinan He 외

Video generation has witnessed significant advancements, yet evaluating these models remains a challenge. A comprehensive evaluation benchmark for video generation is indispensable for two reasons: 1) Existing metrics do…

BenchmarkingImage GenerationImage to Video GenerationVideo Generation