paper-with-me

VidChapters-7M

홈페이지 · 논문 5편

VidChapters-7M is a dataset of 817K user-chaptered videos including 7M chapters in total. VidChapters-7M is automatically created from videos online in a scalable manner by scraping user-annotated chapters and hence without any additional manual annotation. It is designed for training and evaluating models for video chapter generation with or without ground-truth boundaries, and video chapter grounding, as well as for video-language pretraining.

VideosTexts English

벤치마크

Language-Based Temporal Localization on VidChapters-7M 결과 4개
Dense Video Captioning on VidChapters-7M 결과 2개
Video Chaptering on VidChapters-7M 결과 2개
Video Captioning on VidChapters-7M 결과 1개