paper-with-me

CrossTask

홈페이지 · 논문 54편

CrossTask dataset contains instructional videos, collected for 83 different tasks. For each task an ordered list of steps with manual descriptions is provided. The dataset is divided in two parts: 18 primary and 65 related tasks. Videos for the primary tasks are collected manually and provided with annotations for temporal step boundaries. Videos for the related tasks are collected automatically and don't have annotations. Source: CrossTask Image Source: https://arxiv.org/pdf/1903.08225v2.pdf

VideosTexts

벤치마크

Temporal Action Localization on CrossTask 결과 28개