TVR
TV show Retrieval
홈페이지 · 논문 34편
A new multimodal retrieval dataset. TVR requires systems to understand both videos and their associated subtitle (dialogue) texts, making it more realistic. The dataset contains 109K queries collected on 21.8K videos from 6 TV shows of diverse genres, where each query is associated with a tight temporal window. Source: [TVR: A Large-Scale Dataset for Video-Subtitle Moment Retrieval](/paper/tvr-a-large-scale-dataset-for-video-subtitle)
벤치마크
Video Retrieval on TVR
결과 4개