paper-with-me

Papers

SVD: A Large-Scale Short Video Dataset for Near-Duplicate Video Retrieval

2019-10-01 · ICCV 2019 10 · Qing-Yuan Jiang, Yi He, Gen Li, Jian Lin, Lei Li, Wu-Jun Li

With the explosive growth of video data in real applications, near-duplicate video retrieval (NDVR) has become indispensable and challenging, especially for short videos. However, all existing NDVR datasets are introduced for long videos. Furthermore, most of them are small-scale and lack of diversity due to the high cost of collecting and labeling near-duplicate videos. In this paper, we introduce a large-scale short video dataset, called SVD, for the NDVR task. SVD contains over 500,000 short videos and over 30,000 labeled videos of near-duplicates. We use multiple video mining techniques to construct positive/negative pairs. Furthermore, we design temporal and spatial transformations to mimic user-attack behavior in real applications for constructing more difficult variants of SVD. Experiments show that existing state-of-the-art NDVR methods, including real-value based and hashing based methods, fail to achieve satisfactory performance on this challenging dataset. The release of SVD dataset will foster research and system engineering in the NDVR area. The SVD dataset is available at https://svdbase.github.io.

📄 PDF Abstract BibTeX

Code (0)

등록된 구현이 없습니다.

Tasks

DiversityRetrievalVideo Retrieval

Similar Papers 제목 키워드 기반

CBVS: A Large-Scale Chinese Image-Text Benchmark for Real-World Short Video Search Scenarios

2024-01-19 · Xiangshuo Qiao, Xianxin Li, Xiaozhe Qu, Jie Zhang 외

Vision-Language Models pre-trained on large-scale image-text datasets have shown superior performance in downstream tasks such as image retrieval. Most of the images for pre-training are presented in the form of open dom…

Common Sense ReasoningImage Retrieval

Short-video Propagation Influence Rating: A New Real-world Dataset and A New Large Graph Model

2025-03-31 · Dizhan Xue, Jing Cui, Shengsheng Qian, Chuanrui Hu 외

Short-video platforms have gained immense popularity, captivating the interest of millions, if not billions, of users globally. Recently, researchers have highlighted the significance of analyzing the propagation of shor…

Video Propagation

HVM-1: Large-scale video models pretrained with nearly 5000 hours of human-like video data

2024-07-25 · A. Emin Orhan

We introduce Human-like Video Models (HVM-1), large-scale video models pretrained with nearly 5000 hours of curated human-like video data (mostly egocentric, temporally extended, continuous video recordings), using the s…

CREATE: A Benchmark for Chinese Short Video Retrieval and Title Generation

2021-11-16 · ACL ARR November 2021 11 · Anonymous

Previous works of video captioning aim to objectively describe the video's actual content, lack of subjective and attractive expression, limiting its practical application scenarios. Video titling is intended to achieve …

RetrievalVideo CaptioningVideo Retrieval

CREATE: A Benchmark for Chinese Short Video Retrieval and Title Generation

2022-03-31 · Ziqi Zhang, Yuxin Chen, Zongyang Ma, Zhongang Qi 외

Previous works of video captioning aim to objectively describe the video's actual content, which lacks subjective and attractive expression, limiting its practical application scenarios. Video titling is intended to achi…

RetrievalVideo CaptioningVideo Retrieval