Papers Text-to-video search
“Text-to-video search” 태그가 달린 논문 4편 · 필터 해제
Cross-modal Search Method of Technology Video based on Adversarial Learning and Feature Fusion
Technology videos contain rich multi-modal information. In cross-modal information search, the data features of different modalities cannot be compared directly, so the semantic gap between different modalities is a key …
Cross-Modal RetrievalRetrievalText-to-video searchBridging Video-text Retrieval with Multiple Choice Questions
Pre-training a model to learn transferable video-text representation for retrieval has attracted a lot of attention in recent years. Previous dominant works mainly adopt two separate encoders for efficient retrieval, but…
Action RecognitionLinear evaluationMultiple-choiceRetrieval+8Multilingual Multimodal Pre-training for Zero-Shot Cross-Lingual Transfer of Vision-Language Models
This paper studies zero-shot cross-lingual transfer of vision-language models. Specifically, we focus on multilingual text-to-video search and propose a Transformer-based model that learns contextualized multilingual mul…
Cross-Lingual TransferImage RetrievalText-to-video searchZero-Shot Cross-Lingual TransferMultilingual Multimodal Pretraining for Zero-Shot Cross-Lingual Transfer of Vision-Language Models
This paper studies zero-shot cross-lingual transfer of vision-language models. Specifically, we focus on multilingual text-to-video search and propose a Transformer-based model that learns contextualized multilingual mul…
Cross-Lingual TransferImage RetrievalText-to-video searchZero-Shot Cross-Lingual Transfer