Learning High-level Image Representation for Image Retrieval via Multi-Task DNN using Clickthrough Data
Image retrieval refers to finding relevant images from an image database for a query, which is considered difficult for the gap between low-level representation of images and high-level representation of queries. Recently further developed Deep Neural Network sheds light on automatically learning high-level image representation from raw pixels. In this paper, we proposed a multi-task DNN learned for image retrieval, which contains two parts, i.e., query-sharing layers for image representation computation and query-specific layers for relevance estimation. The weights of multi-task DNN are learned on clickthrough data by Ring Training. Experimental results on both simulated and real dataset show the effectiveness of the proposed method.
Code (0)
등록된 구현이 없습니다.
Tasks
Image RetrievalRetrievalSimilar Papers 제목 키워드 기반
Beyond Instance-Level Image Retrieval: Leveraging Captions to Learn a Global Visual Representation for Semantic Retrieval
Querying with an example image is a simple and intuitive interface to retrieve information from a visual database. Most of the research in image retrieval has focused on the task of instance-level image retrieval, where …
Image RetrievalRetrievalSemantic RetrievalSemantic Similarity+1Cross-modal Scene Graph Matching for Relationship-aware Image-Text Retrieval
Image-text retrieval of natural scenes has been a popular research topic. Since image and text are heterogeneous cross-modal data, one of the key challenges is how to learn comprehensive yet unified representations to ex…
Graph MatchingImage-text RetrievalRetrievalText RetrievalContent-Based Image Retrieval Based on Late Fusion of Binary and Local Descriptors
One of the challenges in Content-Based Image Retrieval (CBIR) is to reduce the semantic gaps between low-level features and high-level semantic concepts. In CBIR, the images are represented in the feature space and the p…
Content-Based Image RetrievalImage RetrievalRetrievalTranslationSingle Shot Scene Text Retrieval
Textual information found in scene images provides high level semantic information about the image and its context and it can be leveraged for better scene understanding. In this paper we address the problem of scene tex…
Image RetrievalRetrievalScene UnderstandingText RetrievalCross-modal Image Retrieval with Deep Mutual Information Maximization
In this paper, we study the cross-modal image retrieval, where the inputs contain a source image plus some text that describes certain modifications to this image and the desired image. Prior work usually uses a three-st…
Cross-Modal RetrievalImage RetrievalMetric LearningRetrieval+1