paper-with-me

Papers Image Retrieval with Multi-Modal Query

“Image Retrieval with Multi-Modal Query” 태그가 달린 논문 10편 · 필터 해제

Collaborative Group: Composed Image Retrieval via Consensus Learning from Noisy Annotations

2023-06-03 · Xu Zhang, Zhedong Zheng, Linchao Zhu, Yi Yang

Composed image retrieval extends content-based image retrieval systems by enabling users to search using reference images and captions that describe their intention. Despite great progress in developing image-text compos…

Content-Based Image RetrievalImage RetrievalImage Retrieval with Multi-Modal QueryRetrieval+1

Composed Image Retrieval with Text Feedback via Multi-grained Uncertainty Regularization

2022-11-14 · Yiyang Chen, Zhedong Zheng, Wei Ji, Leigang Qu 외

We investigate composed image retrieval with text feedback. Users gradually look for the target of interest by moving from coarse to fine-grained feedback. However, existing methods merely focus on the latter, i.e., fine…

Composed Image Retrieval (CoIR)Image RetrievalImage Retrieval with Multi-Modal QueryRetrieval

Compositional Learning of Image-Text Query for Image Retrieval

2020-06-19 · Muhammad Umer Anwaar, Egor Labintcev, Martin Kleinsteuber

In this paper, we investigate the problem of retrieving images from a database based on a multi-modal (image-text) query. Specifically, the query text prompts some modification in the query image and the task is to retri…

Image RetrievalImage Retrieval with Multi-Modal QueryMetric Learning+1

Composing Text and Image for Image Retrieval - An Empirical Odyssey

2018-12-18 · CVPR 2019 6 · Nam Vo, Lu Jiang, Chen Sun, Kevin Murphy 외

In this paper, we study the task of image retrieval, where the input query is specified in the form of an image plus some text that describes desired modifications to the input image. For example, we may present an image…

Image RetrievalImage Retrieval with Multi-Modal QueryRetrieval

Attributes as Operators: Factorizing Unseen Attribute-Object Compositions

2018-03-27 · ECCV 2018 9 · Tushar Nagarajan, Kristen Grauman

We present a new approach to modeling visual attributes. Prior work casts attributes in a similar role as objects, learning a latent representation where properties (e.g., sliced) are recognized by classifiers much in th…

AttributeCompositional Zero-Shot LearningImage Retrieval with Multi-Modal QueryObject

FiLM: Visual Reasoning with a General Conditioning Layer

2017-09-22 · Ethan Perez, Florian Strub, Harm de Vries, Vincent Dumoulin 외

We introduce a general-purpose conditioning method for neural networks called FiLM: Feature-wise Linear Modulation. FiLM layers influence neural network computation via a simple, feature-wise affine transformation based …

Image Retrieval with Multi-Modal QueryVisual Question Answering (VQA)Visual Question Answering (VQA) Split AVisual Question Answering (VQA) Split B+1

Automatic Spatially-aware Fashion Concept Discovery

2017-08-03 · ICCV 2017 10 · Xintong Han, Zuxuan Wu, Phoenix X. Huang, Xiao Zhang 외

This paper proposes an automatic spatially-aware concept discovery approach using weakly labeled image-text data from shopping websites. We first fine-tune GoogleNet by jointly modeling clothing images and their correspo…

AttributeClusteringImage Retrieval with Multi-Modal QueryRetrieval

A simple neural network module for relational reasoning

2017-06-05 · NeurIPS 2017 12 · Adam Santoro, David Raposo, David G. T. Barrett, Mateusz Malinowski 외

Relational reasoning is a central component of generally intelligent behavior, but has proven difficult for neural networks to learn. In this paper we describe how to use Relation Networks (RNs) as a simple plug-and-play…

Image Retrieval with Multi-Modal QueryQuestion AnsweringRelational ReasoningVisual Question Answering+1

Image Question Answering using Convolutional Neural Network with Dynamic Parameter Prediction

2015-11-18 · CVPR 2016 6 · Hyeonwoo Noh, Paul Hongsuck Seo, Bohyung Han

We tackle image question answering (ImageQA) problem by learning a convolutional neural network (CNN) with a dynamic parameter layer whose weights are determined adaptively based on questions. For the adaptive parameter …

Image Retrieval with Multi-Modal QueryParameter PredictionPredictionQuestion Answering+1

Show and Tell: A Neural Image Caption Generator

2014-11-17 · CVPR 2015 6 · Oriol Vinyals, Alexander Toshev, Samy Bengio, Dumitru Erhan

Automatically describing the content of an image is a fundamental problem in artificial intelligence that connects computer vision and natural language processing. In this paper, we present a generative model based on a …

Image CaptioningImage Retrieval with Multi-Modal QuerySentenceText Generation+2
1–10 / 10