paper-with-me

홈 › Papers

Image Search with Text Feedback by Additive Attention Compositional Learning

2022-03-08 · Yuxin Tian, Shawn Newsam, Kofi Boakye

Effective image retrieval with text feedback stands to impact a range of real-world applications, such as e-commerce. Given a source image and text feedback that describes the desired modifications to that image, the goal is to retrieve the target images that resemble the source yet satisfy the given modifications by composing a multi-modal (image-text) query. We propose a novel solution to this problem, Additive Attention Compositional Learning (AACL), that uses a multi-modal transformer-based architecture and effectively models the image-text contexts. Specifically, we propose a novel image-text composition module based on additive attention that can be seamlessly plugged into deep neural networks. We also introduce a new challenging benchmark derived from the Shopping100k dataset. AACL is evaluated on three large-scale datasets (FashionIQ, Fashion200k, and Shopping100k), each with strong baselines. Extensive experiments show that AACL achieves new state-of-the-art results on all three datasets.

📄 PDF Abstract BibTeX arXiv:2203.03809

Code (0)

등록된 구현이 없습니다.

Tasks

Image RetrievalRetrieval

Methods 이 논문이 사용한 방법론

Tanh Activation 설명 없음

Similar Papers 제목 키워드 기반

Image Search With Text Feedback by Visiolinguistic Attention Learning

2020-06-01 · CVPR 2020 6 · Yanbei Chen, Shaogang Gong, Loris Bazzani

Image search with text feedback has promising impacts in various real-world applications, such as e-commerce and internet search. Given a reference image and text feedback from user, the goal is to retrieve images that n…

AttributeDeep AttentionImage RetrievalMultimodal Deep Learning

MCTSteg: A Monte Carlo Tree Search-based Reinforcement Learning Framework for Universal Non-additive Steganography

2021-03-25 · Xianbo Mo, Shunquan Tan, Bin Li, Jiwu Huang

Recent research has shown that non-additive image steganographic frameworks effectively improve security performance through adjusting distortion distribution. However, as far as we know, all of the existing non-additive…

Self-Learning

SAC: Semantic Attention Composition for Text-Conditioned Image Retrieval

2020-09-03 · Surgan Jandial, Pinkesh Badjatiya, Pranit Chawla, Ayush Chopra 외

The ability to efficiently search for images is essential for improving the user experiences across various products. Incorporating user feedback, via multi-modal inputs, to navigate visual search can help tailor retriev…

Image RetrievalNavigateRetrieval

Robust Model Predictive Control with Polytopic Model Uncertainty through System Level Synthesis

2022-03-21 · Shaoru Chen, Victor M. Preciado, Manfred Morari, Nikolai Matni

We propose a robust model predictive control (MPC) method for discrete-time linear systems with polytopic model uncertainty and additive disturbances. Optimizing over linear time-varying (LTV) state feedback controllers …

modelModel Predictive Control

AttentionCode: Ultra-Reliable Feedback Codes for Short-Packet Communications

2022-05-30 · Yulin Shao, Emre Ozfatura, Alberto Perotti, Branislav Popovic 외

Ultra-reliable short-packet communication is a major challenge in future wireless networks with critical applications. To achieve ultra-reliable communications beyond 99.999%, this paper envisions a new interaction-based…