paper-with-me

홈 › Papers

Order-Embeddings of Images and Language

2015-11-19 · Ivan Vendrov, Ryan Kiros, Sanja Fidler, Raquel Urtasun

Hypernymy, textual entailment, and image captioning can be seen as special cases of a single visual-semantic hierarchy over words, sentences, and images. In this paper we advocate for explicitly modeling the partial order structure of this hierarchy. Towards this goal, we introduce a general method for learning ordered representations, and show how it can be applied to a variety of tasks involving images and language. We show that the resulting representations improve performance over current approaches for hypernym prediction and image-caption retrieval.

📄 PDF Abstract BibTeX arXiv:1511.06361

Code (2)

iesl/geometric_graph_embedding pytorch
ivendrov/order-embedding

Tasks

Cross-Modal RetrievalImage CaptioningNatural Language InferenceRetrieval

Similar Papers 제목 키워드 기반

UAEM-ITAM at SemEval-2022 Task 5: Vision-Language Approach to Recognize Misogynous Content in Memes

2022-07-01 · SemEval (NAACL) 2022 7 · Edgar Roman-Rangel, Jorge Fuentes-Pacheco, Jorge Hermosillo Valadez

In the context of the Multimedia Automatic Misogyny Identification (MAMI) competition 2022, we developed a framework for extracting lexical-semantic features from text and combine them with semantic descriptions of image…

Dimensionality Reduction

Order embeddings and character-level convolutions for multimodal alignment

2017-06-03 · Jônatas Wehrmann, Anderson Mattjie, Rodrigo C. Barros

With the novel and fast advances in the area of deep neural networks, several challenging image-based tasks have been recently approached by researchers in pattern recognition and computer vision. In this paper, we addre…

RetrievalSemantic correspondenceWord Embeddings

Bridging Languages through Images with Deep Partial Canonical Correlation Analysis

2018-07-01 · ACL 2018 7 · Guy Rotman, Ivan Vuli{\'c}, Roi Reichart

We present a deep neural network that leverages images to improve bilingual text embeddings. Relying on bilingual image tags and descriptions, our approach conditions text embedding induction on the shared visual informa…

Image DescriptionImage RetrievalQuestion AnsweringRepresentation Learning+3

Visual Grounding of Inter-lingual Word-Embeddings

2022-09-08 · Wafaa Mohammed, Hassan Shahmohammadi, Hendrik P. A. Lensch, R. Harald Baayen

Visual grounding of Language aims at enriching textual representations of language with multiple sources of visual knowledge such as images and videos. Although visual grounding is an area of intense research, inter-ling…

Visual GroundingWord EmbeddingsWord Similarity

Retrieving Similar E-Commerce Images Using Deep Learning

2019-01-11 · Rishab Sharma, Anirudha Vishvakarma

In this paper, we propose a deep convolutional neural network for learning the embeddings of images in order to capture the notion of visual similarity. We present a deep siamese architecture that when trained on positiv…

Deep LearningFine-Grained Visual RecognitionImage RetrievalProduct Recommendation+2