paper-with-me

Papers

Captioning Images with Novel Objects via Online Vocabulary Expansion

2020-03-06 · Mikihiro Tanaka, Tatsuya Harada

In this study, we introduce a low cost method for generating descriptions from images containing novel objects. Generally, constructing a model, which can explain images with novel objects, is costly because of the following: (1) collecting a large amount of data for each category, and (2) retraining the entire system. If humans see a small number of novel objects, they are able to estimate their properties by associating their appearance with known objects. Accordingly, we propose a method that can explain images with novel objects without retraining using the word embeddings of the objects estimated from only a small number of image features of the objects. The method can be integrated with general image-captioning models. The experimental results show the effectiveness of our approach.

📄 PDF Abstract BibTeX arXiv:2003.03305

Code (0)

등록된 구현이 없습니다.

Tasks

Image CaptioningWord Embeddings

Similar Papers 제목 키워드 기반

Guided Open Vocabulary Image Captioning with Constrained Beam Search

2016-12-02 · EMNLP 2017 9 · Peter Anderson, Basura Fernando, Mark Johnson, Stephen Gould

Existing image captioning models do not generalize well to out-of-domain images containing novel scenes or objects. This limitation severely hinders the use of these models in real world applications dealing with images …

Image CaptioningTAGWord Embeddings

Pointing Novel Objects in Image Captioning

2019-04-25 · CVPR 2019 6 · Yehao Li, Ting Yao, Yingwei Pan, Hongyang Chao 외

Image captioning has received significant attention with remarkable improvements in recent advances. Nevertheless, images in the wild encapsulate rich knowledge and cannot be sufficiently described with models built on i…

DecoderImage CaptioningObjectObject Recognition+1

Auto-Vocabulary 3D Object Detection

2025-12-18 · Haomeng Zhang, Kuan-Chuan Peng, Suhas Lohit, Raymond A. Yeh arxiv

Open-vocabulary 3D object detection methods are able to localize 3D boxes of classes unseen during training. Despite the name, existing methods rely on user-specified classes both at training and inference. We propose to…

3D Object DetectionImage Captioning

NOC-REK: Novel Object Captioning with Retrieved Vocabulary from External Knowledge

2022-03-28 · CVPR 2022 1 · Duc Minh Vo, Hong Chen, Akihiro Sugimoto, Hideki Nakayama

Novel object captioning aims at describing objects absent from training data, with the key ingredient being the provision of object vocabulary to the model. Although existing methods heavily rely on an object detection m…

Caption GenerationObjectobject-detectionObject Detection+1

Good News, Everyone! Context driven entity-aware captioning for news images

2019-04-02 · CVPR 2019 6 · Ali Furkan Biten, Lluis Gomez, Marçal Rusiñol, Dimosthenis Karatzas

Current image captioning systems perform at a merely descriptive level, essentially enumerating the objects in the scene and their relations. Humans, on the contrary, interpret images by integrating several sources of pr…

ArticlesDescriptiveImage Captioning