paper-with-me

Papers

Automatic Description Generation from Images: A Survey of Models, Datasets, and Evaluation Measures

2016-01-15 · Raffaella Bernardi, Ruket Cakici, Desmond Elliott, Aykut Erdem, Erkut Erdem, Nazli Ikizler-Cinbis, Frank Keller, Adrian Muscat, Barbara Plank

Automatic description generation from natural images is a challenging problem that has recently received a large amount of interest from the computer vision and natural language processing communities. In this survey, we classify the existing approaches based on how they conceptualize this problem, viz., models that cast description as either generation problem or as a retrieval problem over a visual or multimodal representational space. We provide a detailed review of existing models, highlighting their advantages and disadvantages. Moreover, we give an overview of the benchmark image datasets and the evaluation measures that have been developed to assess the quality of machine-generated image descriptions. Finally we extrapolate future directions in the area of automatic image description generation.

📄 PDF Abstract BibTeX arXiv:1601.03896

Code (0)

등록된 구현이 없습니다.

Tasks

Image DescriptionRetrieval

Similar Papers 제목 키워드 기반

Exploring the Behavior of Classic REG Algorithms in the Description of Characters in 3D Images

2017-09-01 · WS 2017 9 · Gonzalo M{\'e}ndez, Raquel Herv{\'a}s, Susana Bautista, Adri{\'a}n Rabad{\'a}n 외

Describing people and characters can be very useful in different contexts, such as computational narrative or image description for the visually impaired. However, a review of the existing literature shows that the autom…

Image DescriptionReferring ExpressionReferring expression generationSurvey+1

Video Description: A Survey of Methods, Datasets and Evaluation Metrics

2018-06-01 · Nayyer Aafaq, Ajmal Mian, Wei Liu, Syed Zulqarnain Gilani 외

Video description is the automatic generation of natural language sentences that describe the contents of a given video. It has applications in human-robot interaction, helping the visually impaired and video subtitling.…

DiversityLanguage ModelingLanguage ModellingVideo Description

A Survey on Deep Learning and Explainability for Automatic Report Generation from Medical Images

2020-10-20 · Pablo Messina, Pablo Pino, Denis Parra, Alvaro Soto 외

Every year physicians face an increasing demand of image-based diagnosis from patients, a problem that can be addressed with recent artificial intelligence methods. In this context, we survey works in the area of automat…

Medical Report GenerationSurvey

GAN Computers Generate Arts? A Survey on Visual Arts, Music, and Literary Text Generation using Generative Adversarial Network

2021-08-09 · Sakib Shahriar

"Art is the lie that enables us to realize the truth." - Pablo Picasso. For centuries, humans have dedicated themselves to producing arts to convey their imagination. The advancement in technology and deep learning in pa…

Generative Adversarial NetworkText Generation

A Comprehensive Review on Recent Methods and Challenges of Video Description

2020-11-30 · Alok Singh, Thoudam Doren Singh, Sivaji Bandyopadhyay

Video description involves the generation of the natural language description of actions, events, and objects in the video. There are various applications of video description by filling the gap between languages and vis…

Machine TranslationSurveyVideo DescriptionVideo-Guided Machine Translation