paper-with-me

홈 › Papers

Event Recognition with Automatic Album Detection based on Sequential Processing, Neural Attention and Image Captioning

2019-11-25 · Andrey V. Savchenko

In this paper a new formulation of event recognition task is examined: it is required to predict event categories in a gallery of images, for which albums (groups of photos corresponding to a single event) are unknown. We propose the novel two-stage approach. At first, features are extracted in each photo using the pre-trained convolutional neural network. These features are classified individually. The scores of the classifier are used to group sequential photos into several clusters. Finally, the features of photos in each group are aggregated into a single descriptor using neural attention mechanism. This algorithm is optionally extended to improve the accuracy for classification of each image in an album. In contrast to conventional fine-tuning of convolutional neural networks (CNN) we proposed to use image captioning, i.e., generative model that converts images to textual descriptions. They are one-hot encoded and summarized into sparse feature vector suitable for learning of arbitrary classifier. Experimental study with Photo Event Collection and Multi-Label Curation of Flickr Events Dataset demonstrates that our approach is 9-20% more accurate than event recognition on single photos. Moreover, proposed method has 13-16% lower error rate than classification of groups of photos obtained with hierarchical clustering. It is experimentally shown that the image captions trained on Conceptual Captions dataset can be classified more accurately than the features from object detector, though they both are obviously not as rich as the CNN-based features. However, it is possible to combine our approach with conventional CNNs in an ensemble to provide the state-of-the-art results for several event datasets.

📄 PDF Abstract BibTeX arXiv:1911.11010

Code (0)

등록된 구현이 없습니다.

Tasks

ClusteringImage Captioning

Similar Papers 제목 키워드 기반

Recognizing and Curating Photo Albums via Event-Specific Image Importance

2017-07-19 · Yufei Wang, Zhe Lin, Xiaohui Shen, Radomir Mech 외

Automatic organization of personal photos is a problem with many real world ap- plications, and can be divided into two main tasks: recognizing the event type of the photo collection, and selecting interesting images fro…

PredictionVocal Bursts Type Prediction

PETA: Photo Albums Event Recognition using Transformers Attention

2021-09-26 · Tamar Glaser, Emanuel Ben-Baruch, Gilad Sharir, Nadav Zamir 외

In recent years the amounts of personal photos captured increased significantly, giving rise to new challenges in multi-image understanding and high-level image understanding. Event recognition in personal photo albums p…

Event-Specific Image Importance

2016-06-01 · CVPR 2016 6 · Yufei Wang, Zhe Lin, Xiaohui Shen, Radomir Mech 외

When creating a photo album of an event, people typically select a few important images to keep or share. There is some consistency in the process of choosing the important images, and discarding the unimportant ones. Mo…

Focal Visual-Text Attention for Memex Question Answering

2018-12-14 · IEEE Transactions on Pattern Analysis and Machine Intelligence (TPAMI) 2018 12 · Junwei Liang, Lu Jiang, Liangliang Cao, Yannis Kalantidis 외

Recent insights on language and vision with neural networks have been successfully applied to simple single-image visual question answering. However, to tackle real-life question answering problems on multimedia collecti…

Memex Question AnsweringQuestion AnsweringVisual Question AnsweringVisual Question Answering (VQA)

Supervised Learning Models for Early Detection of Albuminuria Risk in Type-2 Diabetes Mellitus Patients

2023-09-28 · Arief Purnama Muharram, Dicky Levenus Tahapary, Yeni Dwi Lestari, Randy Sarayar 외

Diabetes, especially T2DM, continues to be a significant health problem. One of the major concerns associated with diabetes is the development of its complications. Diabetic nephropathy, one of the chronic complication o…

Attribute