paper-with-me

Papers

RNN Fisher Vectors for Action Recognition and Image Annotation

2015-12-12 · Guy Lev, Gil Sadeh, Benjamin Klein, Lior Wolf

Recurrent Neural Networks (RNNs) have had considerable success in classifying and predicting sequences. We demonstrate that RNNs can be effectively used in order to encode sequences and provide effective representations. The methodology we use is based on Fisher Vectors, where the RNNs are the generative probabilistic models and the partial derivatives are computed using backpropagation. State of the art results are obtained in two central but distant tasks, which both rely on sequences: video action recognition and image annotation. We also show a surprising transfer learning result from the task of image annotation to the task of video action recognition.

📄 PDF Abstract BibTeX arXiv:1512.03958

Code (0)

등록된 구현이 없습니다.

Tasks

Action RecognitionTemporal Action LocalizationTransfer Learning

Similar Papers 제목 키워드 기반

Fisher Vectors Derived from Hybrid Gaussian-Laplacian Mixture Models for Image Annotation

2014-11-26 · Benjamin Klein, Guy Lev, Gil Sadeh, Lior Wolf

In the traditional object recognition pipeline, descriptors are densely sampled over an image, pooled into a high dimensional non-linear representation and then passed to a classifier. In recent years, Fisher Vectors hav…

Image RetrievalObject RecognitionSentence

Hyper-Fisher Vectors for Action Recognition

2015-09-28 · Sanath Narayan, Kalpathi R. Ramakrishnan

In this paper, a novel encoding scheme combining Fisher vector and bag-of-words encodings has been proposed for recognizing action in videos. The proposed Hyper-Fisher vector encoding is sum of local Fisher vectors which…

Action RecognitionTemporal Action Localization

Associating Neural Word Embeddings With Deep Image Representations Using Fisher Vectors

2015-06-01 · CVPR 2015 6 · Benjamin Klein, Guy Lev, Gil Sadeh, Lior Wolf

In recent years, the problem of associating a sentence with an image has gained a lot of attention. This work continues to push the envelope and makes further progress in the performance of image annotation and image sea…

Image RetrievalSentenceWord Embeddings

They are wearing a mask! Identification of Subjects Wearing a Surgical Mask from their Speech by means of x-vectors and Fisher Vectors

2020-08-23 · José Vicente Egas-López

Challenges based on Computational Paralinguistics in the INTERSPEECH Conference have always had a good reception among the attendees owing to its competitive academic and research demands. This year, the INTERSPEECH 2020…

Speaker Recognition

An end-to-end generative framework for video segmentation and recognition

2015-09-07 · Hilde Kuehne, Juergen Gall, Thomas Serre

We describe an end-to-end generative approach for the segmentation and recognition of human activities. In this approach, a visual representation based on reduced Fisher Vectors is combined with a structured temporal mod…

Video SegmentationVideo Semantic Segmentation