More About VLAD: A Leap From Euclidean to Riemannian Manifolds
This paper takes a step forward in image and video coding by extending the well-known Vector of Locally Aggregated Descriptors (VLAD) onto an extensive space of curved Riemannian manifolds. We provide a comprehensive mathematical framework that formulates the aggregation problem of such manifold data into an elegant solution. In particular, we consider structured descriptors from visual data, namely Region Covariance Descriptors and linear subspaces that reside on the manifold of Symmetric Positive Definite matrices and the Grassmannian manifolds, respectively. Through rigorous experimental validation, we demonstrate the superior performance of this novel Riemannian VLAD descriptor on several visual classification tasks including video-based face recognition, dynamic scene recognition, and head pose classification.
Code (0)
등록된 구현이 없습니다.
Tasks
ClassificationFace RecognitionGeneral ClassificationScene RecognitionSimilar Papers 제목 키워드 기반
Learning Mid-level Words on Riemannian Manifold for Action Recognition
Human action recognition remains a challenging task due to the various sources of video data and large intra-class variations. It thus becomes one of the key issues in recent research to explore effective and robust repr…
Action RecognitionClusteringTemporal Action LocalizationWhen VLAD met Hilbert
Vectors of Locally Aggregated Descriptors (VLAD) have emerged as powerful image/video representations that compete with or even outperform state-of-the-art approaches on many challenging visual recognition tasks. In this…
General ClassificationLearning Euclidean-to-Riemannian Metric for Point-to-Set Classification
In this paper, we focus on the problem of point-to-set classification, where single points are matched against sets of correlated points. Since the points commonly lie in Euclidean space while the sets are typically mode…
ClassificationGeneral ClassificationMetric LearningCross Euclidean-to-Riemannian Metric Learning with Application to Face Recognition from Video
Riemannian manifolds have been widely employed for video representations in visual classification tasks including video-based face recognition. The success mainly derives from learning a discriminant Riemannian metric wh…
Face RecognitionMetric LearningAdaptive Log-Euclidean Metrics for SPD Matrix Learning
Symmetric Positive Definite (SPD) matrices have received wide attention in machine learning due to their intrinsic capacity to encode underlying structural correlation in data. Many successful Riemannian metrics have bee…