Efficient Large-Scale Face Clustering Using an Online Mixture of Gaussians
In this work, we address the problem of large-scale online face clustering: given a continuous stream of unknown faces, create a database grouping the incoming faces by their identity. The database must be updated every time a new face arrives. In addition, the solution must be efficient, accurate and scalable. For this purpose, we present an online gaussian mixture-based clustering method (OGMC). The key idea of this method is the proposal that an identity can be represented by more than just one distribution or cluster. Using feature vectors (f-vectors) extracted from the incoming faces, OGMC generates clusters that may be connected to others depending on their proximity and their robustness. Every time a cluster is updated with a new sample, its connections are also updated. With this approach, we reduce the dependency of the clustering process on the order and the size of the incoming data and we are able to deal with complex data distributions. Experimental results show that the proposed approach outperforms state-of-the-art clustering methods on large-scale face clustering benchmarks not only in accuracy, but also in efficiency and scalability.
Code (0)
등록된 구현이 없습니다.
Tasks
ClusteringFace ClusteringSimilar Papers 제목 키워드 기반
AN ONLINE ALGORITHM FOR CONSTRAINED FACE CLUSTERING IN VIDEOS
We address the problem of face clustering in long, real world videos.This is a challenging task because faces in such videos exhibit wid evariability in scale, pose, illumination, expressions, and may also be par…
ClusteringFace ClusteringOnline ClusteringMemoized Online Variational Inference for Dirichlet Process Mixture Models
Variational inference algorithms provide the most effective framework for large-scale training of Bayesian nonparametric models. Stochastic online approaches are promising, but are sensitive to the chosen learning rate …
ClusteringDenoisingImage ClusteringVariational InferenceScalable Clustering: Large Scale Unsupervised Learning of Gaussian Mixture Models with Outliers
Clustering is a widely used technique with a long and rich history in a variety of areas. However, most existing algorithms do not scale well to large datasets, or are missing theoretical guarantees of convergence. This …
ClusteringDeep Clustering Based on a Mixture of Autoencoders
In this paper we propose a Deep Autoencoder MIxture Clustering (DAMIC) algorithm based on a mixture of deep autoencoders where each cluster is represented by an autoencoder. A clustering network transforms the data into …
ClusteringDeep ClusteringRetraining-Free Merging of Sparse MoE via Hierarchical Clustering
Sparse Mixture-of-Experts (SMoE) models represent a significant advancement in large language model (LLM) development through their efficient parameter utilization. These models achieve substantial performance improvemen…
ClusteringLanguage ModelingLanguage ModellingLarge Language Model+1