paper-with-me

Papers

Benchmarking Multimodal Large Language Models for Face Recognition

2025-10-16 · Hatef Otroshi Shahreza, Sébastien Marcel arxiv

Multimodal large language models (MLLMs) have achieved remarkable performance across diverse vision-and-language tasks. However, their potential in face recognition remains underexplored. In particular, the performance of open-source MLLMs needs to be evaluated and compared with existing face recognition models on standard benchmarks with similar protocol. In this work, we present a systematic benchmark of state-of-the-art MLLMs for face recognition on several face recognition datasets, including LFW, CALFW, CPLFW, CFP, AgeDB and RFW. Experimental results reveal that while MLLMs capture rich semantic cues useful for face-related tasks, they lag behind specialized models in high-precision recognition scenarios in zero-shot applications. This benchmark provides a foundation for advancing MLLM-based face recognition, offering insights for the design of next-generation models with higher accuracy and generalization. The source code of our benchmark is publicly available in the project page.

📄 PDF Abstract BibTeX arXiv:2510.14866

Code (0)

등록된 구현이 없습니다.

Tasks

Face Recognition

Similar Papers 제목 키워드 기반

Demographic Fairness in Multimodal LLMs: A Benchmark of Gender and Ethnicity Bias in Face Verification

2026-03-26 · Ünsal Öztürk, Hatef Otroshi Shahreza, Sébastien Marcel arxiv

Multimodal Large Language Models (MLLMs) have recently been explored as face verification systems that determine whether two face images are of the same person. Unlike dedicated face recognition systems, MLLMs approach t…

Face VerificationFace Recognition

Benchmarking Multimodal Sentiment Analysis

2017-07-29 · Erik Cambria, Devamanyu Hazarika, Soujanya Poria, Amir Hussain 외

We propose a framework for multimodal sentiment analysis and emotion recognition using convolutional neural network-based feature extraction from text and visual modalities. We obtain a performance improvement of 10% ove…

BenchmarkingEmotion RecognitionMultimodal Sentiment AnalysisSentiment Analysis

PsOCR: Benchmarking Large Multimodal Models for Optical Character Recognition in Low-resource Pashto Language

2025-05-15 · Ijazul Haq, Yingjie Zhang, Irfan Ali Khan

This paper evaluates the performance of Large Multimodal Models (LMMs) on Optical Character Recognition (OCR) in the low-resource Pashto language. Natural Language Processing (NLP) in Pashto faces several challenges due …

BenchmarkingOptical Character RecognitionOptical Character Recognition (OCR)

Evaluating Multimodal Large Language Models for Heterogeneous Face Recognition

2026-01-21 · Hatef Otroshi Shahreza, Anjith George, Sébastien Marcel arxiv

Multimodal Large Language Models (MLLMs) have recently demonstrated strong performance on a wide range of vision-language tasks, raising interest in their potential use for biometric applications. In this paper, we condu…

Heterogeneous Face Recognition

On the use of automatically generated synthetic image datasets for benchmarking face recognition

2021-06-08 · Laurent Colbois, Tiago de Freitas Pereira, Sébastien Marcel

The availability of large-scale face datasets has been key in the progress of face recognition. However, due to licensing issues or copyright infringement, some datasets are not available anymore (e.g. MS-Celeb-1M). Rece…

BenchmarkingFace Recognition