paper-with-me

홈 › Papers

Cross-modal Multi-task Learning for Graphic Recognition of Caricature Face

2020-03-10 · Zuheng Ming, Jean-Christophe Burie, Muhammad Muzzamil Luqman

Face recognition of realistic visual images has been well studied and made a significant progress in the recent decade. Unlike the realistic visual images, the face recognition of the caricatures is far from the performance of the visual images. This is largely due to the extreme non-rigid distortions of the caricatures introduced by exaggerating the facial features to strengthen the characters. The heterogeneous modalities of the caricatures and the visual images result the caricature-visual face recognition is a cross-modal problem. In this paper, we propose a method to conduct caricature-visual face recognition via multi-task learning. Rather than the conventional multi-task learning with fixed weights of tasks, this work proposes an approach to learn the weights of tasks according to the importance of tasks. The proposed multi-task learning with dynamic tasks weights enables to appropriately train the hard task and easy task instead of being stuck in the over-training easy task as conventional methods. The experimental results demonstrate the effectiveness of the proposed dynamic multi-task learning for cross-modal caricature-visual face recognition. The performances on the datasets CaVI and WebCaricature show the superiority over the state-of-art methods.

📄 PDF Abstract BibTeX arXiv:2003.05787

Code (0)

등록된 구현이 없습니다.

Tasks

CaricatureFace RecognitionMulti-Task Learning

Similar Papers 제목 키워드 기반

Demographic and Linguistic Bias Evaluation in Omnimodal Language Models

2026-04-11 · Alaa Elobaid arxiv

This paper provides a comprehensive evaluation of demographic and linguistic biases in omnimodal language models that process text, images, audio, and video within a single framework. Although these models are being wide…

Language IdentificationActivity Recognition

Multi-GAT: A Graphical Attention-based Hierarchical Multimodal Representation Learning Approach for Human Activity Recognition

2021-04-01 · IEEE ROBOTICS AND AUTOMATION LETTERS 2021 4 · Md Mofijul Islam, Tariq Iqbal

Recognizing human activities is one of the crucial capabilities that a robot needs to have to be useful around people. Although modern robots are equipped with various types of sensors, human activity recognition (HAR) s…

Activity RecognitionHuman Activity RecognitionMixture-of-ExpertsMultimodal Activity Recognition+1

Enhancing Cross-lingual Transfer via Phonemic Transcription Integration

2023-07-10 · Hoang H. Nguyen, Chenwei Zhang, Tao Zhang, Eugene Rohrbaugh 외

Previous cross-lingual transfer methods are restricted to orthographic representation learning via textual scripts. This limitation hampers cross-lingual transfer and is biased towards languages sharing similar well-know…

Cross-Lingual Transfernamed-entity-recognitionNamed Entity RecognitionPart-Of-Speech Tagging+1

Privacy Enhanced Multimodal Neural Representations for Emotion Recognition

2019-10-29 · Mimansa Jaiswal, Emily Mower Provost

Many mobile applications and virtual conversational agents now aim to recognize and adapt to emotions. To enable this, data are transmitted from users' devices and stored on central servers. Yet, these data contain sensi…

Emotion Recognition

Demographic Fairness in Multimodal LLMs: A Benchmark of Gender and Ethnicity Bias in Face Verification

2026-03-26 · Ünsal Öztürk, Hatef Otroshi Shahreza, Sébastien Marcel arxiv

Multimodal Large Language Models (MLLMs) have recently been explored as face verification systems that determine whether two face images are of the same person. Unlike dedicated face recognition systems, MLLMs approach t…

Face VerificationFace Recognition