paper-with-me

Papers

Mutual Information Regularized Identity-aware Facial ExpressionRecognition in Compressed Video

2020-10-20 · Xiaofeng Liu, Linghao Jin, Xu Han, Jane You

How to extract effective expression representations that invariant to the identity-specific attributes is a long-lasting problem for facial expression recognition (FER). Most of the previous methods process the RGB images of a sequence, while we argue that the off-the-shelf and valuable expression-related muscle movement is already embedded in the compression format. In this paper, we target to explore the inter-subject variations eliminated facial expression representation in the compressed video domain. In the up to two orders of magnitude compressed domain, we can explicitly infer the expression from the residual frames and possibly extract identity factors from the I frame with a pre-trained face recognition network. By enforcing the marginal independence of them, the expression feature is expected to be purer for the expression and be robust to identity shifts. Specifically, we propose a novel collaborative min-min game for mutual information (MI) minimization in latent space. We do not need the identity label or multiple expression samples from the same person for identity elimination. Moreover, when the apex frame is annotated in the dataset, the complementary constraint can be further added to regularize the feature-level game. In testing, only the compressed residual frames are required to achieve expression prediction. Our solution can achieve comparable or better performance than the recent decoded image-based methods on the typical FER benchmarks with about 3 times faster inference.

📄 PDF Abstract BibTeX arXiv:2010.10637

Code (0)

등록된 구현이 없습니다.

Tasks

Face RecognitionFacial Expression RecognitionFacial Expression Recognition (FER)

Similar Papers 제목 키워드 기반

Seeing Your Speech Style: A Novel Zero-Shot Identity-Disentanglement Face-based Voice Conversion

2024-09-01 · Yan Rong, Li Liu

Face-based Voice Conversion (FVC) is a novel task that leverages facial images to generate the target speaker's voice style. Previous work has two shortcomings: (1) suffering from obtaining facial embeddings that are wel…

Contrastive LearningDisentanglementDiversityVoice Conversion

FlowFace: Semantic Flow-guided Shape-aware Face Swapping

2022-12-06 · Hao Zeng, Wei zhang, Changjie Fan, Tangjie Lv 외

In this work, we propose a semantic flow-guided two-stage framework for shape-aware face swapping, namely FlowFace. Unlike most previous methods that focus on transferring the source inner facial features but neglect fac…

Face Swapping

Identity-Decoupled Anonymization for Visual Evidence in Multi-modal Retrieval-Augmented Generation

2026-04-26 · Zehua Cheng, Wei Dai, Jiahao Sun arxiv

Multi-modal retrieval-augmented generation (MRAG) systems retrieve visual evidence from large image corpora to ground the responses of large multi-modal models, yet the retrieved images frequently contain human faces who…

Face Recognition

AniTalker: Animate Vivid and Diverse Talking Faces through Identity-Decoupled Facial Motion Encoding

2024-05-06 · Tao Liu, Feilong Chen, Shuai Fan, Chenpeng Du 외

The paper introduces AniTalker, an innovative framework designed to generate lifelike talking faces from a single portrait. Unlike existing models that primarily focus on verbal cues such as lip synchronization and fail …

Metric LearningSelf-Supervised Learning

Beyond Facial Consistency: Personalized Person Image Generation with Holistic Identity Preservation

2026-07-28 · Yuxuan Xiao, Shanshan Zhang, Jian Yang, Shengcai Liao arxiv

Personalized person image generation requires preserving subject identity across both local facial details and broader appearance cues. Existing methods typically emphasize only one level of identity information, leading…

Image Generation