paper-with-me

홈 › Papers

Improving Bag-of-Visual-Words Towards Effective Facial Expressive Image Classification

2018-09-30 · Dawood Al Chanti, Alice Caplier

Bag-of-Visual-Words (BoVW) approach has been widely used in the recent years for image classification purposes. However, the limitations regarding optimal feature selection, clustering technique, the lack of spatial organization of the data and the weighting of visual words are crucial. These factors affect the stability of the model and reduce performance. We propose to develop an algorithm based on BoVW for facial expression analysis which goes beyond those limitations. Thus the visual codebook is built by using k-Means++ method to avoid poor clustering. To exploit reliable low level features, we search for the best feature detector that avoids locating a large number of keypoints which do not contribute to the classification process. Then, we propose to compute the relative conjunction matrix in order to preserve the spatial order of the data by coding the relationships among visual words. In addition, a weighting scheme that reflects how important a visual word is with respect to a given image is introduced. We speed up the learning process by using histogram intersection kernel by Support Vector Machine to learn a discriminative classifier. The efficiency of the proposed algorithm is compared with standard bag of visual words method and with bag of visual words method with spatial pyramid. Extensive experiments on the CK+, the MMI and the JAFFE databases show good average recognition rates. Likewise, the ability to recognize spontaneous and non-basic expressive states is investigated using the DynEmo database.

📄 PDF Abstract BibTeX arXiv:1810.00360

Code (0)

등록된 구현이 없습니다.

Tasks

Clusteringfeature selectionGeneral Classificationimage-classificationImage Classification

Methods 이 논문이 사용한 방법론

SPEED The monocular depth estimation (MDE) is the task of estimating depth from a single frame. This information is an essential knowledge in many computer vision tasks such as scene…

Similar Papers 제목 키워드 기반

Quantified Facial Temporal-Expressiveness Dynamics for Affect Analysis

2020-10-28 · Md Taufeeq Uddin, Shaun Canavan

The quantification of visual affect data (e.g. face images) is essential to build and monitor automated affect modeling systems efficiently. Considering this, this work proposes quantified facial Temporal-expressiveness …

AVI-Talking: Learning Audio-Visual Instructions for Expressive 3D Talking Face Generation

2024-02-25 · Yasheng Sun, Wenqing Chu, Hang Zhou, Kaisiyuan Wang 외

While considerable progress has been made in achieving accurate lip synchronization for 3D speech-driven talking face generation, the task of incorporating expressive facial detail synthesis aligned with the speaker's sp…

Face GenerationHallucinationTalking Face Generation

X-Actor: Emotional and Expressive Long-Range Portrait Acting from Audio

2025-08-04 · Chenxu Zhang, Zenan Li, Hongyi Xu, You Xie 외 arxiv

We present X-Actor, a novel audio-driven portrait animation framework that generates lifelike, emotionally expressive talking head videos from a single reference image and an input audio clip. Unlike prior methods that e…

DreamActor-M1: Holistic, Expressive and Robust Human Image Animation with Hybrid Guidance

2025-04-02 · Yuxuan Luo, Zhengkun Rong, Lizhen Wang, Longhao Zhang 외

While recent image-based human animation methods achieve realistic body and facial motion synthesis, critical gaps remain in fine-grained holistic controllability, multi-scale adaptability, and long-term temporal coheren…

Human AnimationImage AnimationMotion Synthesis

Facial Expression Recognition by De-Expression Residue Learning

2018-06-01 · CVPR 2018 6 · Huiyuan Yang, Umur Ciftci, Lijun Yin

A facial expression is a combination of an expressive component and a neutral component of a person. In this paper, we propose to recognize facial expressions by extracting information of the expressive component through…

Facial Expression RecognitionFacial Expression Recognition (FER)