paper-with-me

홈 › Papers

AHMSA-Net: Adaptive Hierarchical Multi-Scale Attention Network for Micro-Expression Recognition

2025-01-05 · Lijun Zhang, Yifan Zhang, Weicheng Tang, Xinzhi Sun, Xiaomeng Wang, Zhanshan Li

Micro-expression recognition (MER) presents a significant challenge due to the transient and subtle nature of the motion changes involved. In recent years, deep learning methods based on attention mechanisms have made some breakthroughs in MER. However, these methods still suffer from the limitations of insufficient feature capture and poor dynamic adaptation when coping with the instantaneous subtle movement changes of micro-expressions. Therefore, in this paper, we design an Adaptive Hierarchical Multi-Scale Attention Network (AHMSA-Net) for MER. Specifically, we first utilize the onset and apex frames of the micro-expression sequence to extract three-dimensional (3D) optical flow maps, including horizontal optical flow, vertical optical flow, and optical flow strain. Subsequently, the optical flow feature maps are inputted into AHMSA-Net, which consists of two parts: an adaptive hierarchical framework and a multi-scale attention mechanism. Based on the adaptive downsampling hierarchical attention framework, AHMSA-Net captures the subtle changes of micro-expressions from different granularities (fine and coarse) by dynamically adjusting the size of the optical flow feature map at each layer. Based on the multi-scale attention mechanism, AHMSA-Net learns micro-expression action information by fusing features from different scales (channel and spatial). These two modules work together to comprehensively improve the accuracy of MER. Additionally, rigorous experiments demonstrate that the proposed method achieves competitive results on major micro-expression databases, with AHMSA-Net achieving recognition accuracy of up to 78.21% on composite databases (SMIC, SAMM, CASMEII) and 77.08% on the CASME^{}3 database.

📄 PDF Abstract BibTeX arXiv:2501.02539

Code (0)

등록된 구현이 없습니다.

Tasks

Micro Expression RecognitionMicro-Expression RecognitionOptical Flow Estimation

Methods 이 논문이 사용한 방법론

Softmax The Softmax output function transforms a previous layer's output into a vector of probabilities. It is commonly used for multiclass classification. Given an input vector $x$…
Attention 설명 없음

Similar Papers 제목 키워드 기반

AMANet: Advancing SAR Ship Detection with Adaptive Multi-Hierarchical Attention Network

2024-01-24 · Xiaolin Ma, Junkai Cheng, Aihua Li, Yuhua Zhang 외

Recently, methods based on deep learning have been successfully applied to ship detection for synthetic aperture radar (SAR) images. Despite the development of numerous ship detection methodologies, detecting small and c…

object-detectionObject DetectionSAR Ship Detection

Hierarchical Multi-scale Attention Networks for Action Recognition

2017-08-25 · Shi-Yang Yan, Jeremy S. Smith, Wenjin Lu, Bai-Ling Zhang

Recurrent Neural Networks (RNNs) have been widely used in natural language processing and computer vision. Among them, the Hierarchical Multi-scale RNN (HM-RNN), a kind of multi-scale hierarchical RNN proposed recently, …

Action RecognitionHard AttentionTemporal Action Localization

Pose-adaptive Hierarchical Attention Network for Facial Expression Recognition

2019-05-24 · Yuanyuan Liu, Jiyao Peng, Jiabei Zeng, Shiguang Shan

Multi-view facial expression recognition (FER) is a challenging task because the appearance of an expression varies in poses. To alleviate the influences of poses, recent methods either perform pose normalization or lear…

Facial Expression RecognitionFacial Expression Recognition (FER)

Hierarchical Attention Fusion for Geo-Localization

2021-02-18 · Liqi Yan, Yiming Cui, Yingjie Chen, Dongfang Liu

Geo-localization is a critical task in computer vision. In this work, we cast the geo-localization as a 2D image retrieval task. Current state-of-the-art methods for 2D geo-localization are not robust to locate a scene w…

geo-localizationImage RetrievalRetrieval

Hierarchical Point Attention for Indoor 3D Object Detection

2023-01-06 · Manli Shu, Le Xue, Ning Yu, Roberto Martín-Martín 외

3D object detection is an essential vision technique for various robotic systems, such as augmented reality and domestic robots. Transformers as versatile network architectures have recently seen great success in 3D poin…

3D Object DetectionObjectobject-detectionObject Detection