paper-with-me

홈 › Papers

DEFT-LLM: Disentangled Expert Feature Tuning for Micro-Expression Recognition

2025-11-14 · Ren Zhang, Huilai Li, Chao qi, Guoliang Xu, Tianyu Zhou, Wei wei, Jianqin Yin arxiv

Micro expression recognition (MER) is crucial for inferring genuine emotion. Applying a multimodal large language model (MLLM) to this task enables spatio-temporal analysis of facial motion and provides interpretable descriptions. However, there are still two core challenges: (1) The entanglement of static appearance and dynamic motion cues prevents the model from focusing on subtle motion; (2) Textual labels in existing MER datasets do not fully correspond to underlying facial muscle movements, creating a semantic gap between text supervision and physical motion. To address these issues, we propose DEFT-LLM, which achieves motion semantic alignment by multi-expert disentanglement. We first introduce Uni-MER, a motion-driven instruction dataset designed to align text with local facial motion. Its construction leverages dual constraints from optical flow and Action Unit (AU) labels to ensure spatio-temporal consistency and reasonable correspondence to the movements. We then design an architecture with three experts to decouple facial dynamics into independent and interpretable representations (structure, dynamic textures, and motion-semantics). By integrating the instruction-aligned knowledge from Uni-MER into DEFT-LLM, our method injects effective physical priors for micro expressions while also leveraging the cross modal reasoning ability of large language models, thus enabling precise capture of subtle emotional cues. Experiments on multiple challenging MER benchmarks demonstrate state-of-the-art performance, as well as a particular advantage in interpretable modeling of local facial motion.

📄 PDF Abstract BibTeX arXiv:2511.10948

Code (0)

등록된 구현이 없습니다.

Tasks

Micro-Expression Recognition

Similar Papers 제목 키워드 기반

DEFT 2018: Attention s\'elective pour classification de microblogs (DEFT 2018 : Selective Attention for Microblogging Classification )

2018-05-01 · JEPTALNRECITAL 2018 5 · Charles-Emmanuel Dias, Clara de Forsan de Gainon Gabriac, Patrick Gallinari, Vincent Guigue

Dans le cadre de l{'}atelier DEFT 2018 nous nous sommes int{\'e}ress{\'e}s {\`a} la classification de microblogs (ici, des tweets) r{\'e}dig{\'e}s en fran{\c{c}}ais. Ici, nous proposons une m{\'e}thode se basant sur un r…

ClassificationGeneral Classification

Deft Scheduling of Dynamic Cloud Workflows with Varying Deadlines via Mixture-of-Experts

2026-05-31 · Ya Shen, Gang Chen, Hui Ma, Mengjie Zhang arxiv

Workflow scheduling in cloud computing demands the intelligent allocation of dynamically arriving, graph-structured workflows with varying deadlines onto ever-changing virtual machine resources. However, existing deep re…

Reinforcement Learning

Vision-Language Models are Strong Noisy Label Detectors

2024-09-29 · Tong Wei, Hao-Tian Li, Chun-Shu Li, Jiang-Xin Shi 외

Recent research on fine-tuning vision-language models has demonstrated impressive performance in various downstream tasks. However, the challenge of obtaining accurately labeled data in real-world applications poses a si…

Denoisingimage-classificationImage Classificationparameter-efficient fine-tuning

DEFT: Dexterous Fine-Tuning for Real-World Hand Policies

2023-10-30 · Aditya Kannan, Kenneth Shaw, Shikhar Bahl, Pragna Mannam 외

Dexterity is often seen as a cornerstone of complex manipulation. Humans are able to perform a host of skills with their hands, from making food to operating tools. In this paper, we investigate these challenges, especia…

DEFT2018 : recherche d'information et analyse de sentiments dans des tweets concernant les transports en \^Ile de France (DEFT2018 : Information Retrieval and Sentiment Analysis in Tweets about Public Transportation in \^Ile de France Region )

2018-05-01 · JEPTALNRECITAL 2018 5 · Patrick Paroubek, Cyril Grouin, Patrice Bellot, Vincent Claveau 외

Cet article pr{\'e}sente l{'}{\'e}dition 2018 de la campagne d{'}{\'e}valuation DEFT (D{\'e}fi Fouille de Textes). A partir d{'}un corpus de tweets, quatre t{\^a}ches ont {\'e}t{\'e} propos{\'e}es : identifier les tweets…

Information RetrievalRetrievalSentiment Analysis