paper-with-me

홈 › Papers

MuSACo: Multimodal Subject-Specific Selection and Adaptation for Expression Recognition with Co-Training

2025-08-17 · Muhammad Osama Zeeshan, Natacha Gillet, Alessandro Lameiras Koerich, Marco Pedersoli, Francois Bremond, Eric Granger arxiv

Personalized expression recognition (ER) involves adapting a machine learning model to subject-specific data for improved recognition of expressions with considerable interpersonal variability. Subject-specific ER can benefit significantly from multi-source domain adaptation (MSDA) methods, where each domain corresponds to a specific subject to improve model accuracy and robustness. Despite promising results, state-of-the-art MSDA approaches often overlook multimodal information or blend sources into a single domain, limiting subject diversity and failing to explicitly capture unique subject-specific characteristics. To address these limitations, we introduce MuSACo, a multimodal subject-specific selection and adaptation method for ER based on co-training. It leverages complementary information across multiple modalities and multiple source domains for subject-specific adaptation. This makes MuSACo particularly relevant for affective computing applications in digital health, such as patient-specific assessment for stress or pain, where subject-level nuances are crucial. MuSACo selects source subjects relevant to the target and generates pseudo-labels using the dominant modality for class-aware learning, in conjunction with a class-agnostic loss to learn from less confident target samples. Finally, source features from each modality are aligned, while only confident target features are combined. Experimental results on challenging multimodal ER datasets: BioVid, StressID, and BAH show that MuSACo outperforms UDA (blending) and state-of-the-art MSDA methods.

📄 PDF Abstract BibTeX arXiv:2508.12522

Code (0)

등록된 구현이 없습니다.

Tasks

Domain Adaptation

Similar Papers 제목 키워드 기반

MusaCoder: Native GPU Kernel Generation with Full-Stack Training on Moore Threads GPU

2026-06-03 · Kun Cheng, Songshuo Lu, Sicong Liao, Tankun Li 외 arxiv

Native GPU kernel generation turns high-level tensor programs into executable, efficient low-level code. Existing Large Language Models (LLMs) struggle with this task, while execution-based reinforcement learning suffers…

Reinforcement Learning

UMBRAE: Unified Multimodal Brain Decoding

2024-04-10 · Weihao Xia, Raoul de Charette, Cengiz Öztireli, Jing-Hao Xue

We address prevailing challenges of the brain-powered research, departing from the observation that the literature hardly recover accurate spatial information and require subject-specific models. To address these challen…

Brain DecodingLanguage ModelingLanguage ModellingLarge Language Model+1

Driver Drowsiness Estimation from EEG Signals Using Online Weighted Adaptation Regularization for Regression (OwARR)

2017-02-09 · Dongrui Wu, Vernon J. Lawhern, Stephen Gordon, Brent J. Lance 외

One big challenge that hinders the transition of brain-computer interfaces (BCIs) from laboratory settings to real-life applications is the availability of high-performance and robust learning algorithms that can effecti…

Domain AdaptationEEGElectroencephalogram (EEG)regression+1

Stage-Aware Adaptation and Distribution Calibration for Subject-Driven Personalized Text-to-Image Generation

2026-07-08 · Wenyan Xu, Alizer Wong arxiv

Subject-driven personalized text-to-image generation requires a pretrained diffusion model to acquire a specific subject from a few reference images while preserving subject identity, following novel text prompts, and ma…

Text-to-Image Generation

Adaptive User-Centered Multimodal Interaction towards Reliable and Trusted Automotive Interfaces

2022-11-07 · Amr Gomaa

With the recently increasing capabilities of modern vehicles, novel approaches for interaction emerged that go beyond traditional touch-based and voice command approaches. Therefore, hand gestures, head pose, eye gaze, a…

multimodal interaction