paper-with-me

Papers

M3TCM: Multi-modal Multi-task Context Model for Utterance Classification in Motivational Interviews

2024-04-04 · Sayed Muddashir Hossain, Jan Alexandersson, Philipp Müller

Accurate utterance classification in motivational interviews is crucial to automatically understand the quality and dynamics of client-therapist interaction, and it can serve as a key input for systems mediating such interactions. Motivational interviews exhibit three important characteristics. First, there are two distinct roles, namely client and therapist. Second, they are often highly emotionally charged, which can be expressed both in text and in prosody. Finally, context is of central importance to classify any given utterance. Previous works did not adequately incorporate all of these characteristics into utterance classification approaches for mental health dialogues. In contrast, we present M3TCM, a Multi-modal, Multi-task Context Model for utterance classification. Our approach for the first time employs multi-task learning to effectively model both joint and individual components of therapist and client behaviour. Furthermore, M3TCM integrates information from the text and speech modality as well as the conversation context. With our novel approach, we outperform the state of the art for utterance classification on the recently introduced AnnoMI dataset with a relative improvement of 20% for the client- and by 15% for therapist utterance classification. In extensive ablation studies, we quantify the improvement resulting from each contribution.

📄 PDF Abstract BibTeX arXiv:2404.03312

Code (0)

등록된 구현이 없습니다.

Tasks

ClassificationMulti-Task Learning

Similar Papers 제목 키워드 기반

MSCTD: A Multimodal Sentiment Chat Translation Dataset

2022-02-28 · ACL 2022 5 · Yunlong Liang, Fandong Meng, Jinan Xu, Yufeng Chen 외

Multimodal machine translation and textual chat translation have received considerable attention in recent years. Although the conversation in its natural form is usually multimodal, there still lacks work on multimodal …

Machine TranslationMultimodal Machine TranslationSentiment AnalysisTranslation

Situated and Interactive Multimodal Conversations

2020-06-02 · COLING 2020 8 · Seungwhan Moon, Satwik Kottur, Paul A. Crook, Ankita De 외

Next generation virtual assistants are envisioned to handle multimodal inputs (e.g., vision, memories of previous interactions, in addition to the user's utterances), and perform multimodal actions (e.g., displaying a ro…

Response Generation

Multimodal Sentiment Analysis using Hierarchical Fusion with Context Modeling

2018-06-16 · N. Majumder, D. Hazarika, A. Gelbukh, E. Cambria 외

Multimodal sentiment analysis is a very actively growing field of research. A promising area of opportunity in this field is to improve the multimodal fusion mechanism. We present a novel feature fusion strategy that pro…

Multimodal Emotion RecognitionMultimodal Sentiment AnalysisSentiment Analysis

MicroEmo: Time-Sensitive Multimodal Emotion Recognition with Micro-Expression Dynamics in Video Dialogues

2024-07-23 · Liyun Zhang

Multimodal Large Language Models (MLLMs) have demonstrated remarkable multimodal emotion recognition capabilities, integrating multimodal cues from visual, acoustic, and linguistic contexts in the video to recognize huma…

Emotion RecognitionMultimodal Emotion Recognition

Modeling Text-visual Mutual Dependency for Multi-modal Dialog Generation

2021-05-30 · Shuhe Wang, Yuxian Meng, Xiaofei Sun, Fei Wu 외

Multi-modal dialog modeling is of growing interest. In this work, we propose frameworks to resolve a specific case of multi-modal dialog generation that better mimics multi-modal dialog generation in the real world, wher…