paper-with-me

홈 › Papers

Multimodal Modeling of Task-Mediated Confusion

2022-07-01 · NAACL (ACL) 2022 7 · Camille Mince, Skye Rhomberg, Cecilia Alm, Reynold Bailey, Alex Ororbia

In order to build more human-like cognitive agents, systems capable of detecting various human emotions must be designed to respond appropriately. Confusion, the combination of an emotional and cognitive state, is under-explored. In this paper, we build upon prior work to develop models that detect confusion from three modalities: video (facial features), audio (prosodic features), and text (transcribed speech features). Our research improves the data collection process by allowing for continuous (as opposed to discrete) annotation of confusion levels. We also craft models based on recurrent neural networks (RNNs) given their ability to predict sequential data. In our experiments, we find that text and video modalities are the most important in predicting confusion while the explored audio features are relatively unimportant predictors of confusion in our data.

📄 PDF Abstract BibTeX

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Task Confusion and Catastrophic Forgetting in Class-Incremental Learning: A Mathematical Framework for Discriminative and Generative Modelings

2024-10-28 · Milad Khademi Nori, Il-Min Kim

In class-incremental learning (class-IL), models must classify all previously seen classes at test time without task-IDs, leading to task confusion. Despite being a key challenge, task confusion lacks a theoretical under…

class-incremental learningClass Incremental LearningIncremental Learning

Distribution-Level Memory Recall for Continual Learning: Preserving Knowledge and Avoiding Confusion

2024-08-04 · Shaoxu Cheng, Kanglei Geng, Chiyuan He, Zihuan Qiu 외

Continual Learning (CL) aims to enable Deep Neural Networks (DNNs) to learn new data without forgetting previously learned knowledge. The key to achieving this goal is to avoid confusion at the feature level, i.e., avoid…

Continual Learning

Opening the Black Box: Preliminary Insights into Affective Modeling in Multimodal Foundation Models

2026-01-22 · Zhen Zhang, Runhao Zeng, Sicheng Zhao, Xiping Hu arxiv

Understanding where and how emotions are represented in large-scale foundation models remains an open problem, particularly in multimodal affective settings. Despite the strong empirical performance of recent affective m…

MAMBO-NET: Multi-Causal Aware Modeling Backdoor-Intervention Optimization for Medical Image Segmentation Network

2025-05-28 · Ruiguo Yu, Yiyang Zhang, Yuan Tian, Yujie Diao 외

Medical image segmentation methods generally assume that the process from medical image to segmentation is unbiased, and use neural networks to establish conditional probability models to complete the segmentation task. …

Causal InferenceImage SegmentationMedical Image SegmentationSegmentation+1

Confusion2vec 2.0: Enriching Ambiguous Spoken Language Representations with Subwords

2021-02-03 · Prashanth Gurunath Shivakumar, Panayiotis Georgiou, Shrikanth Narayanan

Word vector representations enable machines to encode human language for spoken language understanding and processing. Confusion2vec, motivated from human speech production and perception, is a word vector representation…

Automatic Speech RecognitionAutomatic Speech Recognition (ASR)Intent DetectionNatural Language Understanding+4