paper-with-me

홈 › Papers

Addressing Data Scarcity in Multimodal User State Recognition by Combining Semi-Supervised and Supervised Learning

2022-02-08 · Hendric Voß, Heiko Wersing, Stefan Kopp

Detecting mental states of human users is crucial for the development of cooperative and intelligent robots, as it enables the robot to understand the user's intentions and desires. Despite their importance, it is difficult to obtain a large amount of high quality data for training automatic recognition algorithms as the time and effort required to collect and label such data is prohibitively high. In this paper we present a multimodal machine learning approach for detecting dis-/agreement and confusion states in a human-robot interaction environment, using just a small amount of manually annotated data. We collect a data set by conducting a human-robot interaction study and develop a novel preprocessing pipeline for our machine learning approach. By combining semi-supervised and supervised architectures, we are able to achieve an average F1-score of 81.1\% for dis-/agreement detection with a small amount of labeled data and a large unlabeled data set, while simultaneously increasing the robustness of the model compared to the supervised approach.

📄 PDF Abstract BibTeX arXiv:2202.03775

Code (0)

등록된 구현이 없습니다.

Tasks

BIG-bench Machine Learning

Similar Papers 제목 키워드 기반

ENCLIP: Ensembling and Clustering-Based Contrastive Language-Image Pretraining for Fashion Multimodal Search with Limited Data and Low-Quality Images

2024-11-25 · Prithviraj Purushottam Naik, Rohit Agarwal

Multimodal search has revolutionized the fashion industry, providing a seamless and intuitive way for users to discover and explore fashion items. Based on their preferences, style, or specific attributes, users can sear…

MedCoDi-M: A Multi-Prompt Foundation Model for Multimodal Medical Data Generation

2025-01-08 · Daniele Molino, Francesco Di Feola, Eliodoro Faiella, Deborah Fazzini 외

Artificial Intelligence is revolutionizing medical practice, enhancing diagnostic accuracy and healthcare delivery. However, its adaptation in medical settings still faces significant challenges, related to data availabi…

Contrastive LearningDiagnostic

A Corpus-free State2Seq User Simulator for Task-oriented Dialogue

2019-09-10 · Yutai Hou, Meng Fang, Wanxiang Che, Ting Liu

Recent reinforcement learning algorithms for task-oriented dialogue system absorbs a lot of interest. However, an unavoidable obstacle for training such algorithms is that annotated dialogue corpora are often unavailable…

DiversityReinforcement Learning

PathInsight: Instruction Tuning of Multimodal Datasets and Models for Intelligence Assisted Diagnosis in Histopathology

2024-08-13 · Xiaomin Wu, Rui Xu, Pengchen Wei, Wenkang Qin 외

Pathological diagnosis remains the definitive standard for identifying tumors. The rise of multimodal large models has simplified the process of integrating image analysis with textual descriptions. Despite this advancem…

Image Captioning

MapGlue: Multimodal Remote Sensing Image Matching

2025-03-20 · Peihao Wu, Yongxiang Yao, Wenfei Zhang, Dong Wei 외

Multimodal remote sensing image (MRSI) matching is pivotal for cross-modal fusion, localization, and object detection, but it faces severe challenges due to geometric, radiometric, and viewpoint discrepancies across imag…

object-detectionObject Detection