paper-with-me

홈 › Papers

Affective Feedback Synthesis Towards Multimodal Text and Image Data

2022-03-23 · Puneet Kumar, Gaurav Bhat, Omkar Ingle, Daksh Goyal, Balasubramanian Raman

In this paper, we have defined a novel task of affective feedback synthesis that deals with generating feedback for input text & corresponding image in a similar way as humans respond towards the multimodal data. A feedback synthesis system has been proposed and trained using ground-truth human comments along with image-text input. We have also constructed a large-scale dataset consisting of image, text, Twitter user comments, and the number of likes for the comments by crawling the news articles through Twitter feeds. The proposed system extracts textual features using a transformer-based textual encoder while the visual features have been extracted using a Faster region-based convolutional neural networks model. The textual and visual features have been concatenated to construct the multimodal features using which the decoder synthesizes the feedback. We have compared the results of the proposed system with the baseline models using quantitative and qualitative measures. The generated feedbacks have been analyzed using automatic and human evaluation. They have been found to be semantically similar to the ground-truth comments and relevant to the given text-image input.

📄 PDF Abstract BibTeX arXiv:2203.12692

Code (1)

mintelligence-group/mmfeed 공식 구현

Tasks

ArticlesDecoder

Similar Papers 제목 키워드 기반

Synthesizing Sentiment-Controlled Feedback For Multimodal Text and Image Data

2024-02-12 · Puneet Kumar, Sarthak Malik, Balasubramanian Raman, Xiaobai Li

The ability to generate sentiment-controlled feedback in response to multimodal inputs comprising text and images addresses a critical gap in human-computer interaction. This capability allows systems to provide empathet…

DecoderMarketingQuestion AnsweringSentiment Analysis+3

Recent Trends of Multimodal Affective Computing: A Survey from NLP Perspective

2024-09-11 · Guimin Hu, Yi Xin, Weimin Lyu, Haojian Huang 외

Multimodal affective computing (MAC) has garnered increasing attention due to its broad applications in analyzing human behaviors and intentions, especially in text-dominated multimodal affective computing field. This su…

Aspect-Based Sentiment AnalysisEmotion RecognitionEmotion Recognition in ConversationMultimodal Emotion Recognition+3

Dual-Model Prediction of Affective Engagement and Vocal Attractiveness from Speaker Expressiveness in Video Learning

2026-03-19 · Hung-Yue Suen, Kuo-En Hung, Fan-Hsun Tseng arxiv

This paper outlines a machine learning-enabled speaker-centric Emotion AI approach capable of predicting audience-affective engagement and vocal attractiveness in asynchronous video-based learning, relying solely on spea…

MMAFFBen: A Multilingual and Multimodal Affective Analysis Benchmark for Evaluating LLMs and VLMs

2025-05-30 · Zhiwei Liu, Lingfei Qian, Qianqian Xie, Jimin Huang 외

Large language models and vision-language models (which we jointly call LMs) have transformed NLP and CV, demonstrating remarkable potential across various fields. However, their capabilities in affective analysis (i.e. …

Emotion ClassificationSentiment Analysis

MRAC Track 1: 2nd Workshop on Multimodal, Generative and Responsible Affective Computing

2024-09-11 · Shreya Ghosh, Zhixi Cai, Abhinav Dhall, Dimitrios Kollias 외

With the rapid advancements in multimodal generative technology, Affective Computing research has provoked discussion about the potential consequences of AI systems equipped with emotional intelligence. Affective Computi…

Emotional Intelligence