paper-with-me

Papers

Emotion and Intention Guided Multi-Modal Learning for Sticker Response Selection

2025-11-16 · Yuxuan Hu, Jian Chen, Yuhao Wang, Zixuan Li, Jing Xiong, Pengyue Jia, Wei Wang, Chengming Li, Xiangyu Zhao arxiv

Stickers are widely used in online communication to convey emotions and implicit intentions. The Sticker Response Selection (SRS) task aims to select the most contextually appropriate sticker based on the dialogue. However, existing methods typically rely on semantic matching and model emotional and intentional cues separately, which can lead to mismatches when emotions and intentions are misaligned. To address this issue, we propose Emotion and Intention Guided Multi-Modal Learning (EIGML). This framework is the first to jointly model emotion and intention, effectively reducing the bias caused by isolated modeling and significantly improving selection accuracy. Specifically, we introduce Dual-Level Contrastive Framework to perform both intra-modality and inter-modality alignment, ensuring consistent representation of emotional and intentional features within and across modalities. In addition, we design an Intention-Emotion Guided Multi-Modal Fusion module that integrates emotional and intentional information progressively through three components: Emotion-Guided Intention Knowledge Selection, Intention-Emotion Guided Attention Fusion, and Similarity-Adjusted Matching Mechanism. This design injects rich, effective information into the model and enables a deeper understanding of the dialogue, ultimately enhancing sticker selection performance. Experimental results on two public SRS datasets show that EIGML consistently outperforms state-of-the-art baselines, achieving higher accuracy and a better understanding of emotional and intentional features. Code is provided in the supplementary materials.

📄 PDF Abstract BibTeX arXiv:2511.17587

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

TGCA-PVT: Topic-Guided Context-Aware Pyramid Vision Transformer for Sticker Emotion Recognition

2024-10-28 · MM '24: Proceedings of the 32nd ACM International Conference on Multimedia 2024 10 · Jian Chen, Wei Wang, Yuzhu Hu, Junxin Chen 외

Online chatting has become an essential aspect of our daily interactions, with stickers emerging as a prevalent tool for conveying emotions more vividly than plain text. While conventional image emotion recognition focus…

Emotion Recognition

MGHFT: Multi-Granularity Hierarchical Fusion Transformer for Cross-Modal Sticker Emotion Recognition

2025-07-25 · Jian Chen, Yuxuan Hu, Haifeng Lu, Wei Wang 외 arxiv

Although pre-trained visual models with text have demonstrated strong capabilities in visual feature extraction, sticker emotion understanding remains challenging due to its reliance on multi-view information, such as ba…

Contrastive LearningEmotion Recognition

Towards Exploiting Sticker for Multimodal Sentiment Analysis in Social Media: A New Dataset and Baseline

2022-10-01 · COLING 2022 10 · Feng Ge, Weizhao Li, Haopeng Ren, Yi Cai

Sentiment analysis in social media is challenging since posts are short of context. As a popular way to express emotion on social media, stickers related to these posts can supplement missing sentiments and help identify…

Multimodal Sentiment AnalysisSentiment Analysis

STICKERCONV: Generating Multimodal Empathetic Responses from Scratch

2024-01-20 · Yiqun Zhang, Fanheng Kong, Peidong Wang, Shuang Sun 외

Stickers, while widely recognized for enhancing empathetic communication in online interactions, remain underexplored in current empathetic dialogue research, notably due to the challenge of a lack of comprehensive datas…

2kEmpathetic Response GenerationResponse Generation

Selecting Stickers in Open-Domain Dialogue through Multitask Learning

2022-09-16 · Findings (ACL) 2022 5 · Zhexin Zhang, Yeshuang Zhu, Zhengcong Fei, Jinchao Zhang 외

With the increasing popularity of online chatting, stickers are becoming important in our online communication. Selecting appropriate stickers in open-domain dialogue requires a comprehensive understanding of both dialog…