paper-with-me

홈 › Papers

EmoVerse: A MLLMs-Driven Emotion Representation Dataset for Interpretable Visual Emotion Analysis

2025-11-16 · Yijie Guo, Dexiang Hong, Weidong Chen, Zihan She, Cheng Ye, Xiaojun Chang, Zhendong Mao arxiv

Visual Emotion Analysis (VEA) aims to bridge the affective gap between visual content and human emotional responses. Despite its promise, progress in this field remains limited by the lack of open-source and interpretable datasets. Most existing studies assign a single discrete emotion label to an entire image, offering limited insight into how visual elements contribute to emotion. In this work, we introduce EmoVerse, a large-scale open-source dataset that enables interpretable visual emotion analysis through multi-layered, knowledge-graph-inspired annotations. By decomposing emotions into Background-Attribute-Subject (B-A-S) triplets and grounding each element to visual regions, EmoVerse provides word-level and subject-level emotional reasoning. With over 219k images, the dataset further includes dual annotations in Categorical Emotion States (CES) and Dimensional Emotion Space (DES), facilitating unified discrete and continuous emotion representation. A novel multi-stage pipeline ensures high annotation reliability with minimal human effort. Finally, we introduce an interpretable model that maps visual cues into DES representations and provides detailed attribution explanations. Together, the dataset, pipeline, and model form a comprehensive foundation for advancing explainable high-level emotion understanding.

📄 PDF Abstract BibTeX arXiv:2511.12554

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

EmoVerse: Exploring Multimodal Large Language Models for Sentiment and Emotion Understanding

2024-12-11 · Ao Li, Longwei Xu, Chen Ling, Jinghui Zhang 외

Sentiment and emotion understanding are essential to applications such as human-computer interaction and depression detection. While Multimodal Large Language Models (MLLMs) demonstrate robust general capabilities, they …

Depression DetectionEmotion-Cause Pair ExtractionEmotion RecognitionFacial Expression Recognition+3

Epistemoverse: Toward an AI-Driven Knowledge Metaverse for Intellectual Heritage Preservation

2025-12-13 · Predrag K. Nikolić, Robert Prentner arxiv

Large language models (LLMs) have often been characterized as "stochastic parrots" that merely reproduce fragments of their training data. This study challenges that assumption by demonstrating that, when placed in an ap…

EmotionHallucer: Evaluating Emotion Hallucinations in Multimodal Large Language Models

2025-05-16 · Bohao Xing, Xin Liu, Guoying Zhao, Chengyu Liu 외

Emotion understanding is a critical yet challenging task. Recent advances in Multimodal Large Language Models (MLLMs) have significantly enhanced their capabilities in this area. However, MLLMs often suffer from hallucin…

Hallucination

Multimodal Large Language Models Meet Multimodal Emotion Recognition and Reasoning: A Survey

2025-09-29 · Yuntao Shou, Tao Meng, Wei Ai, Keqin Li arxiv

In recent years, large language models (LLMs) have driven major advances in language understanding, marking a significant step toward artificial general intelligence (AGI). With increasing demands for higher-level semant…

Multimodal Emotion Recognition

MultiEmo-Bench: Multi-label Visual Emotion Analysis for Multi-modal Large Language Models

2026-05-14 · Tianwei Chen, Takuya Furusawa, Yuki Hirakawa, Ryotaro Shimizu 외 arxiv

This paper introduces a multi-label visual emotion analysis benchmark dataset for comprehensively evaluating the ability of multimodal large language models (MLLMs) to predict the emotions evoked by images. Recent user s…