paper-with-me

홈 › Papers

Towards LLM-centric Affective Visual Customization via Efficient and Precise Emotion Manipulating

2026-02-20 · Jiamin Luo, Xuqian Gu, Jingjing Wang, Jiahong Lu arxiv

Previous studies on visual customization primarily rely on the objective alignment between various control signals (e.g., language, layout and canny) and the edited images, which largely ignore the subjective emotional contents, and more importantly lack general-purpose foundation models for affective visual customization. With this in mind, this paper proposes an LLM-centric Affective Visual Customization (L-AVC) task, which focuses on generating images within modifying their subjective emotions via Multimodal LLM. Further, this paper contends that how to make the model efficiently align emotion conversion in semantics (named inter-emotion semantic conversion) and how to precisely retain emotion-agnostic contents (named exter-emotion semantic retaining) are rather important and challenging in this L-AVC task. To this end, this paper proposes an Efficient and Precise Emotion Manipulating approach for editing subjective emotions in images. Specifically, an Efficient Inter-emotion Converting (EIC) module is tailored to make the LLM efficiently align emotion conversion in semantics before and after editing, followed by a Precise Exter-emotion Retaining (PER) module to precisely retain the emotion-agnostic contents. Comprehensive experimental evaluations on our constructed L-AVC dataset demonstrate the great advantage of the proposed EPEM approach to the L-AVC task over several state-of-the-art baselines. This justifies the importance of emotion information for L-AVC and the effectiveness of EPEM in efficiently and precisely manipulating such information.

📄 PDF Abstract BibTeX arXiv:2602.18016

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Use of Affective Visual Information for Summarization of Human-Centric Videos

2021-07-08 · Berkay Köprü, Engin Erzin

Increasing volume of user-generated human-centric video content and their applications, such as video retrieval and browsing, require compact representations that are addressed by the video summarization literature. Curr…

Emotion RecognitionRetrievalSupervised Video SummarizationVideo Retrieval+1

Benchmarking Dynamic Affective Reasoning: A Viewer-Centric Video Emotion Dataset

2026-07-11 · Zhiyan Zhang, Peipei Song, Jinpeng Hu, Jingyang Jia 외 arxiv

Video emotion analysis is typically framed as a static classification problem, treating each clip as an independent labeled unit. However, such a formulation overlooks a key psychological fact: emotions change as a resul…

Emotion Classification

GRACE: Boosting Video MLLMs with Grounded Action-Centric Evidence for Viewer Sentiment Prediction

2026-06-15 · Ruoxuan Yang, Tieyuan Chen, Xiaofeng Huang, Haibing Yin 외 arxiv

Viewer sentiment prediction in video advertisements aims to infer the latent affective response evoked in the audience. To bridge the gap between what is shown and what is felt, models must deduce hidden viewer emotions …

KEVER^2: Knowledge-Enhanced Visual Emotion Reasoning and Retrieval

2025-05-30 · Fanhang Man, Xiaoyue Chen, Huandong Wang, Baining Zhao 외

Understanding what emotions images evoke in their viewers is a foundational goal in human-centric visual computing. While recent advances in vision-language models (VLMs) have shown promise for visual emotion analysis (V…

Emotion RecognitionRetrieval

Recognition of Advertisement Emotions with Application to Computational Advertising

2019-04-03 · Abhinav Shukla, Shruti Shriya Gullapuram, Harish Katti, Mohan Kankanhalli 외

Advertisements (ads) often contain strong affective content to capture viewer attention and convey an effective message to the audience. However, most computational affect recognition (AR) approaches examine ads via the …

EEGElectroencephalogram (EEG)Multi-Task Learning