paper-with-me

홈 › Papers

Engineering Deep Representations for Modeling Aesthetic Perception

2016-05-25 · Yanxiang Chen, Yuxing Hu, Luming Zhang, Ping Li, Chao Zhang

Many aesthetic models in computer vision suffer from two shortcomings: 1) the low descriptiveness and interpretability of those hand-crafted aesthetic criteria (i.e., nonindicative of region-level aesthetics), and 2) the difficulty of engineering aesthetic features adaptively and automatically toward different image sets. To remedy these problems, we develop a deep architecture to learn aesthetically-relevant visual attributes from Flickr1, which are localized by multiple textual attributes in a weakly-supervised setting. More specifically, using a bag-ofwords (BoW) representation of the frequent Flickr image tags, a sparsity-constrained subspace algorithm discovers a compact set of textual attributes (e.g., landscape and sunset) for each image. Then, a weakly-supervised learning algorithm projects the textual attributes at image-level to the highly-responsive image patches at pixel-level. These patches indicate where humans look at appealing regions with respect to each textual attribute, which are employed to learn the visual attributes. Psychological and anatomical studies have shown that humans perceive visual concepts hierarchically. Hence, we normalize these patches and feed them into a five-layer convolutional neural network (CNN) to mimick the hierarchy of human perceiving the visual attributes. We apply the learned deep features on image retargeting, aesthetics ranking, and retrieval. Both subjective and objective experimental results thoroughly demonstrate the competitiveness of our approach.

📄 PDF Abstract BibTeX arXiv:1605.07699

Code (0)

등록된 구현이 없습니다.

Tasks

AttributeImage RetargetingRetrievalWeakly-supervised Learning

Methods 이 논문이 사용한 방법론

Interpretability 설명 없음

Similar Papers 제목 키워드 기반

QUASAR: QUality and Aesthetics Scoring with Advanced Representations

2024-03-11 · Sergey Kastryulin, Denis Prokopenko, Artem Babenko, Dmitry V. Dylov

This paper introduces a new data-driven, non-parametric method for image quality and aesthetics assessment, surpassing existing approaches and requiring no prompt engineering or fine-tuning. We eliminate the need for exp…

Prompt Engineering

Representing Beauty: Towards a Participatory but Objective Latent Aesthetics

2025-10-03 · Alexander Michael Rusnak arxiv

What does it mean for a machine to recognize beauty? While beauty remains a culturally and experientially compelling but philosophically elusive concept, deep learning systems increasingly appear capable of modeling aest…

AesBench: An Expert Benchmark for Multimodal Large Language Models on Image Aesthetics Perception

2024-01-16 · Yipo Huang, Quan Yuan, Xiangfei Sheng, Zhichao Yang 외

With collective endeavors, multimodal large language models (MLLMs) are undergoing a flourishing development. However, their performances on image aesthetics perception remain indeterminate, which is highly desired in re…

MLLM Evaluation: Aesthetics

AesExpert: Towards Multi-modality Foundation Model for Image Aesthetics Perception

2024-04-15 · Yipo Huang, Xiangfei Sheng, Zhichao Yang, Quan Yuan 외

The highly abstract nature of image aesthetics perception (IAP) poses significant challenge for current multimodal large language models (MLLMs). The lack of human-annotated multi-modality aesthetic data further exacerba…

Enhancing Image Aesthetics with Dual-Conditioned Diffusion Models Guided by Multimodal Perception

2026-03-12 · Xinyu Nan, Ning Wang, Yuyao Zhai, Mei Yang arxiv

Image aesthetic enhancement aims to perceive aesthetic deficiencies in images and perform corresponding editing operations, which is highly challenging and requires the model to possess creativity and aesthetic perceptio…

Image Editing