paper-with-me

홈 › Papers

AesCanvas: A Large-Scale Dataset and Benchmark for Aesthetic Critique and Contextual Suitability

2026-08-27 · Xuanwei Hu, Haoyu Dong, Kejun Wu, Tianyi Liu, Jianjun Gao arxiv

Recent advances in Multimodal Large Language Models (MLLMs) have extended Image Aesthetic Assessment (IAA) beyond scalar scores toward interpretable critique and guidance. Yet existing benchmarks mainly assess intrinsic visual quality or fixed domain criteria, leaving open whether an appealing image is appropriate for a specific purpose, audience, cultural setting, or domain convention. We introduce AesCanvas, a unified suite with two complementary components: CritiqueCanvas with 519,136 instruction-response pairs from 54,300 images supports long-form, multi-dimensional critique across photography, painting, and virtual imagery, whereas ContextCanvas with 301 expert-reviewed use scenarios evaluates contextual aesthetic suitability in realistic use scenarios. Under a unified protocol, we evaluate closed-source frontier, open-weight general, and aesthetic-specific MLLMs. Results reveal a clear separation between critique generation and context-sensitive judgment: reference-based lexical and semantic metrics only partially capture critique quality, while aesthetic specialists remain competitive on selected critique metrics yet substantially lag strong general-purpose MLLMs on ContextCanvas. Further analyses show that aesthetic specialization does not reliably transfer to contextual suitability and that model decisions may fail to track or ground themselves in decisive contextual visual cues. These findings establish culturally situated, evidence-grounded suitability as a distinct objective for aesthetic modeling.

📄 PDF Abstract BibTeX arXiv:2608.26713

Code (2)

Tavish9/awesome-daily-AI-arxiv ★ 113
arxivsub/arXivSub_daily_arxiv ★ 4

Similar Papers 제목 키워드 기반

MADB: A Large-Scale Music Aesthetics Dataset with Professional and Multi-Dimensional Annotations

2026-07-08 · Sirui Zhang, Tianle Wang, Xinyi Tong, Peiyang Yu 외 arxiv

Music aesthetic assessment is a challenging yet underexplored problem, requiring models to capture fine-grained, multi-dimensional human perceptual judgments. Progress in this area has been limited by the lack of large-s…

Advancing Comprehensive Aesthetic Insight with Multi-Scale Text-Guided Self-Supervised Learning

2024-12-16 · Yuti Liu, Shice Liu, Junyuan Gao, PengTao Jiang 외

Image Aesthetic Assessment (IAA) is a vital and intricate task that entails analyzing and assessing an image's aesthetic values, and identifying its highlights and areas for improvement. Traditional methods of IAA often …

In-Context LearningSelf-Supervised LearningZero-Shot Learning

Code Aesthetics with Agentic Reward Feedback

2025-10-27 · Bang Xiao, Lingjie Jiang, Shaohan Huang, Tengchao Lv 외 arxiv

Large Language Models (LLMs) have become valuable assistants for developers in code-related tasks. While LLMs excel at traditional programming tasks such as code generation and bug fixing, they struggle with visually-ori…

Reinforcement LearningCode Generation

Venus: Benchmarking and Empowering Multimodal Large Language Models for Aesthetic Guidance and Cropping

2026-02-27 · Tianxiang Du, Hulingxiao He, Yuxin Peng arxiv

The widespread use of smartphones has made photography ubiquitous, yet a clear gap remains between ordinary users and professional photographers, who can identify aesthetic issues and provide actionable shooting guidance…

Aesthetic Attributes Assessment of Images with AMANv2 and DPC-CaptionsV2

2022-08-09 · Xinghui Zhou, Xin Jin, Jianwen Lv, Heng Huang 외

Image aesthetic quality assessment is popular during the last decade. Besides numerical assessment, nature language assessment (aesthetic captioning) has been proposed to describe the generally aesthetic impression of an…

AttributeImage Captioning