paper-with-me

홈 › Papers

Beyond Absolute Scores: Relative Edit-induced Difference for Generalizable Image Aesthetic Assessment

2026-06-04 · Qifei Jia, Xintong Yao, Yasen Zhang, Minghao Li, Yajie Chai, Qiming Lu, Baoyue Shen, Runyu Shi, Ying Huang, Yue Zhang arxiv

Traditional Image Aesthetic Assessment (IAA) methods mainly rely on regressing absolute Mean Opinion Scores (MOS). However, such a paradigm overlooks the inherently dynamic nature of human aesthetic perception, which relies on subconscious comparison against implicit visual references. Consequently, the lack of causal reasoning regarding aesthetic differences prevents models from learning generalizable aesthetic principles, thus limiting their generalization across diverse scenarios. In this work, we rethink the IAA task and propose Relative Edit-induced Difference Aesthetic learning (RED-Aes), a novel framework that leverages controllable image editing models to simulate the human aesthetic reasoning process. Instead of fitting absolute score distributions, RED-Aes explicitly learns the visual factors that drive aesthetic changes. To support this paradigm, we construct the RED-20k dataset, which comprises editing-based image pairs, quantitative aesthetic differences, and Chain-of-Thought (CoT) reasoning. Furthermore, we introduce a three-stage training strategy guided by a relative ranking consistency reward, optimizing the model solely via relative supervision. Extensive experiments demonstrate that RED-Aes achieves state-of-the-art performance on multiple public benchmarks, exhibiting superior generalization capabilities.

📄 PDF Abstract BibTeX arXiv:2606.05778

Code (0)

등록된 구현이 없습니다.

Tasks

Image Editing

Similar Papers 제목 키워드 기반

High Quality Diffusion Distillation on a Single GPU with Relative and Absolute Position Matching

2025-03-26 · Guoqiang Zhang, Kenta Niwa, J. P. Lewis, Cedric Mesnage 외

We introduce relative and absolute position matching (RAPM), a diffusion distillation method resulting in high quality generation that can be trained efficiently on a single GPU. Recent diffusion distillation research ha…

GPUImage GenerationPositionText to Image Generation+1

Toward Fine-grained Facial Expression Manipulation

2020-04-07 · ECCV 2020 8 · Jun Ling, Han Xue, Li Song, Shuhui Yang 외

Facial expression manipulation aims at editing facial expression with a given condition. Previous methods edit an input image under the guidance of a discrete emotion label or absolute condition (e.g., facial action unit…

Facial Expression TranslationImage-to-Image Translation

MARS-RA: Rank Aggregation for Credit Assignment via Multimodal Comparisons in Embodied Multi-Agent Cooperation

2026-07-30 · Dawei Wang, Di Zhao, Xinyuan Liu, Marci Chi Ma 외 arxiv

Credit assignment is a fundamental challenge in cooperative multi-agent reinforcement learning, particularly in embodied AI settings characterized by limited and delayed feedback as well as dynamically changing numbers o…

Multi-agent Reinforcement Learning

Enhancing LLM-based Search Agents via Contribution Weighted Group Relative Policy Optimization

2026-04-15 · Junzhe Wang, Zhiheng Xi, Yajie Yang, Hao Luo 외 arxiv

Search agents extend Large Language Models (LLMs) beyond static parametric knowledge by enabling access to up-to-date and long-tail information unavailable during pretraining. While reinforcement learning has been widely…

Reinforcement Learning

Pushing the Limits of Translation Quality Estimation

2017-01-01 · TACL 2017 1 · Andr{\'e} F. T. Martins, Marcin Junczys-Dowmunt, Fabio N. Kepler, Ram{\'o}n Astudillo 외

Translation quality estimation is a task of growing importance in NLP, due to its potential to reduce post-editing human effort in disruptive ways. However, this potential is currently limited by the relatively low accur…

Automatic Post-EditingSentenceTranslation