paper-with-me

Papers

CREward: A Type-Specific Creativity Reward Model

2025-11-25 · Jiyeon Han, Ali Mahdavi-Amiri, Hao Zhang, Haedong Jeong arxiv

Creativity is a complex phenomenon. When it comes to representing and assessing creativity, treating it as a single undifferentiated quantity would appear naive and underwhelming. In this work, we learn the \emph{first type-specific creativity reward model}, coined CREward, which spans three creativity ``axes," geometry, material, and texture, to allow us to view creativity through the lens of the image formation pipeline. To build our reward model, we first conduct a human benchmark evaluation to capture human perception of creativity for each type across various creative images. We then analyze the correlation between human judgments and predictions by large vision-language models (LVLMs), confirming that LVLMs exhibit strong alignment with human perception. Building on this observation, we collect LVLM-generated labels to train our CREward model that is applicable to both evaluation and generation of creative images. We explore three applications of CREward: creativity assessment, explainable creativity, and creative sample acquisition for both human design inspiration and guiding creative generation through low-rank adaptation.

📄 PDF Abstract BibTeX arXiv:2511.19995

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

DocReward: A Document Reward Model for Structuring and Stylizing

2025-10-13 · Junpeng Liu, Yuzhong Zhao, Bowen Cao, Jiayu Ding 외 arxiv

Recent agentic workflows automate professional document generation but focus narrowly on textual quality, overlooking structural and stylistic professionalism, which is equally critical for readability. This gap stems ma…

Reinforcement Learning

Rewarding Structural Conformance of Reasoning using Process Mining

2025-10-29 · Yongjae Lee, Taekhyun Park, Sunghyun Sim, Hyerim Bae arxiv

Recent advances in sparse reward policy gradient methods have enabled effective reinforcement learning (RL)-based language model post-training. However, for reasoning tasks such as mathematical problem solving, binarized…

Reinforcement LearningMathematical Reasoning

LogicReward: Incentivizing LLM Reasoning via Step-Wise Logical Supervision

2025-12-20 · Jundong Xu, Hao Fei, Huichi Zhou, Xin Quan 외 arxiv

Although LLMs exhibit strong reasoning capabilities, existing training methods largely depend on outcome-based feedback, which can produce correct answers with flawed reasoning. Prior work introduces supervision on inter…

Natural Language InferenceLogical Reasoning

R3: Robust Rubric-Agnostic Reward Models

2025-05-19 · David Anugraha, Zilu Tang, Lester James V. Miranda, Hanyang Zhao 외

Reward models are essential for aligning language model outputs with human preferences, yet existing approaches often lack both controllability and interpretability. These models are typically optimized for narrow object…

Language ModelingLanguage Modelling

mR3: Multilingual Rubric-Agnostic Reward Reasoning Models

2025-10-01 · David Anugraha, Shou-Yi Hung, Zilu Tang, Annie En-Shiun Lee 외 arxiv

Evaluation using Large Language Model (LLM) judges has been widely adopted in English and shown to be effective for automatic evaluation. However, their performance does not generalize well to non-English settings, and i…