paper-with-me

홈 › Papers

From Competition to Synergy: Unlocking Reinforcement Learning for Subject-Driven Image Generation

2025-10-21 · Ziwei Huang, Ying Shu, Hao Fang, Quanyu Long, Wenya Wang, Qiushi Guo, Tiezheng Ge, Leilei Gan arxiv

Subject-driven image generation models face a fundamental trade-off between identity preservation (fidelity) and prompt adherence (editability). While online reinforcement learning (RL), specifically GPRO, offers a promising solution, we find that a naive application of GRPO leads to competitive degradation, as the simple linear aggregation of rewards with static weights causes conflicting gradient signals and a misalignment with the temporal dynamics of the diffusion process. To overcome these limitations, we propose Customized-GRPO, a novel framework featuring two key innovations: (i) Synergy-Aware Reward Shaping (SARS), a non-linear mechanism that explicitly penalizes conflicted reward signals and amplifies synergistic ones, providing a sharper and more decisive gradient. (ii) Time-Aware Dynamic Weighting (TDW), which aligns the optimization pressure with the model's temporal dynamics by prioritizing prompt-following in the early, identity preservation in the later. Extensive experiments demonstrate that our method significantly outperforms naive GRPO baselines, successfully mitigating competitive degradation. Our model achieves a superior balance, generating images that both preserve key identity features and accurately adhere to complex textual prompts.

📄 PDF Abstract BibTeX arXiv:2510.18263

Code (0)

등록된 구현이 없습니다.

Tasks

Reinforcement LearningImage Generation

Similar Papers 제목 키워드 기반

Muscle Synergy Priors Enhance Biomechanical Fidelity in Predictive Musculoskeletal Locomotion Simulation

2026-03-11 · Ilseung Park, Eunsik Choi, Jangwhan Ahn, Jooeun Ahn arxiv

Human locomotion emerges from high-dimensional neuromuscular control, making predictive musculoskeletal simulation challenging. We present a physiology-informed reinforcement-learning framework that constrains control us…

Reinforcement Learning

Shaping Initial State Prevents Modality Competition in Multi-modal Fusion: A Two-stage Scheduling Framework via Fast Partial Information Decomposition

2025-09-25 · Jiaqi Tang, Yinsong Xu, Yang Liu, Qingchao Chen arxiv

Multi-modal fusion often suffers from modality competition during joint training, where one modality dominates the learning process, leaving others under-optimized. Overlooking the critical impact of the model's initial …

CityLearn: Standardizing Research in Multi-Agent Reinforcement Learning for Demand Response and Urban Energy Management

2020-12-18 · Jose R Vazquez-Canteli, Sourav Dey, Gregor Henze, Zoltan Nagy

Rapid urbanization, increasing integration of distributed renewable energy resources, energy storage, and electric vehicles introduce new challenges for the power grid. In the US, buildings represent about 70% of the tot…

energy managementManagementMulti-agent Reinforcement LearningOpenAI Gym+1

Phage-antibiotic synergy inhibited by temperate and chronic virus competition

2021-04-19 · Kylie J Landa, Lauren M Mossman, Rachel J Whitaker, Zoi Rapti 외

As antibiotic resistance grows more frequent for common bacterial infections, alternative treatment strategies such as phage therapy have become more widely studied in the medical field. While many studies have explored …

SMaRT: Select, Mix, and ReinvenT -- A Strategy Fusion Framework for LLM-Driven Reasoning and Planning

2025-10-20 · Nikhil Verma, Manasa Bharadwaj, Wonjun Jang, Harmanpreet Singh 외 arxiv

Large Language Models (LLMs) have redefined complex task automation with exceptional generalization capabilities. Despite these advancements, state-of-the-art methods rely on single-strategy prompting, missing the synerg…