paper-with-me

Papers

Referenceless User Controllable Semantic Image Synthesis

2023-06-18 · Jonghyun Kim, Gen Li, Joongkyu Kim

Despite recent progress in semantic image synthesis, complete control over image style remains a challenging problem. Existing methods require reference images to feed style information into semantic layouts, which indicates that the style is constrained by the given image. In this paper, we propose a model named RUCGAN for user controllable semantic image synthesis, which utilizes a singular color to represent the style of a specific semantic region. The proposed network achieves reference-free semantic image synthesis by injecting color as user-desired styles into each semantic layout, and is able to synthesize semantic images with unusual colors. Extensive experimental results on various challenging datasets show that the proposed method outperforms existing methods, and we further provide an interactive UI to demonstrate the advantage of our approach for style controllability.

📄 PDF Abstract BibTeX arXiv:2306.10646

Code (1)

benjaminjonghyun/rucgan 공식 구현

Tasks

Image Generation

Similar Papers 제목 키워드 기반

Context Matters for Image Descriptions for Accessibility: Challenges for Referenceless Evaluation Metrics

2022-05-21 · Elisa Kreiss, Cynthia Bennett, Shayan Hooshmand, Eric Zelikman 외

Few images on the Web receive alt-text descriptions that would make them accessible to blind and low vision (BLV) users. Image-based NLG systems have progressed to the point where they can begin to address this persisten…

Motion-Conditioned Diffusion Model for Controllable Video Synthesis

2023-04-27 · Tsai-Shien Chen, Chieh Hubert Lin, Hung-Yu Tseng, Tsung-Yi Lin 외

Recent advancements in diffusion models have greatly improved the quality and diversity of synthesized content. To harness the expressive power of diffusion models, researchers have explored various controllable mechanis…

DiversitymodelMotion Synthesis

Controllable Multi-domain Semantic Artwork Synthesis

2023-08-19 · Yuantian Huang, Satoshi Iizuka, Edgar Simo-Serra, Kazuhiro Fukui

We present a novel framework for multi-domain synthesis of artwork from semantic layouts. One of the main limitations of this challenging task is the lack of publicly available segmentation datasets for art synthesis. To…

Generative Adversarial Network

Controllable Generation of Large-Scale 3D Urban Layouts with Semantic and Structural Guidance

2025-09-28 · Mengyuan Niu, Xinxin Zhuo, Ruizhe Wang, Yuyue Huang 외 arxiv

Urban modeling is essential for city planning, scene synthesis, and gaming. Existing image-based methods generate diverse layouts but often lack geometric continuity and scalability, while graph-based methods capture str…

High-Fidelity Guided Image Synthesis with Latent Diffusion Models

2022-11-30 · CVPR 2023 1 · Jaskirat Singh, Stephen Gould, Liang Zheng

Controllable image synthesis with user scribbles has gained huge public interest with the recent advent of text-conditioned latent diffusion models. The user scribbles control the color composition while the text prompt …

Image GenerationVocal Bursts Intensity Prediction