paper-with-me

홈 › Papers

Controllable Texture Tiling with Transformed RoPE-Enhanced Diffusion Models

2026-06-22 · Junrong Huang, Zhiyuan Zhang, Rui Tang, Hongbo Fu, Jnig Liao arxiv

Realistic integration of user-specified textures into scene images is a fundamental task in computer graphics and image editing. While existing material transfer and reference-guided inpainting methods can edit surface appearances, they often fail to address the specific requirements of texture tiling. This task necessitates precisely repeating a reference pattern according to user-defined parameters such as frequency, orientation, and scale. Furthermore, current generative approaches often struggle to maintain the structural fidelity of the reference texture, limited by either destructive pixel-level resampling or the lack of fine-grained spatial information in semantic image encoders, and they frequently fail to preserve the coherent lighting and geometry of the original scene. In this paper, we propose a novel framework for controllable and high-fidelity texture tiling based on Diffusion Transformers. Our approach introduces two key technical innovations to decouple spatial manipulation from content generation. First, we propose a Coordinate-Transformed Rotary Embedding mechanism. By applying 2D affine transformations directly to the relative positional embeddings between the target latent and the image condition, we achieve precise control over tiling patterns without explicit pixel warping, thereby utilizing the full information of the reference condition without degradation. Second, a Disjoint Attention Mask is employed to shield reference features from semantic leakage. This preserves structural integrity while seamlessly blending the synthesized texture with the scene's original lighting and geometry. Extensive experiments demonstrate that our method outperforms state-of-the-art baselines in both control accuracy and texture fidelity.

📄 PDF Abstract BibTeX arXiv:2606.22945

Code (0)

등록된 구현이 없습니다.

Tasks

Image Editing

Similar Papers 제목 키워드 기반

Structure-Transformed Texture-Enhanced Network for Person Image Synthesis

2021-01-01 · ICCV 2021 10 · Munan Xu, Yuanqi Chen, Shan Liu, Thomas H. Li 외

Pose-guided virtual try-on task aims to modify the fashion item based on pose transfer task. These two tasks that belong to person image synthesis have strong correlations and similarities. However, existing methods …

Image GenerationPose TransferVirtual Try-on

User-Controllable Multi-Texture Synthesis with Generative Adversarial Networks

2019-04-09 · Aibek Alanov, Max Kochurov, Denis Volkhonskiy, Daniil Yashkov 외

We propose a novel multi-texture synthesis model based on generative adversarial networks (GANs) with a user-controllable mechanism. The user control ability allows to explicitly specify the texture which should be gener…

DescriptiveTexture Synthesis

Texture Image Synthesis Using Spatial GAN Based on Vision Transformers

2025-02-03 · Elahe Salari, Zohreh Azimifar

Texture synthesis is a fundamental task in computer vision, whose goal is to generate visually realistic and structurally coherent textures for a wide range of applications, from graphics to scientific simulations. While…

Generative Adversarial NetworkImage GenerationSSIMTexture Synthesis

Dynamic Neural Textures: Generating Talking-Face Videos with Continuously Controllable Expressions

2022-04-13 · Zipeng Ye, Zhiyao Sun, Yu-Hui Wen, Yanan sun 외

Recently, talking-face video generation has received considerable attention. So far most methods generate results with neutral expressions or expressions that are implicitly determined by neural networks in an uncontroll…

Video Generation

TextSR: Diffusion Super-Resolution with Multilingual OCR Guidance

2025-05-29 · Keren Ye, Ignacio Garcia Dorado, Michalis Raptis, Mauricio Delbracio 외

While recent advancements in Image Super-Resolution (SR) using diffusion models have shown promise in improving overall image quality, their application to scene text images has revealed limitations. These models often s…

Image Super-ResolutionOptical Character RecognitionOptical Character Recognition (OCR)Super-Resolution+1