Infinite Texture: Text-guided High Resolution Diffusion Texture Synthesis
We present Infinite Texture, a method for generating arbitrarily large texture images from a text prompt. Our approach fine-tunes a diffusion model on a single texture, and learns to embed that statistical distribution in the output domain of the model. We seed this fine-tuning process with a sample texture patch, which can be optionally generated from a text-to-image model like DALL-E 2. At generation time, our fine-tuned diffusion model is used through a score aggregation strategy to generate output texture images of arbitrary resolution on a single GPU. We compare synthesized textures from our method to existing work in patch-based and deep learning texture synthesis methods. We also showcase two applications of our generated textures in 3D rendering and texture transfer.
Code (0)
등록된 구현이 없습니다.
Tasks
GPUTexture SynthesisMethods 이 논문이 사용한 방법론
Similar Papers 제목 키워드 기반
Local Padding in Patch-Based GANs for Seamless Infinite-Sized Texture Synthesis
Texture models based on Generative Adversarial Networks (GANs) use zero-padding to implicitly encode positional information of the image features. However, when extending the spatial input to generate images at large siz…
DiversityGPUSuper-ResolutionTexture SynthesisWeak Texture Information Map Guided Image Super-resolution with Deep Residual Networks
Single image super-resolution (SISR) is an image processing task which obtains high-resolution (HR) image from a low-resolution (LR) image. Recently, due to the capability in feature extraction, a series of deep learning…
Image Super-ResolutionSuper-ResolutionTEXGen: a Generative Diffusion Model for Mesh Textures
While high-quality texture maps are essential for realistic 3D asset rendering, few studies have explored learning directly in the texture space, especially on large-scale datasets. In this work, we depart from the conve…
modelTexture SynthesisText-guided High-definition Consistency Texture Model
With the advent of depth-to-image diffusion models, text-guided generation, editing, and transfer of realistic textures are no longer difficult. However, due to the limitations of pre-trained diffusion models, they can o…
modelparameter-efficient fine-tuningtext-guided-generationVocal Bursts Intensity PredictionEvTexture++: Event-Driven Texture Enhancement for Video Super-Resolution
Event-based vision has drawn increasing attention owing to its distinctive properties, including ultra-high temporal resolution and extreme dynamic range. Recent works have introduced it to video super-resolution (VSR) t…
Video Super-ResolutionEvent-based vision