GLIDE
2000년 도입 · 논문 28편에서 사용
GLIDE is a generative model based on text-guided diffusion models for more photorealistic image generation. Guided diffusion is applied to text-conditional image synthesis and the model is able to handle free-form prompts. The diffusion model uses a text encoder to condition on natural language descriptions. The model is provided with editing capabilities in addition to zero-shot generation, allowing for iterative improvement of model samples to match more complex prompts. The model is fine-tuned to perform image inpainting.
출처: GLIDE: Towards Photorealistic Image Generation and Editing with Text-Guided Diffusion Models
소개 논문: GLIDE: Towards Photorealistic Image Generation and Editing with Text-Guided Diffusion Models
Image Generation Models · Computer VisionMulti-Modal Methods · Computer Vision