paper-with-me

Papers

Space Narrative: Generating Images and 3D Scenes of Chinese Garden from Text using Deep Learning

2023-11-01 · Jiaxi Shi1, Hao Hua1

The consistent mapping from poems to paintings is essential for the research and restoration of traditional Chinese gardens. But the lack of firsthand ma-terial is a great challenge to the reconstruction work. In this paper, we pro-pose a method to generate garden paintings based on text descriptions using deep learning method. Our image-text pair dataset consists of more than one thousand Ming Dynasty Garden paintings and their inscriptions and post-scripts. A latent text-to-image diffusion model learns the mapping from de-scriptive texts to garden paintings of the Ming Dynasty, and then the text description of Jichang Garden guides the model to generate new garden paintings. The cosine similarity between the guide text and the generated image is the evaluation criterion for the generated images. Our dataset is used to fine-tune the pre-trained diffusion model using Low-Rank Adapta-tion of Large Language Models (LoRA). We also transformed the generated images into a panorama and created a free-roam scene in Unity 3D. Our post-trained model is capable of generating garden images in the style of Ming Dynasty landscape paintings based on textual descriptions. The gener-ated images are compatible with three-dimensional presentation in Unity 3D.

📄 PDF Abstract BibTeX arXiv:2311.00339

Code (0)

등록된 구현이 없습니다.

Tasks

Unity

Methods 이 논문이 사용한 방법론

Diffusion Diffusion models generate samples by gradually removing noise from a signal, and their training objective can be expressed as a reweighted variational lower-bound…

Similar Papers 제목 키워드 기반

Stories for Images-in-Sequence by using Visual and Narrative Components

2018-05-15 · Marko Smilevski, Ilija Lalkovski, Gjorgji Madjarov

Recent research in AI is focusing towards generating narrative stories about visual scenes. It has the potential to achieve more human-like understanding than just basic description generation of images- in-sequence. In …

Sentence

BiT-MCTS: A Theme-based Bidirectional MCTS Approach to Chinese Fiction Generation

2026-03-15 · Zhaoyi Li, Xu Zhang, Xiaojun Wan arxiv

Generating long-form linear fiction from open-ended themes remains a major challenge for large language models, which frequently fail to guarantee global structure and narrative diversity when using premise-based or line…

HoloCine: Holistic Generation of Cinematic Multi-Shot Long Video Narratives

2025-10-23 · Yihao Meng, Hao Ouyang, Yue Yu, Qiuyu Wang 외 arxiv

State-of-the-art text-to-video models excel at generating isolated clips but fall short of creating the coherent, multi-shot narratives, which are the essence of storytelling. We bridge this "narrative gap" with HoloCine…

DiffuVST: Narrating Fictional Scenes with Global-History-Guided Denoising Models

2023-12-12 · Shengguang Wu, Mei Yuan, Qi Su

Recent advances in image and video creation, especially AI-based image synthesis, have led to the production of numerous visual scenes that exhibit a high level of abstractness and diversity. Consequently, Visual Storyte…

DenoisingDiversityImage GenerationImage to text+3

LLMs Behind the Scenes: Enabling Narrative Scene Illustration

2025-09-26 · Melissa Roemmele, John Joon Young Chung, Taewook Kim, Yuqian Sun 외 arxiv

Generative AI has established the opportunity to readily transform content from one medium to another. This capability is especially powerful for storytelling, where visual illustrations can illuminate a story originally…