paper-with-me

홈 › Papers

STORYANCHORS: Generating Consistent Multi-Scene Story Frames for Long-Form Narratives

2025-05-13 · Bo wang, Haoyang Huang, Zhiying Lu, Fengyuan Liu, Guoqing Ma, Jianlong Yuan, Yuan Zhang, Nan Duan, Daxin Jiang

This paper introduces StoryAnchors, a unified framework for generating high-quality, multi-scene story frames with strong temporal consistency. The framework employs a bidirectional story generator that integrates both past and future contexts to ensure temporal consistency, character continuity, and smooth scene transitions throughout the narrative. Specific conditions are introduced to distinguish story frame generation from standard video synthesis, facilitating greater scene diversity and enhancing narrative richness. To further improve generation quality, StoryAnchors integrates Multi-Event Story Frame Labeling and Progressive Story Frame Training, enabling the model to capture both overarching narrative flow and event-level dynamics. This approach supports the creation of editable and expandable story frames, allowing for manual modifications and the generation of longer, more complex sequences. Extensive experiments show that StoryAnchors outperforms existing open-source models in key areas such as consistency, narrative coherence, and scene diversity. Its performance in narrative consistency and story richness is also on par with GPT-4o. Ultimately, StoryAnchors pushes the boundaries of story-driven frame generation, offering a scalable, flexible, and highly editable foundation for future research.

📄 PDF Abstract BibTeX arXiv:2505.08350

Code (0)

등록된 구현이 없습니다.

Tasks

DiversityForm

Similar Papers 제목 키워드 기반

DreamStory: Open-Domain Story Visualization by LLM-Guided Multi-Subject Consistent Diffusion

2024-07-17 · Huiguo He, Huan Yang, Zixi Tuo, Yuan Zhou 외

Story visualization aims to create visually compelling images or videos corresponding to textual narratives. Despite recent advances in diffusion models yielding promising results, existing methods still struggle to crea…

DescriptiveStory Visualization

Action2Dialogue: Generating Character-Centric Narratives from Scene-Level Prompts

2025-05-22 · Taewon Kang, Ming C. Lin

Recent advances in scene-based video generation have enabled systems to synthesize coherent visual narratives from structured prompts. However, a crucial dimension of storytelling -- character-driven dialogue and speech …

Dialogue GenerationLarge Language ModelStory GenerationVideo Generation+1

Make-A-Story: Visual Memory Conditioned Consistent Story Generation

2022-11-23 · CVPR 2023 1 · Tanzila Rahman, Hsin-Ying Lee, Jian Ren, Sergey Tulyakov 외

There has been a recent explosion of impressive generative models that can produce high quality images (or videos) conditioned on text descriptions. However, all such approaches rely on conditional sentences that contain…

SentenceStory GenerationStory Visualization

ContextualStory: Consistent Visual Storytelling with Spatially-Enhanced and Storyline Context

2024-07-13 · Sixiao Zheng, Yanwei Fu

Visual storytelling involves generating a sequence of coherent frames from a textual storyline while maintaining consistency in characters and scenes. Existing autoregressive methods, which rely on previous frame-sentenc…

Image GenerationStory ContinuationStory VisualizationText-to-Image Generation+1

InfinityStory: Unlimited Video Generation with World Consistency and Character-Aware Shot Transitions

2026-03-04 · Mohamed Elmoghany, Liangbing Zhao, Xiaoqian Shen, Subhojyoti Mukherjee 외 arxiv

Generating long-form storytelling videos with consistent visual narratives remains a significant challenge in video synthesis. We present a novel framework, dataset, and a model that address three critical limitations: b…

Video Generation