paper-with-me

홈 › Papers

DreamStyle: A Unified Framework for Video Stylization

2026-01-06 · Mengtian Li, Jinshu Chen, Songtao Zhao, Wanquan Feng, Pengqi Tu, Qian He arxiv

Video stylization, an important downstream task of video generation models, has not yet been thoroughly explored. Its input style conditions typically include text, style image, and stylized first frame. Each condition has a characteristic advantage: text is more flexible, style image provides a more accurate visual anchor, and stylized first frame makes long-video stylization feasible. However, existing methods are largely confined to a single type of style condition, which limits their scope of application. Additionally, their lack of high-quality datasets leads to style inconsistency and temporal flicker. To address these limitations, we introduce DreamStyle, a unified framework for video stylization, supporting (1) text-guided, (2) style-image-guided, and (3) first-frame-guided video stylization, accompanied by a well-designed data curation pipeline to acquire high-quality paired video data. DreamStyle is built on a vanilla Image-to-Video (I2V) model and trained using a Low-Rank Adaptation (LoRA) with token-specific up matrices that reduces the confusion among different condition tokens. Both qualitative and quantitative evaluations demonstrate that DreamStyle is competent in all three video stylization tasks, and outperforms the competitors in style consistency and video quality.

📄 PDF Abstract BibTeX arXiv:2601.02785

Code (0)

등록된 구현이 없습니다.

Tasks

Video Generation

Similar Papers 제목 키워드 기반

DreamStyler: Paint by Style Inversion with Text-to-Image Diffusion Models

2023-09-13 · Namhyuk Ahn, Junsoo Lee, Chunggi Lee, Kunhee Kim 외

Recent progresses in large-scale text-to-image models have yielded remarkable accomplishments, finding various applications in art domain. However, expressing unique characteristics of an artwork (e.g. brushwork, colorto…

Image GenerationStyle Transfer

Towards 4D Human Video Stylization

2023-12-07 · Tiantian Wang, Xinxin Zuo, Fangzhou Mu, Jian Wang 외

We present a first step towards 4D (3D and time) human video stylization, which addresses style transfer, novel view synthesis and human animation within a unified framework. While numerous video stylization methods have…

Human AnimationNovel View SynthesisStyle TransferVideo Reconstruction

DiT as Real-Time Rerenderer: Streaming Video Stylization with Autoregressive Diffusion Transformer

2026-04-15 · Hengye Lyu, Zisu Li, Yue Hong, Yueting Weng 외 arxiv

Recent advances in video generation models has significantly accelerated video generation and related downstream tasks. Among these, video stylization holds important research value in areas such as immersive application…

Video Generation

FreeViS: Training-free Video Stylization with Inconsistent References

2025-10-02 · Jiacong Xu, Yiqun Mei, Ke Zhang, Vishal M. Patel arxiv

Video stylization plays a key role in content creation, but it remains a challenging problem. Naïvely applying image stylization frame-by-frame hurts temporal consistency and reduces style richness. Alternatively, traini…

EchoStyle: Unlocking High-Fidelity Video Stylization with Reverse Data Synthesis

2026-06-24 · Huaqiu Li, Jiahao Wang, Sijia Cai, Hualian Sheng 외 arxiv

While image stylization has been studied extensively, video stylization remains a critical and largely unsolved challenge in the field of intelligent content creation. Existing methods, usually utilizing a reference imag…