paper-with-me

Papers

LVCD: Reference-based Lineart Video Colorization with Diffusion Models

2024-09-19 · Zhitong Huang, Mohan Zhang, Jing Liao

We propose the first video diffusion framework for reference-based lineart video colorization. Unlike previous works that rely solely on image generative models to colorize lineart frame by frame, our approach leverages a large-scale pretrained video diffusion model to generate colorized animation videos. This approach leads to more temporally consistent results and is better equipped to handle large motions. Firstly, we introduce Sketch-guided ControlNet which provides additional control to finetune an image-to-video diffusion model for controllable video synthesis, enabling the generation of animation videos conditioned on lineart. We then propose Reference Attention to facilitate the transfer of colors from the reference frame to other frames containing fast and expansive motions. Finally, we present a novel scheme for sequential sampling, incorporating the Overlapped Blending Module and Prev-Reference Attention, to extend the video diffusion model beyond its original fixed-length limitation for long video colorization. Both qualitative and quantitative results demonstrate that our method significantly outperforms state-of-the-art techniques in terms of frame and video quality, as well as temporal consistency. Moreover, our method is capable of generating high-quality, long temporal-consistent animation videos with large motions, which is not achievable in previous works. Our code and model are available at https://luckyhzt.github.io/lvcd.

📄 PDF Abstract BibTeX arXiv:2409.12960

Code (0)

등록된 구현이 없습니다.

Tasks

Colorization

Methods 이 논문이 사용한 방법론

Softmax The Softmax output function transforms a previous layer's output into a vector of probabilities. It is commonly used for multiclass classification. Given an input vector $x$…
Attention 설명 없음
Diffusion Diffusion models generate samples by gradually removing noise from a signal, and their training objective can be expressed as a reweighted variational lower-bound…

Similar Papers 제목 키워드 기반

OmniColor: A Unified Framework for Multi-modal Lineart Colorization

2026-03-29 · Xulu Zhang, Haoqian Du, Xiaoyong Wei, Qing Li arxiv

Lineart colorization is a critical stage in professional content creation, yet achieving precise and flexible results under diverse user constraints remains a significant challenge. To address this, we propose OmniColor,…

Video Colorization with Pre-trained Text-to-Image Diffusion Models

2023-06-02 · Hanyuan Liu, Minshan Xie, Jinbo Xing, Chengze Li 외

Video colorization is a challenging task that involves inferring plausible and temporally consistent colors for grayscale frames. In this paper, we present ColorDiffuser, an adaptation of a pre-trained text-to-image late…

Colorization

AnimeColor: Reference-based Animation Colorization with Diffusion Transformers

2025-07-27 · Yuhong Zhang, Liyao Wang, Han Wang, Danni Wu 외 arxiv

Animation colorization plays a vital role in animation production, yet existing methods struggle to achieve color accuracy and temporal consistency. To address these challenges, we propose \textbf{AnimeColor}, a novel re…

Uni-Animator: Towards Unified Visual Colorization

2026-02-26 · Xinyuan Chen, Yao Xu, Shaowen Wang, Pengjie Song 외 arxiv

We propose Uni-Animator, a novel Diffusion Transformer (DiT)-based framework for unified image and video sketch colorization. Existing sketch colorization methods struggle to unify image and video tasks, suffering from i…

LatentColorization: Latent Diffusion-Based Speaker Video Colorization

2024-05-09 · Rory Ward, Dan Bigioi, Shubhajit Basak, John G. Breslin 외

While current research predominantly focuses on image-based colorization, the domain of video-based colorization remains relatively unexplored. Most existing video colorization techniques operate on a frame-by-frame basi…

Colorization