paper-with-me

Papers

Diffutoon: High-Resolution Editable Toon Shading via Diffusion Models

2024-01-29 · Zhongjie Duan, Chengyu Wang, Cen Chen, Weining Qian, Jun Huang

Toon shading is a type of non-photorealistic rendering task of animation. Its primary purpose is to render objects with a flat and stylized appearance. As diffusion models have ascended to the forefront of image synthesis methodologies, this paper delves into an innovative form of toon shading based on diffusion models, aiming to directly render photorealistic videos into anime styles. In video stylization, extant methods encounter persistent challenges, notably in maintaining consistency and achieving high visual quality. In this paper, we model the toon shading problem as four subproblems: stylization, consistency enhancement, structure guidance, and colorization. To address the challenges in video stylization, we propose an effective toon shading approach called \textit{Diffutoon}. Diffutoon is capable of rendering remarkably detailed, high-resolution, and extended-duration videos in anime style. It can also edit the content according to prompts via an additional branch. The efficacy of Diffutoon is evaluated through quantitive metrics and human evaluation. Notably, Diffutoon surpasses both open-source and closed-source baseline approaches in our experiments. Our work is accompanied by the release of both the source code and example videos on Github (Project page: https://ecnu-cilab.github.io/DiffutoonProjectPage/).

📄 PDF Abstract BibTeX arXiv:2401.16224

Code (0)

등록된 구현이 없습니다.

Tasks

ColorizationImage Generation

Methods 이 논문이 사용한 방법론

Diffusion Diffusion models generate samples by gradually removing noise from a signal, and their training objective can be expressed as a reweighted variational lower-bound…

Similar Papers 제목 키워드 기반

Learning to Incorporate Texture Saliency Adaptive Attention to Image Cartoonization

2022-08-02 · Xiang Gao, Yuqi Zhang, Yingjie Tian

Image cartoonization is recently dominated by generative adversarial networks (GANs) from the perspective of unsupervised image-to-image translation, in which an inherent challenge is to precisely capture and sufficientl…

Image-to-Image TranslationStyle TransferUnsupervised Image-To-Image Translation

CartoonGAN: Generative Adversarial Networks for Photo Cartoonization

2018-06-01 · CVPR 2018 6 · Yang Chen, Yu-Kun Lai, Yong-Jin Liu

In this paper, we propose a solution to transforming photos of real-world scenes into cartoon style images, which is valuable and challenging in computer vision and computer graphics. Our solution belongs to learning bas…

Generative Adversarial NetworkImage-to-Image TranslationReal-to-Cartoon translation

Tree-Structured Shading Decomposition

2023-09-13 · ICCV 2023 1 · Chen Geng, Hong-Xing Yu, Sharon Zhang, Maneesh Agrawala 외

We study inferring a tree-structured representation from a single image for object shading. Prior work typically uses the parametric or measured representation to model shading, which is neither interpretable nor easily …

Object

Instance-guided Cartoon Editing with a Large-scale Dataset

2023-12-04 · Jian Lin, Chengze Li, Xueting Liu, Zhongping Ge

Cartoon editing, appreciated by both professional illustrators and hobbyists, allows extensive creative freedom and the development of original narratives within the cartoon domain. However, the existing literature on ca…

Image SegmentationSegmentationSemantic Segmentation

Creative Flow+ Dataset

2019-06-01 · CVPR 2019 6 · Maria Shugrina, Ziheng Liang, Amlan Kar, Jiaman Li 외

We present the Creative Flow+ Dataset, the first diverse multi-style artistic video dataset richly labeled with per-pixel optical flow, occlusions, correspondences, segmentation labels, normals, and depth. Our dataset in…

3D Character Animation From A Single PhotoDepth EstimationImage AnimationObject Tracking+5