paper-with-me

홈 › Papers

AnimatableDreamer: Text-Guided Non-rigid 3D Model Generation and Reconstruction with Canonical Score Distillation

2023-12-06 · Xinzhou Wang, Yikai Wang, Junliang Ye, Zhengyi Wang, Fuchun Sun, Pengkun Liu, Ling Wang, Kai Sun, Xintong Wang, Bin He

Advances in 3D generation have facilitated sequential 3D model generation (a.k.a 4D generation), yet its application for animatable objects with large motion remains scarce. Our work proposes AnimatableDreamer, a text-to-4D generation framework capable of generating diverse categories of non-rigid objects on skeletons extracted from a monocular video. At its core, AnimatableDreamer is equipped with our novel optimization design dubbed Canonical Score Distillation (CSD), which lifts 2D diffusion for temporal consistent 4D generation. CSD, designed from a score gradient perspective, generates a canonical model with warp-robustness across different articulations. Notably, it also enhances the authenticity of bones and skinning by integrating inductive priors from a diffusion model. Furthermore, with multi-view distillation, CSD infers invisible regions, thereby improving the fidelity of monocular non-rigid reconstruction. Extensive experiments demonstrate the capability of our method in generating high-flexibility text-guided 3D models from the monocular video, while also showing improved reconstruction performance over existing non-rigid reconstruction methods.

📄 PDF Abstract BibTeX arXiv:2312.03795

Code (0)

등록된 구현이 없습니다.

Tasks

3D GenerationDenoisingText to 3D

Methods 이 논문이 사용한 방법론

Diffusion Diffusion models generate samples by gradually removing noise from a signal, and their training objective can be expressed as a reweighted variational lower-bound…

Similar Papers 제목 키워드 기반

AvatarGen: a 3D Generative Model for Animatable Human Avatars

2022-08-01 · Jianfeng Zhang, Zihang Jiang, Dingdong Yang, Hongyi Xu 외

Unsupervised generation of clothed virtual humans with various appearance and animatable poses is important for creating 3D human avatars and other AR/VR applications. Existing methods are either limited to rigid object …

3D Human Reconstruction

MHR-Net: Multiple-Hypothesis Reconstruction of Non-Rigid Shapes from 2D Views

2022-07-19 · Haitian Zeng, Xin Yu, Jiaxu Miao, Yi Yang

We propose MHR-Net, a novel method for recovering Non-Rigid Shapes from Motion (NRSfM). MHR-Net aims to find a set of reasonable reconstructions for a 2D view, and it also selects the most likely reconstruction from the …

MasaCtrl: Tuning-Free Mutual Self-Attention Control for Consistent Image Synthesis and Editing

2023-04-17 · ICCV 2023 1 · Mingdeng Cao, Xintao Wang, Zhongang Qi, Ying Shan 외

Despite the success in large-scale text-to-image generation and text-conditioned image editing, existing methods still struggle to produce consistent generation and editing results. For example, generation approaches usu…

Image GenerationText-based Image EditingText to Image GenerationText-to-Image Generation

Neural Rendering for Stereo 3D Reconstruction of Deformable Tissues in Robotic Surgery

2022-06-30 · Yuehao Wang, Yonghao Long, Siu Hin Fan, Qi Dou

Reconstruction of the soft tissues in robotic surgery from endoscopic stereo videos is important for many applications such as intra-operative navigation and image-guided robotic surgery automation. Previous works on thi…

3D ReconstructionNeural Rendering

LoCAtion: Long-time Collaborative Attention Framework for High Dynamic Range Video Reconstruction

2026-03-15 · Qianyu Zhang, Bolun Zheng, Lingyu Zhu, Aiai Huang 외 arxiv

Prevailing High Dynamic Range (HDR) video reconstruction methods are fundamentally trapped in a fragile alignment-and-fusion paradigm. While explicit spatial alignment can successfully recover fine details in controlled …

Computational EfficiencyVideo ReconstructionVideo Generation