paper-with-me

Papers

HyperDreamer: Hyper-Realistic 3D Content Generation and Editing from a Single Image

2023-12-07 · Tong Wu, Zhibing Li, Shuai Yang, Pan Zhang, Xinggang Pan, Jiaqi Wang, Dahua Lin, Ziwei Liu

3D content creation from a single image is a long-standing yet highly desirable task. Recent advances introduce 2D diffusion priors, yielding reasonable results. However, existing methods are not hyper-realistic enough for post-generation usage, as users cannot view, render and edit the resulting 3D content from a full range. To address these challenges, we introduce HyperDreamer with several key designs and appealing properties: 1) Viewable: 360 degree mesh modeling with high-resolution textures enables the creation of visually compelling 3D models from a full range of observation points. 2) Renderable: Fine-grained semantic segmentation and data-driven priors are incorporated as guidance to learn reasonable albedo, roughness, and specular properties of the materials, enabling semantic-aware arbitrary material estimation. 3) Editable: For a generated model or their own data, users can interactively select any region via a few clicks and efficiently edit the texture with text-based guidance. Extensive experiments demonstrate the effectiveness of HyperDreamer in modeling region-aware materials with high-resolution textures and enabling user-friendly editing. We believe that HyperDreamer holds promise for advancing 3D content creation and finding applications in various domains.

📄 PDF Abstract BibTeX arXiv:2312.04543

Code (0)

등록된 구현이 없습니다.

Tasks

Semantic Segmentation

Methods 이 논문이 사용한 방법론

Diffusion Diffusion models generate samples by gradually removing noise from a signal, and their training objective can be expressed as a reweighted variational lower-bound…

Similar Papers 제목 키워드 기반

HyperEdit: Unlocking Instruction-based Text Editing in LLMs via Hypernetworks

2025-12-14 · Yiming Zeng, Jinghan Cao, Zexin Li, Wanhao Yu 외 arxiv

Instruction-based text editing is increasingly critical for real-world applications such as code editors (e.g., Cursor), but Large Language Models (LLMs) continue to struggle with this task. Unlike free-form generation, …

Text Generation

Enhanced 3D Generation by 2D Editing

2024-12-08 · Haoran Li, Yuli Tian, Yong Liao, Lin Wang 외

Distilling 3D representations from pretrained 2D diffusion models is essential for 3D creative applications across gaming, film, and interior design. Current SDS-based methods are hindered by inefficient information dist…

3D GenerationDenoising

Make-A-Protagonist: Generic Video Editing with An Ensemble of Experts

2023-05-15 · Yuyang Zhao, Enze Xie, Lanqing Hong, Zhenguo Li 외

The text-driven image and video diffusion models have achieved unprecedented success in generating realistic and diverse content. Recently, the editing and variation of existing images and videos in diffusion-based gener…

DenoisingVideo EditingVideo Generation

Unifying Speech Editing Detection and Content Localization via Prior-Enhanced Audio LLMs

2026-01-29 · Jun Xue, Yi Chai, Yanzhen Ren, Jinshen He 외 arxiv

Existing speech editing detection (SED) datasets are predominantly constructed using manual splicing or limited editing operations, resulting in restricted diversity and poor coverage of realistic editing scenarios. Mean…

Text Generation

PICABench: How Far Are We from Physically Realistic Image Editing?

2025-10-20 · Yuandong Pu, Le Zhuo, Songhao Han, Jinbo Xing 외 arxiv

Image editing has achieved remarkable progress recently. Modern editing models could already follow complex instructions to manipulate the original content. However, beyond completing the editing instructions, the accomp…

Image Editing