paper-with-me

홈 › Papers

Language-driven Object Fusion into Neural Radiance Fields with Pose-Conditioned Dataset Updates

2023-09-20 · CVPR 2024 1 · Ka Chun Shum, Jaeyeon Kim, Binh-Son Hua, Duc Thanh Nguyen, Sai-Kit Yeung

Neural radiance field is an emerging rendering method that generates high-quality multi-view consistent images from a neural scene representation and volume rendering. Although neural radiance field-based techniques are robust for scene reconstruction, their ability to add or remove objects remains limited. This paper proposes a new language-driven approach for object manipulation with neural radiance fields through dataset updates. Specifically, to insert a new foreground object represented by a set of multi-view images into a background radiance field, we use a text-to-image diffusion model to learn and generate combined images that fuse the object of interest into the given background across views. These combined images are then used for refining the background radiance field so that we can render view-consistent images containing both the object and the background. To ensure view consistency, we propose a dataset updates strategy that prioritizes radiance field training with camera views close to the already-trained views prior to propagating the training to remaining views. We show that under the same dataset updates strategy, we can easily adapt our method for object insertion using data from text-to-3D models as well as object removal. Experimental results show that our method generates photorealistic images of the edited scenes, and outperforms state-of-the-art methods in 3D reconstruction and neural radiance field blending.

📄 PDF Abstract BibTeX arXiv:2309.11281

Code (1)

kcshum/pose-conditioned-NeRF-object-fusion 공식 구현 pytorch

Tasks

3D ReconstructionObjectText to 3D

Methods 이 논문이 사용한 방법론

Diffusion Diffusion models generate samples by gradually removing noise from a signal, and their training objective can be expressed as a reweighted variational lower-bound…

Similar Papers 제목 키워드 기반

DiLightNet: Fine-grained Lighting Control for Diffusion-based Image Generation

2024-02-19 · Chong Zeng, Yue Dong, Pieter Peers, Youkang Kong 외

This paper presents a novel method for exerting fine-grained lighting control during text-driven diffusion-based image generation. While existing diffusion models already have the ability to generate images under any lig…

Image Generation

A Diffusion Approach to Radiance Field Relighting using Multi-Illumination Synthesis

2024-09-13 · Yohan Poirier-Ginter, Alban Gauthier, Julien Philip, Jean-Francois Lalonde 외

Relighting radiance fields is severely underconstrained for multi-view data, which is most often captured under a single illumination condition; It is especially hard for full scenes containing multiple objects. We intro…

DreamHOI: Subject-Driven Generation of 3D Human-Object Interactions with Diffusion Priors

2024-09-12 · Thomas Hanwen Zhu, Ruining Li, Tomas Jakab

We present DreamHOI, a novel method for zero-shot synthesis of human-object interactions (HOIs), enabling a 3D human model to realistically interact with any given object based on a textual description. This task is comp…

Human-Object Interaction DetectionNeRF

ViFu: Multiple 360$^\circ$ Objects Reconstruction with Clean Background via Visible Part Fusion

2024-04-15 · Tianhan Xu, Takuya Ikeda, Koichi Nishiwaki

In this paper, we propose a method to segment and recover a static, clean background and multiple 360$^\circ$ objects from observations of scenes at different timestamps. Recent works have used neural radiance fields to …

Novel View SynthesisSynthetic Data Generation

Infrared and visible Image Fusion with Language-driven Loss in CLIP Embedding Space

2024-02-26 · Yuhao Wang, Lingjuan Miao, Zhiqiang Zhou, Lei Zhang 외

Infrared-visible image fusion (IVIF) has attracted much attention owing to the highly-complementary properties of the two image modalities. Due to the lack of ground-truth fused images, the fusion output of current deep-…

Infrared And Visible Image Fusion