paper-with-me

Papers

MVInpainter: Learning Multi-View Consistent Inpainting to Bridge 2D and 3D Editing

2024-08-15 · Chenjie Cao, Chaohui Yu, Fan Wang, xiangyang xue, Yanwei Fu

Novel View Synthesis (NVS) and 3D generation have recently achieved prominent improvements. However, these works mainly focus on confined categories or synthetic 3D assets, which are discouraged from generalizing to challenging in-the-wild scenes and fail to be employed with 2D synthesis directly. Moreover, these methods heavily depended on camera poses, limiting their real-world applications. To overcome these issues, we propose MVInpainter, re-formulating the 3D editing as a multi-view 2D inpainting task. Specifically, MVInpainter partially inpaints multi-view images with the reference guidance rather than intractably generating an entirely novel view from scratch, which largely simplifies the difficulty of in-the-wild NVS and leverages unmasked clues instead of explicit pose conditions. To ensure cross-view consistency, MVInpainter is enhanced by video priors from motion components and appearance guidance from concatenated reference key&value attention. Furthermore, MVInpainter incorporates slot attention to aggregate high-level optical flow features from unmasked regions to control the camera movement with pose-free training and inference. Sufficient scene-level experiments on both object-centric and forward-facing datasets verify the effectiveness of MVInpainter, including diverse tasks, such as multi-view object removal, synthesis, insertion, and replacement. The project page is https://ewrfcas.github.io/MVInpainter/.

📄 PDF Abstract BibTeX arXiv:2408.08000

Code (0)

등록된 구현이 없습니다.

Tasks

3D GenerationNovel View SynthesisOptical Flow Estimation

Methods 이 논문이 사용한 방법론

Softmax The Softmax output function transforms a previous layer's output into a vector of probabilities. It is commonly used for multiclass classification. Given an input vector $x$…
Attention 설명 없음
Focus 설명 없음
Inpainting Train a convolutional neural network to generate the contents of an arbitrary image region conditioned on its surroundings.

Similar Papers 제목 키워드 기반

Generative Object Insertion in Gaussian Splatting with a Multi-View Diffusion Model

2024-09-25 · Hongliang Zhong, Can Wang, Jingbo Zhang, Jing Liao

Generating and inserting new objects into 3D content is a compelling approach for achieving versatile scene recreation. Existing methods, which rely on SDS optimization or single-view inpainting, often struggle to produc…

3D ReconstructionObject

IMFine: 3D Inpainting via Geometry-guided Multi-view Refinement

2025-03-06 · CVPR 2025 1 · Zhihao Shi, Dong Huo, Yuhongze Zhou, Kejia Yin 외

Current 3D inpainting and object removal methods are largely limited to front-facing scenes, facing substantial challenges when applied to diverse, "unconstrained" scenes where the camera orientation and trajectory are u…

3D InpaintingImage InpaintingTest-time Adaptation

MVIP-NeRF: Multi-view 3D Inpainting on NeRF Scenes via Diffusion Prior

2024-05-05 · CVPR 2024 1 · Honghua Chen, Chen Change Loy, Xingang Pan

Despite the emergence of successful NeRF inpainting methods built upon explicit RGB and depth 2D inpainting supervisions, these methods are inherently constrained by the capabilities of their underlying 2D inpainters. Th…

3D InpaintingNeRF

Geometry-Aware Diffusion Models for Multiview Scene Inpainting

2025-02-18 · Ahmad Salimi, Tristan Aumentado-Armstrong, Marcus A. Brubaker, Konstantinos G. Derpanis

In this paper, we focus on 3D scene inpainting, where parts of an input image set, captured from different viewpoints, are masked out. The main challenge lies in generating plausible image completions that are geometrica…

3D InpaintingNeRF

3D Gaussian Inpainting with Depth-Guided Cross-View Consistency

2025-02-17 · CVPR 2025 1 · Sheng-Yu Huang, Zi-Ting Chou, Yu-Chiang Frank Wang

When performing 3D inpainting using novel-view rendering methods like Neural Radiance Field (NeRF) or 3D Gaussian Splatting (3DGS), how to achieve texture and geometry consistency across camera views has been a challenge…

3DGS3D InpaintingNeRF