paper-with-me

Papers

DepR: Depth Guided Single-view Scene Reconstruction with Instance-level Diffusion

2025-07-30 · Qingcheng Zhao, Xiang Zhang, Haiyang Xu, Zeyuan Chen, Jianwen Xie, Yuan Gao, Zhuowen Tu arxiv

We propose DepR, a depth-guided single-view scene reconstruction framework that integrates instance-level diffusion within a compositional paradigm. Instead of reconstructing the entire scene holistically, DepR generates individual objects and subsequently composes them into a coherent 3D layout. Unlike previous methods that use depth solely for object layout estimation during inference and therefore fail to fully exploit its rich geometric information, DepR leverages depth throughout both training and inference. Specifically, we introduce depth-guided conditioning to effectively encode shape priors into diffusion models. During inference, depth further guides DDIM sampling and layout optimization, enhancing alignment between the reconstruction and the input image. Despite being trained on limited synthetic data, DepR achieves state-of-the-art performance and demonstrates strong generalization in single-view scene reconstruction, as shown through evaluations on both synthetic and real-world datasets.

📄 PDF Abstract BibTeX arXiv:2507.22825

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Deep Reinforcement Learning of Volume-guided Progressive View Inpainting for 3D Point Scene Completion from a Single Depth Image

2019-03-10 · CVPR 2019 6 · Xiaoguang Han, Zhaoxuan Zhang, Dong Du, Mingdai Yang 외

We present a deep reinforcement learning method of progressive view inpainting for 3D point scene completion under volume guidance, achieving high-quality scene reconstruction from only a single depth image with severe o…

Deep Reinforcement LearningReinforcement Learning

Tabletop Transparent Scene Reconstruction via Epipolar-Guided Optical Flow with Monocular Depth Completion Prior

2023-10-15 · Xiaotong Chen, Zheming Zhou, Zhuo Deng, Omid Ghasemalizadeh 외

Reconstructing transparent objects using affordable RGB-D cameras is a persistent challenge in robotic perception due to inconsistent appearances across views in the RGB domain and inaccurate depth readings in each singl…

3D ReconstructionDepth CompletionOptical Flow EstimationSemantic Segmentation+1

WonderWorld: Interactive 3D Scene Generation from a Single Image

2024-06-13 · CVPR 2025 1 · Hong-Xing Yu, Haoyi Duan, Charles Herrmann, William T. Freeman 외

We present WonderWorld, a novel framework for interactive 3D scene generation that enables users to interactively specify scene contents and layout and see the created scenes in low latency. The major challenge lies in a…

Depth EstimationGPUNavigateScene Generation

Pseudo Label-Guided Multi Task Learning for Scene Understanding

2021-01-01 · Sunkyung Kim, Hyesong Choi, Dongbo Min

Multi-task learning (MTL) for scene understanding has been actively studied by exploiting correlation of multiple tasks. This work focuses on improving the performance of the MTL network that infers depth and semantic se…

Depth EstimationMonocular Depth EstimationMulti-Task LearningPseudo Label+4

Reference-guided Controllable Inpainting of Neural Radiance Fields

2023-04-19 · ICCV 2023 1 · Ashkan Mirzaei, Tristan Aumentado-Armstrong, Marcus A. Brubaker, Jonathan Kelly 외

The popularity of Neural Radiance Fields (NeRFs) for view synthesis has led to a desire for NeRF editing tools. Here, we focus on inpainting regions in a view-consistent and controllable manner. In addition to the typica…

NeRF