paper-with-me

홈 › Papers

SceneMaker: Open-set 3D Scene Generation with Decoupled De-occlusion and Pose Estimation Model

2025-12-11 · Yukai Shi, Weiyu Li, Zihao Wang, Hongyang Li, Xingyu Chen, Ping Tan, Lei Zhang arxiv

We propose a decoupled 3D scene generation framework called SceneMaker in this work. Due to the lack of sufficient open-set de-occlusion and pose estimation priors, existing methods struggle to simultaneously produce high-quality geometry and accurate poses under severe occlusion and open-set settings. To address these issues, we first decouple the de-occlusion model from 3D object generation, and enhance it by leveraging image datasets and collected de-occlusion datasets for much more diverse open-set occlusion patterns. Then, we propose a unified pose estimation model that integrates global and local mechanisms for both self-attention and cross-attention to improve accuracy. Besides, we construct an open-set 3D scene dataset to further extend the generalization of the pose estimation model. Comprehensive experiments demonstrate the superiority of our decoupled framework on both indoor and open-set scenes. Our codes and datasets is released at https://idea-research.github.io/SceneMaker/.

📄 PDF Abstract BibTeX arXiv:2512.10957

Code (0)

등록된 구현이 없습니다.

Tasks

Scene GenerationPose Estimation

Similar Papers 제목 키워드 기반

Decoupled Diffusion Sparks Adaptive Scene Generation

2025-04-14 · Yunsong Zhou, Naisheng Ye, William Ljungbergh, Tianyu Li 외

Controllable scene generation could reduce the cost of diverse data collection substantially for autonomous driving. Prior works formulate the traffic layout generation as predictive progress, either by denoising entire …

Autonomous DrivingData AugmentationDenoisingLayout Generation+1

SeeThrough3D: Occlusion Aware 3D Control in Text-to-Image Generation

2026-02-26 · Vaibhav Agrawal, Rishubh Parihar, Pradhaan Bhat, Ravi Kiran Sarvadevabhatla 외 arxiv

We identify occlusion reasoning as a fundamental yet overlooked aspect for 3D layout-conditioned generation. It is essential for synthesizing partially occluded objects with depth-consistent geometry and scale. While exi…

Text-to-Image Generation

DDS: Decoupled Dynamic Scene-Graph Generation Network

2023-01-18 · A S M Iftekhar, Raphael Ruschel, Satish Kumar, Suya You 외

Scene-graph generation involves creating a structural representation of the relationships between objects in a scene by predicting subject-object-relation triplets from input data. Existing methods show poor performance …

Graph GenerationObjectScene Graph Generation

SoDa2: Single-Stage Open-Set Domain Adaptation via Decoupled Alignment for Cross-Scene Hyperspectral Image Classification

2026-05-05 · Yiwen Liu, Minghua Wang, Jing Yao, Xin Zhao 외 arxiv

Cross-scene hyperspectral image (HSI) classification stands as a fundamental research topic in remote sensing, with extensive applications spanning various fields. Owing to the inclusion of unknown categories in the targ…

Hyperspectral Image ClassificationDomain Adaptation

WaterGen: Decoupling Scene and Medium in Underwater Image Generation

2026-06-30 · Jiayi Wu, Tianfu Wang, Tianyi Xiong, Dehao Yuan 외 arxiv

Underwater computer vision tasks, such as detection, restoration, and segmentation, are limited by the scarcity of large-scale and diverse training data. We introduce WaterGen, a method for generating large-scale, realis…

Semantic SegmentationScene GenerationImage Generation