paper-with-me

Papers

SceneExpander: Text-Guided 3D Scene Expansion via Free-Form View Insertion

2026-03-28 · Zijian He, Renjie Liu, Yihao Wang, Weizhi Zhong, Huan Yuan, Kun Gai, Guangrun Wang, Guanbin Li arxiv

World building with 3D scene representations is increasingly important for content creation, simulation, and interactive experiences, yet real workflows are inherently iterative: creators repeatedly extend existing scenes under user control. Motivated by this gap, we study text-guided 3D scene expansion via free-form view insertion. Starting from a real scene captured by multi-view images, a user specifies a text expansion intent, which a generative model materializes as an inserted view extending the scene coverage. Unlike simple object editing or style transfer within a fixed scene, the inserted view may be 3D-misaligned with the original reconstruction, introducing geometric shifts, hallucinated content, or view-dependent artifacts that disrupt global multi-view consistency. To address this challenge, we propose SceneExpander, which applies test-time adaptation to a parametric feed-forward 3D reconstruction model with two complementary distillation signals: anchor distillation stabilizes the captured scene using geometric cues from the captured views, while inserted-view self-distillation retains insertion-supported predictions to accommodate the misaligned view. Experiments on ETH scenes and online data demonstrate improved expansion behavior and reconstruction quality under misalignment.

📄 PDF Abstract BibTeX arXiv:2603.27084

Code (0)

등록된 구현이 없습니다.

Tasks

Test-time Adaptation3D ReconstructionStyle Transfer

Similar Papers 제목 키워드 기반

LayerPano3D: Layered 3D Panorama for Hyper-Immersive Scene Generation

2024-08-23 · Shuai Yang, Jing Tan, Mengchen Zhang, Tong Wu 외

3D immersive scene generation is a challenging yet critical task in computer vision and graphics. A desired virtual 3D scene should 1) exhibit omnidirectional view consistency, and 2) allow for free exploration in comple…

Scene Generation

CBNet: A Plug-and-Play Network for Segmentation-Based Scene Text Detection

2022-12-05 · Xi Zhao, Wei Feng, Zheng Zhang, Jingjing Lv 외

Recently, segmentation-based methods are quite popular in scene text detection, which mainly contain two steps: text kernel segmentation and expansion. However, the segmentation process only considers each pixel independ…

Scene Text DetectionSegmentationText Detection

Salient Object-Aware Background Generation using Text-Guided Diffusion Models

2024-04-15 · Amir Erfan Eshratifar, Joao V. B. Soares, Kapil Thadani, Shaunak Mishra 외

Generating background scenes for salient objects plays a crucial role across various domains including creative design and e-commerce, as it enhances the presentation and context of subjects by integrating them into tail…

Object

Free-DyGS: Camera-Pose-Free Scene Reconstruction based on Gaussian Splatting for Dynamic Surgical Videos

2024-09-02 · Qian Li, Shuojue Yang, Daiyun Shen, Yueming Jin

Reconstructing endoscopic videos is crucial for high-fidelity visualization and the efficiency of surgical operations. Despite the importance, existing 3D reconstruction methods encounter several challenges, including st…

3D ReconstructionComputational Efficiency

Text-to-seed generation: Training-free open-vocabulary seeded semantic segmentation via re-purposing diffusion as text-guided seed generator

2026-08-27 · Kumju Jo, Heesun Jung, Sungyong Baik arxiv

Open-vocabulary semantic segmentation (OVSS) aims to segment image regions corresponding to arbitrary text queries. Although the Segment Anything Model (SAM) is a powerful foundation model for segmentation, its standalon…

Semantic Segmentation