paper-with-me

홈 › Papers

TIMI: Training-Free Image-to-3D Multi-Instance Generation with Spatial Fidelity

2026-03-02 · Xiao Cai, Pengpeng Zeng, Ji Zhang, Heng Tao Shen, Jingkuan Song, Lianli Gao arxiv

Precise spatial fidelity in Image-to-3D multi-instance generation is critical for downstream real-world applications. Recent work attempts to address this by fine-tuning pre-trained Image-to-3D (I23D) models on multi-instance datasets, which incurs substantial training overhead and struggles to guarantee spatial fidelity. In fact, we observe that pre-trained I23D models already possess meaningful spatial priors, which remain underutilized as evidenced by instance entanglement issues. Motivated by this, we propose TIMI, a novel Training-free framework for Image-to-3D Multi-Instance generation that achieves high spatial fidelity. Specifically, we first introduce an Instance-aware Separation Guidance (ISG) module, which facilitates instance disentanglement during the early denoising stage. Next, to stabilize the guidance introduced by ISG, we devise a Spatial-stabilized Geometry-adaptive Update (SGU) module that promotes the preservation of the geometric characteristics of instances while maintaining their relative relationships. Extensive experiments demonstrate that our method yields better performance in terms of both global layout and distinct local instances compared to existing multi-instance methods, without requiring additional training and with faster inference speed.

📄 PDF Abstract BibTeX arXiv:2603.01371

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Regression-free Blind Image Quality Assessment with Content-Distortion Consistency

2023-07-18 · Xiaoqi Wang, Jian Xiong, Hao Gao, Weisi Lin

The optimization objective of regression-based blind image quality assessment (IQA) models is to minimize the mean prediction error across the training dataset, which can lead to biased parameter estimation due to potent…

Image Quality AssessmentNo-Reference Image Quality Assessmentparameter estimationregression+2

Label-Free Synthetic Pretraining of Object Detectors

2022-08-08 · Hei Law, Jia Deng

We propose a new approach, Synthetic Optimized Layout with Instance Detection (SOLID), to pretrain object detectors with synthetic images. Our "SOLID" approach consists of two main components: (1) generating synthetic im…

Object

Training-Free Instance-Aware 3D Scene Reconstruction and Diffusion-Based View Synthesis from Sparse Images

2026-03-22 · Jiatong Xia, Lingqiao Liu arxiv

We introduce a novel, training-free system for reconstructing, understanding, and rendering 3D indoor scenes from a sparse set of unposed RGB images. Unlike traditional radiance field approaches that require dense views …

An Instance-Aware Prompting Framework for Training-free Camouflaged Object Segmentation

2025-08-09 · Chao Yin, Jide Li, Hang Yao, Xiaoqiang Li arxiv

Training-free Camouflaged Object Segmentation (COS) seeks to segment camouflaged objects without task-specific training, by automatically generating visual prompts to guide the Segment Anything Model (SAM). However, exis…

Camouflaged Object Segmentation

Learning with Free Object Segments for Long-Tailed Instance Segmentation

2022-02-22 · Cheng Zhang, Tai-Yu Pan, Tianle Chen, Jike Zhong 외

One fundamental challenge in building an instance segmentation model for a large number of classes in complex scenes is the lack of training examples, especially for rare objects. In this paper, we explore the possibilit…

Instance SegmentationObjectSemantic Segmentation