paper-with-me

홈 › Papers

ASSIST-3D: Adapted Scene Synthesis for Class-Agnostic 3D Instance Segmentation

2025-12-10 · Shengchao Zhou, Jiehong Lin, Jiahui Liu, Shizhen Zhao, Chirui Chang, Xiaojuan Qi arxiv

Class-agnostic 3D instance segmentation tackles the challenging task of segmenting all object instances, including previously unseen ones, without semantic class reliance. Current methods struggle with generalization due to the scarce annotated 3D scene data or noisy 2D segmentations. While synthetic data generation offers a promising solution, existing 3D scene synthesis methods fail to simultaneously satisfy geometry diversity, context complexity, and layout reasonability, each essential for this task. To address these needs, we propose an Adapted 3D Scene Synthesis pipeline for class-agnostic 3D Instance SegmenTation, termed as ASSIST-3D, to synthesize proper data for model generalization enhancement. Specifically, ASSIST-3D features three key innovations, including 1) Heterogeneous Object Selection from extensive 3D CAD asset collections, incorporating randomness in object sampling to maximize geometric and contextual diversity; 2) Scene Layout Generation through LLM-guided spatial reasoning combined with depth-first search for reasonable object placements; and 3) Realistic Point Cloud Construction via multi-view RGB-D image rendering and fusion from the synthetic scenes, closely mimicking real-world sensor data acquisition. Experiments on ScanNetV2, ScanNet++, and S3DIS benchmarks demonstrate that models trained with ASSIST-3D-generated data significantly outperform existing methods. Further comparisons underscore the superiority of our purpose-built pipeline over existing 3D scene synthesis approaches.

📄 PDF Abstract BibTeX arXiv:2512.09364

Code (0)

등록된 구현이 없습니다.

Tasks

Synthetic Data Generation3D Instance SegmentationSpatial Reasoning

Similar Papers 제목 키워드 기반

From depth image to semantic scene synthesis through point cloud classification and labeling: Application to assistive systems

2020-08-09 · Chayma Zatout, Slimane Larabi

The aim of this work is to provide a semantic scene synthesis from depth image. First, depth image is segmented and each segment is classified in the context of assistive systems using a deep learning network. Second, in…

Point Cloud Classification

Painting 3D Nature in 2D: View Synthesis of Natural Scenes from a Single Semantic Mask

2023-02-14 · CVPR 2023 1 · Shangzhan Zhang, Sida Peng, Tianrun Chen, Linzhan Mou 외

We introduce a novel approach that takes a single semantic mask as input to synthesize multi-view consistent color images of natural scenes, trained with a collection of single images from the Internet. Prior works on 3D…

3D-Aware Image SynthesisImage Generation

MVT: Mask-Grounded Vision-Language Models for Taxonomy-Aligned Land-Cover Tagging

2025-09-23 · Siyi Chen, Kai Wang, Weicong Pang, Ruiming Yang 외 arxiv

Land-cover understanding in remote sensing increasingly demands class-agnostic systems that generalize across datasets while remaining spatially precise and interpretable. We study a geometry-first discovery-and-interpre…

Domain Adaptation

Observation-Assisted Heuristic Synthesis of Covert Attackers Against Unknown Supervisors

2021-03-20 · Liyong Lin, Ruochen Tai, Yuting Zhu, Rong Su

In this work, we address the problem of synthesis of covert attackers in the setup where the model of the plant is available, but the model of the supervisor is unknown, to the adversary. To compensate the lack of knowle…

DSCENet: Dynamic Screening and Clinical-Enhanced Multimodal Fusion for MPNs Subtype Classification

2024-07-11 · Yuan Zhang, Yaolei Qi, Xiaoming Qi, Yongyue Wei 외

The precise subtype classification of myeloproliferative neoplasms (MPNs) based on multimodal information, which assists clinicians in diagnosis and long-term treatment plans, is of great clinical significance. However, …

Diagnosticwhole slide images