paper-with-me

홈 › Papers

Prompt-guided Scene Generation for 3D Zero-Shot Learning

2022-09-29 · Majid Nasiri, Ali Cheraghian, Townim Faisal Chowdhury, Sahar Ahmadi, Morteza Saberi, Shafin Rahman

Zero-shot learning on 3D point cloud data is a related underexplored problem compared to its 2D image counterpart. 3D data brings new challenges for ZSL due to the unavailability of robust pre-trained feature extraction models. To address this problem, we propose a prompt-guided 3D scene generation and supervision method that augments 3D data to learn the network better, exploring the complex interplay of seen and unseen objects. First, we merge point clouds of two 3D models in certain ways described by a prompt. The prompt acts like the annotation describing each 3D scene. Later, we perform contrastive learning to train our proposed architecture in an end-to-end manner. We argue that 3D scenes can relate objects more efficiently than single objects because popular language models (like BERT) can achieve high performance when objects appear in a context. Our proposed prompt-guided scene generation method encapsulates data augmentation and prompt-based annotation/captioning to improve 3D ZSL performance. We have achieved state-of-the-art ZSL and generalized ZSL performance on synthetic (ModelNet40, ModelNet10) and real-scanned (ScanOjbectNN) 3D object datasets.

📄 PDF Abstract BibTeX arXiv:2209.14690

Code (0)

등록된 구현이 없습니다.

Tasks

Contrastive LearningData AugmentationScene GenerationZero-Shot Learning

Methods 이 논문이 사용한 방법론

Contrastive Learning 설명 없음

Similar Papers 제목 키워드 기반

GenZI: Zero-Shot 3D Human-Scene Interaction Generation

2023-11-29 · CVPR 2024 1 · Lei LI, Angela Dai

Can we synthesize 3D humans interacting with scenes without learning from any 3D human-scene interaction data? We propose GenZI, the first zero-shot approach to generating 3D human-scene interactions. Key to GenZI is our…

ZeroScene: A Zero-Shot Framework for 3D Scene Generation from a Single Image and Controllable Texture Editing

2025-09-28 · Xiang Tang, Ruotong Li, Xiaopeng Fan arxiv

In the field of 3D content generation, single image scene reconstruction methods still struggle to simultaneously ensure the quality of individual assets and the coherence of the overall scene in complex environments, wh…

Scene GenerationImage Generation

Zero-Shot Learning in Industrial Scenarios: New Large-Scale Benchmark, Challenges and Baseline

2026-06-06 · Zekai Zhang, Qinghui Chen, Maomao Xiong, Shijiao Ding 외 arxiv

Large Visual Language Models (LVLMs) have achieved remarkable success in vision tasks. However, the significant differences between industrial and natural scenes make applying LVLMs challenging. Existing LVLMs rely on us…

Zero-Shot LearningDomain Adaptation

AnchoredDream: Zero-Shot 360° Indoor Scene Generation from a Single View via Geometric Grounding

2026-01-23 · Runmao Yao, Junsheng Zhou, Zhen Dong, Yu-Shen Liu arxiv

Single-view indoor scene generation plays a crucial role in a range of real-world applications. However, generating a complete 360° scene from a single image remains a highly ill-posed and challenging problem. Recent app…

Depth EstimationScene Generation

ZeroHSI: Zero-Shot 4D Human-Scene Interaction by Video Generation

2024-12-24 · Hongjie Li, Hong-Xing Yu, Jiaman Li, Jiajun Wu

Human-scene interaction (HSI) generation is crucial for applications in embodied AI, virtual reality, and robotics. Yet, existing methods cannot synthesize interactions in unseen environments such as in-the-wild scenes o…

Human-Object Interaction DetectionVideo Generation