paper-with-me

Papers

Interact3D: Compositional 3D Generation of Interactive Objects

2026-03-17 · Hui Shan, Keyang Luo, Ming Li, Sizhe Zheng, Yanwei Fu, Zhen Chen, Xiangru Huang arxiv

Recent breakthroughs in 3D generation have enabled the synthesis of high-fidelity individual assets. However, generating 3D compositional objects from single images--particularly under occlusions--remains challenging. Existing methods often degrade geometric details in hidden regions and fail to preserve the underlying object-object spatial relationships (OOR). We present a novel framework Interact3D designed to generate physically plausible interacting 3D compositional objects. Our approach first leverages advanced generative priors to curate high-quality individual assets with a unified 3D guidance scene. To physically compose these assets, we then introduce a robust two-stage composition pipeline. Based on the 3D guidance scene, the primary object is anchored through precise global-to-local geometric alignment (registration), while subsequent geometries are integrated using a differentiable Signed Distance Field (SDF)-based optimization that explicitly penalizes geometry intersections. To reduce challenging collisions, we further deploy a closed-loop, agentic refinement strategy. A Vision-Language Model (VLM) autonomously analyzes multi-view renderings of the composed scene, formulates targeted corrective prompts, and guides an image editing module to iteratively self-correct the generation pipeline. Extensive experiments demonstrate that Interact3D successfully produces promising collsion-aware compositions with improved geometric fidelity and consistent spatial relationships.

📄 PDF Abstract BibTeX arXiv:2603.16085

Code (0)

등록된 구현이 없습니다.

Tasks

3D GenerationImage Editing

Similar Papers 제목 키워드 기반

Inclusive Interactive Collisions for Multi-View Consistent Compositional 3D Generation

2026-06-23 · Chang Liu, Mingwen Shao, Xiang Lv, Xinyuan Chen 외 arxiv

Recent breakthroughs in 3D generation have advanced notably with the development of text-to-image diffusion model. However, existing methods remain two practical challenges: (1) They primarily generate single 3D object, …

Scene Generation3D Generation

HiScene: Creating Hierarchical 3D Scenes with Isometric View Generation

2025-04-17 · Wenqi Dong, Bangbang Yang, Zesong Yang, Yuan Li 외

Scene-level 3D generation represents a critical frontier in multimedia and computer graphics, yet existing approaches either suffer from limited object categories or lack editing flexibility for interactive applications.…

3D GenerationImage GenerationObject

ASSIST: Interactive Scene Nodes for Scalable and Realistic Indoor Simulation

2023-11-10 · Zhide Zhong, Jiakai Cao, Songen Gu, Sirui Xie 외

We present ASSIST, an object-wise neural radiance field as a panoptic representation for compositional and realistic simulation. Central to our approach is a novel scene node data structure that stores the information of…

Panoptic Segmentation

CP4D: Compositional Physics-aware 4D Scene Generation

2026-06-08 · Hanxin Zhu, Cong Wang, Tianyu He, Long Chen 외 arxiv

4D generation (\textit{i.e.}, dynamic 3D generation) has recently emerged as a rapidly growing research frontier due to its powerful spatiotemporal modeling capabilities. However, despite notable advances, existing appro…

Scene GenerationMotion Synthesis3D Generation

Unsupervised Skill Discovery for Robotic Manipulation through Automatic Task Generation

2024-10-07 · Paul Jansonnie, Bingbing Wu, Julien Perez, Jan Peters

Learning skills that interact with objects is of major importance for robotic manipulation. These skills can indeed serve as an efficient prior for solving various manipulation tasks. We propose a novel Skill Learning ap…

Hierarchical Reinforcement Learning