paper-with-me

홈 › Papers

RelaxFlow: Text-Driven Amodal 3D Generation

2026-03-05 · Jiayin Zhu, Guoji Fu, Xiaolu Liu, Qiyuan He, Yicong Li, Angela Yao arxiv

Image-to-3D generation faces inherent semantic ambiguity under occlusion, where partial observation alone is often insufficient to determine object category. In this work, we formalize text-driven amodal 3D generation, where text prompts steer the completion of unseen regions while strictly preserving input observation. Crucially, we identify that these objectives demand distinct control granularities: rigid control for the observation versus relaxed structural control for the prompt. To this end, we propose RelaxFlow, a training-free dual-branch framework that decouples control granularity via a Multi-Prior Consensus Module and a Relaxation Mechanism. Theoretically, we prove that our relaxation is equivalent to applying a low-pass filter on the generative vector field, which suppresses high-frequency instance details to isolate geometric structure that accommodates the observation. To facilitate evaluation, we introduce two diagnostic benchmarks, ExtremeOcc-3D and AmbiSem-3D. Extensive experiments demonstrate that RelaxFlow successfully steers the generation of unseen regions to match the prompt intent without compromising visual fidelity.

📄 PDF Abstract BibTeX arXiv:2603.05425

Code (0)

등록된 구현이 없습니다.

Tasks

3D Generation

Similar Papers 제목 키워드 기반

Unveiling the Invisible: Reasoning Complex Occlusions Amodally with AURA

2025-03-13 · Zhixuan Li, Hyunse Yoon, SangHoon Lee, Weisi Lin

Amodal segmentation aims to infer the complete shape of occluded objects, even when the occluded region's appearance is unavailable. However, current amodal segmentation methods lack the capability to interact with users…

Dataset GenerationReasoning SegmentationSegmentation

Amodal Cityscapes: A New Dataset, its Generation, and an Amodal Semantic Segmentation Challenge Baseline

2022-06-01 · Jasmin Breitenstein, Tim Fingscheidt

Amodal perception terms the ability of humans to imagine the entire shapes of occluded objects. This gives humans an advantage to keep track of everything that is going on, especially in crowded situations. Typical perce…

SegmentationSemantic Segmentation

SynTable: A Synthetic Data Generation Pipeline for Unseen Object Amodal Instance Segmentation of Cluttered Tabletop Scenes

2023-07-14 · Zhili Ng, Haozhe Wang, Zhengshen Zhang, Francis Tay Eng Hock 외

In this work, we present SynTable, a unified and flexible Python-based dataset generator built using NVIDIA's Isaac Sim Replicator Composer for generating high-quality synthetic datasets for unseen object amodal instance…

Amodal Instance SegmentationDataset GenerationInstance SegmentationSemantic Segmentation+1

AmodalSVG: Amodal Image Vectorization via Semantic Layer Peeling

2026-04-13 · Juncheng Hu, Ziteng Xue, Guotao Liang, Anran Qi 외 arxiv

We introduce AmodalSVG, a new framework for amodal image vectorization that produces semantically organized and geometrically complete SVG representations from natural images. Existing vectorization methods operate under…

TAO-Amodal: A Benchmark for Tracking Any Object Amodally

2023-12-19 · Cheng-Yen Hsieh, Kaihua Chen, Achal Dave, Tarasha Khurana 외

Amodal perception, the ability to comprehend complete object structures from partial visibility, is a fundamental skill, even for infants. Its significance extends to applications like autonomous driving, where a clear u…

Amodal TrackingAutonomous DrivingBenchmarkingData Augmentation+1