paper-with-me

홈 › Papers

Graph Canvas for Controllable 3D Scene Generation

2024-11-27 · Libin Liu, Shen Chen, Sen Jia, Jingzhe Shi, Zhongyu Jiang, Can Jin, Wu Zongkai, Jenq-Neng Hwang, Lei LI

Spatial intelligence is foundational to AI systems that interact with the physical world, particularly in 3D scene generation and spatial comprehension. Current methodologies for 3D scene generation often rely heavily on predefined datasets, and struggle to adapt dynamically to changing spatial relationships. In this paper, we introduce GraphCanvas3D, a programmable, extensible, and adaptable framework for controllable 3D scene generation. Leveraging in-context learning, GraphCanvas3D enables dynamic adaptability without the need for retraining, supporting flexible and customizable scene creation. Our framework employs hierarchical, graph-driven scene descriptions, representing spatial elements as graph nodes and establishing coherent relationships among objects in 3D environments. Unlike conventional approaches, which are constrained in adaptability and often require predefined input masks or retraining for modifications, GraphCanvas3D allows for seamless object manipulation and scene adjustments on the fly. Additionally, GraphCanvas3D supports 4D scene generation, incorporating temporal dynamics to model changes over time. Experimental results and user studies demonstrate that GraphCanvas3D enhances usability, flexibility, and adaptability for scene generation. Our code and models are available on the project website: https://github.com/ILGLJ/Graph-Canvas.

📄 PDF Abstract BibTeX arXiv:2412.00091

Code (0)

등록된 구현이 없습니다.

Tasks

In-Context LearningScene Generation

Similar Papers 제목 키워드 기반

MotionCanvas: Cinematic Shot Design with Controllable Image-to-Video Generation

2025-02-06 · Jinbo Xing, Long Mai, Cusuh Ham, Jiahui Huang 외

This paper presents a method that allows users to design cinematic video shots in the context of image-to-video generation. Shot design, a critical aspect of filmmaking, involves meticulously planning both camera movemen…

Image to Video GenerationVideo EditingVideo Generation

The World is Your Canvas: Painting Promptable Events with Reference Images, Trajectories, and Text

2025-12-18 · Hanlin Wang, Hao Ouyang, Qiuyu Wang, Yue Yu 외 arxiv

We present WorldCanvas, a framework for promptable world events that enables rich, user-directed simulation by combining text, trajectories, and reference images. Unlike text-only approaches and existing trajectory-contr…

Visual Grounding

VideoCanvas: Unified Video Completion from Arbitrary Spatiotemporal Patches via In-Context Conditioning

2025-10-09 · Minghong Cai, Qiulin Wang, Zongli Ye, Wenze Liu 외 arxiv

Existing controllable video generation methods are typically designed for rigid, task-specific settings, such as first-frame image-to-video, inpainting, or interpolation, treating spatio-temporal control as a set of isol…

Video Generation

CANVAS: Continuity-Aware Narratives via Visual Agentic Storyboarding

2026-04-15 · Ishani Mondal, Yiwen Song, Mihir Parmar, Palash Goyal 외 arxiv

Long-form visual storytelling requires maintaining continuity across shots, including consistent characters, stable environments, and smooth scene transitions. While existing generative models can produce strong individu…

Visual Storytelling

CogCanvas: A Benchmark for Evaluating Multi-Subject Reference-Based Image Generation

2026-06-14 · Long-Bao Nguyen, Quang-Khai Tran, Tam V. Nguyen, Minh-Triet Tran 외 arxiv

Multi-subject reference-based image generation requires jointly preserving multiple human identities, binding per-person objects and fashion items, and respecting a specified background scene, a regime where current diff…

Image Generation