paper-with-me

Papers

Controllable Dynamic 3D Shape Generation via 3D Trajectories and Text

2026-06-03 · Jaeyeong Kim, Ines Kim, Jahyeok Koo, Seungryong Kim arxiv

We introduce T2Mo, a feed-forward framework for controllable dynamic 3D shape generation conditioned on 3D trajectories and text. Due to the inherent ambiguity of language, generating precisely intended motions using text alone remains challenging. To address this, we adopt 3D trajectories as controllable spatial guidance, specifying the exact paths along which selected points should move. By combining both, T2Mo generates object motions that spatially adhere to the given trajectories while globally reflecting the text semantics. To robustly handle trajectory inputs with arbitrary configurations, ranging from dense to sparse and unevenly distributed, we further propose a shape-grounded trajectory embedding that maps an input trajectory set into a shape-aware token set covering the entire object. We conduct extensive comparisons against text-based baselines and cascaded video-based baselines that combine trajectory-guided video generation with video-to-dynamic mesh generation. Quantitative and qualitative evaluations, along with user studies, demonstrate that our approach produces motions that more faithfully follow the given prompts with higher expressiveness while preserving motion quality.

📄 PDF Abstract BibTeX arXiv:2606.05162

Code (0)

등록된 구현이 없습니다.

Tasks

Video Generation

Similar Papers 제목 키워드 기반

IM-Zero: Instance-level Motion Controllable Video Generation in a Zero-shot Manner

2025-01-01 · CVPR 2025 1 · YuYang Huang, Yabo Chen, Li Ding, Xiaopeng Zhang 외

Controllability of video generation has been recently concerned in addition to the quality of generated videos. The main challenge to controllable video generation is to synthesize videos based on user-specified inst…

Motion GenerationText-to-Video GenerationVideo Generation

DiffLOB: Diffusion Models for Counterfactual Generation in Limit Order Books

2026-02-03 · Zhuohan Wang, Carmine Ventre arxiv

Modern generative models for limit order books (LOBs) can reproduce realistic market dynamics, but remain fundamentally passive: they either model what typically happens without accounting for hypothetical future market …

The World is Your Canvas: Painting Promptable Events with Reference Images, Trajectories, and Text

2025-12-18 · Hanlin Wang, Hao Ouyang, Qiuyu Wang, Yue Yu 외 arxiv

We present WorldCanvas, a framework for promptable world events that enables rich, user-directed simulation by combining text, trajectories, and reference images. Unlike text-only approaches and existing trajectory-contr…

Visual Grounding

VHOI: Controllable Video Generation of Human-Object Interactions from Sparse Trajectories via Motion Densification

2025-12-10 · Wanyue Zhang, Lin Geng Foo, Thabo Beeler, Rishabh Dabral 외 arxiv

Synthesizing realistic human-object interactions (HOI) in video is challenging due to the complex, instance-specific interaction dynamics of both humans and objects. Incorporating controllability in video generation furt…

Video Generation

Prim2Room: Layout-Controllable Room Mesh Generation from Primitives

2024-09-09 · Chengzeng Feng, Jiacheng Wei, Cheng Chen, Yang Li 외

We propose Prim2Room, a novel framework for controllable room mesh generation leveraging 2D layout conditions and 3D primitive retrieval to facilitate precise 3D layout specification. Diverging from existing methods that…

DiversityRetrieval