paper-with-me

홈 › Papers

Gaussian See, Gaussian Do: Semantic 3D Motion Transfer from Multiview Video

2025-11-18 · Yarin Bekor, Gal Michael Harari, Or Perel, Or Litany arxiv

We present Gaussian See, Gaussian Do, a novel approach for semantic 3D motion transfer from multiview video. Our method enables rig-free, cross-category motion transfer between objects with semantically meaningful correspondence. Building on implicit motion transfer techniques, we extract motion embeddings from source videos via condition inversion, apply them to rendered frames of static target shapes, and use the resulting videos to supervise dynamic 3D Gaussian Splatting reconstruction. Our approach introduces an anchor-based view-aware motion embedding mechanism, ensuring cross-view consistency and accelerating convergence, along with a robust 4D reconstruction pipeline that consolidates noisy supervision videos. We establish the first benchmark for semantic 3D motion transfer and demonstrate superior motion fidelity and structural consistency compared to adapted baselines. Code and data for this paper available at https://gsgd-motiontransfer.github.io/

📄 PDF Abstract BibTeX arXiv:2511.14848

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

UfM*: Uncertainty from Motion* for DNN Depth Estimation Using Gaussians

2026-05-21 · Soumya Sudhakar, Sertac Karaman, Vivienne Sze arxiv

Reliable uncertainty estimation is critical for deploying monocular depth deep neural networks (DNNs) in safety-critical robotic systems. Conventional uncertainty methods such as ensembles and sampling-based approaches r…

Depth Estimation

Multiview Geometric Regularization of Gaussian Splatting for Accurate Radiance Fields

2025-06-16 · Jungeon Kim, Geonsoo Park, Seungyong Lee

Recent methods, such as 2D Gaussian Splatting and Gaussian Opacity Fields, have aimed to address the geometric inaccuracies of 3D Gaussian Splatting while retaining its superior rendering quality. However, these approach…

PLA4D: Pixel-Level Alignments for Text-to-4D Gaussian Splatting

2024-05-30 · Qiaowei Miao, JinSheng Quan, Kehan Li, Yawei Luo

Previous text-to-4D methods have leveraged multiple Score Distillation Sampling (SDS) techniques, combining motion priors from video-based diffusion models (DMs) with geometric priors from multiview DMs to implicitly gui…

3D GenerationContrastive LearningText to 3D

World from Motion: Generative Dynamic Gaussian Reconstruction from Monocular Video

2026-07-01 · Liyuan Zhu, Shengyu Huang, Amrita Mazumdar, Tianye Li 외 arxiv

We present World from Motion, a method for generating freely renderable dynamic 3D Gaussian representations from monocular videos. Our approach conditions a video model on dense, pixel-aligned renderings that encode appe…

GAGS: Granularity-Aware Feature Distillation for Language Gaussian Splatting

2024-12-18 · Yuning Peng, Haiping Wang, YuAn Liu, Chenglu Wen 외

3D open-vocabulary scene understanding, which accurately perceives complex semantic properties of objects in space, has gained significant attention in recent years. In this paper, we propose GAGS, a framework that disti…

Scene UnderstandingSemantic SegmentationVisual Grounding