paper-with-me

홈 › Papers

Splannequin: Freezing Monocular Mannequin-Challenge Footage with Dual-Detection Splatting

2025-12-04 · Hao-Jen Chien, Yi-Chuan Huang, Chung-Ho Wu, Wei-Lun Chao, Yu-Lun Liu arxiv

Synthesizing high-fidelity frozen 3D scenes from monocular Mannequin-Challenge (MC) videos is a unique problem distinct from standard dynamic scene reconstruction. Instead of focusing on modeling motion, our goal is to create a frozen scene while strategically preserving subtle dynamics to enable user-controlled instant selection. To achieve this, we introduce a novel application of dynamic Gaussian splatting: the scene is modeled dynamically, which retains nearby temporal variation, and a static scene is rendered by fixing the model's time parameter. However, under this usage, monocular capture with sparse temporal supervision introduces artifacts like ghosting and blur for Gaussians that become unobserved or occluded at weakly supervised timestamps. We propose Splannequin, an architecture-agnostic regularization that detects two states of Gaussian primitives, hidden and defective, and applies temporal anchoring. Under predominantly forward camera motion, hidden states are anchored to their recent well-observed past states, while defective states are anchored to future states with stronger supervision. Our method integrates into existing dynamic Gaussian pipelines via simple loss terms, requires no architectural changes, and adds zero inference overhead. This results in markedly improved visual quality, enabling high-fidelity, user-selectable frozen-time renderings, validated by a 96% user preference. Project page: https://chien90190.github.io/splannequin/

📄 PDF Abstract BibTeX arXiv:2512.05113

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Learning the Depths of Moving People by Watching Frozen People

2019-04-25 · CVPR 2019 6 · Zhengqi Li, Tali Dekel, Forrester Cole, Richard Tucker 외

We present a method for predicting dense depth in scenarios where both a monocular camera and people in the scene are freely moving. Existing methods for recovering depth for dynamic, non-rigid objects from monocular vid…

Depth EstimationDepth Prediction

From Mannequin to Human: A Pose-Aware and Identity-Preserving Video Generation Framework for Lifelike Clothing Display

2025-10-19 · Xiangyu Mu, Dongliang Zhou, Jie Hou, Haijun Zhang 외 arxiv

Mannequin-based clothing displays offer a cost-effective alternative to real-model showcases for online fashion presentation, but lack realism and expressive detail. To overcome this limitation, we introduce a new task c…

Video Generation

Personalized 3D mannequin reconstruction based on 3D scanning

2018-04-16 · Pengpeng Hu, Duan Li, Ge Wu, Taku Komura 외

Personalized customization is a new manufacturing trend in high-end products (e.g. senior custom clothing). Traditional apparel customization (made-to-measure & bespoken) highly depends on experienced tailors. A person…

End-to-end depth from motion with stabilized monocular videos

2018-09-12 · Clément Pinard, Laure Chevalley, Antoine Manzanera, David Filliat

We propose a depth map inference system from monocular videos based on a novel dataset for navigation that mimics aerial footage from gimbal stabilized monocular camera in rigid scenes. Unlike most navigation datasets, t…

Depth EstimationDepth Prediction

GeoFill: Reference-Based Image Inpainting with Better Geometric Understanding

2022-01-20 · Yunhan Zhao, Connelly Barnes, Yuqian Zhou, Eli Shechtman 외

Reference-guided image inpainting restores image pixels by leveraging the content from another single reference image. The primary challenge is how to precisely place the pixels from the reference image into the hole reg…

3D geometryDepth EstimationImage InpaintingMonocular Depth Estimation