paper-with-me

홈 › Papers

MoRe: Motion-aware Feed-forward 4D Reconstruction Transformer

2026-03-05 · Juntong Fang, Zequn Chen, Weiqi Zhang, Donglin Di, Xuancheng Zhang, Chengmin Yang, Yu-Shen Liu arxiv

Reconstructing dynamic 4D scenes remains challenging due to the presence of moving objects that corrupt camera pose estimation. Existing optimization methods alleviate this issue with additional supervision, but they are mostly computationally expensive and impractical in real-time applications. To address these limitations, we propose MoRe, a feedforward 4D reconstruction network that efficiently recovers dynamic 3D scenes from monocular videos. Built upon a strong static reconstruction backbone, MoRe employs an attention-forcing strategy to disentangle dynamic motion from static structure. To further enhance robustness, we fine-tune the model on large-scale, diverse datasets encompassing both dynamic and static scenes. Moreover, our grouped causal attention captures temporal dependencies and adapts to varying token lengths across frames, ensuring temporally coherent geometry reconstruction. Extensive experiments on multiple benchmarks demonstrate that MoRe achieves high-quality dynamic reconstructions with exceptional efficiency.

📄 PDF Abstract BibTeX arXiv:2603.05078

Code (0)

등록된 구현이 없습니다.

Tasks

Camera Pose Estimation

Similar Papers 제목 키워드 기반

DynamicVGGT: Learning Dynamic Point Maps for 4D Scene Reconstruction in Autonomous Driving

2026-03-09 · Zhuolin He, Jing Li, Guanghao Li, Xiaolei Chen 외 arxiv

Dynamic scene reconstruction in autonomous driving remains a fundamental challenge due to significant temporal variations, moving objects, and complex scene dynamics. Existing feed-forward 3D models have demonstrated str…

Autonomous Driving

StreetForward: Perceiving Dynamic Street with Feedforward Causal Attention

2026-03-20 · Zhongrui Yu, Zhao Wang, Yijia Xie, Yida Wang 외 arxiv

Feedforward reconstruction is crucial for autonomous driving applications, where rapid scene reconstruction enables efficient utilization of large-scale driving datasets in closed-loop simulation and other downstream tas…

Novel View SynthesisAutonomous DrivingDepth Estimation

Giving Faces Their Feelings Back: Explicit Emotion Control for Feedforward Single-Image 3D Head Avatars

2026-04-16 · Yicheng Gong, Jiawei Zhang, Liqiang Liu, Yanwen Wang 외 arxiv

We present a framework for explicit emotion control in feed-forward, single-image 3D head avatar reconstruction. Unlike existing pipelines where emotion is implicitly entangled with geometry or appearance, we treat emoti…

NemoSplat: Feed-Forward 4D Gaussian Splatting for Media-Aware Underwater Reconstruction

2026-08-24 · Xiaopeng Guo, Wai Chung Tse, Yipeng Zhu, Hanwen Zhang 외 arxiv

Reconstructing photorealistic scenes in unconstrained underwater environments remains challenging due to severe media-induced light scattering and unpredictable dynamic objects. Recent feed-forward visual foundation mode…

Dynamic ReconstructionNovel View Synthesis

Feed-Forward Bullet-Time Reconstruction of Dynamic Scenes from Monocular Videos

2024-12-04 · Hanxue Liang, Jiawei Ren, Ashkan Mirzaei, Antonio Torralba 외

Recent advancements in static feed-forward scene reconstruction have demonstrated significant progress in high-quality novel view synthesis. However, these models often struggle with generalizability across diverse envir…

Novel View Synthesis