paper-with-me

Papers

Deep Non-rigid Structure-from-Motion: A Sequence-to-Sequence Translation Perspective

2022-04-10 · Hui Deng, Tong Zhang, Yuchao Dai, Jiawei Shi, Yiran Zhong, Hongdong Li

Directly regressing the non-rigid shape and camera pose from the individual 2D frame is ill-suited to the Non-Rigid Structure-from-Motion (NRSfM) problem. This frame-by-frame 3D reconstruction pipeline overlooks the inherent spatial-temporal nature of NRSfM, i.e., reconstructing the whole 3D sequence from the input 2D sequence. In this paper, we propose to model deep NRSfM from a sequence-to-sequence translation perspective, where the input 2D frame sequence is taken as a whole to reconstruct the deforming 3D non-rigid shape sequence. First, we apply a shape-motion predictor to estimate the initial non-rigid shape and camera motion from a single frame. Then we propose a context modeling module to model camera motions and complex non-rigid shapes. To tackle the difficulty in enforcing the global structure constraint within the deep framework, we propose to impose the union-of-subspace structure by replacing the self-expressiveness layer with multi-head attention and delayed regularizers, which enables end-to-end batch-wise training. Experimental results across different datasets such as Human3.6M, CMU Mocap and InterHand prove the superiority of our framework.

📄 PDF Abstract BibTeX arXiv:2204.04730

Code (0)

등록된 구현이 없습니다.

Tasks

3D ReconstructionTranslation

Methods 이 논문이 사용한 방법론

Softmax The Softmax output function transforms a previous layer's output into a vector of probabilities. It is commonly used for multiclass classification. Given an input vector $x$…
Linear Layer A Linear Layer is a projection $\mathbf{XW + b}$.

Similar Papers 제목 키워드 기반

Nonrigid Structure from Motion in Trajectory Space

2008-12-01 · NeurIPS 2008 12 · Ijaz Akhter, Yaser Sheikh, Sohaib Khan, Takeo Kanade

Existing approaches to nonrigid structure from motion assume that the instantaneous 3D shape of a deforming object is a linear combination of basis shapes, which have to be estimated anew for each video sequence. In cont…

Deep Non-rigid Structure-from-Motion Revisited: Canonicalization and Sequence Modeling

2024-12-10 · Hui Deng, Jiawei Shi, Zhen Qin, Yiran Zhong 외

Non-Rigid Structure-from-Motion (NRSfM) is a classic 3D vision problem, where a 2D sequence is taken as input to estimate the corresponding 3D sequence. Recently, the deep neural networks have greatly advanced the task o…

Structure from Recurrent Motion: From Rigidity to Recurrency

2018-04-18 · CVPR 2018 6 · Xiu Li, Hongdong Li, Hanbyul Joo, Yebin Liu 외

This paper proposes a new method for Non-Rigid Structure-from-Motion (NRSfM) from a long monocular video sequence observing a non-rigid object performing recurrent and possibly repetitive dynamic action. Departing from t…

Clustering

SfM-Net: Learning of Structure and Motion from Video

2017-04-25 · Sudheendra Vijayanarasimhan, Susanna Ricco, Cordelia Schmid, Rahul Sukthankar 외

We propose SfM-Net, a geometry-aware neural network for motion estimation in videos that decomposes frame-to-frame pixel motion in terms of scene and object depth, camera motion and 3D object rotations and translations. …

Motion EstimationObjectOptical Flow Estimation

Injecting Multimodal Information into Rigid Protein Docking via Bi-level Optimization

2023-09-21 · NeurIPS 2023 11

The structure of protein-protein complexes is critical for understanding binding dynamics, biological mechanisms, and intervention strategies. Rigid protein docking, a fundamental problem in this field, aims to predict t…