MaskFlow: Object-Aware Motion Estimation
We introduce a novel motion estimation method, MaskFlow, that is capable of estimating accurate motion fields, even in very challenging cases with small objects, large displacements and drastic appearance changes. In addition to lower-level features, that are used in other Deep Neural Network (DNN)-based motion estimation methods, MaskFlow draws from object-level features and segmentations. These features and segmentations are used to approximate the objects' translation motion field. We propose a novel and effective way of incorporating the incomplete translation motion field into a subsequent motion estimation network for refinement and completion. We also produced a new challenging synthetic dataset with motion field ground truth, and also provide extra ground truth for the object-instance matchings and corresponding segmentation masks. We demonstrate that MaskFlow outperforms state of the art methods when evaluated on our new challenging dataset, whilst still producing comparable results on the popular FlyingThings3D benchmark dataset.
Code (0)
등록된 구현이 없습니다.
Tasks
Motion EstimationObjectTranslationSimilar Papers 제목 키워드 기반
MaskFlownet: Asymmetric Feature Matching with Learnable Occlusion Mask
Feature warping is a core technique in optical flow estimation; however, the ambiguity caused by occluded areas during warping is a major problem that remains unsolved. In this paper, we propose an asymmetric occlusion-a…
Optical Flow EstimationMaskFlow: Discrete Flows For Flexible and Efficient Long Video Generation
Generating long, high-quality videos remains a challenge due to the complex interplay of spatial and temporal dynamics and hardware limitations. In this work, we introduce \textbf{MaskFlow}, a unified video generation fr…
Video GenerationAirDOS: Dynamic SLAM benefits from Articulated Objects
Dynamic Object-aware SLAM (DOS) exploits object-level information to enable robust motion estimation in dynamic environments. Existing methods mainly focus on identifying and excluding dynamic objects from the optimizati…
Camera Pose EstimationMotion EstimationObjectPose EstimationDO3D: Self-supervised Learning of Decomposed Object-aware 3D Motion and Depth from Monocular Videos
Although considerable advancements have been attained in self-supervised depth estimation from monocular videos, most existing methods often treat all objects in a video as static entities, which however violates the dyn…
Depth EstimationDisentanglementMotion DisentanglementMotion Estimation+3TARS: Traffic-Aware Radar Scene Flow Estimation
Scene flow provides crucial motion information for autonomous driving. Recent LiDAR scene flow models utilize the rigid-motion assumption at the instance level, assuming objects are rigid bodies. However, these instance-…
Autonomous Drivingobject-detectionObject DetectionScene Flow Estimation+1