paper-with-me

Papers

Discriminately Treating Motion Components Evolves Joint Depth and Ego-Motion Learning

2025-11-03 · Mengtan Zhang, Zizhan Guo, Hongbo Zhao, Yi Feng, Zuyi Xiong, Yue Wang, Shaoyi Du, Hanli Wang, Rui Fan arxiv

Unsupervised learning of depth and ego-motion, two fundamental 3D perception tasks, has made significant strides in recent years. However, most methods treat ego-motion as an auxiliary task, either mixing all motion types or excluding depth-independent rotational motions in supervision. Such designs limit the incorporation of strong geometric constraints, reducing reliability and robustness under diverse conditions. This study introduces a discriminative treatment of motion components, leveraging the geometric regularities of their respective rigid flows to benefit both depth and ego-motion estimation. Given consecutive video frames, network outputs first align the optical axes and imaging planes of the source and target cameras. Optical flows between frames are transformed through these alignments, and deviations are quantified to impose geometric constraints individually on each ego-motion component, enabling more targeted refinement. These alignments further reformulate the joint learning process into coaxial and coplanar forms, where depth and each translation component can be mutually derived through closed-form geometric relationships, introducing complementary constraints that improve depth robustness. DiMoDE, a general depth and ego-motion joint learning framework incorporating these designs, achieves state-of-the-art performance on multiple public datasets and a newly collected diverse real-world dataset, particularly under challenging conditions. Our source code will be publicly available at mias.group/DiMoDE upon publication.

📄 PDF Abstract BibTeX arXiv:2511.01502

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

RefineVAD: Semantic-Guided Feature Recalibration for Weakly Supervised Video Anomaly Detection

2025-11-17 · Junhee Lee, ChaeBeen Bang, MyoungChul Kim, MyeongAh Cho arxiv

Weakly-Supervised Video Anomaly Detection aims to identify anomalous events using only video-level labels, balancing annotation efficiency with practical applicability. However, existing methods often oversimplify the an…

Weakly-supervised Video Anomaly Detection

Improving Human Motion Plausibility with Body Momentum

2025-09-11 · Ha Linh Nguyen, Tze Ho Elden Tse, Angela Yao arxiv

Many studies decompose human motion into local motion in a frame attached to the root joint and global motion of the root joint in the world frame, treating them separately. However, these two components are not independ…

Recognize Actions by Disentangling Components of Dynamics

2018-06-01 · CVPR 2018 6 · Yue Zhao, Yuanjun Xiong, Dahua Lin

Despite the remarkable progress in action recognition over the past several years, existing methods remain limited in efficiency and effectiveness. The methods treating appearance and motion as separate streams are usual…

Action RecognitionOptical Flow EstimationRepresentation LearningTemporal Action Localization

SDD-4DGS: Static-Dynamic Aware Decoupling in Gaussian Splatting for 4D Scene Reconstruction

2025-03-12 · Dai Sun, Huhao Guan, Kun Zhang, Xike Xie 외

Dynamic and static components in scenes often exhibit distinct properties, yet most 4D reconstruction methods treat them indiscriminately, leading to suboptimal performance in both cases. This work introduces SDD-4DGS, t…

4D reconstruction

Disjoint principal component analysis by constrained binary particle swarm optimization

2020-04-22 · John Ramírez-Figueroa, Carlos Martín-Barreiro, Ana B. Nieto-Librero, Victor Leiva-Sánchez 외

In this paper, we propose an alternative method to the disjoint principal component analysis. The method consists of a principal component analysis with constraints, which allows us to determine disjoint components that …

Stochastic Optimization