paper-with-me

홈 › Papers

RPEFlow: Multimodal Fusion of RGB-PointCloud-Event for Joint Optical Flow and Scene Flow Estimation

2023-09-26 · ICCV 2023 1 · Zhexiong Wan, Yuxin Mao, Jing Zhang, Yuchao Dai

Recently, the RGB images and point clouds fusion methods have been proposed to jointly estimate 2D optical flow and 3D scene flow. However, as both conventional RGB cameras and LiDAR sensors adopt a frame-based data acquisition mechanism, their performance is limited by the fixed low sampling rates, especially in highly-dynamic scenes. By contrast, the event camera can asynchronously capture the intensity changes with a very high temporal resolution, providing complementary dynamic information of the observed scenes. In this paper, we incorporate RGB images, Point clouds and Events for joint optical flow and scene flow estimation with our proposed multi-stage multimodal fusion model, RPEFlow. First, we present an attention fusion module with a cross-attention mechanism to implicitly explore the internal cross-modal correlation for 2D and 3D branches, respectively. Second, we introduce a mutual information regularization term to explicitly model the complementary information of three modalities for effective multimodal feature learning. We also contribute a new synthetic dataset to advocate further research. Experiments on both synthetic and real datasets show that our model outperforms the existing state-of-the-art by a wide margin. Code and dataset is available at https://npucvr.github.io/RPEFlow.

📄 PDF Abstract BibTeX arXiv:2309.15082

Code (1)

danqu130/RPEFlow 공식 구현 pytorch

Tasks

Optical Flow EstimationScene Flow Estimation

Similar Papers 제목 키워드 기반

Simulation-to-Reality domain adaptation for offline 3D object annotation on pointclouds with correlation alignment

2022-02-06 · Weishuang Zhang, B Ravi Kiran, Thomas Gauthier, Yanis Mazouz 외

Annotating objects with 3D bounding boxes in LiDAR pointclouds is a costly human driven process in an autonomous driving perception system. In this paper, we present a method to semi-automatically annotate real-world poi…

Autonomous DrivingDomain AdaptationObjectobject-detection+1

Faraway-Frustum: Dealing with Lidar Sparsity for 3D Object Detection using Fusion

2020-11-03 · Haolin Zhang, Dongfang Yang, Ekim Yurtsever, Keith A. Redmill 외

Learned pointcloud representations do not generalize well with an increase in distance to the sensor. For example, at a range greater than 60 meters, the sparsity of lidar pointclouds reaches to a point where even humans…

3D Object DetectionObjectobject-detectionObject Detection+1

Learning task-specific features for 3D pointcloud graph creation

2022-09-02 · Elías Abad-Rocamora, Javier Ruiz-Hidalgo

Processing 3D pointclouds with Deep Learning methods is not an easy task. A common choice is to do so with Graph Neural Networks, but this framework involves the creation of edges between points, which are explicitly not…

World Reconstruction From Inconsistent Views

2026-03-17 · Lukas Höllein, Matthias Nießner arxiv

Video diffusion models generate high-quality and diverse worlds; however, individual frames often lack 3D consistency across the output sequence, which makes the reconstruction of 3D worlds difficult. To this end, we pro…

3D Reconstruction

DeSPITE: Exploring Contrastive Deep Skeleton-Pointcloud-IMU-Text Embeddings for Advanced Point Cloud Human Activity Understanding

2025-06-16 · Thomas Kreutz, Max Mühlhäuser, Alejandro Sanchez Guinea

Despite LiDAR (Light Detection and Ranging) being an effective privacy-preserving alternative to RGB cameras to perceive human activities, it remains largely underexplored in the context of multi-modal contrastive pre-tr…

Activity RecognitionHuman Activity RecognitionMoment RetrievalPerson Re-Identification+2