paper-with-me

Papers

LiCamPose: Combining Multi-View LiDAR and RGB Cameras for Robust Single-frame 3D Human Pose Estimation

2023-12-11 · Zhiyu Pan, Zhicheng Zhong, Wenxuan Guo, Yifan Chen, Jianjiang Feng, Jie zhou

Several methods have been proposed to estimate 3D human pose from multi-view images, achieving satisfactory performance on public datasets collected under relatively simple conditions. However, there are limited approaches studying extracting 3D human skeletons from multimodal inputs, such as RGB and point cloud data. To address this gap, we introduce LiCamPose, a pipeline that integrates multi-view RGB and sparse point cloud information to estimate robust 3D human poses via single frame. We demonstrate the effectiveness of the volumetric architecture in combining these modalities. Furthermore, to circumvent the need for manually labeled 3D human pose annotations, we develop a synthetic dataset generator for pretraining and design an unsupervised domain adaptation strategy to train a 3D human pose estimator without manual annotations. To validate the generalization capability of our method, LiCamPose is evaluated on four datasets, including two public datasets, one synthetic dataset, and one challenging self-collected dataset named BasketBall, covering diverse scenarios. The results demonstrate that LiCamPose exhibits great generalization performance and significant application potential. The code, generator, and datasets will be made available upon acceptance of this paper.

📄 PDF Abstract BibTeX arXiv:2312.06409

Code (0)

등록된 구현이 없습니다.

Tasks

3D Human Pose EstimationDomain AdaptationPose EstimationUnsupervised Domain Adaptation

Similar Papers 제목 키워드 기반

Multi-LVI-SAM: A Robust LiDAR-Visual-Inertial Odometry for Multiple Fisheye Cameras

2025-09-06 · Xinyu Zhang, Kai Huang, Junqiao Zhao, Zihan Yuan 외 arxiv

We propose a multi-camera LiDAR-visual-inertial odometry framework, Multi-LVI-SAM, which fuses data from multiple fisheye cameras, LiDAR and inertial sensors for highly accurate and robust state estimation. To enable eff…

Pose Estimation

A Flexible Multi-view Multi-modal Imaging System for Outdoor Scenes

2023-02-21 · Meng Zhang, Wenxuan Guo, Bohao Fan, Yifan Chen 외

Multi-view imaging systems enable uniform coverage of 3D space and reduce the impact of occlusion, which is beneficial for 3D object detection and tracking accuracy. However, existing imaging systems built with multi-vie…

3D Object DetectionObjectobject-detectionObject Detection

LiDAR-as-Camera for End-to-End Driving

2022-06-30 · Ardi Tampuu, Romet Aidla, Jan Are van Gent, Tambet Matiisen

The core task of any autonomous driving system is to transform sensory inputs into driving commands. In end-to-end driving, this is achieved via a neural network, with one or multiple cameras as the most commonly used in…

Autonomous Driving

3D Multi-Object Tracking Employing MS-GLMB Filter for Autonomous Driving

2024-10-19 · Linh Van Ma, Muhammad Ishfaq Hussain, Kin-Choong Yow, Moongu Jeon

The MS-GLMB filter offers a robust framework for tracking multiple objects through the use of multi-sensor data. Building on this, the MV-GLMB and MV-GLMB-AB filters enhance the MS-GLMB capabilities by employing cameras …

3D Multi-Object TrackingAutonomous DrivingMulti-Object TrackingObject+1

BEV@DC: Bird's-Eye View Assisted Training for Depth Completion

2023-01-01 · CVPR 2023 1 · Wending Zhou, Xu Yan, Yinghong Liao, Yuankai Lin 외

Depth completion plays a crucial role in autonomous driving, in which cameras and LiDARs are two complementary sensors. Recent approaches attempt to exploit spatial geometric constraints hidden in LiDARs to enhance i…

Autonomous DrivingDepth Completion