paper-with-me

홈 › Papers

From Camera to World: A Plug-and-Play Module for Human Mesh Transformation

2025-12-17 · Changhai Ma, Ziyu Wu, Yunkang Zhang, Qijun Ying, Boyan Liu, Xiaohui Cai arxiv

Reconstructing accurate 3D human meshes in the world coordinate system from in-the-wild images remains challenging due to the lack of camera rotation information. While existing methods achieve promising results in the camera coordinate system by assuming zero camera rotation, this simplification leads to significant errors when transforming the reconstructed mesh to the world coordinate system. To address this challenge, we propose Mesh-Plug, a plug-and-play module that accurately transforms human meshes from camera coordinates to world coordinates. Our key innovation lies in a human-centered approach that leverages both RGB images and depth maps rendered from the initial mesh to estimate camera rotation parameters, eliminating the dependency on environmental cues. Specifically, we first train a camera rotation prediction module that focuses on the human body's spatial configuration to estimate camera pitch angle. Then, by integrating the predicted camera parameters with the initial mesh, we design a mesh adjustment module that simultaneously refines the root joint orientation and body pose. Extensive experiments demonstrate that our framework outperforms state-of-the-art methods on the benchmark datasets SPEC-SYN and SPEC-MTP.

📄 PDF Abstract BibTeX arXiv:2512.15212

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

EvPlug: Learn a Plug-and-Play Module for Event and Image Fusion

2023-12-28 · Jianping Jiang, Xinyu Zhou, Peiqi Duan, Boxin Shi

Event cameras and RGB cameras exhibit complementary characteristics in imaging: the former possesses high dynamic range (HDR) and high temporal resolution, while the latter provides rich texture and color information. Th…

3D Hand Pose EstimationHand Pose Estimationobject-detectionObject Detection+2

Uni3C: Unifying Precisely 3D-Enhanced Camera and Human Motion Controls for Video Generation

2025-04-21 · Chenjie Cao, Jingkai Zhou, Shikai Li, Jingyun Liang 외

Camera and human motion controls have been extensively studied for video generation, but existing approaches typically address them separately, suffering from limited data with high-quality annotations for both aspects. …

Video Generation

GS-Net: Generalizable Plug-and-Play 3D Gaussian Splatting Module

2024-09-17 · Yichen Zhang, Zihan Wang, Jiali Han, Peilin Li 외

3D Gaussian Splatting (3DGS) integrates the strengths of primitive-based representations and volumetric rendering techniques, enabling real-time, high-quality rendering. However, 3DGS models typically overfit to single-s…

3DGS

SynCamMaster: Synchronizing Multi-Camera Video Generation from Diverse Viewpoints

2024-12-10 · Jianhong Bai, Menghan Xia, Xintao Wang, Ziyang Yuan 외

Recent advancements in video diffusion models have shown exceptional abilities in simulating real-world dynamics and maintaining 3D consistency. This progress inspires us to investigate the potential of these models to e…

4D reconstructionVideo Generation

Glass Surface Segmentation with an RGB-D Camera via Weighted Feature Fusion for Service Robots

2025-08-03 · Henghong Lin, Zihan Zhu, Tao Wang, Anastasia Ioannou 외 arxiv

We address the problem of glass surface segmentation with an RGB-D camera, with a focus on effectively fusing RGB and depth information. To this end, we propose a Weighted Feature Fusion (WFF) module that dynamically and…