paper-with-me

Papers

Real-Time Multi-View 3D Human Pose Estimation using Semantic Feedback to Smart Edge Sensors

2021-06-28 · Simon Bultmann, Sven Behnke

We present a novel method for estimation of 3D human poses from a multi-camera setup, employing distributed smart edge sensors coupled with a backend through a semantic feedback loop. 2D joint detection for each camera view is performed locally on a dedicated embedded inference processor. Only the semantic skeleton representation is transmitted over the network and raw images remain on the sensor board. 3D poses are recovered from 2D joints on a central backend, based on triangulation and a body model which incorporates prior knowledge of the human skeleton. A feedback channel from backend to individual sensors is implemented on a semantic level. The allocentric 3D pose is backprojected into the sensor views where it is fused with 2D joint detections. The local semantic model on each sensor can thus be improved by incorporating global context information. The whole pipeline is capable of real-time operation. We evaluate our method on three public datasets, where we achieve state-of-the-art results and show the benefits of our feedback architecture, as well as in our own setup for multi-person experiments. Using the feedback signal improves the 2D joint detections and in turn the estimated 3D poses.

📄 PDF Abstract BibTeX arXiv:2106.14729

Code (1)

AIS-Bonn/SmartEdgeSensor3DHumanPose 공식 구현

Tasks

3D Human Pose Estimation3D Multi-Person Pose EstimationMulti-view 3D Human Pose EstimationPose Estimation

Similar Papers 제목 키워드 기반

NeuralHumanFVV: Real-Time Neural Volumetric Human Performance Rendering using RGB Cameras

2021-03-13 · CVPR 2021 1 · Xin Suo, Yuheng Jiang, Pei Lin, Yingliang Zhang 외

4D reconstruction and rendering of human activities is critical for immersive VR/AR experience.Recent advances still fail to recover fine geometry and texture results with the level of detail present in the input images …

4D reconstructionMulti-Task Learning

iButter: Neural Interactive Bullet Time Generator for Human Free-viewpoint Rendering

2021-08-12 · Liao Wang, Ziyu Wang, Pei Lin, Yuheng Jiang 외

Generating ``bullet-time'' effects of human free-viewpoint videos is critical for immersive visual effects and VR/AR experience. Recent neural advances still lack the controllable and interactive bullet-time design abili…

NeRFVideo Generation

Holoported Characters: Real-time Free-viewpoint Rendering of Humans from Sparse RGB Cameras

2023-12-12 · CVPR 2024 1 · Ashwath Shetty, Marc Habermann, Guoxing Sun, Diogo Luvizon 외

We present the first approach to render highly realistic free-viewpoint videos of a human actor in general apparel, from sparse multi-view recording to display, in real-time at an unprecedented 4K resolution. At inferenc…

4k

GIGA: Generalizable Sparse Image-driven Gaussian Avatars

2025-04-08 · Anton Zubekhin, Heming Zhu, Paulo Gotardo, Thabo Beeler 외

Driving a high-quality and photorealistic full-body human avatar, from only a few RGB cameras, is a challenging problem that has become increasingly relevant with emerging virtual reality technologies. To democratize suc…

Human4DiT: 360-degree Human Video Generation with 4D Diffusion Transformer

2024-05-27 · Ruizhi Shao, Youxin Pang, Zerong Zheng, Jingxiang Sun 외

We present a novel approach for generating 360-degree high-quality, spatio-temporally coherent human videos from a single image. Our framework combines the strengths of diffusion transformers for capturing global correla…

Video Generation