paper-with-me

홈 › Papers

ACRNet: Attention Cube Regression Network for Multi-view Real-time 3D Human Pose Estimation in Telemedicine

2022-10-11 · Boce Hu, Chenfei Zhu, Xupeng Ai, Sunil K. Agrawal

Human pose estimation (HPE) for 3D skeleton reconstruction in telemedicine has long received attention. Although the development of deep learning has made HPE methods in telemedicine simpler and easier to use, addressing low accuracy and high latency remains a big challenge. In this paper, we propose a novel multi-view Attention Cube Regression Network (ACRNet), which regresses the 3D position of joints in real time by aggregating informative attention points on each cube surface. More specially, a cube whose each surface contains uniformly distributed attention points with specific coordinate values is first created to wrap the target from the main view. Then, our network regresses the 3D position of each joint by summing and averaging the coordinates of attention points on each surface after being weighted. To verify our method, we first tested ACRNet on the open-source ITOP dataset; meanwhile, we collected a new multi-view upper body movement dataset (UBM) on the trunk support trainer (TruST) to validate the capability of our model in real rehabilitation scenarios. Experimental results demonstrate the superiority of ACRNet compared with other state-of-the-art methods. We also validate the efficacy of each module in ACRNet. Furthermore, Our work analyzes the performance of ACRNet under the medical monitoring indicator. Because of the high accuracy and running speed, our model is suitable for real-time telemedicine settings. The source code is available at https://github.com/BoceHu/ACRNet

📄 PDF Abstract BibTeX arXiv:2210.05130

Code (0)

등록된 구현이 없습니다.

Tasks

3D Human Pose EstimationPose EstimationPositionregression

Similar Papers 제목 키워드 기반

Aggregated Network for Massive MIMO CSI Feedback

2021-01-17 · Zhilin Lu, Hongyi He, Zhengyang Duan, Jintao Wang 외

In frequency division duplexing (FDD) mode, it is necessary to send the channel state information (CSI) from user equipment to base station. The downlink CSI is essential for the massive multiple-input multiple-output (M…

compressed sensingvalid

VirtualCube: An Immersive 3D Video Communication System

2021-12-13 · Yizhong Zhang, Jiaolong Yang, Zhen Liu, Ruicheng Wang 외

The VirtualCube system is a 3D video conference system that attempts to overcome some limitations of conventional technologies. The key ingredient is VirtualCube, an abstract representation of a real-world cubicle instru…

3D geometryDepth Estimation

CubeDiff: Repurposing Diffusion-Based Image Models for Panorama Generation

2025-01-28 · Nikolai Kalischek, Michael Oechsle, Fabian Manhardt, Philipp Henzler 외

We introduce a novel method for generating 360{\deg} panoramas from text prompts or images. Our approach leverages recent advances in 3D generation by employing multi-view diffusion models to jointly synthesize the six f…

3D Generation

CubeComposer: Spatio-Temporal Autoregressive 4K 360° Video Generation from Perspective Video

2026-03-04 · Lingen Li, Guangzhi Wang, Xiaoyu Li, Zhaoyang Zhang 외 arxiv

Generating high-quality 360° panoramic videos from perspective input is one of the crucial applications for virtual reality (VR), whereby high-resolution videos are especially important for immersive experience. Existing…

Video Generation

Binarized Aggregated Network with Quantization: Flexible Deep Learning Deployment for CSI Feedback in Massive MIMO System

2021-05-01 · Zhilin Lu, Xudong Zhang, Hongyi He, Jintao Wang 외

Massive multiple-input multiple-output (MIMO) is one of the key techniques to achieve better spectrum and energy efficiency in 5G system. The channel state information (CSI) needs to be fed back from the user equipment t…

BinarizationQuantization