paper-with-me

홈 › Papers

Panoptic Studio: A Massively Multiview System for Social Interaction Capture

2016-12-09 · Hanbyul Joo, Tomas Simon, Xulong Li, Hao liu, Lei Tan, Lin Gui, Sean Banerjee, Timothy Godisart, Bart Nabbe, Iain Matthews, Takeo Kanade, Shohei Nobuhara, Yaser Sheikh

We present an approach to capture the 3D motion of a group of people engaged in a social interaction. The core challenges in capturing social interactions are: (1) occlusion is functional and frequent; (2) subtle motion needs to be measured over a space large enough to host a social group; (3) human appearance and configuration variation is immense; and (4) attaching markers to the body may prime the nature of interactions. The Panoptic Studio is a system organized around the thesis that social interactions should be measured through the integration of perceptual analyses over a large variety of view points. We present a modularized system designed around this principle, consisting of integrated structural, hardware, and software innovations. The system takes, as input, 480 synchronized video streams of multiple people engaged in social activities, and produces, as output, the labeled time-varying 3D structure of anatomical landmarks on individuals in the space. Our algorithm is designed to fuse the "weak" perceptual processes in the large number of views by progressively generating skeletal proposals from low-level appearance cues, and a framework for temporal refinement is also presented by associating body parts to reconstructed dense 3D trajectory stream. Our system and method are the first in reconstructing full body motion of more than five people engaged in social interactions without using markers. We also empirically demonstrate the impact of the number of views in achieving this goal.

📄 PDF Abstract BibTeX arXiv:1612.03153

Code (2)

CMU-Perceptual-Computing-Lab/panoptic-toolbox 공식 구현
https://gitlab.com/Percipiote/skelda

Similar Papers 제목 키워드 기반

Panoptic Studio: A Massively Multiview System for Social Motion Capture

2015-12-01 · ICCV 2015 12 · Hanbyul Joo, Hao liu, Lei Tan, Lin Gui 외

We present an approach to capture the 3D structure and motion of a group of people engaged in a social interaction. The core challenges in capturing social interactions are: (1) occlusion is functional and frequent; (2) …

HUMBI: A Large Multiview Dataset of Human Body Expressions and Benchmark Challenge

2021-09-30 · Jae Shin Yoon, Zhixuan Yu, Jaesik Park, Hyun Soo Park

This paper presents a new large multiview dataset called HUMBI for human body expressions with natural clothing. The goal of HUMBI is to facilitate modeling view-specific appearance and geometry of five primary body sign…

HUMBI: A Large Multiview Dataset of Human Body Expressions

2018-12-01 · CVPR 2020 6 · Zhixuan Yu, Jae Shin Yoon, In Kyu Lee, Prashanth Venkatesh 외

This paper presents a new large multiview dataset called HUMBI for human body expressions with natural clothing. The goal of HUMBI is to facilitate modeling view-specific appearance and geometry of gaze, face, hand, body…

PanORama: Multiview Consistent Panoptic Segmentation in Operating Rooms

2026-03-20 · Tuna Gürbüz, Ege Özsoy, Tony Danjun Wang, Nassir Navab arxiv

Operating rooms (ORs) are cluttered, dynamic, highly occluded environments, where reliable spatial understanding is essential for situational awareness during complex surgical workflows. Achieving spatial understanding f…

Panoptic Segmentation

How2Sign: A Large-scale Multimodal Dataset for Continuous American Sign Language

2020-08-18 · CVPR 2021 1 · Amanda Duarte, Shruti Palaskar, Lucas Ventura, Deepti Ghadiyaram 외

One of the factors that have hindered progress in the areas of sign language recognition, translation, and production is the absence of large annotated datasets. Towards this end, we introduce How2Sign, a multimodal and …

Sign Language ProductionSign Language TranslationTranslation