paper-with-me

홈 › Papers

GestureLens: Visual Analysis of Gestures in Presentation Videos

2022-04-19 · Haipeng Zeng, Xingbo Wang, Yong Wang, Aoyu Wu, Ting Chuen Pong, Huamin Qu

Appropriate gestures can enhance message delivery and audience engagement in both daily communication and public presentations. In this paper, we contribute a visual analytic approach that assists professional public speaking coaches in improving their practice of gesture training through analyzing presentation videos. Manually checking and exploring gesture usage in the presentation videos is often tedious and time-consuming. There lacks an efficient method to help users conduct gesture exploration, which is challenging due to the intrinsically temporal evolution of gestures and their complex correlation to speech content. In this paper, we propose GestureLens, a visual analytics system to facilitate gesture-based and content-based exploration of gesture usage in presentation videos. Specifically, the exploration view enables users to obtain a quick overview of the spatial and temporal distributions of gestures. The dynamic hand movements are firstly aggregated through a heatmap in the gesture space for uncovering spatial patterns, and then decomposed into two mutually perpendicular timelines for revealing temporal patterns. The relation view allows users to explicitly explore the correlation between speech content and gestures by enabling linked analysis and intuitive glyph designs. The video view and dynamic view show the context and overall dynamic movement of the selected gestures, respectively. Two usage scenarios and expert interviews with professional presentation coaches demonstrate the effectiveness and usefulness of GestureLens in facilitating gesture exploration and analysis of presentation videos.

📄 PDF Abstract BibTeX arXiv:2204.08894

Code (0)

등록된 구현이 없습니다.

Methods 이 논문이 사용한 방법론

Heatmap 설명 없음

Similar Papers 제목 키워드 기반

Transfer Learning of Deep Spatiotemporal Networks to Model Arbitrarily Long Videos of Seizures

2021-06-22 · Fernando Pérez-García, Catherine Scott, Rachel Sparks, Beate Diehl 외

Detailed analysis of seizure semiology, the symptoms and signs which occur during a seizure, is critical for management of epilepsy patients. Inter-rater reliability using qualitative visual analysis is often poor for se…

Action RecognitionManagementTemporal Action LocalizationTransfer Learning

Deep Neural Network approaches for Analysing Videos of Music Performances

2022-05-05 · Foteini Simistira Liwicki, Richa Upadhyay, Prakash Chandra Chhipa, Killian Murphy 외

This paper presents a framework to automate the labelling process for gestures in musical performance videos with a 3D Convolutional Neural Network (CNN). While this idea was proposed in a previous study, this paper intr…

Joint Surgical Gesture and Task Classification with Multi-Task and Multimodal Learning

2018-05-02 · Duygu Sarikaya, Khurshid A. Guru, Jason J. Corso

We propose a novel multi-modal and multi-task architecture for simultaneous low level gesture and surgical task classification in Robot Assisted Surgery (RAS) videos.Our end-to-end architecture is based on the principles…

General ClassificationMulti-Task Learning

Self-Supervised Learning of Deviation in Latent Representation for Co-speech Gesture Video Generation

2024-09-26 · Huan Yang, Jiahui Chen, Chaofan Ding, Runhua Shi 외

Gestures are pivotal in enhancing co-speech communication. While recent works have mostly focused on point-level motion transformation or fully supervised motion representations through data-driven approaches, we explore…

Self-Supervised LearningSSIMVideo Generation

From Phase Grounding to Intelligent Surgical Narratives

2026-03-05 · Ethan Peterson, Huixin Zhan arxiv

Video surgery timelines are an important part of tool-assisted surgeries, as they allow surgeons to quickly focus on key parts of the procedure. Current methods involve the surgeon filling out a post-operation (OP) repor…