paper-with-me

홈 › Papers

Toward Scalable Co-located Practical Learning: Assisting with Computer Vision and Multimodal Analytics

2026-03-14 · Xinyu Li, Linxuan Zhao, Yueqiao Jin, Yuchen Liu, Jin Zhou, Roberto Martinez-Maldonado, Dragan Gasevic, Lixiang Yan arxiv

Co-located practical learning leaves evidence in visible actions around patients, task resources and room zones, but these traces are often recovered through live observation or retrospective video review. Fixed wide-angle video could reduce sensing burden, yet a debriefing pipeline must do more than detect behaviours: it must maintain detection after small camera-position shifts, relate the detector-derived behaviour trace to instructor-labelled outcomes and preserve room-zone context. This study evaluates a fixed-camera pipeline in repeated nursing simulation. Using a harmonised six-code taxonomy, we tested YOLO26 target-only training and two-stage source-to-target adaptation across two same-room side-view data sources. We then converted detections from 51 instructor-labelled sessions into one-second behaviour and behaviour-zone traces for rate, ordered-network, transition-network and sequence analyses. Two-stage adaptation improved mean mAP50 from 0.815 to 0.848 for the 2021 target view and from 0.690 to 0.855 for the smaller 2022 target view; with a balanced target quota of \(N = 22\), the 2022 model reached 0.850 mAP50. In the detector-derived behaviour trace analyses, higher phone use characterised low task-performance sessions. Zone labels changed the interpretation of patient interaction: primary patient-care-zone interaction was stronger in higher-performance sessions, while secondary-zone interaction was stronger in lower-performance sessions. Ordered and transition network models showed that ordered room-zone relations contributed beyond behaviour frequency, with the strongest task-performance classifier using zoned and co-presence features. The resulting trace is most appropriate for searchable simulation debriefing, where instructors inspect detected moments rather than receive automated assessment scores.

📄 PDF Abstract BibTeX arXiv:2603.13679

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Fast and Robust Structural Damage Analysis of Civil Infrastructure Using UAV Imagery

2021-10-10 · Alon Oring

The usage of Unmanned Aerial Vehicles (UAVs) in the context of structural health inspection is recently gaining tremendous popularity. Camera mounted UAVs enable the fast acquisition of a large number of images often use…

object-detectionObject DetectionRetrieval

Implications of Computer Vision Driven Assistive Technologies Towards Individuals with Visual Impairment

2019-05-20 · Linda Wang, Alexander Wong

Computer vision based technology is becoming ubiquitous in society. One application area that has seen an increase in computer vision is assistive technologies, specifically for those with visual impairment. Research has…

Comparing computer vision analysis of signed language video with motion capture recordings

2012-05-01 · LREC 2012 5 · Matti Karppa, Tommi Jantunen, Ville Viitaniemi, Jorma Laaksonen 외

We consider a non-intrusive computer-vision method for measuring the motion of a person performing natural signing in video recordings. The quality and usefulness of the method is compared to a traditional marker-based m…

Evaluating Scalable Bayesian Deep Learning Methods for Robust Computer Vision

2019-06-04 · Fredrik K. Gustafsson, Martin Danelljan, Thomas B. Schön

While deep neural networks have become the go-to approach in computer vision, the vast majority of these models fail to properly capture the uncertainty inherent in their predictions. Estimating this predictive uncertain…

Deep LearningDepth CompletionSemantic Segmentation

SpiritSight Agent: Advanced GUI Agent with One Look

2025-03-05 · CVPR 2025 1 · Zhiyuan Huang, Ziming Cheng, Junting Pan, Zhaohui Hou 외

Graphical User Interface (GUI) agents show amazing abilities in assisting human-computer interaction, automating human user's navigation on digital devices. An ideal GUI agent is expected to achieve high accuracy, low la…