A Multimodal Corpus of Expert Gaze and Behavior during Phonetic Segmentation Tasks
Phonetic segmentation is the process of splitting speech into distinct phonetic units. Human experts routinely perform this task manually by analyzing auditory and visual cues using analysis software, which is an extremely time-consuming process. Methods exist for automatic segmentation, but these are not always accurate enough. In order to improve automatic segmentation, we need to model it as close to the manual segmentation as possible. This corpus is an effort to capture the human segmentation behavior by recording experts performing a segmentation task. We believe that this data will enable us to highlight the important aspects of manual segmentation, which can be used in automatic segmentation to improve its accuracy.
Code (1)
Tasks
SegmentationSimilar Papers 제목 키워드 기반
Deep semantic gaze embedding and scanpath comparison for expertise classification during OPT viewing
Modeling eye movement indicative of expertise behavior is decisive in user evaluation. However, it is indisputable that task semantics affect gaze behavior. We present a novel approach to gaze scanpath comparison that in…
General ClassificationMutual Gaze and Linguistic Repetition in a Multimodal Corpus
This paper investigates the correlation between mutual gaze and linguistic repetition, a form of alignment, which we take as evidence of mutual understanding. We focus on a multimodal corpus made of three-party conversat…
Mutual GazeGEM: Context-Aware Gaze EstiMation with Visual Search Behavior Matching for Chest Radiograph
Gaze estimation is pivotal in human scene comprehension tasks, particularly in medical diagnostic analysis. Eye-tracking technology facilitates the recording of physicians' ocular movements during image interpretation, t…
DiagnosticGaze Estimationgraph constructionA Multimodal Corpus for Mutual Gaze and Joint Attention in Multiparty Situated Interaction
Toward Scalable and Transparent Multimodal Analytics to Study Standard Medical Procedures: Linking Hand Movement, Proximity, and Gaze Data
This study employed multimodal learning analytics (MMLA) to analyze behavioral dynamics during the ABCDE procedure in nursing education, focusing on gaze entropy, hand movement velocities, and proximity measures. Utilizi…