paper-with-me

Papers

Robust Self Supervised Speech Embeddings for Child-Adult Classification in Interactions involving Children with Autism

2023-07-31 · Rimita Lahiri, Tiantian Feng, Rajat Hebbar, Catherine Lord, So Hyun Kim, Shrikanth Narayanan

We address the problem of detecting who spoke when in child-inclusive spoken interactions i.e., automatic child-adult speaker classification. Interactions involving children are richly heterogeneous due to developmental differences. The presence of neurodiversity e.g., due to Autism, contributes additional variability. We investigate the impact of additional pre-training with more unlabelled child speech on the child-adult classification performance. We pre-train our model with child-inclusive interactions, following two recent self-supervision algorithms, Wav2vec 2.0 and WavLM, with a contrastive loss objective. We report 9 - 13% relative improvement over the state-of-the-art baseline with regards to classification F1 scores on two clinical interaction datasets involving children with Autism. We also analyze the impact of pre-training under different conditions by evaluating our model on interactions involving different subgroups of children based on various demographic factors.

📄 PDF Abstract BibTeX arXiv:2307.16398

Code (0)

등록된 구현이 없습니다.

Tasks

Classification

Similar Papers 제목 키워드 기반

Improving Children's Speech Recognition by Fine-tuning Self-supervised Adult Speech Representations

2022-11-14 · Renee Lu, Mostafa Shahin, Beena Ahmed

Children's speech recognition is a vital, yet largely overlooked domain when building inclusive speech technologies. The major challenge impeding progress in this domain is the lack of adequate child speech corpora; howe…

Self-Supervised Learningspeech-recognitionSpeech Recognition

Meta-learning for robust child-adult classification from speech

2019-10-28

Computational modeling of naturalistic conversations in clinical applications has seen growing interest in the past decade. An important use-case involves child-adult interactions within the autism diagnosis and interven…

ClassificationMeta-Learningspeaker-diarizationSpeaker Diarization+1

BabyHuBERT: Multilingual Self-Supervised Learning for Segmenting Speakers in Child-Centered Long-Form Recordings

2025-09-18 · Théo Charlot, Tarek Kunze, Maxime Poli, Alejandrina Cristia 외 arxiv

Child-centered daylong recordings are essential for studying early language development, but existing speech models trained on clean adult data perform poorly due to acoustic and linguistic differences. We introduce Baby…

Self-Supervised Learning

Analysis of Self-Supervised Speech Models on Children's Speech and Infant Vocalizations

2024-02-10 · Jialu Li, Mark Hasegawa-Johnson, Nancy L. McElwain

To understand why self-supervised learning (SSL) models have empirically achieved strong performances on several speech-processing downstream tasks, numerous studies have focused on analyzing the encoded information of t…

Phoneme RecognitionSelf-Supervised Learning

Adaptation of Whisper models to child speech recognition

2023-07-24 · Rishabh Jain, Andrei Barcovschi, Mariam Yiwere, Peter Corcoran 외

Automatic Speech Recognition (ASR) systems often struggle with transcribing child speech due to the lack of large child speech datasets required to accurately train child-friendly ASR models. However, there are huge amou…

Automatic Speech RecognitionAutomatic Speech Recognition (ASR)speech-recognitionSpeech Recognition