paper-with-me

Papers

Employing self-supervised learning models for cross-linguistic child speech maturity classification

2025-06-10 · Theo Zhang, Madurya Suresh, Anne S. Warlaumont, Kasia Hitczenko, Alejandrina Cristia, Margaret Cychosz

Speech technology systems struggle with many downstream tasks for child speech due to small training corpora and the difficulties that child speech pose. We apply a novel dataset, SpeechMaturity, to state-of-the-art transformer models to address a fundamental classification task: identifying child vocalizations. Unlike previous corpora, our dataset captures maximally ecologically-valid child vocalizations across an unprecedented sample, comprising children acquiring 25+ languages in the U.S., Bolivia, Vanuatu, Papua New Guinea, Solomon Islands, and France. The dataset contains 242,004 labeled vocalizations, magnitudes larger than previous work. Models were trained to distinguish between cry, laughter, mature (consonant+vowel), and immature speech (just consonant or vowel). Models trained on the dataset outperform state-of-the-art models trained on previous datasets, achieved classification accuracy comparable to humans, and were robust across rural and urban settings.

📄 PDF Abstract BibTeX arXiv:2506.08999

Code (1)

spoglab-stanford/w2v2-pro-sm 공식 구현 pytorch

Tasks

Self-Supervised Learningvalid

Similar Papers 제목 키워드 기반

BabyHuBERT: Multilingual Self-Supervised Learning for Segmenting Speakers in Child-Centered Long-Form Recordings

2025-09-18 · Théo Charlot, Tarek Kunze, Maxime Poli, Alejandrina Cristia 외 arxiv

Child-centered daylong recordings are essential for studying early language development, but existing speech models trained on clean adult data perform poorly due to acoustic and linguistic differences. We introduce Baby…

Self-Supervised Learning

Layer-Wise Analysis of Self-Supervised Representations for Age and Gender Classification in Children's Speech

2025-08-14 · Abhijit Sinha, Harishankar Kumar, Mohit Joshi, Hemant Kumar Kathania 외 arxiv

Children's speech presents challenges for age and gender classification due to high variability in pitch, articulation, and developmental traits. While self-supervised learning (SSL) models perform well on adult speech t…

Age And Gender ClassificationSelf-Supervised Learning

Towards few-shot isolated word reading assessment

2025-07-16 · Reuben Smit, Retief Louw, Herman Kamper arxiv

We explore an ASR-free method for isolated word reading assessment in low-resource settings. Our few-shot approach compares input child speech to a small set of adult-provided reference templates. Inputs and templates ar…

Improving Children's Speech Recognition by Fine-tuning Self-supervised Adult Speech Representations

2022-11-14 · Renee Lu, Mostafa Shahin, Beena Ahmed

Children's speech recognition is a vital, yet largely overlooked domain when building inclusive speech technologies. The major challenge impeding progress in this domain is the lack of adequate child speech corpora; howe…

Self-Supervised Learningspeech-recognitionSpeech Recognition

Child-Centric Voice Anonymization in Single and Multi-Speaker Speech via Domain-Adapted SSL Models

2026-06-29 · Pranav Tushar, Xiao Xiao Miao, Rong Tong arxiv

Voice anonymization aims to protect speaker identity while preserving linguistic content and speech usability. However, most anonymization systems are developed on adult speech, leading to degraded performance when appli…

Self-Supervised LearningDomain Adaptation