paper-with-me

Papers

Perceptual and acoustic analysis of voice similarities between parents and young children

2019-09-01 · WS (NoDaLiDa) 2019 9 · Evgeniia Rykova, Stefan Werner

Human voice provides the means for verbal communication and forms a part of personal identity. Due to genetic and environmental factors, a voice of a child should resemble the voice of her parent(s), but voice similarities between parents and young children are underresearched. Read-aloud speech of Finnish-speaking and Russian-speaking parent-child pairs was subject to perceptual and multi-step instrumental and statistical analysis. Finnish-speaking listeners could not discriminate family pairs auditorily in an XAB paradigm, but the Russian-speaking listeners’ mean accuracy of answers reached 72.5%. On average, in both language groups family-internal f0 similarities were stronger than family-external, with parents showing greater family-internal similarities than children. Auditory similarities did not reflect acoustic similarities in a straightforward way.

📄 PDF Abstract BibTeX

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Unified Acoustic Representations for Screening Neurological and Respiratory Pathologies from Voice

2025-08-28 · Ran Piao, Yuan Lu, Hareld Kemps, Tong Xia 외 arxiv

Voice-based health assessment offers unprecedented opportunities for scalable, non-invasive disease screening, yet existing approaches typically focus on single conditions and fail to leverage the rich, multi-faceted inf…

Towards an Interpretable Representation of Speaker Identity via Perceptual Voice Qualities

2023-10-04 · Robin Netzorg, Bohan Yu, Andrea Guzman, Peter Wu 외

Unlike other data modalities such as text and vision, speech does not lend itself to easy interpretation. While lay people can understand how to describe an image or sentence via perception, non-expert descriptions of sp…

Sentence

Acoustic and perceptual differences between standard and accented speech and their voice clones

2026-04-02 · Tianle Yang, Chengzhe Sun, Phil Rose, Siwei Lyu arxiv

Voice cloning is often evaluated in terms of overall quality, but less is known about accent preservation and its perceptual consequences. We compare standard and heavily accented Mandarin speech and their voice clones u…

When Fine-Tuning Fails and when it Generalises: Role of Data Diversity and Mixed Training in LLM-based TTS

2026-03-11 · Anupam Purwar, Aditya Choudhary arxiv

Large language models are increasingly adopted as semantic backbones for neural text-to-speech systems. However, frozen LLM representations are insufficient for modeling speaker specific acoustic and perceptual character…

Speech Synthesis along Perceptual Voice Quality Dimensions

2025-01-15 · Frederik Rautenberg, Michael Kuhlmann, Fritz Seebauer, Jana Wiechmann 외

While expressive speech synthesis or voice conversion systems mainly focus on controlling or manipulating abstract prosodic characteristics of speech, such as emotion or accent, we here address the control of perceptual …

Expressive Speech SynthesisSpeech Synthesistext-to-speechText to Speech+1