paper-with-me

홈 › Papers

Does My Dog ''Speak'' Like Me? The Acoustic Correlation between Pet Dogs and Their Human Owners

2023-09-21 · Jieyi Huang, Chunhao Zhang, YuFei Wang, Mengyue Wu, Kenny Zhu

How hosts language influence their pets' vocalization is an interesting yet underexplored problem. This paper presents a preliminary investigation into the possible correlation between domestic dog vocal expressions and their human host's language environment. We first present a new dataset of Shiba Inu dog vocals from YouTube, which provides 7500 clean sound clips, including their contextual information of these vocals and their owner's speech clips with a carefully-designed data processing pipeline. The contextual information includes the scene category in which the vocal was recorded, the dog's location and activity. With a classification task and prominent factor analysis, we discover significant acoustic differences in the dog vocals from the two language environments. We further identify some acoustic features from dog vocalizations that are potentially correlated to their host language patterns.

📄 PDF Abstract BibTeX arXiv:2309.13085

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Articulatory-WaveNet: Autoregressive Model For Acoustic-to-Articulatory Inversion

2020-06-22 · Narjes Bozorg, Michael T. Johnson

This paper presents Articulatory-WaveNet, a new approach for acoustic-to-articulator inversion. The proposed system uses the WaveNet speech synthesis architecture, with dilated causal convolutional layers using previous …

Speech Synthesis

An objective evaluation of the effects of recording conditions and speaker characteristics in multi-speaker deep neural speech synthesis

2021-06-03 · Beata Lorincz, Adriana Stan, Mircea Giurgiu

Multi-speaker spoken datasets enable the creation of text-to-speech synthesis (TTS) systems which can output several voice identities. The multi-speaker (MSPK) scenario also enables the use of fewer training samples per …

Speaker VerificationSpeech Synthesistext-to-speechText to Speech+1

Speaker-Independent Acoustic-to-Articulatory Speech Inversion

2023-02-14 · Peter Wu, Li-Wei Chen, Cheol Jun Cho, Shinji Watanabe 외

To build speech processing methods that can handle speech as naturally as humans, researchers have explored multiple ways of building an invertible mapping from speech to an interpretable space. The articulatory space is…

Resynthesis

Dynamic Layer Normalization for Adaptive Neural Acoustic Modeling in Speech Recognition

2017-07-19 · Taesup Kim, Inchul Song, Yoshua Bengio

Layer normalization is a recently introduced technique for normalizing the activities of neurons in deep neural networks to improve the training speed and stability. In this paper, we introduce a new layer normalization …

speech-recognitionSpeech Recognition

Acoustic Analysis of Native (L1) Bengali Speakers’ Phonological Realization of English Lexical Stress Contrast

2020-12-01 · ICON 2020 12 · Shambhu Nath Saha, Shyamal Kr. Das Mandal

Acoustically, English lexical stress is multidimensional and involving manipulation of duration, intensity, fundamental frequency (F0) and vowel quality. The current study investigates the acquisition of English lexical …

Sentence