paper-with-me

홈 › Papers

Korean aegyo speech shows systematic F1 increase to signal childlike qualities

2026-04-28 · Ji-eun Kim, Volker Dellwo arxiv

Korean aegyo is a socially recognized childlike speaking style used predominantly in romantic interactions among adults. This study examined vowel space modification in aegyo by analyzing formant frequencies from twelve Seoul Korean speakers who produced identical scripts in aegyo and non-aegyo styles. Results show that aegyo speech features a significant increase in F1 values across vowels and selective fronting of front vowels, leading to vowel space expansion but mainly a shift to higher F1. These findings suggest that adult speakers stylize childlike speech by imitating the shorter vocal tract of children, mainly through global vowel lowering and partial fronting.

📄 PDF Abstract BibTeX arXiv:2604.25133

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Automated Pronunciation Evaluation for Korean Toddler Speech using Speech Diarization and Self-Supervised Learning

2026-06-08 · Diane Myung-kyung Woodbridge, Jee Hyun Suh arxiv

Speech sound disorders affect approximately 44% of Korean pediatric communication disorder cases, yet automated assessment tools for Korean toddler speech remain underdeveloped. This paper presents an end-to-end pipeline…

Self-Supervised LearningRepresentation LearningSpeaker Diarization

Analyzing Error Propagation in Korean Spoken QA with ASR-LLM Cascades

2026-05-17 · Donghyuk Jung, Youngwon Choi arxiv

We analyze how automatic speech recognition (ASR) errors propagate through ASR--LLM cascades in Korean spoken question answering (SQA), focusing on downstream semantic failures that conventional ASR metrics cannot fully …

Speech RecognitionQuestion Answering

Learning pronunciation from a foreign language in speech synthesis networks

2018-10-22 · Anonymous

Although there are more than 65,000 languages in the world, the pronunciations of many phonemes sound similar across the languages. When people learn a foreign language, their pronunciation often reflect their native lan…

Speech Synthesis

Raon-Speech Technical Report

2026-04-08 · Beomsoo Kim, Changho Choi, Dohyun Kim, Dongki Lee 외 arxiv

We present Raon-Speech, a top-performing 9B-parameter speech language model (SpeechLM) for English and Korean speech understanding, answering, and generation, and Raon-SpeechChat, a high-performing full-duplex extension …

Knowledge DistillationQuestion Answering

Multi-task Learning is Not Enough: Representational Entanglement in Dual-output Second Language Speech Recognition

2026-06-04 · Seung Hwan Cho, Young-Min Kim arxiv

Second-language (L2) speech recognition often requires transcriptions of pronunciations and intended meanings. Multi-task learning (MTL) is a natural approach because it assumes that shared representations benefit both o…

Multi-Task LearningSpeech Recognition