paper-with-me

홈 › Papers

The Power of Prosody and Prosody of Power: An Acoustic Analysis of Finnish Parliamentary Speech

2023-05-25 · Martti Vainio, Antti Suni, Juraj Šimko, Sofoklis Kakouros

Parliamentary recordings provide a rich source of data for studying how politicians use speech to convey their messages and influence their audience. This provides a unique context for studying how politicians use speech, especially prosody, to achieve their goals. Here we analyzed a corpus of parliamentary speeches in the Finnish parliament between the years 2008-2020 and highlight methodological considerations related to the robustness of signal based features with respect to varying recording conditions and corpus design. We also present results of long term changes pertaining to speakers' status with respect to their party being in government or in opposition. Looking at large scale averages of fundamental frequency - a robust prosodic feature - we found systematic changes in speech prosody with respect opposition status and the election term. Reflecting a different level of urgency, members of the parliament have higher f0 at the beginning of the term or when they are in opposition.

📄 PDF Abstract BibTeX arXiv:2305.16040

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Prosody-Enhanced Acoustic Pre-training and Acoustic-Disentangled Prosody Adapting for Movie Dubbing

2025-03-15 · CVPR 2025 1 · Zhedong Zhang, Liang Li, Chenggang Yan, Chunshan Liu 외

Movie dubbing describes the process of transforming a script into speech that aligns temporally and emotionally with a given movie clip while exemplifying the speaker's voice demonstrated in a short reference audio clip.…

Emotion Recognition

Prosody-controllable spontaneous TTS with neural HMMs

2022-11-24 · Harm Lameris, Shivam Mehta, Gustav Eje Henter, Joakim Gustafson 외

Spontaneous speech has many affective and pragmatic functions that are interesting and challenging to model in TTS. However, the presence of reduced articulation, fillers, repetitions, and other disfluencies in spontaneo…

Diversityvalid

DiFlowDubber: Discrete Flow Matching for Automated Video Dubbing via Cross-Modal Alignment and Synchronization

2026-03-15 · Ngoc-Son Nguyen, Thanh V. T. Tran, Jeongsoo Choi, Hieu-Nghia Huynh-Nguyen 외 arxiv

Video dubbing requires content accuracy, expressive prosody, high-quality acoustics, and precise lip synchronization, yet existing approaches struggle on all four fronts. To address these issues, we propose DiFlowDubber,…

Prosody Transfer in Neural Text to Speech Using Global Pitch and Loudness Features

2019-11-21 · Siddharth Gururani, Kilol Gupta, Dhaval Shah, Zahra Shakeri 외

This paper presents a simple yet effective method to achieve prosody transfer from a reference speech signal to synthesized speech. The main idea is to incorporate well-known acoustic correlates of prosody such as pitch …

text-to-speechText to Speech

ProMode: A Speech Prosody Model Conditioned on Acoustic and Textual Inputs

2025-08-12 · Eray Eren, Qingju Liu, Hyeongwoo Kim, Pablo Garrido 외 arxiv

Prosody conveys rich emotional and semantic information of the speech signal as well as individual idiosyncrasies. We propose a stand-alone model that maps text-to-prosodic features such as F0 and energy and can be used …