paper-with-me

Papers

VoiceCoach: Interactive Evidence-based Training for Voice Modulation Skills in Public Speaking

2020-01-22 · Xingbo Wang, Haipeng Zeng, Yong Wang, Aoyu Wu, Zhida Sun, Xiaojuan Ma, Huamin Qu

The modulation of voice properties, such as pitch, volume, and speed, is crucial for delivering a successful public speech. However, it is challenging to master different voice modulation skills. Though many guidelines are available, they are often not practical enough to be applied in different public speaking situations, especially for novice speakers. We present VoiceCoach, an interactive evidence-based approach to facilitate the effective training of voice modulation skills. Specifically, we have analyzed the voice modulation skills from 2623 high-quality speeches (i.e., TED Talks) and use them as the benchmark dataset. Given a voice input, VoiceCoach automatically recommends good voice modulation examples from the dataset based on the similarity of both sentence structures and voice modulation skills. Immediate and quantitative visual feedback is provided to guide further improvement. The expert interviews and the user study provide support for the effectiveness and usability of VoiceCoach.

📄 PDF Abstract BibTeX arXiv:2001.07876

Code (1)

xingbow/TED_dataset

Tasks

Sentence

Similar Papers 제목 키워드 기반

MM-RealSR: Metric Learning based Interactive Modulation for Real-World Super-Resolution

2022-05-10 · Chong Mou, Yanze Wu, Xintao Wang, Chao Dong 외

Interactive image restoration aims to restore images by adjusting several controlling coefficients, which determine the restoration strength. Existing methods are restricted in learning the controllable functions under t…

Image RestorationMetric LearningSuper-Resolution

Interactive Reinforcement Learning for Table Balancing Robot

2021-08-01 · ACL (splurobonlp) 2021 8 · Haein Jeon, Yewon Kim, Bo-Yeong Kang

With the development of robotics, the use of robots in daily life is increasing, which has led to the need for anyone to easily train robots to improve robot use. Interactive reinforcement learning(IARL) is a method for …

Automatic Speech RecognitionAutomatic Speech Recognition (ASR)Deep Reinforcement Learningreinforcement-learning+5

A Transversal Study of Fundamental Frequency Contours in Parkinsonian Voices

2024-02-09 · Pablo Rodriguez-Perez, Ruben Fraile, Miguel Garcia-Escrig, Nicolas Saenz-Lechon 외

A transversal study of the pitch variability of parkinsonian voices in read speech is presented. 30 patients suffering from Parkinson's disease (PD) and 32 healthy speakers were recorded while reading a text without voic…

Generative Moment Matching Network-based Random Modulation Post-filter for DNN-based Singing Voice Synthesis and Neural Double-tracking

2019-02-09 · Hiroki Tamaru, Yuki Saito, Shinnosuke Takamichi, Tomoki Koriyama 외

This paper proposes a generative moment matching network (GMMN)-based post-filter that provides inter-utterance pitch variation for deep neural network (DNN)-based singing voice synthesis. The natural pitch variation of …

Singing Voice Synthesis

Mic Drop or Data Flop? Evaluating the Fitness for Purpose of AI Voice Interviewers for Data Collection within Quantitative & Qualitative Research Contexts

2025-09-01 · Shreyas Tirumala, Nishant Jain, Danny D. Leybzon, Trent D. Buskirk arxiv

Transformer-based Large Language Models (LLMs) have paved the way for "AI interviewers" that can administer voice-based surveys with respondents in real-time. This position paper reviews emerging evidence to understand w…

Speech Recognition