paper-with-me

홈 › Papers

The ISCSLP 2024 Conversational Voice Clone (CoVoC) Challenge: Tasks, Results and Findings

2024-10-31 · Kangxiang Xia, Dake Guo, Jixun Yao, Liumeng Xue, Hanzhao Li, Shuai Wang, Zhao Guo, Lei Xie, Qingqing Zhang, Lei Luo, Minghui Dong, Peng Sun

The ISCSLP 2024 Conversational Voice Clone (CoVoC) Challenge aims to benchmark and advance zero-shot spontaneous style voice cloning, particularly focusing on generating spontaneous behaviors in conversational speech. The challenge comprises two tracks: an unconstrained track without limitation on data and model usage, and a constrained track only allowing the use of constrained open-source datasets. A 100-hour high-quality conversational speech dataset is also made available with the challenge. This paper details the data, tracks, submitted systems, evaluation results, and findings.

📄 PDF Abstract BibTeX arXiv:2411.00064

Code (0)

등록된 구현이 없습니다.

Tasks

Voice Cloning

Similar Papers 제목 키워드 기반

TSUP Speaker Diarization System for Conversational Short-phrase Speaker Diarization Challenge

2022-10-26 · Bowen Pang, Huan Zhao, Gaosheng Zhang, Xiaoyue Yang 외

This paper describes the TSUP team's submission to the ISCSLP 2022 conversational short-phrase speaker diarization (CSSD) challenge which particularly focuses on short-phrase conversations with a new evaluation metric ca…

Action DetectionActivity Detectionspeaker-diarizationSpeaker Diarization

LeVoice ASR Systems for the ISCSLP 2022 Intelligent Cockpit Speech Recognition Challenge

2022-10-14 · Yan Jia, Mi Hong, Jingyu Hou, Kailong Ren 외

This paper describes LeVoice automatic speech recognition systems to track2 of intelligent cockpit speech recognition challenge 2022. Track2 is a speech recognition task without limits on the scope of model size. Our mai…

Automatic Speech RecognitionAutomatic Speech Recognition (ASR)Data AugmentationDecoder+7

CloneMem: Benchmarking Long-Term Memory for AI Clones

2026-01-11 · Sen Hu, Zhiyu Zhang, Yuxiang Wei, Xueran Han 외 arxiv

AI Clones aim to simulate an individual's thoughts and behaviors to enable long-term, personalized interaction, placing stringent demands on memory systems to model experiences, emotions, and opinions over time. Existing…

Securing Voice-driven Interfaces against Fake (Cloned) Audio Attacks

2019-02-18

Voice cloning technologies have found applications in a variety of areas ranging from personalized speech interfaces to advertisement, robotics, and so on. Existing voice cloning systems are capable of learning speaker c…

Speech SynthesisVoice Cloning

Voice "Cloning" is Style Transfer

2026-05-15 · Kaitlyn Zhou, Federico Bianchi, Martijn Bartelds, Anna Pot 외 arxiv

Artificially generated speech is increasingly embedded in everyday life. Voice cloning in particular enables applications where identity preservation is important, such as completing a recording, dubbing in a new languag…

Style Transfer