WaveGrad
2000년 도입 · 논문 7편에서 사용
WaveGrad is a conditional model for waveform generation through estimating gradients of the data density. This model is built on the prior work on score matching and diffusion probabilistic models. It starts from Gaussian white noise and iteratively refines the signal via a gradient-based sampler conditioned on the mel-spectrogram. WaveGrad is non-autoregressive, and requires only a constant number of generation steps during inference. It can use as few as 6 iterations to generate high fidelity audio samples.
출처: WaveGrad: Estimating Gradients for Waveform Generation
소개 논문: WaveGrad: Estimating Gradients for Waveform Generation
Generative Audio Models · Audio