A 2019 Guide to Speech Synthesis with Deep Learning
The authors of this paper are from Google. They present a neural network for generating raw audio waves. Their model is fully probabilistic and autoregressive, and it generates state-of-the-art text-to-speech results for both English and Mandarin. WaveNet is an audio generative model based on the PixelCNN. In this generative model, each audio sample is conditioned on the previous audio sample.
Aug-30-2019, 08:49:07 GMT
- Technology: