← SBrT2021
Conversão Texto-Fala para o Português Brasileiro Utilizando Tacotron 2 com Vocoder Griffin-Lim
Redes neuraisSíntese de vozTacotron 2Português brasileiro
Resumo
This paper presents the training of a state-of-the-art neural network model, Tacotron-2, using a open-source voice dataset from the Common Voice project. Results from training the model from scratch and by applying transfer learning of a pre-trained english model were evaluated. The results show that it is possible to train the model with limited data resources and the model trained from scratch had less synthesis errors.