Multi-Tacotron-Voice-Cloning icon indicating copy to clipboard operation
Multi-Tacotron-Voice-Cloning copied to clipboard

How to run on new voices?

Open sravanidn opened this issue 3 years ago • 1 comments

Hello, Amazing work. I am running inference using your models on 2080 gpu. your example is perfect. But when I give a new audio clip (in English) and make it say the same Russian sentence, the output audio isn't good. There's lot of noise, and cloning is not even of good quality.

My question is:

  1. Can I use pretrained models(from this repo) to clone a new speaker, and make it speak Russian? or Should I train every thing(g2p, encoder, synthesizer, vocoder) on new speaker(assuming I obtain hours of this speaker's audio)? Please advise.

Thanks, S

sravanidn avatar Sep 01 '21 07:09 sravanidn

You need to train the model yourself on much larger datasets, I'm doing that now.

fancat-programer avatar Sep 30 '21 16:09 fancat-programer