Multi-Tacotron-Voice-Cloning How to run on new voices?

How to run on new voices?

Open sravanidn opened this issue 3 years ago • 1 comments

Hello, Amazing work. I am running inference using your models on 2080 gpu. your example is perfect. But when I give a new audio clip (in English) and make it say the same Russian sentence, the output audio isn't good. There's lot of noise, and cloning is not even of good quality.

My question is:

Can I use pretrained models(from this repo) to clone a new speaker, and make it speak Russian? or Should I train every thing(g2p, encoder, synthesizer, vocoder) on new speaker(assuming I obtain hours of this speaker's audio)? Please advise.

Thanks, S

Sep 01 '21 07:09 sravanidn

You need to train the model yourself on much larger datasets, I'm doing that now.

Sep 30 '21 16:09 fancat-programer

Multi-Tacotron-Voice-Cloning Multi-Tacotron-Voice-Cloning copied to clipboard

How to run on new voices?

Multi-Tacotron-Voice-Cloning
Multi-Tacotron-Voice-Cloning copied to clipboard