HN user

TheWellKnownEIP

4 karma
Posts0
Comments5
View on HN
No posts found.

The repo itself has good tutorials under the notebooks/ folder to get started with training and generating synthesized voices. Check out "Tutorial_2_train_your_first_TTS_model" you could start there.

FYI the format that they expect in metadata.csv has changed over time, it used to be "filename|transcribed text" and now it expects "filename|speaker name|transcribed text" but that's not reflected in the notebook.

Coincidentally I've just started playing around with Coqui TTS for training on my own experimental datasets. I was naive enough to think I could get it to run on Windows instead of Linux, I would suggest you save yourselves the time and start from Linux if you're giving it a go!

I had the impression that the optimal swipes had an equal amount of dots on either side with the swipe going across as many open gaps (edges) of the shape as possible.