Clone a voice in 5 seconds to generate arbitrary speech in real time (github.com) 50 points by roseway4 7y ago ↗ HN
[–] allana 7y ago ↗ How does output compare to Mozilla TTS or Google's new customizable TTS: https://sample-efficient-adaptive-tts.github.io/demo/The latter requires a few minutes of audio to tune it to clone a voice. [–] zamadatix 7y ago ↗ I'd say the 10 second trained clips towards the bottom sound much better than the clips from the YouTube video of this project.
[–] zamadatix 7y ago ↗ I'd say the 10 second trained clips towards the bottom sound much better than the clips from the YouTube video of this project.
[–] rasz 7y ago ↗ Ironically demo YT clip is "relative_loudness": "-17.739dB", aka cant hear it. Not great for someone doing sound stuff.
3 comments
[ 3.0 ms ] story [ 18.5 ms ] threadThe latter requires a few minutes of audio to tune it to clone a voice.