15 comments

[ 0.26 ms ] story [ 23.5 ms ] thread
This is cool! What model are you using?
Why did you pick such creepy voices? They are very childlike and weird.
They just sound like anime voices
They're definitely anime-style voices, though most of the female ones are very annoying. I couldn't stand to watch an anime with them.
You know, I never did find out why people dub anime with such unnatural voices.
(comment deleted)
Great, Working fast and cool. You can also try to adding some more voices from different parts of the world.
All but one of the voices sounds tinny. The timing and cadence was good though.

Is this intended for anime autodubbing? There's only one voice (Rowan) that is remotely traditional broadcaster-style.

(comment deleted)
As with seemingly all AI these days - it seems to fail with prosody and simply speaks the very next word with zero regard to the nuance or cadence that author intended, or an understanding of any the words being spoken.

When AI achieves the ability to deliver some of Shakespeare’s greatest soliloquies or monologues, then I’ll pay attention.

Can you change the voice? Also, is it possible to vary the emotion in one audio?