Voicv

Voicv clones a speaker's voice from a 10 to 30 second sample, then uses that model for text-to-speech, speech-to-text, and talking avatar videos. You uplo