Skip to content

Comment on AI speech generator 'reaches human parity' – but it's too dangerous to releaseparent

Comments

Try StyleTTS2. You will still have to experiment with the settings a little to get the right level of adherence to the reference speaker’s voice and the emotion content.

Without looking at this, are you sure that this can do speech to speech? Maybe my flaw in searching has been disregarding anything that's called "text to speech" as not also "speech to speech"?

Ah my bad, you’re right. I think it doesn’t do V2V directly, but can use reference audio to guide TTS.

AboutSource Built by g1lg1l

Hackerly is an independent reader for Hacker News, built on the public HN API. Not affiliated with Y Combinator.