Skip to content

Comment on Show HN: Cloning the Voices of Alan Rickman, Carl Sagan and Robin Williams

Comments

Pretty cool. Rough and fuzzy around the edges, some of the usual speech synthesis cadence issues. Any insight into the techniques you're using to do this?

This uses http://hts.sp.nitech.ac.jp (HMM-based Speech Synthesis System). Yes, it is expected to be robotic sounding since real audio streams are not used during synthesis. Some improvements can be made I believe (larger training set and some post processing) - but don't think it will change dramatically. Thanks for trying.

AboutSource Built by g1lg1l

Hackerly is an independent reader for Hacker News, built on the public HN API. Not affiliated with Y Combinator.