According to your description, it should be the problem of spectrum generation.
AudioFlux itself does not directly provide TTS function, but it can be used to analyze spectrum problems, or try different spectrum types. The effect of using erb/bar type spectrum is better than mel spectrum.
Comments
According to your description, it should be the problem of spectrum generation.
AudioFlux itself does not directly provide TTS function, but it can be used to analyze spectrum problems, or try different spectrum types. The effect of using erb/bar type spectrum is better than mel spectrum.
I don’t need help making the TTS, but I need to know when it’s a weird generation.
Based on your description this may be what I need but I clearly have a lot to learn