Post #4194856
2026-01-13 16:34 UTC
@fastfinge@fed.interfree.ca My results shows that a dedicated IPC server performs faster, E.G the synthDriver is ready for use in just 5 seconds; response is good even in longer sentences, but this can be attributed to the 4.2m model I'm using. And when I ran the model through a streaming vocoder, response is surprisingly realtime, suitable for screen reader. As for voice rate, I'm using a modification of the good "audiostretchy" pip package. I can't give more details ATM, but I hope this helps in your research
Replies (1)
-
@fastfinge@fed.interfree.ca 2026-01-13 17:04
@rmcpantoja@mastodon.social Also, I'd love to hear if and when you release anything!