RU version is available. Content is displayed in original English for accuracy.
Advertisement
Advertisement
⚡ Community Insights
Discussion Sentiment
92% Positive
Analyzed from 369 words in the discussion.
Trending Topics
#tts#quality#speech#voice#model#inflect#micro#amazing#small#text

Discussion (27 Comments)Read Original on HackerNews
here my implementation with speech dispatcher and server: https://github.com/skorotkiewicz/inflect-speechd
thanks for shearing!
> Complete local text-to-waveform speech synthesis under 10M parameters.
In case, like me, you hoped "complete" voice might mean both stt and tts. Not to speak poorly of it, just clarifying.
> English only, with one fixed male voice. This is not zero-shot voice cloning.
(And then a bunch of statements on limitations that I read as 'quality can be spotty but if you play with it it should be fine') But like. In <10M params I'm not judging:)
https://inflect-tts.geronimo-labs.com
code: https://github.com/geronimi73/inflect-tts
On my iPhone 14 Pro the page crashes after 2-3 plays. I wonder if it uses too much memory?
IMHO, its at about the same quality level of historic TTS tools.