Inflect-Micro-v2: complete voice in 9.36M parameters

nateb2022 69 points 5 comments July 26, 2026
huggingface.co · View on Hacker News

Discussion Highlights (4 comments)

tmaly

This is impressive. I wish there were a voice clone option.

jsomedon

amazing quality for such small size!

yjftsjthsd-h

Couple highlights: > Complete local text-to-waveform speech synthesis under 10M parameters. In case, like me, you hoped "complete" voice might mean both stt and tts. Not to speak poorly of it, just clarifying. > English only, with one fixed male voice. This is not zero-shot voice cloning. (And then a bunch of statements on limitations that I read as 'quality can be spotty but if you play with it it should be fine') But like. In <10M params I'm not judging:)

modinfo

This is amazing, the quality blow my mind for such small model! I just replaced my old onnx model with yours! here my implementation with speech dispatcher and server: https://github.com/skorotkiewicz/inflect-speechd thanks for shearing!

Semantic search powered by Rivestack pgvector
14,850 stories · 138,743 chunks indexed