Inflect-Micro-v2: complete voice in 9.36M parameters
nateb2022
69 points
5 comments
July 26, 2026
Related Discussions
Found 5 related stories in 348.9ms across 14,850 title embeddings via pgvector HNSW
- VibeVoice: Open-source frontier voice AI tosh · 345 pts · April 28, 2026 · 49% similar
- Advancing voice intelligence with new models in the API meetpateltech · 33 pts · May 07, 2026 · 47% similar
- Show HN: On-device transcriber that's 97% accurate at identifying speakers marshalla · 14 pts · June 05, 2026 · 47% similar
- Turn your singing voice into printable notes (in the browser) busssard · 36 pts · July 14, 2026 · 44% similar
- Speaking of Voxtral Palmik · 18 pts · March 26, 2026 · 44% similar
Discussion Highlights (4 comments)
tmaly
This is impressive. I wish there were a voice clone option.
jsomedon
amazing quality for such small size!
yjftsjthsd-h
Couple highlights: > Complete local text-to-waveform speech synthesis under 10M parameters. In case, like me, you hoped "complete" voice might mean both stt and tts. Not to speak poorly of it, just clarifying. > English only, with one fixed male voice. This is not zero-shot voice cloning. (And then a bunch of statements on limitations that I read as 'quality can be spotty but if you play with it it should be fine') But like. In <10M params I'm not judging:)
modinfo
This is amazing, the quality blow my mind for such small model! I just replaced my old onnx model with yours! here my implementation with speech dispatcher and server: https://github.com/skorotkiewicz/inflect-speechd thanks for shearing!