Inflect-Micro-v2: complete voice in 9.36M parameters
nateb2022
69 points
5 comments
July 26, 2026
Related Discussions
Found 5 related stories in 44.0ms across 3,370 title embeddings via pgvector HNSW
- Turn your singing voice into printable notes (in the browser) busssard · 36 pts · July 14, 2026 · 44% similar
- Show HN: Ski – Voice Coding for Claude Code, Codex and More – On-Device – Free jomon003 · 13 pts · July 30, 2026 · 43% similar
- Model 4: Our best model yet for local dictation and cleanup on Mac oakst · 11 pts · July 20, 2026 · 43% similar
- Show HN: Yap – OSS on-device voice dictation for macOS with no model to download pancomplex · 42 pts · July 27, 2026 · 43% similar
- Apple's new SpeechAnalyzer API, benchmarked against Whisper and its predecessor get-inscribe · 511 pts · July 13, 2026 · 42% similar
Discussion Highlights (4 comments)
tmaly
This is impressive. I wish there were a voice clone option.
jsomedon
amazing quality for such small size!
yjftsjthsd-h
Couple highlights: > Complete local text-to-waveform speech synthesis under 10M parameters. In case, like me, you hoped "complete" voice might mean both stt and tts. Not to speak poorly of it, just clarifying. > English only, with one fixed male voice. This is not zero-shot voice cloning. (And then a bunch of statements on limitations that I read as 'quality can be spotty but if you play with it it should be fine') But like. In <10M params I'm not judging:)
modinfo
This is amazing, the quality blow my mind for such small model! I just replaced my old onnx model with yours! here my implementation with speech dispatcher and server: https://github.com/skorotkiewicz/inflect-speechd thanks for shearing!