Show HN: FnScribe – Open-source, offline dictation for macOS
modagent
20 points
14 comments
August 28, 2026
Related Discussions
Found 5 related stories in 63.9ms across 4,827 title embeddings via pgvector HNSW
- Show HN: Qwen Scribe – local transcription and dictation for Apple Silicon sidclaw · 79 pts · July 29, 2026 · 72% similar
- Show HN: Yap – OSS on-device voice dictation for macOS with no model to download pancomplex · 42 pts · July 27, 2026 · 69% similar
- Show HN: Orate – On-device neural text-to-speech queue for Mac bherms · 11 pts · July 21, 2026 · 65% similar
- Show HN: OfflineTTS — Free browser-based TTS & STT that runs locally twainyoung · 11 pts · July 19, 2026 · 60% similar
- Show HN: Minute – Offline meeting notes on macOS with Whisper and llama.cpp mraza007 · 12 pts · July 28, 2026 · 59% similar
Discussion Highlights (6 comments)
modagent
I built FnScribe because I'm pretty privacy concious and wanted a wispr flow-like app that kept everything local and on device. Currently works for Mac (sillicon and intel). It's dead simple. Hold the fn key, speak and release. I use a quantized Wisper small.en model for transcription. It inserts the text into the active application. There's also a hands-free model for longer dictation. Audio transcription is kept in memory. There's no account or transcription history. Clipboard contents are restored after it inserts it. GPLv3, Mac-only, English only..still in alpha. Hope you enjoy it! Would love some feedback.
itsdesmond
I see that it changes the menu bar icon to indicate status. How is this communicated when windows are full screen and the menu bar isn’t visible?
djx22
You may want to look into different models that are more accurate and maybe an AEC layer to remove background noise. Or at the very least a RNN de-noiser on the mic channel. Also, you may want to stream audio to the model instead of holding it all in memory and transcribe at the very end as that can potentially allow you to take the app much further than it is now. I see Claude implemented a very crude upsampling/downsampling algorithm, which is what LLMs usually do when prompted to handle such a problem. But I would suggest restraining the model from implementing DSP processing on their own and instead use battle tested libraries. You can use rubato's FFT Resampler. Audio processing is genuinely a hard engineering problem, LLMs usually don't get it right. If you decide to get deep into it, the knowledge you'll get is very rewarding.
zackify
https://voxtype.io/ https://tryvoiceink.com/ There's so many of these, at this point I've seen 10 clones make the front page each time as if there never existed local only options before. Also whisper is pretty outdated vs parakeet
ahaferburg
Are you aware of https://handy.computer/ ?
vivzkestrel
- tonnes of these get released daily but I ll give you the guys on HN an awesome idea - havent seen a single one on HN yet in the last year (i read HN twice a day like brushing my teeth) - When I input my voice into the mic, I want an AI voice as output converting my words in real time in AI voice - Use case: gaming, I have a terrible voice and dont want to do a voice over with that but at the same time I would love to if I could - Know any github projects capable of pulling this off? maybe direct integration as an OBS plugin would make it godtier