It gets better as you use it, finetuning on your recordings overnight. Fully local, no data leaves your PC, source available, free.
I believe the future of software is software you actually own and that adapts to your personal needs.
github: https://github.com/Hugo0/voiceio
How finetuning works:
1. mp3 recordings of your voice are stored on your disk (if you choose so).
2. A big Model (Whisper large-v3-turbo by default, configurable) re-transcribes them when when your PC is idle overnight.
3. A LoRA fine-tunes your base model.
`voiceio learn schedule on` runs all of this automatically. For repeated mistakes there's also an instant fix: `voiceio vocab add <word>` steers the decoder right away, and `voiceio corrections add wrong right` handles anything left.
Beyond Whisper you can run Parakeet or a whisper.cpp server (both experimental). Feedback appreciated, have tested this with a small set of people so far.