A subscriber cancelled and asked for more accurate real-time transcription. So I benchmarked every local ASR model I could get running against the one I ship. Mine shows words about two seconds behind and drops fewer…
1 comment
marshalla4 days ago
Ouch. A subscriber cancelled with a P.S. saying the real-time transcription left something to be desired and that I should use GPU models with more accuracy.
Long story short: I got a batching model to behave more like a streaming model, ~2s updates, and it runs on the ANE so it’s efficient annnd it’s more accurate than all feasible alternatives annnd it only took me like a month. Annnd this is a meeting transcription / notetaker app for macOS.
I used a bot to help me generate precise language around the numbers for this post. Forgive me.
Read the full thread on Hacker News →
Related stories
- The Verge · 0 points · 6 days ago
- Hacker News · 2 points · 6 days ago
- Hacker News · 1 points · 4 days ago
- DEV Community · 38 points · 14 days ago
- Gemini 3.8 with Live Avatarcloud.google.comHacker News · 1 points · 6 days ago
- Hacker News · 1 points · 3 days ago