A subscriber cancelled and asked for more accurate real-time transcription. So I benchmarked every local ASR model I could get running against the one I ship. Mine shows words about two seconds behind and drops fewer…

2 points•marshalla•4 days ago•1 comment•

1 comment

marshalla4 days ago
Ouch. A subscriber cancelled with a P.S. saying the real-time transcription left something to be desired and that I should use GPU models with more accuracy.

Long story short: I got a batching model to behave more like a streaming model, ~2s updates, and it runs on the ANE so it’s efficient annnd it’s more accurate than all feasible alternatives annnd it only took me like a month. Annnd this is a meeting transcription / notetaker app for macOS.

I used a bot to help me generate precise language around the numbers for this post. Forgive me.

Read the full thread on Hacker News →

Related stories