Today we overhauled the core engine for commercial-grade speed and accuracy. We migrated to Faster-Whisper, which gives us flawless auto-punctuation and capitalization. To achieve a zero-latency feel, we built a background worker that streams and processes audio chunks every 500ms while you hold the hotkey, meaning the final text drops instantly on release. Finally, we engineered a dynamic disfluency filter that automatically strips out filler words like umms and ahhs before typing. Up next: connecting a local LLM to turn spoken commands into OS-level keyboard shortcuts!
Comments 0
No comments yet. Be the first!
Sign in to join the conversation.