The situation
Raw transcription is only part of useful dictation. Spoken text still needs punctuation, formatting, and the right names—and it needs to arrive in the application where the work is happening. Anna brings that complete interaction together.
What we built
Hold a hotkey, speak, and release. Anna captures audio, transcribes it, cleans up the text, and inserts the result into the focused app. We combined existing speech models with a fast language model, then built the native desktop behavior, onboarding, billing, and update system around them.