Switching to cloud
Open the Models page from the sidebar and go to the speech-to-text section. Switch from “On my device” to “Cloud”. Choose a provider, paste your API key, and pick a model. The local model stays on disk but is unloaded while cloud is active. The same group also appears in Settings → Dictation under “Where transcription runs”.Providers
Eight providers ship configured. You can also point the Custom entry at any OpenAI-compatible transcription server.
API keys are stored in your system keychain. Each provider has a “Get a key” link in Settings that goes to its key page.
The model field is free-text, so a model released after your build still works. You can also load the full list from the provider.
Streaming
ElevenLabs and Deepgram support live streaming: text arrives while you are still talking, so releasing the keys only flushes the last few words. Turn this on with the Transcribe as I speak switch. For ElevenLabs, you need a realtime model (likescribe_v2_realtime). The batch model (scribe_v2) only transcribes after you stop.
If the stream fails, SpeakoFlow falls back to a full batch transcription. The batch and realtime models can produce slightly different wording.