Screen access
Every screenshot that goes to the model is shown in the conversation, in both active modes, so there is always a visible record of when the assistant looked.
What actually gets sent
SpeakoFlow captures the monitor under your mouse cursor, so on a multi-monitor desk you get the screen you are actually working on. Only a compact thumbnail is stored in the conversation and in History. The full-resolution frame is sent to the model once and never written to disk. Thumbnails show inline in the panel and can be clicked to enlarge.When the frame is grabbed for a voice question
When the frame is grabbed for a voice question
When I ask, When the message sends. Default: When I ask. Shown in Manual mode only.This changes the timing for voice questions only, where there is a real gap between starting to ask and finishing.
- When I ask (default). Grabs the frame the moment you press the hotkey, so it captures what you were looking at when you started talking, not whatever is on screen after you finish.
- When the message sends. Grabs it after you stop talking and the speech is transcribed.
How much image quality each provider gets
How much image quality each provider gets
The frame is JPEG-compressed down to a budget picked from your provider, because “send the sharpest possible image” and “the request succeeds” are not the same goal.
What happens if your model cannot read images
What happens if your model cannot read images
The panel says so and names the problem rather than silently dropping the screenshot. Switch to a vision-capable model, or ask again with screen vision off.