Skip to main content
OpenWhispr has an AI agent you reach two different ways, depending on what you want:
  • The voice agent types the result of an instruction at your cursor. Say “write a short thank-you note” and the note appears.
  • The chat agent opens a conversation window that can search your notes, write and edit them, search the web, and read your calendar.
They share settings but behave quite differently.

The voice agent

Press a key, speak an instruction, get the result typed at your cursor.

The chat agent

A conversation overlay with tools that work on your notes.

Your agent's name

Saying its name during dictation hands the rest over to the agent.

Which mode is running?

Dictation, agent, translation and meetings, told apart.

Two ways to reach the voice agent

This is the part that surprises people. The voice agent runs when either:
  1. You press the Voice Agent Hotkey, or
  2. You say the agent’s name at the start of an ordinary dictation.
The second is on by default, and the default name is OpenWhispr — so it’s possible to trigger the agent without meaning to. Your agent’s name covers how the detection works and how to rename or disable it.

Where the settings live

Everything except the hotkeys is under SettingsLanguage Models under AI Models: The hotkeys themselves are under SettingsHotkeys under AppVoice Agent Hotkey and Chat Agent Hotkey are separate entries.

Where it can run

Both agents run on whichever provider you choose: OpenWhispr Cloud with no setup, your own API key with OpenAI, Anthropic, Gemini or Groq, a local model on your own device, a self-hosted endpoint, or an enterprise provider such as AWS Bedrock, Azure OpenAI or Google Vertex. A local model keeps everything on your machine. See cloud vs local and enterprise providers.