Skip to main content

LLM providers and prompts

How to connect is in Connect an LLM. This page collects the differences that help you choose.

Providers​

ProviderModelsNotes
Amazon BedrockNova Micro, Lite, Pro, Premier; Claude Haiku 4.5, Sonnet 5, Opus 5One key for Nova and Claude. Round trips for short input were Haiku 1.9s, Sonnet 2.7s, Opus 3.3s (us-east-1).
Anthropic ClaudeFour Claude modelsFaster than the same models through Bedrock: Haiku 1.3s, Sonnet 1.7s, Opus 2.7s.
Google GeminiDefault 3.5 Flash-LiteThe most stable on the free tier. Often leaves out periods.
OpenAI (GPT)GPT modelsUses the Codex CLI sign-in (codex login).
Groq CloudGroq modelsUses the same key as Groq STT.
Local MLXA downloaded modelText never leaves your Mac.

These timings were measured in one development environment and vary with network and region.

Correction modes​

ModeWhat it doesOriginal protection
Standard (STT Correction)Fixes misrecognitions, spacing, and English term spellingIf more than 50% of words change, the original is used.
Filler RemovalRemoves fillersSame as above.
StructuredOrganizes into lists and paragraphsNot applied, since changing the shape is expected.
CustomUses the prompt template you chooseNot applied.

After correction, English terms that were in the original but disappeared are restored, and the Dictionary corrections are applied once more.

Prompt templates​

Edit templates in Settings → LLM. There are three built-in templates:

TemplateUse
General correctionPolishes dictation.
LLM promptTurns a request you described out loud into a structure that works well when pasted into an LLM.
Meeting notesUsed for the final pass of the minutes output.
  • Set a default template for dictation and one for minutes.
  • Preview results with sample input.
  • Deleting a template asks for confirmation. Deleted built-in templates come back with Restore built-ins; user templates cannot be recovered.
  • Updating the app does not change built-in template text you already have. To get the new text, reset the template to its default in the editor.

What is sent to the LLM​

Correction sends two things:

  1. A system prompt: the mode or template text, input-handling instructions, Dictionary corrections, and the active word list
  2. The text to correct

Sogon doesn't send previous utterances or conversation history, and it doesn't send screen region images to the LLM either.