Back to docsCore concepts

Models

Picking the reasoning engine for a run — the per-thread model picker, plan limits, and provider routing.

Last updated October 5, 2026 · 4 min read

A model is the reasoning engine behind a run. Okou handles the tools, the context, and the procedure; the model does the thinking. Which one is in play affects quality, speed, and how many credits a run consumes.

The model lives on the chat thread

The model picker sits beside the chat composer and controls the model for that chat. The selection is stored on the thread, not on the agent and not on the workflow.

  • A new chat starts from your member default, then the workspace default if you haven't set one. The system default is Auto.
  • Changing the model mid-chat applies to the next run in the thread; a run already in progress finishes on the model it started with.
  • An automation chat picks up its model when it is first created, and later firings reuse it.
  • In Slack, the /okou model command binds the model for the next thread you start — see Slack.

Agents and workflows deliberately have no model setting of their own. The thread is the single place to look.

What's available

The current catalogue, with each model's strengths and relative cost, is on the Models page. It covers reasoning and image models and changes as models and provider routes become available.

Availability depends on the current catalogue, your plan, workspace policy, and provider route:

  • Free can use the built-in models enabled for Free in the current catalogue.
  • Paid plans can access additional models, subject to the catalogue, workspace policy, and available provider connections.

Workspace admins choose which of those models members can actually pick, and which one is the workspace default, under Settings → Models.

Where the inference runs

A model can be routed more than one way, and the route is set by workspace policy rather than by the picker:

  • Built-in — Okou runs the model and it bills against workspace credits. Nothing to connect.
  • Your own subscription — each member connects their own Claude or Codex account and their runs go through it.
  • A shared provider key or gateway — an admin supplies the credential once for the workspace.

An eligible, connected personal Claude or Codex subscription can be used on Free and paid plans. Shared organization API keys and custom gateways require a paid plan. A subscription run uses that subscription and does not fall back to built-in credits. Tool calls, generation, and storage still consume Okou credits.

The picker chooses a model. It does not override the route policy has assigned to that model.

Choosing one

  • Match the model to the task, not to the price list. A cheap model on a long agentic run can cost more than an expensive one, because it takes more steps to get there.
  • Use a cheap model for volume. Triage, classification, and bulk pre-filtering rarely need a frontier model.
  • Use a frontier model for judgement. Long-horizon work, code, and anything where a wrong answer is expensive.
  • Set a member default for the model you want most chats to start on, rather than changing the picker every time.

Troubleshooting

What you seeWhat to do
Selected model is not availableThe thread's model isn't usable under the current workspace policy or provider connection. Pick another model in the composer, or reconnect the provider under Settings → Models.
A model you expect is missingCheck the current catalogue, your plan, workspace policy, and the available provider route under Settings → Models.
Credits draining faster than expectedCheck which model the thread is on. Rates differ substantially between tiers.

What's next

  • See Credits & billing for how model choice shows up on the bill.
  • See Chat for the thread the model is attached to.