OpenInspect
Models and agent harnesses

Choosing a model and reasoning effort

Enable models for your workspace, pick a model and effort in the composer, and understand which credential each provider needs.

Last reviewed View as MarkdownEdit on GitHubGive feedback

This page shows you where models are turned on, how to pick a model and reasoning effort for a session, and what each provider needs before a session can use it.

Settings, Models: the Anthropic section with a toggle per model and an Enable all link
Settings › Models decides which models appear in the model picker.

Enable models in Settings › Models

Settings › Models lists every model in the catalog, grouped by provider, under the heading Enabled Models. Toggle a model on or off, or use Enable all and Disable all on a group. Changes save automatically and apply to the model selector across the web app and the Slack bot. You cannot disable the last enabled model.

GroupEnabled by default
AnthropicYes
OpenAIYes
OpenCode ZenNo
OpenCode GoNo
xAI / SuperGrokNo
Z.AI Coding PlanNo
DeepSeekNo

The deployment default model is Claude Sonnet 4.6 (anthropic/claude-sonnet-4-6). It is used when nothing else (a session, an integration setting, or an automation) picks a model.

Pick a model and effort in the composer

The composer shows one control that reads <harness>: <model> <effort>. Open it to see up to three rows:

RowWhat it does
AgentChooses the harness (OpenCode or Claude Agent). Shown only when you create a session; the harness is fixed afterwards. See Agent harnesses.
ModelLists the enabled models the chosen harness can run, grouped by provider, with each model's display name and one-line description.
EffortLists Default plus the efforts the model supports (None, Low, Medium, High, XHigh, Max). Hidden for models without reasoning controls.

Default effort means the model's own default from the table below. Choosing a specific effort applies it to the prompts you send.

You can change the model and effort for each follow-up; the harness cannot change. The list therefore only contains models the session's harness can run. If a message override names a model the harness cannot run, the prompt is rejected with an error such as Model "openai/gpt-5.5" cannot run on the Claude Agent harness. It is never silently replaced with another model.

Model catalog

Reasoning efforts and defaults below come from the shared model catalog. A model with no reasoning efforts has no Effort row in the composer.

Anthropic

Runs on both harnesses. Credential: ANTHROPIC_API_KEY as a global secret, or on Claude Agent sessions a connected Claude account (see Provider accounts).

Model IDDisplay nameReasoning effortsDefault effort
anthropic/claude-haiku-4-5Claude Haiku 4.5high, maxmax
anthropic/claude-sonnet-4-5Claude Sonnet 4.5high, maxmax
anthropic/claude-sonnet-4-6Claude Sonnet 4.6low, medium, high, maxhigh
anthropic/claude-sonnet-5Claude Sonnet 5low, medium, high, xhigh, maxhigh
anthropic/claude-opus-4-5Claude Opus 4.5high, maxmax
anthropic/claude-opus-4-6Claude Opus 4.6low, medium, high, maxhigh
anthropic/claude-opus-4-7Claude Opus 4.7low, medium, high, xhigh, maxhigh
anthropic/claude-opus-4-8Claude Opus 4.8low, medium, high, xhigh, maxhigh
anthropic/claude-opus-5Claude Opus 5low, medium, high, xhigh, maxhigh
anthropic/claude-opus-5-5Claude Opus 5.5low, medium, high, xhigh, maxhigh
anthropic/claude-fable-5Claude Fable 5low, medium, high, xhigh, maxhigh
anthropic/claude-fable-5-1Claude Fable 5.1low, medium, high, xhigh, maxhigh

OpenAI

Runs on OpenCode only. Credential: a connected ChatGPT account, or OPENAI_API_KEY in the session's secret scope when the session uses API-key mode.

Model IDDisplay nameReasoning effortsDefault effort
openai/gpt-5.4GPT 5.4none, low, medium, high, xhighNot set
openai/gpt-5.5GPT 5.5none, low, medium, high, xhighNot set
openai/gpt-5.6-solGPT 5.6 Solnone, low, medium, high, xhighmedium
openai/gpt-5.6-terraGPT 5.6 Terranone, low, medium, high, xhighmedium
openai/gpt-5.6-lunaGPT 5.6 Lunanone, low, medium, high, xhigh, maxmedium
openai/gpt-6-astraGPT-6 Astralow, medium, high, xhigh, maxmedium
openai/gpt-6-solGPT-6 Solnone, low, medium, high, xhigh, maxmedium
openai/gpt-6-lunaGPT-6 Lunanone, low, medium, high, xhigh, maxmedium
openai/gpt-5.3-codexGPT 5.3 Codexlow, medium, high, xhighhigh
openai/gpt-5.3-codex-sparkGPT 5.3 Codex Sparklow, medium, high, xhighhigh

"Not set" means the composer shows Default and sends no effort unless you choose one.

xAI / SuperGrok

Runs on OpenCode only. Opt-in group. Credential: a connected SuperGrok account, or XAI_API_KEY in API-key mode. Entitlement to a given Grok model is decided by xAI for the connected account.

Model IDDisplay nameReasoning effortsDefault effort
xai/grok-4.5Grok 4.5low, medium, highhigh
xai/grok-4.6Grok 4.6low, medium, high, xhighhigh
xai/grok-4.7Grok 4.7low, medium, high, xhighhigh
xai/grok-build-0.1Grok Build 0.1Not configurableN/A

OpenCode Zen

Runs on OpenCode only. Opt-in group. Credential: OPENCODE_API_KEY as a global or repository secret. Zen is billed per token.

Model IDDisplay nameReasoning effortsDefault effort
opencode/kimi-k2.5Kimi K2.5Not supportedN/A
opencode/kimi-k2.6Kimi K2.6Not supportedN/A
opencode/kimi-k3Kimi K3Not supportedN/A
opencode/minimax-m2.5MiniMax M2.5Not supportedN/A
opencode/qwen3.7-maxQwen3.7 MaxNot supportedN/A
opencode/glm-5GLM 5Not supportedN/A
opencode/glm-5.1GLM 5.1Not supportedN/A
opencode/glm-5.2GLM 5.2Not supportedN/A

OpenCode Go

Runs on OpenCode only. Opt-in group. Credential: the same OPENCODE_API_KEY secret as Zen, but the key must carry an active OpenCode Go subscription. Go bills against the subscription's rolling allowance instead of per token, so a key without a Go subscription fails on opencode-go/* models while opencode/* models keep working.

Rolling allowance and unattended sessions

Go usage is capped on three rolling windows: 20% of the monthly allowance per 5 hours, 50% per week, and 100% per month. Allowances differ per model. An unattended session pinned to a Go model can exhaust its window and fail mid-run, so keep Go models off Slack, GitHub, Linear, and automation launches unless you accept that.

Model IDDisplay nameReasoning effortsDefault effort
opencode-go/grok-4.6Grok 4.6Not supportedN/A
opencode-go/gpt-5.6-lunaGPT 5.6 LunaNot supportedN/A
opencode-go/glm-5.3-flashGLM 5.3 FlashNot supportedN/A
opencode-go/glm-5.3GLM 5.3Not supportedN/A
opencode-go/glm-5.2GLM 5.2Not supportedN/A
opencode-go/glm-5.1GLM 5.1Not supportedN/A
opencode-go/kimi-k3Kimi K3Not supportedN/A
opencode-go/kimi-k2.7-codeKimi K2.7 CodeNot supportedN/A
opencode-go/kimi-k2.6Kimi K2.6Not supportedN/A
opencode-go/longcat-2.0LongCat 2.0Not supportedN/A
opencode-go/deepseek-v4.1-flashDeepSeek V4.1 FlashNot supportedN/A
opencode-go/deepseek-v4-proDeepSeek V4 ProNot supportedN/A
opencode-go/deepseek-v4-flashDeepSeek V4 FlashNot supportedN/A
opencode-go/deepseek-v4-flash-vision-expDeepSeek V4 Flash Vision ExpNot supportedN/A
opencode-go/mimo-v2.5MiMo V2.5Not supportedN/A
opencode-go/mimo-v2.5-proMiMo V2.5 ProNot supportedN/A
opencode-go/minimax-m3MiniMax M3Not supportedN/A
opencode-go/minimax-m2.7MiniMax M2.7Not supportedN/A
opencode-go/muse-spark-1.3-contributorMuse Spark 1.3 ContributorNot supportedN/A
opencode-go/muse-spark-1.2-contributorMuse Spark 1.2 ContributorNot supportedN/A
opencode-go/qwen3.8-maxQwen3.8 MaxNot supportedN/A
opencode-go/qwen3.8-flashQwen3.8 FlashNot supportedN/A
opencode-go/qwen3.7-maxQwen3.7 MaxNot supportedN/A
opencode-go/qwen3.7-plusQwen3.7 PlusNot supportedN/A
opencode-go/qwen3.6-plusQwen3.6 PlusNot supportedN/A
opencode-go/hy4-previewHy4 PreviewNot supportedN/A
opencode-go/hy3Hy3Not supportedN/A

Go also publishes minimax-m2.5, but the pinned OpenCode release does not resolve it on the Go gateway, so it is left out of the catalog. Use opencode/minimax-m2.5 on Zen instead.

Z.AI Coding Plan

Runs on OpenCode only. Opt-in group. Credential: ZHIPU_API_KEY as a global or repository secret.

Model IDDisplay nameReasoning effortsDefault effort
zai-coding-plan/glm-5.2GLM 5.2Not supportedN/A
zai-coding-plan/glm-5.3GLM 5.3Not supportedN/A

DeepSeek

Runs on OpenCode only. Opt-in group. Credential: DEEPSEEK_API_KEY as a global or repository secret.

Model IDDisplay nameReasoning effortsDefault effort
deepseek/deepseek-v4-flashDeepSeek V4 FlashNot supportedN/A
deepseek/deepseek-v4-proDeepSeek V4 ProNot supportedN/A

The same model behind several gateways

Some models appear more than once because they are reachable through more than one provider, each billed against a different credential. opencode-go/grok-4.6 and xai/grok-4.6 are the same Grok model behind two gateways. opencode-go/glm-5.2, opencode/glm-5.2, and zai-coding-plan/glm-5.2 are the same GLM model behind three. Pick the entry whose billing you want.

Defaults for integrations and automations

Sessions started by Slack, GitHub, and Linear always run on the OpenCode harness. Automations run on the harness chosen in the automation editor (OpenCode by default). Each surface picks its model from these settings:

SurfaceWhere the default lives
GitHubSettings › Integrations › GitHub: Default model and Default reasoning effort, plus per-repository Repository Overrides. Falls back to the deployment default model.
LinearSettings › Integrations › Linear: Default model and effort and Repository Overrides. A model:* issue label wins when Allow model labels is on, then a Linear user preference (when allowed), then the repository override or global default, then the deployment default.
SlackEach Slack user sets Model and Reasoning effort in the app's Home tab. A request can start with !model and !reasoning flags to override them; on a request that starts a session the flags become that session's defaults.
AutomationsThe automation editor picks a model and effort per automation. The list is limited to enabled models the automation's harness can run.

All of these selectors show only models enabled in Settings › Models. Reasoning values must be ones the selected model supports.

@OpenInspect !model anthropic/claude-sonnet-4-6 !reasoning max investigate the flaky test

Troubleshooting

Next steps

On this page