- during first-boot setup, at the AI provider step;
- afterwards, in Settings → AI Provider, headed Hermes models.
Providers in the panel
Sign in
Anthropic (Claude) · OpenAI (Codex) · GitHub Copilot · Nous Portal
Paste an API key
OpenRouter · Anthropic · Google Gemini · z.ai / GLM · Kimi
The panel is a deliberate shortlist, not the whole of what Hermes supports.
Hermes knows a much longer provider registry; anything not listed here stays
reachable from the Hermes dashboard on the desktop. Configure it there and
it shows up in the ClawBox panel and chat header afterwards.
Connecting a provider
1
Pick the provider
Select its row. The panel loads the models that provider offers.
2
Sign in, or paste a key
For a sign-in provider, use Sign in ↗. That opens Hermes’ own
authentication page in a new tab — the flow is Hermes’, not a ClawBox
imitation of it. Finish there and come back; the panel notices and refreshes.For a key provider, paste the key into the API key field. The key is
stored by Hermes on the device and is never displayed back to you.
3
Choose a default model
Pick from Default model. The list is what that provider actually offers
on your account, so it is short until you have signed in or added a key.
4
Save
Save model & provider. The key, if you entered one, is stored first —
it is what unlocks the model list.
The on-device model
A Hermes ClawBox can also run a model locally, with nothing leaving the box. It is configured separately, in Settings → Local AI, because it works the same way on both editions.- It appears in the provider and chat pickers as Gemma 4 (on-device).
- It runs with on-demand standby: the model wakes on the first request and sleeps again afterwards to give the RAM back. The first message after a quiet period is therefore slower than the ones that follow.
- It has no reasoning-effort control — the backend does not take one, so the chat header hides the dial while it is selected.
- Turning it on makes it available; it does not take over. Making it your default is a separate, explicit choice.
ClawBox Connect (Jetson Orin Nano, 8 GB) runs a small model comfortably and is
best paired with a cloud provider for heavier work. See
Hardware.
Reasoning effort
Hermes takes a reasoning level per request. In the ClawBox chat header the brain icon offers eight: None, Minimal, Low, Medium, High, X-High, Max, Ultra. That control is a per-conversation override — it does not rewrite the device default. Two details worth knowing:- Not every provider accepts every level. Choosing a provider that tops out lower clamps your selection down to the nearest level it supports rather than failing.
- Providers with no reasoning dial at all — the on-device model among them — hide the control entirely.

