Models

Any model, and switching without losing anything.

Read as Markdown

Your agent runs on the model you choose. Its memory, archive and conversations live in your own database, not with the model, so you can switch whenever you like and it carries on where it was. By default that database is on your agent's machine; you can also connect a database server you manage elsewhere.

Providers

HowOptions
API keyAnthropic, OpenAI, Google, xAI, DeepSeek, Moonshot, Qwen, MiniMax, Z.AI, Xiaomi MiMo, OpenRouter
SubscriptionClaude, ChatGPT, Grok: connect by signing in
A model server you controlOllama, or another server with the same interface

The installer connects the first one. Add more in the admin app under Integrations, Inference. Each addition is tested before it is saved. Your provider bills you directly; nothing passes through us.

Two roles

Each agent scope has two default models:

  • Reasoning: the model your agent thinks and talks with.
  • Utility: a smaller, cheaper model for background work, such as condensing long conversations and indexing names.

Set them in the admin app under Models: r for reasoning, u for utility. New spaces start with those defaults. When you change them, the app can apply the change to existing spaces too; otherwise their settings stay as they were.

If the utility model stops answering, your agent moves to the next one that works, and the admin app shows it. It does not switch back automatically when the first model recovers: select it again in the admin app. A reasoning model is never used as the utility fallback.

One conversation on another model

Each space can run on its own model. Send /model in that conversation to pick one, or set it in the admin app under Spaces. /thinking, /effort, /budget and /maxoutput adjust how that model works, for that conversation only.