Models
Custodos offers AI models from several providers in one interface. This page shows how to choose a model in the chat, what the “Auto” entry means and how to tell what a model can do.
Last checked on 28 September 2026
On this page
Choosing a model
You choose the model in the chat's message field. You can change it at any time, even in the middle of a chat. Every reply shows which model wrote it.
- 1
Open the model picker
Click the name of the current model in the message field, or Choose model. On a phone the button shows only a logo, without the name. - 2
Compare models
Hover over an entry or move to it with the keyboard. An info card about that model appears beside the list. - 3
Pick the model
Click the model or Auto. The current choice is highlighted in the list, and your next message goes to this model.
Auto: your workspace default
Auto sits at the top of the model picker. Behind it is the default model an admin has set for new chats. If nobody has set one, Custodos uses its own default model. For that it picks a model that can use tools, as long as your workspace allows one.
The line under Auto tells you which model it currently stands for. If you have no special requirements, stay on Auto. Another model can be worth it for very long documents, for demanding analyses or when you want to compare two replies. To compare, use Regenerate with … on the latest reply, see Chat basics.
The model groups
The model picker sorts the models into groups. The group tells you who developed the model and where it runs.
| Group | What it contains | Where it runs |
|---|---|---|
| Frontier | The most capable models from major US providers, such as Claude, Gemini and GPT | in data centres in the EU or Switzerland |
| Open models | Models with open weights, such as those from Mistral, Qwen, MiniMax and GLM | on European infrastructure |
What the info card shows
The info card beside the list describes the model your mouse or keyboard is on. On narrow screens it is hidden.
- Provider and name, plus a short description of what the model is good for.
- Credit usage in three levels: Low, Medium or High, with the approximate credits per reply underneath. The figure assumes one question and one reply of average length. How credits work is explained under Credits and throttling.
- Context window in tokens: how much text the model can handle at once in a chat, attachments included.
- Hosting location: the country or region and the infrastructure provider the model runs on.
- Notes when a model thinks before answering or cannot use tools (see below).
Models without tools
Some models cannot use tools. They answer questions and read attached files, but they do not search the web or the Company Brain, and they do not create or edit files. The info card says so with the sentence No web or brain search, and it creates no files.
When a Brain is attached to the chat, these models are greyed out in the picker and cannot be chosen. The info card still explains why.
Models that think before answering
Some models think before they answer. Until the first text appears, the chat shows Thinking… for them. The thinking is billed like output, so these replies use more credits than their length suggests. The info card points this out for these models.
Available chat models
The table lists the chat models in the Custodos catalogue as of 28 September 2026. Which of them you see depends on your workspace's processing region and on what your admin has allowed. The model picker is always the authority.
| Model | Group | Reads images | Tools | Context window |
|---|---|---|---|---|
| Claude Opus 5 (Anthropic) | Frontier | yes | yes | 1M tokens |
| Claude Sonnet 5 (Anthropic) | Frontier | yes | yes | 1M tokens |
| Claude 3 Haiku (Anthropic) | Frontier | yes | yes | 200,000 tokens |
| Gemini 3.7 Flash (Google) | Frontier | yes | yes | 1M tokens |
| Gemini 3.5 Flash-Lite (Google) | Frontier | yes | yes | 1M tokens |
| Gemini 2.5 Pro (Google) | Frontier | yes | yes | 200,000 tokens |
| GPT-5.6 Sol, Terra and Luna (OpenAI) | Frontier | yes | yes | 272,000 tokens |
| GPT-5.4 Mini (OpenAI) | Frontier | yes | yes | 400,000 tokens |
| Mistral Pixtral Large (Mistral AI) | Open models | yes | no | 128,000 tokens |
| Qwen3 235B and Qwen3 Coder 30B (Qwen) | Open models | no | no | 256,000 tokens |
| MiniMax M2.5 (MiniMax) | Open models | no | yes | 192,000 tokens |
| GLM 4.7 Flash (Z.ai) | Open models | no | no | 128,000 tokens |
Who decides which models there are
Admins use the Models area to set the region the workspace is processed in, which models are allowed and which model new chats get. Members choose freely among the allowed models in the chat. How this works is described under Managing models.
If no model is allowed yet, the picker shows No model enabled yet. Ask an admin to enable one.
Read next
Still have a question? Write to us.
