The right model per task
Chat, code, embedding and image: for each task you pick the model that fits – open or frontier, from twelve providers.
Models
Open models such as Qwen, Mistral and gpt-oss run on your appliance, fully offline if you wish. Frontier models from OpenAI and Anthropic are yours EU hosted when your plan includes them.
Mistral Large 3
Mistral Medium 3.5
Mistral Small 4
Ministral 3 8B
Devstral Small 2
Magistral Small 1.2
Codestral
Mistral Embed
Qwen3.6-35B-A3B
Qwen3.6-27B
Qwen3-Coder-30B-A3B
Qwen3-30B-A3B-2507
Qwen3-14B
Qwen3-8B
Qwen3 Embedding 8B
DeepSeek V4-Pro
DeepSeek V4-Flash
Claude Fable 5.1
Claude Fable 5
Claude Opus 5.5
Claude Opus 5
Claude Opus 4.8
Claude Sonnet 5
Claude Haiku 4.5
Gemini 3.1 Pro
Gemini 3.5 Flash
Gemini 3.1 Flash-Lite
Gemini Embedding 2
Gemma 4
Gemma 3 27B
GPT-5.6 Sol
GPT-5.6 Terra
GPT-5.6 Luna
GPT-5.5
GPT-5.4
GPT-5.4 mini
GPT-5.1
GPT-5
GPT Image 2
gpt-oss-120b
gpt-oss-20b
Llama 4 Scout
Llama 3.3 70B
Llama 3.1 8B
Phi-4
Phi-4-reasoning-plus
Llama 3.3 Nemotron Super 49B
Nemotron Nano 9B v2
GLM-5.2
Kimi K2.6
MiniMax-M3
51 of 51 models
Models marked available offline also run on your appliance, without internet. OpenAI, Anthropic and Mistral AI compute in EU tenants; global models only once approved per tenant. Which models you can use is set by your plan and your offer.
Missing a model?
We add further open models to your profile after a licence and quality review. Tell us which model you need.
Chat, code, embedding and image: for each task you pick the model that fits – open or frontier, from twelve providers.
Run on your appliance, fully offline if you wish, or in European data centres.
Your applications talk to one stable interface. Which model answers is configured and logged on every request.
Context window, maximum output and operating route are out in the open. An open model only changes its build once you approve an update.
Operating routes
A model name says nothing about where your request is processed. So you choose the operating route first; the models that fit follow from it.
Open models such as Qwen3.6, Mistral Small 4, Gemma 4, Llama and gpt-oss run on the DGX appliance in your building. No external model interface.
The same platform in European data centres, with no hardware of your own. Plus models from OpenAI, Anthropic and Mistral AI in EU tenants.
Models that run neither offline nor in the EU, such as Gemini, DeepSeek, GLM, Kimi and MiniMax. Once your administrator has enabled them, your users decide for themselves whether to use them.
Common questions
You decide. On the appliance an open model such as Qwen3.6 or gpt-oss-120b does the work; which one depends on your tasks and how many people work at once, and is recorded in the offer. In European On Demand you can add frontier models.
No. The models run as pinned versions and are not further trained on your content. What the platform learns from your documents sits as sourced knowledge in your knowledge layer, not in model weights.
Yes, if it has open weights and a licence that allows commercial use. We review licence, quality and memory needs on the reference appliance and then add it to your profile.
In European On Demand, yes, depending on the plan. OpenAI and Anthropic models run in the vendors' EU tenants; what goes there is the single request with selected excerpts, not your collection. Gemini computes globally; once your administrator has enabled it, users decide for themselves whether to use it. On the appliance, frontier models are an approved exception.
As reviewed packages. A new version goes through the same qualification as the first model and only replaces the old build once you approve it. Nothing is fetched automatically.
The right profile
Tell us your tasks and how many people work at once. We propose the model profile and record it in the offer.