The right AI model for every task
Siesta AI routes every task to the right model and lets teams optimize for quality, cost, or balanced performance without changing the user experience.
See routing optionsChoose the outcome, not the model
Set routing policies for quality, cost, or balanced performance. Siesta AI evaluates each task and routes it to the most suitable model, while keeping governance, deployment, and regional controls aligned with your Azure environment.
Quality mode
Prioritize stronger reasoning models for complex analysis, coding, and high-stakes work.
Cost mode
Route routine prompts to faster, more efficient models to keep spend under control.
Balanced mode
Use a default policy that balances quality, latency, and cost for everyday work.
Model availability follows Microsoft Foundry Model Router support. Routing mode and selected model subset determine which models are eligible for each request.
All models in one place
Compare providers, regions, and model capabilities for your use case.
These models are eligible for Microsoft Foundry Model Router, so Siesta AI can route each request dynamically based on routing mode and selected model subset.
| Model | Regions | Context window | Best for |
|---|---|---|---|
| OpenAI | |||
|
|
|
128 000 | Multimodal |
|
|
|
128 000 | Speed & Efficiency |
|
|
|
1 047 576 | Coding |
|
|
|
1 047 576 | Speed & Efficiency |
|
|
|
1 047 576 | Speed & Efficiency |
|
|
|
200 000 | Reasoning |
|
|
|
400 000 | Speed & Efficiency |
|
|
|
400 000 | Speed & Efficiency |
|
|
|
400 000 | General Purpose |
|
|
|
128 000 | Conversations |
|
|
|
400 000 | Reasoning |
|
|
|
128 000 | Conversations |
|
|
|
128 000 | Conversations |
|
|
|
400 000 | Speed & Efficiency |
|
|
|
400 000 | Speed & Efficiency |
|
|
|
1 050 000 | General Purpose |
|
|
|
1 050 000 | General Purpose |
|
|
|
131 072 | Enterprise Workflows |
| Anthropic | |||
|
|
|
200 000 | Speed & Efficiency |
|
|
|
200 000 | Coding |
|
|
|
200 000 | Deep Research |
|
|
|
200 000 | Reasoning |
|
|
|
200 000 | Deep Research |
| DeepSeek | |||
|
|
|
128 000 | Data Analysis |
|
|
|
128 000 | Data Analysis |
| Meta | |||
|
|
|
1 000 000 | Long Context |
| xAI | |||
|
|
|
262 000 | Reasoning |
|
|
|
128 000 | Reasoning |
Router eligibility follows Microsoft Foundry Model Router support. Routing mode and selected subset determine which models can be used for each request.
Popular models Siesta AI can deploy or select directly through Azure AI Foundry when a team wants a specific model outside dynamic routing.
| Model | Regions | Context window | Best for |
|---|---|---|---|
|
|
|
1 000 000 | Multimodal |
|
|
|
1 000 000 | Speed & Efficiency |
|
|
|
1 000 000 | Reasoning |
| Mistral | |||
|
|
|
128 000 | General Purpose |
|
|
|
128 000 | General Purpose |
|
|
|
256 000 | Coding |
| Anthropic | |||
|
|
|
200 000 | Coding |
|
|
|
200 000 | Reasoning |
|
|
|
200 000 | General Purpose |
| DeepSeek | |||
|
|
|
128 000 | Reasoning |
|
|
|
128 000 | Speed & Efficiency |
|
|
|
128 000 | Data Analysis |
| xAI | |||
|
|
|
262 000 | Reasoning |
|
|
|
128 000 | Reasoning |
|
|
|
128 000 | Speed & Efficiency |
| Meta | |||
|
|
|
10 000 000 | Long Context |
|
|
|
128 000 | General Purpose |
| Moonshot AI | |||
|
|
|
256 000 | Coding |
|
|
|
256 000 | General Purpose |
|
|
|
256 000 | Long Context |
| Microsoft | |||
|
|
|
16 000 | Speed & Efficiency |
|
|
|
128 000 | Speed & Efficiency |
| Cohere | |||
|
|
|
256 000 | Enterprise Workflows |
Direct models are curated catalog options. They can be deployed or selected directly, but they are not claimed as part of dynamic Model Router selection.
Start building with Siesta AI
See how Siesta AI can power your use case with the right model.
Book a demo