OpenAI compatible API · Attested · Public status
Chutes
Chutes models on TrustedRouter with prices, routes, policy notes, and source links.
Onebase URL to migrate
100sof models and routes
Noneprompt logs by default
chutes
No logs
| Provider | Chutes |
|---|---|
| Models | 13 public models |
| Prepaid routes | 13 |
| BYOK routes | 13 |
| Zero data retention | yes |
| Confidential compute | yes |
| Provider E2EE | no |
| Policy note | Chutes documents no prompt/output storage or training and serves these routes in confidential-compute TEEs. Standard API calls are not marked provider end-to-end encrypted. Policy source |
Measured performance
191 samplesContinuously sampled across Chutes's routed models — p50 TTFT, throughput, and success rate. Unsupported route and probe-configuration rows are separated from provider downtime. No prompt or output content stored.
| p50 TTFT | 5592 ms |
|---|---|
| Throughput | — |
| Uptime | 91.10% |
| Model | p50 TTFT | p50 TTFB | Throughput | Uptime | Config excluded | Samples |
|---|---|---|---|---|---|---|
| qwen/qwen3.5-397b-a17b | 3152 ms | 3152 ms | — | 100.00% | — | 22 |
| mistralai/mistral-nemo | 3760 ms | 3760 ms | — | 100.00% | — | 23 |
| qwen/qwen3.6-27b | 4194 ms | 4194 ms | — | 100.00% | — | 14 |
| google/gemma-4-31b-turbo | 4461 ms | 4461 ms | — | 100.00% | — | 17 |
| moonshotai/kimi-k2.6 | 5445 ms | 5445 ms | — | 87.50% | — | 16 |
| z-ai/glm-5.2 | 5592 ms | 5592 ms | — | 100.00% | — | 14 |
| qwen/qwen3-32b | 6262 ms | 6261 ms | — | 100.00% | — | 21 |
| qwen/qwen3-235b-a22b-thinking-2507 | 7616 ms | 7616 ms | — | 100.00% | — | 13 |
| z-ai/glm-5 | 7638 ms | 7638 ms | — | 50.00% | — | 6 |
| deepseek/deepseek-v3.2 | 7864 ms | 7864 ms | — | 92.31% | — | 13 |
| minimax/minimax-m2.5 | 7886 ms | 7886 ms | — | 100.00% | — | 8 |
| z-ai/glm-5.1 | 14505 ms | 14504 ms | — | 83.33% | — | 12 |
| moonshotai/kimi-k2.5 | 18382 ms | 18382 ms | — | 25.00% | — | 12 |
Chutes performance history · Full provider & model leaderboard.
Provider models
Models served by Chutes.
Each row links to pricing, provider, benchmark, and API pages for the model.
| Model | AI IQ | Context | Endpoints | Prompt | Completion | Routes |
|---|---|---|---|---|---|---|
deepseek/deepseek-v3.2DeepSeek: DeepSeek V3.2 |
IQ 104#54 | 163,840 | 2 | $1.05/1M | $1.05/1M | prepaid BYOK |
google/gemma-4-31b-turbogoogle/gemma-4-31B-turbo |
— | 131,072 | 2 | $0.126/1M | $0.3885/1M | prepaid BYOK |
minimax/minimax-m2.5MiniMax: MiniMax M2.5 |
IQ 106#52 | 204,800 | 2 | $0.1575/1M | $1.26/1M | prepaid BYOK |
mistralai/mistral-nemoMistral: Mistral Nemo |
— | 131,072 | 2 | $0.025725/1M | $0.10269/1M | prepaid BYOK |
moonshotai/kimi-k2.5MoonshotAI: Kimi K2.5 |
IQ 111#39 | 262,144 | 2 | $0.462/1M | $2.1/1M | prepaid BYOK |
moonshotai/kimi-k2.6MoonshotAI: Kimi K2.6 |
IQ 119#19 | 262,144 | 2 | $0.693/1M | $3.675/1M | prepaid BYOK |
qwen/qwen3-235b-a22b-thinking-2507Qwen: Qwen3 235B A22B Thinking 2507 |
— | 262,144 | 2 | $0.313845/1M | $1.255485/1M | prepaid BYOK |
qwen/qwen3-32bQwen: Qwen3 32B |
— | 131,072 | 2 | $0.1092/1M | $0.4368/1M | prepaid BYOK |
qwen/qwen3.5-397b-a17bQwen: Qwen3.5 397B A17B |
— | 262,144 | 2 | $0.4725/1M | $3.15/1M | prepaid BYOK |
qwen/qwen3.6-27bQwen: Qwen3.6 27B |
IQ 112#38 | 262,144 | 2 | $0.315/1M | $2.1/1M | prepaid BYOK |
z-ai/glm-5Z.ai: GLM 5 |
IQ 106#50 | 204,800 | 2 | $0.9975/1M | $2.6775/1M | prepaid BYOK |
z-ai/glm-5.1Z.ai: GLM 5.1 |
IQ 113#30 | 204,800 | 2 | $1.029/1M | $3.234/1M | prepaid BYOK |
z-ai/glm-5.2Z.ai: GLM 5.2 |
IQ 120#16 | 1,048,576 | 2 | $1.47/1M | $4.62/1M | prepaid BYOK |