Chutes
Explore Chutes models on TrustedRouter with current routes, token pricing, policy sources, privacy posture, regional availability, and API support.
Chuteschutes
Confidential| Provider | Chutes |
|---|---|
| Routing status | Active |
| Provider website | https://chutes.ai/ |
| Models | 10 public models |
| Credits routes | 10 |
| Zero data retention | yes All routes |
| Confidential compute | yes |
| Provider E2EE | yes |
| Policy note | TrustedRouter encrypts each request to an attested Chutes workload and verifies Intel TDX plus NVIDIA GPU attestation inside the TrustedRouter enclave before sending content. Verification fails closed. Chutes also documents no prompt/output storage or training. Policy source |
Measured performance
46 samplesContinuously sampled across Chutes's routed models: p50 TTFT, effective throughput, and success rate. Effective throughput uses provider-reported output tokens over complete request time. Unsupported route and probe-configuration rows are separated from provider downtime. No prompt or output content stored.
| p50 TTFT | 8083 ms |
|---|---|
| Effective throughput | 16 tok/s n=3 |
| Uptime | 41.30% |
| Model | p50 TTFT | Effective throughput | Uptime | Config excluded | Availability samples |
|---|---|---|---|---|---|
| qwen/qwen3.5-397b-a17b | 5877 ms | — | 60.00% | — | 5 |
| z-ai/glm-5.1 | 6697 ms | — | 75.00% | — | 4 |
| deepseek/deepseek-v3.2 | 8083 ms | — | 100.00% | — | 7 |
| moonshotai/kimi-k2.6 | 9389 ms | 16 tok/s n=2 | 100.00% | — | 6 |
| google/gemma-4-31b-turbo | — | — | 0.00% | — | 4 |
| mistralai/mistral-nemo | — | — | 0.00% | — | 2 |
| qwen/qwen3-235b-a22b-thinking-2507 | — | — | 0.00% | — | 5 |
| qwen/qwen3-32b | — | — | 0.00% | — | 4 |
| qwen/qwen3.6-27b | — | — | 0.00% | — | 9 |
| z-ai/glm-5.2 | — | 10 tok/s n=1 | — | — | 0 |
Chutes performance history · Full provider & model leaderboard.
Models served by Chutes.
Each row links to pricing, provider, benchmark, and API pages for the model.
| Model | AI IQ | Context | Input | Cached input | Output |
|---|---|---|---|---|---|
deepseek/deepseek-v3.2DeepSeek: DeepSeek V3.2 |
IQ 103#74 | 163,840 | $1.055/1M | $0.1055/1M | $1.055/1M |
google/gemma-4-31b-turbogoogle/gemma-4-31B-turbo |
— | 131,072 | $0.1266/1M | $0.01266/1M | $0.39035/1M |
mistralai/mistral-nemoMistral: Mistral Nemo |
— | 131,072 | $0.025848/1M | $0.01/1M | $0.103179/1M |
moonshotai/kimi-k2.6MoonshotAI: Kimi K2.6 |
IQ 119#30 | 262,144 | $0.6119/1M | $0.06119/1M | $3.587/1M |
qwen/qwen3-235b-a22b-thinking-2507Qwen: Qwen3 235B A22B Thinking 2507 |
— | 131,072 | $0.31534/1M | $0.031534/1M | $1.261464/1M |
qwen/qwen3-32bQwen: Qwen3 32B |
— | 40,960 | $0.10972/1M | $0.010972/1M | $0.43888/1M |
qwen/qwen3.5-397b-a17bQwen: Qwen3.5 397B A17B |
— | 262,144 | $0.47475/1M | $0.047475/1M | $3.165/1M |
qwen/qwen3.6-27bQwen: Qwen3.6 27B |
IQ 111#56 | 262,144 | $0.3165/1M | $0.03165/1M | $2.11/1M |
z-ai/glm-5.1Z.ai: GLM 5.1 |
IQ 114#43 | 200,000 | $1.0339/1M | $0.10339/1M | $3.2494/1M |
z-ai/glm-5.2Z.ai: GLM 5.2 |
IQ 120#27 | 1,048,576 | $1.31875/1M | $0.131875/1M | $4.16725/1M |
Questions
Does Chutes have zero data retention?
TrustedRouter records Chutes as supporting provider-level zero data retention based on the policy source linked on this page. This is a provider policy claim, separate from TrustedRouter's content-stateless real-time gateway and from end-to-end confidential compute.
Is Chutes end-to-end encrypted?
TrustedRouter records Chutes as supporting provider-side confidential compute and end-to-end encrypted inference. The route-specific model page shows whether that protection applies to a particular endpoint.
Which Chutes models are available through TrustedRouter?
This page currently lists 10 public Chutes models, with live pricing, route count, context length, measured performance when available, and links to each model's provider and benchmark pages.