OpenAI compatible API · Attested · Public status

Chutes

Explore Chutes models on TrustedRouter with current routes, token pricing, policy sources, privacy posture, regional availability, and API support.

Verify gateway
Onebase URL to migrate
100sof models and routes
0prompt or output logs. Always.

Chuteschutes

Confidential

All providers

ProviderChutes
Routing statusActive
Provider websitehttps://chutes.ai/
Models10 public models
Credits routes10
Zero data retentionyes All routes
Confidential computeyes
Provider E2EEyes
Policy noteTrustedRouter encrypts each request to an attested Chutes workload and verifies Intel TDX plus NVIDIA GPU attestation inside the TrustedRouter enclave before sending content. Verification fails closed. Chutes also documents no prompt/output storage or training.
Policy source

Measured performance

46 samples

Continuously sampled across Chutes's routed models: p50 TTFT, effective throughput, and success rate. Effective throughput uses provider-reported output tokens over complete request time. Unsupported route and probe-configuration rows are separated from provider downtime. No prompt or output content stored.

p50 TTFT8083 ms
Effective throughput16 tok/s n=3
Uptime41.30%
Modelp50 TTFTEffective throughputUptimeConfig excludedAvailability samples
qwen/qwen3.5-397b-a17b 5877 ms 60.00% 5
z-ai/glm-5.1 6697 ms 75.00% 4
deepseek/deepseek-v3.2 8083 ms 100.00% 7
moonshotai/kimi-k2.6 9389 ms 16 tok/s n=2 100.00% 6
google/gemma-4-31b-turbo 0.00% 4
mistralai/mistral-nemo 0.00% 2
qwen/qwen3-235b-a22b-thinking-2507 0.00% 5
qwen/qwen3-32b 0.00% 4
qwen/qwen3.6-27b 0.00% 9
z-ai/glm-5.2 10 tok/s n=1 0

Chutes performance history · Full provider & model leaderboard.

Provider models

Models served by Chutes.

Each row links to pricing, provider, benchmark, and API pages for the model.

Model AI IQ Context Input Cached input Output
deepseek/deepseek-v3.2
DeepSeek: DeepSeek V3.2
IQ 103#74 163,840 $1.055/1M $0.1055/1M $1.055/1M
google/gemma-4-31b-turbo
google/gemma-4-31B-turbo
131,072 $0.1266/1M $0.01266/1M $0.39035/1M
mistralai/mistral-nemo
Mistral: Mistral Nemo
131,072 $0.025848/1M $0.01/1M $0.103179/1M
moonshotai/kimi-k2.6
MoonshotAI: Kimi K2.6
IQ 119#30 262,144 $0.6119/1M $0.06119/1M $3.587/1M
qwen/qwen3-235b-a22b-thinking-2507
Qwen: Qwen3 235B A22B Thinking 2507
131,072 $0.31534/1M $0.031534/1M $1.261464/1M
qwen/qwen3-32b
Qwen: Qwen3 32B
40,960 $0.10972/1M $0.010972/1M $0.43888/1M
qwen/qwen3.5-397b-a17b
Qwen: Qwen3.5 397B A17B
262,144 $0.47475/1M $0.047475/1M $3.165/1M
qwen/qwen3.6-27b
Qwen: Qwen3.6 27B
IQ 111#56 262,144 $0.3165/1M $0.03165/1M $2.11/1M
z-ai/glm-5.1
Z.ai: GLM 5.1
IQ 114#43 200,000 $1.0339/1M $0.10339/1M $3.2494/1M
z-ai/glm-5.2
Z.ai: GLM 5.2
IQ 120#27 1,048,576 $1.31875/1M $0.131875/1M $4.16725/1M

Questions

Does Chutes have zero data retention?

TrustedRouter records Chutes as supporting provider-level zero data retention based on the policy source linked on this page. This is a provider policy claim, separate from TrustedRouter's content-stateless real-time gateway and from end-to-end confidential compute.

Is Chutes end-to-end encrypted?

TrustedRouter records Chutes as supporting provider-side confidential compute and end-to-end encrypted inference. The route-specific model page shows whether that protection applies to a particular endpoint.

Which Chutes models are available through TrustedRouter?

This page currently lists 10 public Chutes models, with live pricing, route count, context length, measured performance when available, and links to each model's provider and benchmark pages.

Workspace access

Sign in

Choose a sign in method to access your TrustedRouter workspace.

By signing in you agree to the terms of service and privacy policy.