Wafer
Explore Wafer models on TrustedRouter with current routes, token pricing, policy sources, privacy posture, regional availability, and API support.
Waferwafer
No provider claim| Provider | Wafer |
|---|---|
| Provider website | https://wafer.ai/ |
| Models | 7 public models |
| Prepaid routes | 7 |
| BYOK routes | 7 |
| Zero data retention | not claimed |
| Confidential compute | not claimed |
| Provider E2EE | not claimed |
| Policy note | Wafer supports request-scoped ZDR via Wafer-ZDR: required on supported models; model-level support differs, so TrustedRouter keeps provider-level claims conservative. Policy source |
Measured performance
35 samplesContinuously sampled across Wafer's routed models: p50 TTFT, effective throughput, and success rate. Effective throughput uses provider-reported output tokens over complete request time. Unsupported route and probe-configuration rows are separated from provider downtime. No prompt or output content stored.
| p50 TTFT | 2171 ms |
|---|---|
| Effective throughput | 140 tok/s n=5 |
| Uptime | 94.29% |
| Model | p50 TTFT | p50 TTFB | Effective throughput | Uptime | Config excluded | Availability samples |
|---|---|---|---|---|---|---|
| z-ai/glm-5.2 | 2636 ms | 2636 ms | 207 tok/s n=2 | 100.00% | — | 11 |
| z-ai/glm-5.1 | 664 ms | 664 ms | — | 100.00% | — | 2 |
| deepseek/deepseek-v4-flash-0731-fast | 1216 ms | 1216 ms | — | 100.00% | — | 6 |
| z-ai/glm-5.2-fast | 1934 ms | 1933 ms | 140 tok/s n=1 | 100.00% | — | 8 |
| moonshotai/kimi-k2.6 | 2171 ms | 2171 ms | 31 tok/s n=1 | 100.00% | — | 3 |
| moonshotai/kimi-k3 | 3056 ms | 3056 ms | 66 tok/s n=1 | 100.00% | — | 2 |
| moonshotai/kimi-k3-fast | 3672 ms | 3672 ms | — | 33.33% | — | 3 |
Wafer performance history · Full provider & model leaderboard.
Models served by Wafer.
Each row links to pricing, provider, benchmark, and API pages for the model.
| Model | AI IQ | Context | Endpoints | Prompt | Completion | Routes |
|---|---|---|---|---|---|---|
deepseek/deepseek-v4-flash-0731-fastDeepSeek V4 Flash 0731 Fast |
— | 1,000,000 | 2 | $0.294/1M | $0.588/1M | prepaid BYOK |
moonshotai/kimi-k2.6MoonshotAI: Kimi K2.6 |
IQ 119#19 | 262,144 | 2 | $1.197/1M | $5.04/1M | prepaid BYOK |
moonshotai/kimi-k3MoonshotAI: Kimi K3 |
IQ 123#14 | 1,048,576 | 2 | $3.15/1M | $15.75/1M | prepaid BYOK |
moonshotai/kimi-k3-fastKimi-K3-Fast |
— | 1,048,576 | 2 | $4.725/1M | $23.625/1M | prepaid BYOK |
z-ai/glm-5.1Z.ai: GLM 5.1 |
IQ 114#32 | 204,800 | 2 | $1.05/1M | $3.36/1M | prepaid BYOK |
z-ai/glm-5.2Z.ai: GLM 5.2 |
IQ 120#16 | 1,048,576 | 2 | $1.323/1M | $4.158/1M | prepaid BYOK |
z-ai/glm-5.2-fastGLM 5.2 Fast on Fireworks |
— | 1,048,576 | 2 | $2.205/1M | $6.93/1M | prepaid BYOK |
Questions
Does Wafer have zero data retention?
TrustedRouter does not currently mark Wafer as provider-level zero data retention. Use trustedrouter/zdr or provider.min_privacy=zdr to select a different eligible route, and review the linked policy source for changes.
Is Wafer end-to-end encrypted?
TrustedRouter does not currently mark Wafer as end-to-end encrypted at the provider boundary. The TrustedRouter gateway is still attested, but the selected provider normally receives the request in order to run the model. Use trustedrouter/e2e for the stronger route requirement.
Which Wafer models are available through TrustedRouter?
This page currently lists 7 public Wafer models, with live pricing, route count, context length, measured performance when available, and links to each model's provider and benchmark pages.