Wafer
Explore Wafer models on TrustedRouter with current routes, token pricing, policy sources, privacy posture, regional availability, and API support.
Waferwafer
No provider claim| Provider | Wafer |
|---|---|
| Provider website | https://wafer.ai/ |
| Models | 4 public models |
| Prepaid routes | 4 |
| BYOK routes | 4 |
| Zero data retention | not claimed |
| Confidential compute | not claimed |
| Provider E2EE | not claimed |
| Policy note | Wafer supports request-scoped ZDR via Wafer-ZDR: required on supported models; model-level support differs, so TrustedRouter keeps provider-level claims conservative. Policy source |
Measured performance
51 samplesContinuously sampled across Wafer's routed models: p50 TTFT, effective throughput, and success rate. Effective throughput uses provider-reported output tokens over complete request time. Unsupported route and probe-configuration rows are separated from provider downtime. No prompt or output content stored.
| p50 TTFT | 3200 ms |
|---|---|
| Effective throughput | 39 tok/s n=3 |
| Uptime | 76.47% |
| Model | p50 TTFT | p50 TTFB | Effective throughput | Uptime | Config excluded | Availability samples |
|---|---|---|---|---|---|---|
| moonshotai/kimi-k2.6 | 2065 ms | 2064 ms | 39 tok/s n=1 | 100.00% | — | 16 |
| deepseek/deepseek-v4-flash-0731-fast | 3200 ms | 3200 ms | — | 78.57% | — | 14 |
| moonshotai/kimi-k3 | 4276 ms | 4276 ms | 10 tok/s n=1 | 41.67% | — | 12 |
| z-ai/glm-5.2 | 3859 ms | 3859 ms | 196 tok/s n=1 | 77.78% | — | 9 |
Wafer performance history · Full provider & model leaderboard.
Models served by Wafer.
Each row links to pricing, provider, benchmark, and API pages for the model.
| Model | AI IQ | Context | Endpoints | Prompt | Completion | Routes |
|---|---|---|---|---|---|---|
deepseek/deepseek-v4-flash-0731-fastDeepSeek V4 Flash 0731 Fast |
— | 1,000,000 | 2 | $0.2954/1M | $0.5908/1M | prepaid BYOK |
moonshotai/kimi-k2.6MoonshotAI: Kimi K2.6 |
IQ 119#23 | 262,144 | 2 | $1.2027/1M | $5.064/1M | prepaid BYOK |
moonshotai/kimi-k3MoonshotAI: Kimi K3 |
IQ 123#16 | 1,048,576 | 2 | $3.165/1M | $15.825/1M | prepaid BYOK |
z-ai/glm-5.2Z.ai: GLM 5.2 |
IQ 120#19 | 1,048,576 | 2 | $1.3293/1M | $4.1778/1M | prepaid BYOK |
Questions
Does Wafer have zero data retention?
TrustedRouter does not currently mark Wafer as provider-level zero data retention. Use trustedrouter/zdr or provider.min_privacy=zdr to select a different eligible route, and review the linked policy source for changes.
Is Wafer end-to-end encrypted?
TrustedRouter does not currently mark Wafer as end-to-end encrypted at the provider boundary. The TrustedRouter gateway is still attested, but the selected provider normally receives the request in order to run the model. Use trustedrouter/e2e for the stronger route requirement.
Which Wafer models are available through TrustedRouter?
This page currently lists 4 public Wafer models, with live pricing, route count, context length, measured performance when available, and links to each model's provider and benchmark pages.