OpenAI compatible API · Attested · Public status

Wafer

Explore Wafer models on TrustedRouter with current routes, token pricing, policy sources, privacy posture, regional availability, and API support.

Verify gateway
Onebase URL to migrate
100sof models and routes
0prompt or output logs. Always.

Waferwafer

No provider claim

All providers

ProviderWafer
Provider websitehttps://wafer.ai/
Models7 public models
Prepaid routes7
BYOK routes7
Zero data retentionnot claimed
Confidential computenot claimed
Provider E2EEnot claimed
Policy noteWafer supports request-scoped ZDR via Wafer-ZDR: required on supported models; model-level support differs, so TrustedRouter keeps provider-level claims conservative.
Policy source

Measured performance

35 samples

Continuously sampled across Wafer's routed models: p50 TTFT, effective throughput, and success rate. Effective throughput uses provider-reported output tokens over complete request time. Unsupported route and probe-configuration rows are separated from provider downtime. No prompt or output content stored.

p50 TTFT2171 ms
Effective throughput140 tok/s n=5
Uptime94.29%
Modelp50 TTFTp50 TTFBEffective throughputUptimeConfig excludedAvailability samples
z-ai/glm-5.2 2636 ms 2636 ms 207 tok/s n=2 100.00% 11
z-ai/glm-5.1 664 ms 664 ms 100.00% 2
deepseek/deepseek-v4-flash-0731-fast 1216 ms 1216 ms 100.00% 6
z-ai/glm-5.2-fast 1934 ms 1933 ms 140 tok/s n=1 100.00% 8
moonshotai/kimi-k2.6 2171 ms 2171 ms 31 tok/s n=1 100.00% 3
moonshotai/kimi-k3 3056 ms 3056 ms 66 tok/s n=1 100.00% 2
moonshotai/kimi-k3-fast 3672 ms 3672 ms 33.33% 3

Wafer performance history · Full provider & model leaderboard.

Provider models

Models served by Wafer.

Each row links to pricing, provider, benchmark, and API pages for the model.

Model AI IQ Context Endpoints Prompt Completion Routes
deepseek/deepseek-v4-flash-0731-fast
DeepSeek V4 Flash 0731 Fast
1,000,000 2 $0.294/1M $0.588/1M prepaid BYOK
moonshotai/kimi-k2.6
MoonshotAI: Kimi K2.6
IQ 119#19 262,144 2 $1.197/1M $5.04/1M prepaid BYOK
moonshotai/kimi-k3
MoonshotAI: Kimi K3
IQ 123#14 1,048,576 2 $3.15/1M $15.75/1M prepaid BYOK
moonshotai/kimi-k3-fast
Kimi-K3-Fast
1,048,576 2 $4.725/1M $23.625/1M prepaid BYOK
z-ai/glm-5.1
Z.ai: GLM 5.1
IQ 114#32 204,800 2 $1.05/1M $3.36/1M prepaid BYOK
z-ai/glm-5.2
Z.ai: GLM 5.2
IQ 120#16 1,048,576 2 $1.323/1M $4.158/1M prepaid BYOK
z-ai/glm-5.2-fast
GLM 5.2 Fast on Fireworks
1,048,576 2 $2.205/1M $6.93/1M prepaid BYOK

Questions

Does Wafer have zero data retention?

TrustedRouter does not currently mark Wafer as provider-level zero data retention. Use trustedrouter/zdr or provider.min_privacy=zdr to select a different eligible route, and review the linked policy source for changes.

Is Wafer end-to-end encrypted?

TrustedRouter does not currently mark Wafer as end-to-end encrypted at the provider boundary. The TrustedRouter gateway is still attested, but the selected provider normally receives the request in order to run the model. Use trustedrouter/e2e for the stronger route requirement.

Which Wafer models are available through TrustedRouter?

This page currently lists 7 public Wafer models, with live pricing, route count, context length, measured performance when available, and links to each model's provider and benchmark pages.

Workspace access

Sign in

Choose a sign in method to access your TrustedRouter workspace.

By signing in you agree to the terms of service and privacy policy.