Wafer
Explore Wafer models on TrustedRouter with current routes, token pricing, policy sources, privacy posture, regional availability, and API support.
Waferwafer
No provider claimThese privacy labels describe Wafer, the upstream model provider. ZDR is a retention policy; verified confidential inference additionally requires attested provider compute and end-to-end encryption.
| Provider | Wafer |
|---|---|
| Routing status | Active |
| Provider website | https://wafer.ai/ |
| Models | 7 public models |
| Credits routes | 7 |
| Zero data retention | not claimed |
| Verified confidential inference | Not verified |
| Policy note | Wafer supports request-scoped ZDR via Wafer-ZDR: required on supported models; model-level support differs, so TrustedRouter keeps provider-level claims conservative. Policy source |
Measured performance
33 samplesContinuously sampled across Wafer's routed models: p50 TTFT, effective throughput, and success rate. Effective throughput uses provider-reported output tokens over complete request time. Unsupported route and probe-configuration rows are separated from provider downtime. No prompt or output content stored.
| p50 TTFT | 1610 ms |
|---|---|
| Effective throughput | 125 tok/s n=2 |
| Uptime | 100.00% |
| Model | p50 TTFT | Effective throughput | Uptime | Config excluded | Availability samples |
|---|---|---|---|---|---|
| moonshotai/kimi-k3 | 891 ms | 110 tok/s n=1 | 100.00% | — | 2 |
| z-ai/glm-5.3 | 1012 ms | — | 100.00% | — | 5 |
| deepseek/deepseek-v4.1-flash | 1078 ms | — | 100.00% | — | 4 |
| deepseek/deepseek-v4-flash-0423 | 1123 ms | — | 100.00% | — | 2 |
| deepseek/deepseek-v4-flash-0731-fast | 1610 ms | — | 100.00% | — | 5 |
| moonshotai/kimi-k2.6 | 1750 ms | 140 tok/s n=1 | 100.00% | — | 7 |
| z-ai/glm-5.3-flash | 1857 ms | — | 100.00% | — | 8 |
Wafer performance history · Full provider & model leaderboard.
Models served by Wafer.
Each row links to pricing, provider, benchmark, and API pages for the model.
| Model | AI IQ | Context | Input | Cached input | Output |
|---|---|---|---|---|---|
deepseek/deepseek-v4-flash-0423DeepSeek-V4-Flash-0423 |
— | 1,048,576 | $0.07385/1M | $0.0211/1M | $0.26375/1M |
deepseek/deepseek-v4-flash-0731-fastDeepSeek V4 Flash 0731 Fast |
— | 1,000,000 | $0.1055/1M | $0.05275/1M | $0.26375/1M |
deepseek/deepseek-v4.1-flashDeepSeek: DeepSeek V4.1 Flash |
IQ 116#38 | 1,048,576 | $0.3165/1M | $0.01055/1M | $1.266/1M |
moonshotai/kimi-k2.6MoonshotAI: Kimi K2.6 |
IQ 118#35 | 262,144 | $1.2027/1M | $0.20045/1M | $5.064/1M |
moonshotai/kimi-k3MoonshotAI: Kimi K3 |
IQ 121#27 | 1,048,576 | $3.165/1M | $0.3165/1M | $13.45125/1M |
z-ai/glm-5.3Z.ai: GLM 5.3 |
IQ 123#18 | 1,048,576 | $1.25545/1M | $0.2743/1M | $4.642/1M |
z-ai/glm-5.3-flashZ.ai: GLM 5.3 Flash |
IQ 116#40 | 1,048,576 | $0.1055/1M | $0.0211/1M | $0.36925/1M |
Questions
Does Wafer have zero data retention?
TrustedRouter does not currently mark Wafer as provider-level zero data retention. Use trustedrouter/zdr or provider.min_privacy=zdr to select a different eligible route, and review the linked policy source for changes.
Is Wafer end-to-end encrypted?
TrustedRouter does not currently mark Wafer as end-to-end encrypted at the provider boundary. The TrustedRouter gateway is still attested, but the selected provider normally receives the request in order to run the model. Use trustedrouter/e2e for the stronger route requirement.
Which Wafer models are available through TrustedRouter?
This page currently lists 7 public Wafer models, with live pricing, route count, context length, measured performance when available, and links to each model's provider and benchmark pages.