OpenAI compatible API · Attested · Public status
Chutes
Chutes models on TrustedRouter with prices, routes, policy notes, and source links.
Onebase URL to migrate
100sof models and routes
0prompt or output logs. Always.
chutes
No logs
| Provider | Chutes |
|---|---|
| Models | 10 public models |
| Prepaid routes | 10 |
| BYOK routes | 10 |
| Zero data retention | yes |
| Confidential compute | yes |
| Provider E2EE | no |
| Policy note | Chutes documents no prompt/output storage or training and serves these routes in confidential-compute TEEs. Standard API calls are not marked provider end-to-end encrypted. Policy source |
Measured performance
54 samplesContinuously sampled across Chutes's routed models: p50 TTFT, effective throughput, and success rate. Effective throughput uses provider-reported output tokens over complete request time. Unsupported route and probe-configuration rows are separated from provider downtime. No prompt or output content stored.
| p50 TTFT | 2855 ms |
|---|---|
| Effective throughput | 34 tok/s n=3 |
| Uptime | 98.15% |
| Model | p50 TTFT | p50 TTFB | Effective throughput | Uptime | Config excluded | Availability samples |
|---|---|---|---|---|---|---|
| qwen/qwen3-235b-a22b-thinking-2507 | 1379 ms | 1379 ms | — | 100.00% | — | 4 |
| qwen/qwen3.6-27b | 1986 ms | 1986 ms | — | 100.00% | — | 2 |
| z-ai/glm-5.1 | 2027 ms | 2027 ms | — | 100.00% | — | 4 |
| mistralai/mistral-nemo | 2282 ms | 2282 ms | — | 100.00% | 1 router_error |
6 |
| moonshotai/kimi-k2.6 | 2855 ms | 2855 ms | 34 tok/s n=2 | 100.00% | — | 5 |
| qwen/qwen3-32b | 3504 ms | 3503 ms | — | 100.00% | — | 10 |
| deepseek/deepseek-v3.2 | 3517 ms | 3517 ms | — | 100.00% | — | 7 |
| google/gemma-4-31b-turbo | 3553 ms | 3552 ms | — | 100.00% | — | 5 |
| z-ai/glm-5.2 | 4049 ms | 4049 ms | 21 tok/s n=1 | 100.00% | — | 5 |
| qwen/qwen3.5-397b-a17b | 2475 ms | 2475 ms | — | 83.33% | — | 6 |
Chutes performance history · Full provider & model leaderboard.
Provider models
Models served by Chutes.
Each row links to pricing, provider, benchmark, and API pages for the model.
| Model | AI IQ | Context | Endpoints | Prompt | Completion | Routes |
|---|---|---|---|---|---|---|
deepseek/deepseek-v3.2DeepSeek: DeepSeek V3.2 |
IQ 103#56 | 163,840 | 2 | $1.05/1M | $1.05/1M | prepaid BYOK |
google/gemma-4-31b-turbogoogle/gemma-4-31B-turbo |
— | 131,072 | 2 | $0.126/1M | $0.3885/1M | prepaid BYOK |
mistralai/mistral-nemoMistral: Mistral Nemo |
— | 131,072 | 2 | $0.025725/1M | $0.10269/1M | prepaid BYOK |
moonshotai/kimi-k2.6MoonshotAI: Kimi K2.6 |
IQ 119#19 | 262,144 | 2 | $0.693/1M | $3.675/1M | prepaid BYOK |
qwen/qwen3-235b-a22b-thinking-2507Qwen: Qwen3 235B A22B Thinking 2507 |
— | 262,144 | 2 | $0.313845/1M | $1.255485/1M | prepaid BYOK |
qwen/qwen3-32bQwen: Qwen3 32B |
— | 131,072 | 2 | $0.1092/1M | $0.4368/1M | prepaid BYOK |
qwen/qwen3.5-397b-a17bQwen: Qwen3.5 397B A17B |
— | 262,144 | 2 | $0.4725/1M | $3.15/1M | prepaid BYOK |
qwen/qwen3.6-27bQwen: Qwen3.6 27B |
IQ 111#40 | 262,144 | 2 | $0.315/1M | $2.1/1M | prepaid BYOK |
z-ai/glm-5.1Z.ai: GLM 5.1 |
IQ 114#32 | 204,800 | 2 | $1.029/1M | $3.234/1M | prepaid BYOK |
z-ai/glm-5.2Z.ai: GLM 5.2 |
IQ 120#16 | 1,048,576 | 2 | $1.3125/1M | $4.1475/1M | prepaid BYOK |