OpenAI compatible API · Attested · Public status
DigitalOcean Gradient AI
DigitalOcean Gradient AI models on TrustedRouter with prices, routes, policy notes, and source links.
Onebase URL to migrate
100sof models and routes
0prompt or output logs. Always.
digitalocean
No provider claim
| Provider | DigitalOcean Gradient AI |
|---|---|
| Models | 22 public models |
| Prepaid routes | 22 |
| BYOK routes | 22 |
| Zero data retention | not claimed |
| Confidential compute | not claimed |
| Provider E2EE | not claimed |
| Policy note | No provider-ZDR claim is tracked here. DigitalOcean's Gradient AI model and pricing documentation is linked for data-handling review. Policy source |
Measured performance
69 samplesContinuously sampled across DigitalOcean Gradient AI's routed models: p50 TTFT, effective throughput, and success rate. Effective throughput uses provider-reported output tokens over complete request time. Unsupported route and probe-configuration rows are separated from provider downtime. No prompt or output content stored.
| p50 TTFT | 1586 ms |
|---|---|
| Effective throughput | 37 tok/s n=6 |
| Uptime | 98.55% |
| Model | p50 TTFT | p50 TTFB | Effective throughput | Uptime | Config excluded | Availability samples |
|---|---|---|---|---|---|---|
| minimax/minimax-m2.5 | 651 ms | 651 ms | — | 100.00% | — | 2 |
| nvidia/nemotron-3-nano-omni | 962 ms | 961 ms | — | 100.00% | — | 3 |
| xiaomi/mimo-v2.5-pro | 1198 ms | 1198 ms | 64 tok/s n=1 | 100.00% | — | 1 |
| z-ai/glm-5 | 1341 ms | 1341 ms | — | 100.00% | — | 1 |
| qwen/qwen3.5-397b-a17b | 1374 ms | 1374 ms | — | 100.00% | — | 6 |
| qwen/qwen3-coder-flash | 1399 ms | 1399 ms | — | 100.00% | — | 6 |
| meta-llama/llama-4-maverick | 1429 ms | 1429 ms | — | 100.00% | — | 5 |
| moonshotai/kimi-k2.6 | 1474 ms | 1474 ms | 33 tok/s n=2 | 100.00% | — | 4 |
| z-ai/glm-5.1 | 1554 ms | 1553 ms | — | 100.00% | — | 4 |
| nvidia/nemotron-nano-12b-v2-vl | 1586 ms | 1586 ms | — | 100.00% | — | 6 |
| deepseek/deepseek-v3.2 | 1596 ms | 1596 ms | — | 100.00% | — | 3 |
| moonshotai/kimi-k2.5 | 1624 ms | 1623 ms | — | 100.00% | — | 4 |
| meta-llama/llama-3.3-70b-instruct | 1689 ms | 1689 ms | — | 100.00% | — | 1 |
| deepseek/deepseek-r1-distill-llama-70b | 1788 ms | 1788 ms | — | 100.00% | — | 4 |
| nvidia/nemotron-3-super-120b | 2307 ms | 2307 ms | — | 100.00% | — | 3 |
| google/gemma-4-31b-it | 2834 ms | 2834 ms | — | 100.00% | — | 3 |
| mistralai/ministral-3-14b-instruct | 3109 ms | 3109 ms | — | 100.00% | — | 1 |
| deepseek/deepseek-v4-flash | 3423 ms | 3423 ms | 20 tok/s n=1 | 100.00% | — | 2 |
| nvidia/nemotron-3-ultra-550b | 3424 ms | 3424 ms | — | 100.00% | — | 1 |
| deepseek/deepseek-v4-pro | 3489 ms | 3489 ms | 44 tok/s n=1 | 100.00% | — | 6 |
| z-ai/glm-5.2 | 5351 ms | 5351 ms | 41 tok/s n=1 | 66.67% | — | 3 |
DigitalOcean Gradient AI performance history · Full provider & model leaderboard.
Provider models
Models served by DigitalOcean Gradient AI.
Each row links to pricing, provider, benchmark, and API pages for the model.
| Model | AI IQ | Context | Endpoints | Prompt | Completion | Routes |
|---|---|---|---|---|---|---|
deepseek/deepseek-r1-distill-llama-70bDeepSeek: R1 Distill Llama 70B |
— | 8,192 | 2 | $1.0395/1M | $1.0395/1M | prepaid BYOK |
deepseek/deepseek-v3.2DeepSeek: DeepSeek V3.2 |
IQ 103#56 | 163,840 | 2 | $0.44625/1M | $1.428/1M | prepaid BYOK |
deepseek/deepseek-v4-flashDeepSeek: DeepSeek V4 Flash |
IQ 108#46 | 1,048,576 | 2 | $0.1176/1M | $0.2352/1M | prepaid BYOK |
deepseek/deepseek-v4-proDeepSeek: DeepSeek V4 Pro |
IQ 115#30 | 1,048,576 | 2 | $1.4616/1M | $2.9232/1M | prepaid BYOK |
google/gemma-4-31b-itGoogle: Gemma 4 31B |
IQ 101#65 | 262,144 | 2 | $0.189/1M | $0.525/1M | prepaid BYOK |
meta-llama/llama-3.3-70b-instructMeta: Llama 3.3 70B Instruct |
— | 131,072 | 2 | $0.6825/1M | $0.6825/1M | prepaid BYOK |
meta-llama/llama-4-maverickMeta: Llama 4 Maverick |
IQ 90#90 | 1,048,576 | 2 | $0.2625/1M | $0.9135/1M | prepaid BYOK |
minimax/minimax-m2.5MiniMax: MiniMax M2.5 |
IQ 105#54 | 204,800 | 2 | $0.23625/1M | $0.945/1M | prepaid BYOK |
mistralai/ministral-3-14b-instructMinistral 3 14B Instruct |
— | 262,144 | 2 | $0.21/1M | $0.21/1M | prepaid BYOK |
moonshotai/kimi-k2.5MoonshotAI: Kimi K2.5 |
IQ 111#39 | 262,144 | 2 | $0.39375/1M | $2.12625/1M | prepaid BYOK |
moonshotai/kimi-k2.6MoonshotAI: Kimi K2.6 |
IQ 119#19 | 262,144 | 2 | $0.798/1M | $3.36/1M | prepaid BYOK |
nvidia/nemotron-3-nano-omniNemotron Nano 3 Omni |
— | 65,536 | 2 | $0.525/1M | $0.945/1M | prepaid BYOK |
nvidia/nemotron-3-super-120bNemotron-3-Super-120B |
— | 1,000,000 | 2 | $0.2205/1M | $0.47775/1M | prepaid BYOK |
nvidia/nemotron-3-ultra-550bNemotron 3 Ultra |
— | 131,072 | 2 | $0.945/1M | $1.785/1M | prepaid BYOK |
nvidia/nemotron-nano-12b-v2-vlNemotron Nano 12B v2 VL |
— | 128,000 | 2 | $0.21/1M | $0.63/1M | prepaid BYOK |
qwen/qwen3-32bQwen: Qwen3 32B |
— | 131,072 | 2 | $0.2625/1M | $0.5775/1M | prepaid BYOK |
qwen/qwen3-coder-flashQwen3 Coder Flash |
— | 262,144 | 2 | $0.4725/1M | $1.785/1M | prepaid BYOK |
qwen/qwen3.5-397b-a17bQwen: Qwen3.5 397B A17B |
— | 262,144 | 2 | $0.40425/1M | $2.5725/1M | prepaid BYOK |
xiaomi/mimo-v2.5-proXiaomi: MiMo-V2.5-Pro |
IQ 116#27 | 1,050,000 | 2 | $0.63/1M | $3.15/1M | prepaid BYOK |
z-ai/glm-5Z.ai: GLM 5 |
IQ 105#51 | 204,800 | 2 | $0.7875/1M | $2.52/1M | prepaid BYOK |
z-ai/glm-5.1Z.ai: GLM 5.1 |
IQ 114#32 | 204,800 | 2 | $1.02375/1M | $4.515/1M | prepaid BYOK |
z-ai/glm-5.2Z.ai: GLM 5.2 |
IQ 120#16 | 1,048,576 | 2 | $1.1025/1M | $4.62/1M | prepaid BYOK |