OpenAI compatible API · Attested · Public status

DigitalOcean Gradient AI

DigitalOcean Gradient AI models on TrustedRouter with prices, routes, policy notes, and source links.

Verify gateway
Onebase URL to migrate
100sof models and routes
0prompt or output logs. Always.

digitalocean

No provider claim

All providers

ProviderDigitalOcean Gradient AI
Models22 public models
Prepaid routes22
BYOK routes22
Zero data retentionnot claimed
Confidential computenot claimed
Provider E2EEnot claimed
Policy noteNo provider-ZDR claim is tracked here. DigitalOcean's Gradient AI model and pricing documentation is linked for data-handling review.
Policy source

Measured performance

69 samples

Continuously sampled across DigitalOcean Gradient AI's routed models: p50 TTFT, effective throughput, and success rate. Effective throughput uses provider-reported output tokens over complete request time. Unsupported route and probe-configuration rows are separated from provider downtime. No prompt or output content stored.

p50 TTFT1586 ms
Effective throughput37 tok/s n=6
Uptime98.55%
Modelp50 TTFTp50 TTFBEffective throughputUptimeConfig excludedAvailability samples
minimax/minimax-m2.5 651 ms 651 ms 100.00% 2
nvidia/nemotron-3-nano-omni 962 ms 961 ms 100.00% 3
xiaomi/mimo-v2.5-pro 1198 ms 1198 ms 64 tok/s n=1 100.00% 1
z-ai/glm-5 1341 ms 1341 ms 100.00% 1
qwen/qwen3.5-397b-a17b 1374 ms 1374 ms 100.00% 6
qwen/qwen3-coder-flash 1399 ms 1399 ms 100.00% 6
meta-llama/llama-4-maverick 1429 ms 1429 ms 100.00% 5
moonshotai/kimi-k2.6 1474 ms 1474 ms 33 tok/s n=2 100.00% 4
z-ai/glm-5.1 1554 ms 1553 ms 100.00% 4
nvidia/nemotron-nano-12b-v2-vl 1586 ms 1586 ms 100.00% 6
deepseek/deepseek-v3.2 1596 ms 1596 ms 100.00% 3
moonshotai/kimi-k2.5 1624 ms 1623 ms 100.00% 4
meta-llama/llama-3.3-70b-instruct 1689 ms 1689 ms 100.00% 1
deepseek/deepseek-r1-distill-llama-70b 1788 ms 1788 ms 100.00% 4
nvidia/nemotron-3-super-120b 2307 ms 2307 ms 100.00% 3
google/gemma-4-31b-it 2834 ms 2834 ms 100.00% 3
mistralai/ministral-3-14b-instruct 3109 ms 3109 ms 100.00% 1
deepseek/deepseek-v4-flash 3423 ms 3423 ms 20 tok/s n=1 100.00% 2
nvidia/nemotron-3-ultra-550b 3424 ms 3424 ms 100.00% 1
deepseek/deepseek-v4-pro 3489 ms 3489 ms 44 tok/s n=1 100.00% 6
z-ai/glm-5.2 5351 ms 5351 ms 41 tok/s n=1 66.67% 3

DigitalOcean Gradient AI performance history · Full provider & model leaderboard.

Provider models

Models served by DigitalOcean Gradient AI.

Each row links to pricing, provider, benchmark, and API pages for the model.

Model AI IQ Context Endpoints Prompt Completion Routes
deepseek/deepseek-r1-distill-llama-70b
DeepSeek: R1 Distill Llama 70B
8,192 2 $1.0395/1M $1.0395/1M prepaid BYOK
deepseek/deepseek-v3.2
DeepSeek: DeepSeek V3.2
IQ 103#56 163,840 2 $0.44625/1M $1.428/1M prepaid BYOK
deepseek/deepseek-v4-flash
DeepSeek: DeepSeek V4 Flash
IQ 108#46 1,048,576 2 $0.1176/1M $0.2352/1M prepaid BYOK
deepseek/deepseek-v4-pro
DeepSeek: DeepSeek V4 Pro
IQ 115#30 1,048,576 2 $1.4616/1M $2.9232/1M prepaid BYOK
google/gemma-4-31b-it
Google: Gemma 4 31B
IQ 101#65 262,144 2 $0.189/1M $0.525/1M prepaid BYOK
meta-llama/llama-3.3-70b-instruct
Meta: Llama 3.3 70B Instruct
131,072 2 $0.6825/1M $0.6825/1M prepaid BYOK
meta-llama/llama-4-maverick
Meta: Llama 4 Maverick
IQ 90#90 1,048,576 2 $0.2625/1M $0.9135/1M prepaid BYOK
minimax/minimax-m2.5
MiniMax: MiniMax M2.5
IQ 105#54 204,800 2 $0.23625/1M $0.945/1M prepaid BYOK
mistralai/ministral-3-14b-instruct
Ministral 3 14B Instruct
262,144 2 $0.21/1M $0.21/1M prepaid BYOK
moonshotai/kimi-k2.5
MoonshotAI: Kimi K2.5
IQ 111#39 262,144 2 $0.39375/1M $2.12625/1M prepaid BYOK
moonshotai/kimi-k2.6
MoonshotAI: Kimi K2.6
IQ 119#19 262,144 2 $0.798/1M $3.36/1M prepaid BYOK
nvidia/nemotron-3-nano-omni
Nemotron Nano 3 Omni
65,536 2 $0.525/1M $0.945/1M prepaid BYOK
nvidia/nemotron-3-super-120b
Nemotron-3-Super-120B
1,000,000 2 $0.2205/1M $0.47775/1M prepaid BYOK
nvidia/nemotron-3-ultra-550b
Nemotron 3 Ultra
131,072 2 $0.945/1M $1.785/1M prepaid BYOK
nvidia/nemotron-nano-12b-v2-vl
Nemotron Nano 12B v2 VL
128,000 2 $0.21/1M $0.63/1M prepaid BYOK
qwen/qwen3-32b
Qwen: Qwen3 32B
131,072 2 $0.2625/1M $0.5775/1M prepaid BYOK
qwen/qwen3-coder-flash
Qwen3 Coder Flash
262,144 2 $0.4725/1M $1.785/1M prepaid BYOK
qwen/qwen3.5-397b-a17b
Qwen: Qwen3.5 397B A17B
262,144 2 $0.40425/1M $2.5725/1M prepaid BYOK
xiaomi/mimo-v2.5-pro
Xiaomi: MiMo-V2.5-Pro
IQ 116#27 1,050,000 2 $0.63/1M $3.15/1M prepaid BYOK
z-ai/glm-5
Z.ai: GLM 5
IQ 105#51 204,800 2 $0.7875/1M $2.52/1M prepaid BYOK
z-ai/glm-5.1
Z.ai: GLM 5.1
IQ 114#32 204,800 2 $1.02375/1M $4.515/1M prepaid BYOK
z-ai/glm-5.2
Z.ai: GLM 5.2
IQ 120#16 1,048,576 2 $1.1025/1M $4.62/1M prepaid BYOK
Workspace access

Sign in

Choose a sign in method. New email and OAuth accounts include $0.10 in starter credit; wallet-only accounts start at $0.

By signing in you agree to the terms of service and privacy policy.