OpenAI compatible API · Attested · Public status

DigitalOcean Gradient AI performance

Measured TTFT, TTFB, effective throughput, uptime, and sampled model routes for DigitalOcean Gradient AI.

Verify gateway
Onebase URL to migrate
100sof models and routes
0prompt or output logs. Always.

digitalocean

67 samples

Provider overview

Continuously sampled provider performance. TrustedRouter reports unsupported route and probe-configuration rows separately from provider downtime. Prompt and output content is not stored.

p50 TTFT1586 ms
p95 TTFT5364 ms
p50 TTFB1596 ms
Effective throughput33 tok/s n=5
Uptime98.51%

Measured model routes

Modelp50 TTFTp50 TTFBEffective throughputUptimeConfig excludedAvailability samples
minimax/minimax-m2.5 651 ms 651 ms 100.00% 2
nvidia/nemotron-3-nano-omni 962 ms 961 ms 100.00% 2
meta-llama/llama-4-maverick 1164 ms 1164 ms 100.00% 4
xiaomi/mimo-v2.5-pro 1198 ms 1198 ms 100.00% 1
z-ai/glm-5 1341 ms 1341 ms 100.00% 1
qwen/qwen3.5-397b-a17b 1374 ms 1374 ms 100.00% 6
qwen/qwen3-coder-flash 1399 ms 1399 ms 100.00% 6
moonshotai/kimi-k2.6 1474 ms 1474 ms 33 tok/s n=2 100.00% 4
z-ai/glm-5.1 1554 ms 1553 ms 100.00% 4
nvidia/nemotron-nano-12b-v2-vl 1586 ms 1586 ms 100.00% 6
deepseek/deepseek-v3.2 1596 ms 1596 ms 100.00% 3
moonshotai/kimi-k2.5 1624 ms 1623 ms 100.00% 4
meta-llama/llama-3.3-70b-instruct 1689 ms 1689 ms 100.00% 1
deepseek/deepseek-r1-distill-llama-70b 1788 ms 1788 ms 100.00% 4
nvidia/nemotron-3-super-120b 2307 ms 2307 ms 100.00% 3
qwen/qwen3-32b 2709 ms 2708 ms 100.00% 1
google/gemma-4-31b-it 2834 ms 2834 ms 100.00% 3
mistralai/ministral-3-14b-instruct 3109 ms 3109 ms 100.00% 1
deepseek/deepseek-v4-flash 3423 ms 3423 ms 20 tok/s n=1 100.00% 2
deepseek/deepseek-v4-pro 3489 ms 3489 ms 44 tok/s n=1 100.00% 6
z-ai/glm-5.2 5351 ms 5351 ms 41 tok/s n=1 66.67% 3
Workspace access

Sign in

Choose a sign in method. New email and OAuth accounts include $0.10 in starter credit; wallet-only accounts start at $0.

By signing in you agree to the terms of service and privacy policy.