OpenAI compatible API · Attested · Public status

MiniMax: MiniMax M3 Performance

Compare measured TTFT, throughput, uptime, and route health for MiniMax: MiniMax M3 across TrustedRouter providers using metadata-only production probes.

Verify gateway
Onebase URL to migrate
100sof models and routes
0prompt or output logs. Always.

minimax/minimax-m3

open weights Performance

All models

AI IQ IQ 115 #45 public AI IQ rank for minimax-m3
View AI IQ profile

Measured performance

Continuously sampled p50/p95 time-to-first-token (TTFT), effective throughput, and success rate for MiniMax: MiniMax M3. Effective throughput uses provider-reported output tokens over complete request time. Unsupported route and probe-configuration rows are separated from provider downtime, and no prompt or output content is stored.

Providerp50 TTFTp95 TTFTEffective throughputUptimeConfig excludedAvailability samples
together 1599 ms 4198 ms 234 tok/s n=1 100.00% 32
fireworks 1676 ms 3487 ms 100.00% 54
parasail 1794 ms 4837 ms 100.00% 2
minimax 2052 ms 5639 ms 144 tok/s n=1 100.00% 7
telnyx 3284 ms 3580 ms 130 tok/s n=1 100.00% 3
siliconflow 3582 ms 3582 ms 100.00% 1
morph 25067 ms 25067 ms 32 tok/s n=1 20.00% 5
atlas-cloud 119 tok/s n=1 0
deepinfra 19 tok/s n=1 0
featherless 222 tok/s n=1 0
novita 93 tok/s n=1 0
wandb 157 tok/s n=1 0

Full provider & model leaderboard.

Provider diversity

12 routes.

More routes give the auto router more room to fail over around provider 429 and 5xx responses.

Streaming

Gateway overhead is measured separately.

Public status separates TLS/health overhead from full model latency so slow LLMs do not inflate the router metric.

Status

Metadata rollups.

Status samples store latency, outcome, provider, model, route, cost, and region metadata only.

Workspace access

Sign in

Choose a sign in method to access your TrustedRouter workspace.

By signing in you agree to the terms of service and privacy policy.