OpenAI compatible API · Attested · Public status

LLM Provider Latency Benchmarks

Compare measured time-to-first-token, throughput, uptime, and success rates across LLM providers routed continuously through TrustedRouter.

Verify gateway
Onebase URL to migrate
100sof models and routes
0prompt or output logs. Always.
Measured provider latency

Provider speed data from real routed requests.

TrustedRouter publishes metadata-only measurements for time-to-first-token, throughput, uptime, and excluded probe-configuration rows. The goal is to show what the router actually sees, not what a provider claims in a launch post.

  • ✓ Provider and model leaderboards
  • ✓ Per-provider performance pages when enough samples exist
  • ✓ Per-model performance pages when enough samples exist
  • ✓ Prompt and output content never stored for these rollups

Open leaderboard Open status

Signalsmetadata only
{
  "provider": "tinfoil",
  "model": "moonshotai/kimi-k2.6",
  "p50_ttft_ms": 1192,
  "uptime": 0.999,
  "sample_count": 42
}

Provider pages

Model pages

Live catalog evidence

Current routes, prices, privacy, and measured performance.

Catalog facts come from the routes currently configured in TrustedRouter. Performance uses the same cached metadata snapshot as the public leaderboard. Prompts and outputs are not part of these measurements.

556public models
95providers
1765configured routes
309ZDR routes
35provider E2EE routes
5957recent availability samples
Model Providers Context Input Output Privacy Measured route
Anthropic: Claude Opus 4.8anthropic/claude-opus-4.8
3 routes
1,000,000 $5.275/1M $26.375/1M varies 5 cited scores 1735 ms TTFT anthropic · 100.00% available · n=4
OpenAI: GPT-5.5openai/gpt-5.5
4 routes
1,050,000 $5.275/1M $31.65/1M ZDR 3 cited scores Warming up
Google: Gemini 3.5 Flashgoogle/gemini-3.5-flash
+1
7 routes
1,048,576 $1.5825/1M $9.495/1M ZDR 1372 ms TTFT google-vertex · 58 tok/s · 100.00% available · n=8
MoonshotAI: Kimi K2.7 Codemoonshotai/kimi-k2.7-code
+10
19 routes
262,144 $0.70685/1M to $1.00225/1M $3.587/1M to $4.22/1M ZDR 5 cited scores 2918 ms TTFT kimi · 100.00% available · n=23
Z.ai: GLM 5.2z-ai/glm-5.2
+24
44 routes
1,048,576 $0.7174/1M to $2.434729/1M $1.5825/1M to $7.030281/1M E2EE 4 cited scores 4944 ms TTFT baseten · 100.00% available · n=71
MiniMax: MiniMax M3minimax/minimax-m3
+9
21 routes
524,288 $0.24265/1M to $0.633/1M $1.0128/1M to $2.532/1M ZDR 4 cited scores 1599 ms TTFT together · 137 tok/s · 100.00% available · n=71
AnthropicPolicy varies 11 models 2005 ms p50 · n=218
GMI CloudPolicy varies 75 models 4026 ms p50 · n=27
OpenAIZDR on prepaid 45 models 2288 ms p50 · n=201
Atlas CloudPolicy varies 78 models 1766 ms p50 · n=26
Google AI StudioPolicy varies 13 models 804 ms p50 · n=500
Google Vertex AIZDR on prepaid 11 models 1566 ms p50 · n=42
Lightning AIPolicy varies 36 models 2047 ms p50 · n=30
KimiPolicy varies 4 models 6366 ms p50 · n=220

Browse every modelReview provider policiesOpen the full leaderboardSnapshot 2026-09-16T06:07:37.348Z

Questions

Are these vendor claims?

No. The leaderboard is generated from TrustedRouter synthetic probes and runtime metadata, not provider marketing claims.

Do latency probes store prompts or outputs?

No. Status and leaderboard records store provider, model, latency, token, route, cost, and outcome metadata only.

Workspace access

Sign in

Choose a sign in method to access your TrustedRouter workspace.

By signing in you agree to the terms of service and privacy policy.