OpenAI compatible API · Attested · Public status

LLM Provider Latency Benchmarks

Compare measured time-to-first-token, throughput, uptime, and success rates across LLM providers routed continuously through TrustedRouter.

Verify gateway
Onebase URL to migrate
100sof models and routes
0prompt or output logs. Always.
Measured provider latency

Provider speed data from real routed requests.

TrustedRouter publishes metadata-only measurements for time-to-first-token, throughput, uptime, and excluded probe-configuration rows. The goal is to show what the router actually sees, not what a provider claims in a launch post.

  • ✓ Provider and model leaderboards
  • ✓ Per-provider performance pages when enough samples exist
  • ✓ Per-model performance pages when enough samples exist
  • ✓ Prompt and output content never stored for these rollups

Open leaderboard Open status

Signalsmetadata only
{
  "provider": "tinfoil",
  "model": "moonshotai/kimi-k2.6",
  "p50_ttft_ms": 1192,
  "uptime": 0.999,
  "sample_count": 42
}

Provider pages

Model pages

Live catalog evidence

Current routes, prices, privacy, and measured performance.

Catalog facts come from the routes currently configured in TrustedRouter. Performance uses the same cached metadata snapshot as the public leaderboard. Prompts and outputs are not part of these measurements.

555public models
95providers
1758configured routes
313ZDR routes
35provider E2EE routes
5044recent availability samples
Model Providers Context Input Output Privacy Measured route
Anthropic: Claude Opus 4.8anthropic/claude-opus-4.8
3 routes
1,000,000 $5.275/1M $26.375/1M varies 5 cited scores 2308 ms TTFT anthropic · 100.00% available · n=3
OpenAI: GPT-5.5openai/gpt-5.5
4 routes
1,050,000 $5.275/1M $31.65/1M ZDR 3 cited scores Warming up
Google: Gemini 3.5 Flashgoogle/gemini-3.5-flash
+1
7 routes
1,048,576 $1.5825/1M $9.495/1M ZDR 1372 ms TTFT google-vertex · 58 tok/s · 100.00% available · n=9
MoonshotAI: Kimi K2.7 Codemoonshotai/kimi-k2.7-code
+10
19 routes
262,144 $0.70685/1M to $1.13096/1M $3.587/1M to $4.90575/1M ZDR 5 cited scores 2465 ms TTFT kimi · 100.00% available · n=28
Z.ai: GLM 5.2z-ai/glm-5.2
+24
44 routes
1,048,576 $0.7174/1M to $2.434729/1M $1.5825/1M to $7.030281/1M E2EE 4 cited scores 4944 ms TTFT baseten · 100.00% available · n=103
MiniMax: MiniMax M3minimax/minimax-m3
+9
21 routes
524,288 $0.24265/1M to $0.633/1M $1.0128/1M to $2.532/1M ZDR 4 cited scores 1676 ms TTFT fireworks · 100.00% available · n=102
AnthropicPolicy varies 11 models 2078 ms p50 · n=197
GMI CloudPolicy varies 76 models 3924 ms p50 · n=34
OpenAIZDR on prepaid 47 models 1434 ms p50 · n=292
Atlas CloudPolicy varies 78 models 2008 ms p50 · n=27
Google AI StudioPolicy varies 13 models 555 ms p50 · n=499
Google Vertex AIZDR on prepaid 11 models 1566 ms p50 · n=41
Lightning AIPolicy varies 33 models 2198 ms p50 · n=24
KimiPolicy varies 4 models 5055 ms p50 · n=122

Browse every modelReview provider policiesOpen the full leaderboardSnapshot 2026-09-16T11:17:34.213Z

Questions

Are these vendor claims?

No. The leaderboard is generated from TrustedRouter synthetic probes and runtime metadata, not provider marketing claims.

Do latency probes store prompts or outputs?

No. Status and leaderboard records store provider, model, latency, token, route, cost, and outcome metadata only.

Workspace access

Sign in

Choose a sign in method to access your TrustedRouter workspace.

By signing in you agree to the terms of service and privacy policy.