OpenAI compatible API · Attested · Public status

Fireworks AI performance

Review measured TTFT, effective throughput, uptime, and sampled model routes for Fireworks AI on TrustedRouter using metadata-only production probes.

Verify gateway
Onebase URL to migrate
100sof models and routes
0prompt or output logs. Always.

Fireworks AIfireworks

281 samples

Provider overview

Continuously sampled provider performance. TrustedRouter reports unsupported route and probe-configuration rows separately from provider downtime. Prompt and output content is not stored.

p50 TTFT1676 ms
p95 TTFT4072 ms
Effective throughput82 tok/s n=9
Uptime100.00%

Measured model routes

Modelp50 TTFTEffective throughputUptimeConfig excludedAvailability samples
minimax/minimax-m3 1676 ms 100.00% 54
meta-models/muse-glimmer-30b 879 ms 100.00% 1
z-ai/glm-5.3-fast 1181 ms 100.00% 1
moonshotai/kimi-k2.7-code 1203 ms 88 tok/s n=2 100.00% 3
openai/gpt-oss-120b 1327 ms 90 tok/s n=1 100.00% 3
qwen/qwen3.8-max 1501 ms 100.00% 2
deepseek/deepseek-v4-flash-0731 1541 ms 12 tok/s n=1 100.00% 1
moonshotai/kimi-k2.6 1586 ms 48 tok/s n=1 100.00% 2
deepseek/deepseek-v4-pro-0813 1595 ms 59 tok/s n=1 100.00% 1
z-ai/glm-5.3-flash 1658 ms 100.00% 1
moonshotai/kimi-k3-fast 1661 ms 100.00% 2
deepseek/deepseek-v4p1-flash 1760 ms 100.00% 1
deepseek/deepseek-v4-flash-vision-exp 1804 ms 100.00% 1
z-ai/glm-5.3 2853 ms 100.00% 5
nvidia/nemotron-3-ultra-550b-a55b 3618 ms 100.00% 1
moonshotai/kimi-k3 4072 ms 54 tok/s n=1 100.00% 2
z-ai/glm-5.2-fast 161 tok/s n=1 100.00% 200
z-ai/glm-5.2 82 tok/s n=1 0
Workspace access

Sign in

Choose a sign in method to access your TrustedRouter workspace.

By signing in you agree to the terms of service and privacy policy.