OpenAI compatible API · Attested · Public status
Google Vertex AI performance
Measured TTFT, TTFB, effective throughput, uptime, and sampled model routes for Google Vertex AI.
Onebase URL to migrate
100sof models and routes
0prompt or output logs. Always.
google-vertex
52 samples
Continuously sampled provider performance. TrustedRouter reports unsupported route and probe-configuration rows separately from provider downtime. Prompt and output content is not stored.
| p50 TTFT | 2163 ms |
|---|---|
| p95 TTFT | 5354 ms |
| p50 TTFB | 2749 ms |
| Effective throughput | 56 tok/s n=7 |
| Uptime | 100.00% |
Measured model routes
| Model | p50 TTFT | p50 TTFB | Effective throughput | Uptime | Config excluded | Availability samples |
|---|---|---|---|---|---|---|
| google/gemini-2.5-flash-lite | 1102 ms | 1102 ms | 137 tok/s n=2 | 100.00% | — | 3 |
| google/gemini-2.5-flash | 1474 ms | 1474 ms | — | 100.00% | — | 9 |
| google/gemini-3-flash-preview | 1643 ms | 1643 ms | — | 100.00% | — | 4 |
| google/gemini-3.5-flash-lite | 1678 ms | 1678 ms | 42 tok/s n=1 | 100.00% | — | 7 |
| google/gemini-3.5-flash | 2163 ms | 2163 ms | 60 tok/s n=1 | 100.00% | — | 7 |
| google/gemini-3.1-flash-lite | 2581 ms | 2580 ms | — | 100.00% | — | 7 |
| google/gemini-3.1-pro-preview | 3097 ms | 3097 ms | 56 tok/s n=2 | 100.00% | — | 4 |
| google/gemini-3.6-flash | 3110 ms | 3110 ms | 45 tok/s n=1 | 100.00% | — | 3 |
| google/gemini-2.5-pro | 4722 ms | 4722 ms | — | 100.00% | — | 8 |