OpenAI compatible API · Attested · Public status
Google Vertex AI performance
Measured TTFT, TTFB, throughput, uptime, and sampled model routes for Google Vertex AI.
Onebase URL to migrate
100sof models and routes
Noneprompt logs by default
google-vertex
191 samples
Continuously sampled provider performance. TrustedRouter reports unsupported route and probe-configuration rows separately from provider downtime. Prompt and output content is not stored.
| p50 TTFT | 8726 ms |
|---|---|
| p95 TTFT | ms |
| p50 TTFB | ms |
| Throughput | 2 tok/s |
| Uptime | 100.00% |
Measured model routes
| Model | p50 TTFT | p50 TTFB | Throughput | Uptime | Config excluded | Samples |
|---|---|---|---|---|---|---|
| google/gemini-3-flash-preview | 1860 ms | 1860 ms | — | 100.00% | — | 5 |
| google/gemini-3.5-flash | 2108 ms | 2108 ms | — | 100.00% | — | 3 |
| google/gemini-3.6-flash | 3386 ms | 3386 ms | — | 100.00% | — | 5 |
| google/gemini-3.1-pro-preview | 3505 ms | 3505 ms | — | 100.00% | — | 6 |
| google/gemini-2.5-flash | 3981 ms | 3981 ms | 165 tok/s | 100.00% | — | 16 |
| google/gemini-3.1-flash-lite | 6267 ms | 6267 ms | — | 100.00% | — | 4 |
| google/gemini-3.5-flash-lite | 7161 ms | 7160 ms | — | 100.00% | — | 5 |
| google/gemini-2.5-flash-lite | 8726 ms | 8725 ms | 2 tok/s | 100.00% | — | 147 |