OpenAI compatible API · Attested · Public status
Together performance
Review measured TTFT, TTFB, effective throughput, uptime, and sampled model routes for Together on TrustedRouter using metadata-only production probes.
Onebase URL to migrate
100sof models and routes
0prompt or output logs. Always.
Togethertogether
42 samplesContinuously sampled provider performance. TrustedRouter reports unsupported route and probe-configuration rows separately from provider downtime. Prompt and output content is not stored.
| p50 TTFT | 2537 ms |
|---|---|
| p95 TTFT | 6363 ms |
| p50 TTFB | 2537 ms |
| Effective throughput | 45 tok/s n=10 |
| Uptime | 97.62% |
Measured model routes
| Model | p50 TTFT | p50 TTFB | Effective throughput | Uptime | Config excluded | Availability samples |
|---|---|---|---|---|---|---|
| prism-ml/ternary-bonsai-27b | 2537 ms | 2537 ms | — | 91.67% | 1 provider_error |
12 |
| moonshotai/kimi-k2.6 | 880 ms | 880 ms | — | 100.00% | — | 2 |
| minimax/minimax-m3 | 894 ms | 894 ms | 83 tok/s n=2 | 100.00% | — | 4 |
| moonshotai/kimi-k2.7-code | 1024 ms | 1024 ms | — | 100.00% | — | 1 |
| thinkingmachines/inkling | 1083 ms | 1082 ms | 19 tok/s n=1 | 100.00% | — | 2 |
| qwen/qwen-2.5-7b-instruct | 1235 ms | 1235 ms | — | 100.00% | — | 2 |
| nvidia/nemotron-3-ultra-550b-a55b | 1710 ms | 1710 ms | 125 tok/s n=1 | 100.00% | — | 1 |
| moonshotai/kimi-k3 | 2155 ms | 2154 ms | 45 tok/s n=2 | 100.00% | — | 3 |
| google/gemma-3n-e4b-it | 2578 ms | 2578 ms | — | 100.00% | — | 4 |
| openai/gpt-oss-120b | 2728 ms | 2728 ms | 35 tok/s n=1 | 100.00% | — | 2 |
| meta-llama/llama-3.3-70b-instruct | 3305 ms | 2545 ms | — | 100.00% | — | 7 |
| z-ai/glm-5.2 | 4038 ms | 4038 ms | — | 100.00% | — | 1 |
| pearl-ai/gemma-4-31b-it | — | — | — | 100.00% | 3 probe_config_error |
1 |
| deepseek/deepseek-v4-flash-0731 | — | — | 87 tok/s n=1 | — | 1 probe_config_error |
0 |
| google/gemma-4-31b-it | — | — | 17 tok/s n=2 | — | 3 probe_config_error |
0 |