OpenAI compatible API · Attested · Public status
Together
Together models on TrustedRouter with prices, routes, policy notes, and source links.
Onebase URL to migrate
100sof models and routes
Noneprompt logs by default
together
No logs
| Provider | Together |
|---|---|
| Models | 18 public models |
| Prepaid routes | 18 |
| BYOK routes | 18 |
| Zero data retention | yes |
| Confidential compute | not claimed |
| Provider E2EE | not claimed |
| Policy note | Tracked as provider ZDR. Together documents that inference inputs and outputs are not stored by default; temporary prompt caching may be used for performance, and sharing content for training is opt-in. Policy source |
Measured performance
128 samplesContinuously sampled across Together's routed models — p50 TTFT, throughput, and success rate. Unsupported route and probe-configuration rows are separated from provider downtime. No prompt or output content stored.
| p50 TTFT | 5270 ms |
|---|---|
| Throughput | — |
| Uptime | 98.44% |
| Model | p50 TTFT | p50 TTFB | Throughput | Uptime | Config excluded | Samples |
|---|---|---|---|---|---|---|
| moonshotai/kimi-k2.6 | 2432 ms | 2431 ms | — | 100.00% | — | 10 |
| moonshotai/kimi-k2.7-code | 2570 ms | 2570 ms | — | 100.00% | — | 13 |
| openai/gpt-oss-120b | 3349 ms | 3349 ms | — | 88.89% | — | 9 |
| qwen/qwen-2.5-7b-instruct | 3693 ms | 3693 ms | — | 100.00% | — | 9 |
| meta-llama/llama-3.3-70b-instruct | 4483 ms | 4483 ms | — | 100.00% | — | 16 |
| nvidia/nemotron-3-ultra-550b-a55b | 5270 ms | 5269 ms | — | 100.00% | — | 9 |
| openai/gpt-oss-20b | 5579 ms | 5578 ms | — | 100.00% | — | 14 |
| thinkingmachines/inkling | 5994 ms | 5994 ms | — | 100.00% | — | 9 |
| prism-ml/ternary-bonsai-27b | 7821 ms | 7821 ms | — | 100.00% | — | 14 |
| z-ai/glm-5.2 | 8581 ms | 8581 ms | — | 90.91% | — | 11 |
| google/gemma-3n-e4b-it | 9337 ms | 9337 ms | — | 100.00% | — | 14 |
Together performance history · Full provider & model leaderboard.
Provider models
Models served by Together.
Each row links to pricing, provider, benchmark, and API pages for the model.
| Model | AI IQ | Context | Endpoints | Prompt | Completion | Routes |
|---|---|---|---|---|---|---|
deepseek/deepseek-v4-proDeepSeek: DeepSeek V4 Pro |
IQ 114#27 | 1,048,576 | 2 | $1.827/1M | $3.654/1M | prepaid BYOK |
google/gemma-3n-e4b-itGoogle: Gemma 3n 4B |
— | 32,768 | 2 | $0.063/1M | $0.126/1M | prepaid BYOK |
google/gemma-4-31b-itGoogle: Gemma 4 31B |
IQ 101#65 | 262,144 | 2 | $0.4095/1M | $1.0185/1M | prepaid BYOK |
intfloat/multilingual-e5-large-instructMultilingual E5 Large Instruct |
— | 512 | 2 | $0.021/1M | selected route | prepaid BYOK |
meta-llama/llama-3.3-70b-instructMeta: Llama 3.3 70B Instruct |
— | 131,072 | 2 | $1.092/1M | $1.092/1M | prepaid BYOK |
minimax/minimax-m2.7MiniMax: MiniMax M2.7 |
IQ 107#49 | 204,800 | 2 | $0.315/1M | $1.26/1M | prepaid BYOK |
minimax/minimax-m3MiniMax: MiniMax M3 |
IQ 112#37 | 1,048,576 | 2 | $0.315/1M | $1.26/1M | prepaid BYOK |
moonshotai/kimi-k2.6MoonshotAI: Kimi K2.6 |
IQ 119#19 | 262,144 | 2 | $1.26/1M | $4.725/1M | prepaid BYOK |
moonshotai/kimi-k2.7-codeMoonshotAI: Kimi K2.7 Code |
IQ 117#20 | 262,144 | 2 | $0.9975/1M | $4.2/1M | prepaid BYOK |
nvidia/nemotron-3-ultra-550b-a55bNVIDIA: Nemotron 3 Ultra |
— | 512,288 | 2 | $0.63/1M | $3.78/1M | prepaid BYOK |
openai/gpt-oss-120bOpenAI: gpt-oss-120b |
IQ 103#57 | 131,072 | 2 | $0.1575/1M | $0.63/1M | prepaid BYOK |
openai/gpt-oss-20bOpenAI: gpt-oss-20b |
IQ 100#68 | 131,072 | 2 | $0.0525/1M | $0.21/1M | prepaid BYOK |
pearl-ai/gemma-4-31b-itPearl-ai Gemma-4-31B-it-pearl |
IQ 101#65 | 262,144 | 2 | $0.294/1M | $0.903/1M | prepaid BYOK |
prism-ml/ternary-bonsai-27bTernary Bonsai 27B |
— | 262,144 | 2 | $0.01/1M | $0.01/1M | prepaid BYOK |
qwen/qwen-2.5-7b-instructQwen: Qwen2.5 7B Instruct |
— | 32,768 | 2 | $0.315/1M | $0.315/1M | prepaid BYOK |
qwen/qwen3.5-9bQwen: Qwen3.5-9B |
IQ 93#88 | 262,144 | 2 | $0.1785/1M | $0.2625/1M | prepaid BYOK |
thinkingmachines/inklingThinking Machines: Inkling |
IQ 106#51 | 524,288 | 2 | $1.05/1M | $4.2525/1M | prepaid BYOK |
z-ai/glm-5.2Z.ai: GLM 5.2 |
IQ 120#16 | 1,048,576 | 2 | $1.47/1M | $4.62/1M | prepaid BYOK |