Venice
Explore Venice models on TrustedRouter with current routes, token pricing, policy sources, privacy posture, regional availability, and API support.
Venicevenice
No provider claim| Provider | Venice |
|---|---|
| Routing status | Active |
| Provider website | https://venice.ai/ |
| Models | 57 public models |
| Credits routes | 57 |
| Zero data retention | no |
| Confidential compute | no |
| Provider E2EE | no |
| Policy note | Mixed model-specific posture. TrustedRouter cannot independently verify a complete chain from a live Venice endpoint through immutable source and hardware measurements to committed model weights. Venice is therefore not tracked as confidential or E2EE. Exact routes marked Private in Venice's live catalog qualify only as policy-backed ZDR through endpoint-specific records; Anonymized routes do not. Policy source |
Measured performance
47 samplesContinuously sampled across Venice's routed models: p50 TTFT, effective throughput, and success rate. Effective throughput uses provider-reported output tokens over complete request time. Unsupported route and probe-configuration rows are separated from provider downtime. No prompt or output content stored.
| p50 TTFT | 2143 ms |
|---|---|
| Effective throughput | 37 tok/s n=9 |
| Uptime | 91.49% |
| Model | p50 TTFT | Effective throughput | Uptime | Config excluded | Availability samples |
|---|---|---|---|---|---|
| google/gemma-4-uncensored | 817 ms | — | 100.00% | — | 1 |
| minimax/minimax-m3-preview | 1016 ms | — | 100.00% | — | 2 |
| qwen/qwen3-vl-235b-a22b | 1029 ms | — | 100.00% | — | 2 |
| qwen/qwen3-6-35b-a3b | 1083 ms | — | 100.00% | — | 2 |
| qwen/qwen3-235b-a22b-thinking-2507 | 1143 ms | — | 100.00% | — | 1 |
| qwen/qwen3-5-35b-a3b | 1287 ms | — | 100.00% | — | 2 |
| qwen/qwen3-next-80b | 1507 ms | — | 100.00% | — | 1 |
| deepseek/deepseek-v4-pro-0423 | 1582 ms | 32 tok/s n=2 | 100.00% | — | 2 |
| moonshotai/kimi-k3-fast-api | 1661 ms | — | 100.00% | — | 1 |
| moonshotai/kimi-k3 | 1711 ms | 37 tok/s n=2 | 50.00% | — | 4 |
| qwen/qwen3.5-397b-a17b | 1741 ms | — | 100.00% | — | 1 |
| moonshotai/kimi-k2-6 | 1958 ms | — | 100.00% | — | 2 |
| z-ai/glm-5v-turbo | 2133 ms | — | 100.00% | — | 1 |
| qwen/qwen3-235b-a22b-instruct-2507 | 2143 ms | — | 100.00% | — | 1 |
| deepseek/deepseek-v4-flash | 2177 ms | 19 tok/s n=1 | 100.00% | — | 2 |
| deepseek/deepseek-v4-flash-0731-fast | 2196 ms | — | 100.00% | — | 3 |
| z-ai/glm-4.7 | 2665 ms | — | 100.00% | — | 3 |
| qwen/qwen-3-7-max | 2736 ms | — | 100.00% | — | 1 |
| minimax/minimax-m27 | 2738 ms | — | 100.00% | — | 1 |
| deepseek/deepseek-v4-pro | 2981 ms | 26 tok/s n=1 | 100.00% | — | 2 |
| moonshotai/kimi-k2-7-code | 3071 ms | — | 100.00% | — | 1 |
| z-ai/glm-4.6 | 3800 ms | — | 100.00% | — | 2 |
| qwen/qwen-3-8-2-4t-a95b | 3857 ms | — | 100.00% | — | 1 |
| qwen/qwen3-coder-480b-a35b-instruct-turbo | 6191 ms | — | 100.00% | — | 2 |
| deepseek/deepseek-v4-flash-0731 | 6873 ms | 52 tok/s n=1 | 100.00% | — | 1 |
| z-ai/glm-5-turbo | 9299 ms | — | 100.00% | — | 1 |
| meta-llama/llama-3.2-3b | 9639 ms | — | 100.00% | — | 2 |
| deepseek/deepseek-v3.2 | — | — | 0.00% | — | 1 |
| z-ai/glm-4.7-flash | — | — | 0.00% | — | 1 |
| z-ai/glm-5.2 | — | 56 tok/s n=2 | — | — | 0 |
Venice performance history · Full provider & model leaderboard.
Models served by Venice.
Each row links to pricing, provider, benchmark, and API pages for the model.
| Model | AI IQ | Context | Input | Cached input | Output |
|---|---|---|---|---|---|
alibaba/wan-2.7Alibaba Wan 2.7 |
— | 5,000 | selected route | Not published | selected route |
bytedance/seedance-2.0ByteDance Seedance 2.0 |
— | 10,000 | selected route | Not published | selected route |
bytedance/seedance-2.0-fastByteDance Seedance 2.0 Fast |
— | 10,000 | selected route | Not published | selected route |
deepseek/deepseek-v3.2DeepSeek: DeepSeek V3.2 |
IQ 103#73 | 163,840 | $0.34815/1M | $0.1688/1M | $0.5064/1M |
deepseek/deepseek-v4-flashDeepSeek: DeepSeek V4 Flash 0423 |
IQ 113#46 | 1,024,000 | $0.14559/1M | $0.02954/1M | $0.290125/1M |
deepseek/deepseek-v4-flash-0731DeepSeek: DeepSeek V4 Flash 0731 |
— | 1,048,576 | $0.184625/1M | $0.036925/1M | $0.36925/1M |
deepseek/deepseek-v4-flash-0731-fastDeepSeek V4 Flash 0731 Fast |
— | 1,000,000 | $0.36925/1M | $0.092313/1M | $0.7385/1M |
deepseek/deepseek-v4-proDeepSeek: DeepSeek V4 Pro 0423 |
IQ 117#36 | 1,024,000 | $1.74075/1M | $0.34815/1M | $3.482555/1M |
deepseek/deepseek-v4-pro-0423DeepSeek V4 Pro 0423 |
— | 1,024,000 | $1.74075/1M | $0.34815/1M | $3.482555/1M |
google/gemini-omni-flashGoogle Gemini Omni Flash |
— | 1,048,576 | selected route | Not published | selected route |
google/gemma-4-uncensoredGemma 4 Uncensored |
— | 256,000 | $0.171438/1M | Not published | $0.5275/1M |
google/veo-3.1Google Veo 3.1 |
— | 2,500 | selected route | Not published | selected route |
google/veo-3.1-fastGoogle Veo 3.1 Fast |
— | 2,500 | selected route | Not published | selected route |
kling/o3-proKling Video 3.0 Omni Pro |
— | 3,072 | selected route | Not published | selected route |
kling/v3-proKling Video 3.0 Pro |
— | 3,072 | selected route | Not published | selected route |
lightricks/ltx-2.3Lightricks LTX 2.3 |
— | 5,000 | selected route | Not published | selected route |
lightricks/ltx-2.3-fastLightricks LTX 2.3 Fast |
— | 5,000 | selected route | Not published | selected route |
meta-llama/llama-3.2-3bLlama 3.2 3B |
— | 128,000 | $0.15825/1M | Not published | $0.633/1M |
meta-llama/llama-3.3-70bLlama 3.3 70B |
— | 128,000 | $0.7385/1M | Not published | $2.954/1M |
minimax/hailuo-3MiniMax Hailuo 3 (H3) |
— | 7,000 | selected route | Not published | selected route |
minimax/minimax-m25MiniMax M2.5 |
— | 198,000 | $0.28485/1M | $0.03165/1M | $1.00225/1M |
minimax/minimax-m27MiniMax M2.7 |
— | 198,000 | $0.395625/1M | $0.072532/1M | $1.5825/1M |
minimax/minimax-m3-previewMiniMax M3 Preview |
IQ 114#44 | 524,288 | $0.3165/1M | $0.0633/1M | $1.266/1M |
moonshotai/kimi-k2-5Kimi K2.5 |
— | 256,000 | $0.5908/1M | $0.2321/1M | $3.6925/1M |
moonshotai/kimi-k2-6Kimi K2.6 |
— | 256,000 | $0.79125/1M | $0.1688/1M | $3.6925/1M |
moonshotai/kimi-k2-7-codeKimi K2.7 Code |
— | 256,000 | $0.79125/1M | $0.1688/1M | $3.6925/1M |
moonshotai/kimi-k3MoonshotAI: Kimi K3 |
IQ 123#21 | 1,048,576 | $3.95625/1M | $0.395625/1M | $19.78125/1M |
moonshotai/kimi-k3-fast-apiKimi K3 Fast |
— | 1,000,000 | $4.7475/1M | $0.47475/1M | $23.7375/1M |
openai/sora-2OpenAI Sora 2 |
— | 2,500 | selected route | Not published | selected route |
openai/sora-2-proOpenAI Sora 2 Pro |
— | 2,500 | selected route | Not published | selected route |
pixverse/c1PixVerse C1 |
— | 2,500 | selected route | Not published | selected route |
qwen/qwen-3-6-plusQwen 3.6 Plus Uncensored |
— | 1,000,000 | $0.659375/1M | $0.065938/1M | $3.95625/1M |
qwen/qwen-3-7-maxQwen 3.7 Max |
— | 1,000,000 | $2.8485/1M | $0.28485/1M | $8.49275/1M |
qwen/qwen-3-7-plusQwen 3.7 Plus |
— | 1,000,000 | $0.5275/1M | $0.05275/1M | $2.11/1M |
qwen/qwen-3-8-2-4t-a95bQwen 3.8 2.4T |
— | 262,144 | $2.6375/1M | $0.329688/1M | $7.9125/1M |
qwen/qwen-3-8-27bQwen 3.8 27B |
— | 262,144 | $0.47475/1M | Not published | $3.376/1M |
qwen/qwen-3-8-maxQwen 3.8 Max |
— | 1,000,000 | $2.6375/1M | $0.329688/1M | $7.9125/1M |
qwen/qwen3-235b-a22b-instruct-2507Qwen3 235B A22B Instruct 2507 |
— | 131,072 | $0.15825/1M | Not published | $0.79125/1M |
qwen/qwen3-235b-a22b-thinking-2507Qwen: Qwen3 235B A22B Thinking 2507 |
— | 131,072 | $0.47475/1M | Not published | $3.6925/1M |
qwen/qwen3-5-35b-a3bQwen 3.5 35B A3B |
— | 256,000 | $0.329688/1M | $0.164844/1M | $1.31875/1M |
qwen/qwen3-6-35b-a3bQwen 3.6 35B A3B |
— | 256,000 | $0.1055/1M | Not published | $1.055/1M |
qwen/qwen3-coder-480b-a35b-instruct-turboQwen/Qwen3-Coder-480B-A35B-Instruct-Turbo |
— | 262,144 | $0.36925/1M | $0.0422/1M | $1.5825/1M |
qwen/qwen3-next-80bQwen 3 Next 80b |
— | 256,000 | $0.36925/1M | Not published | $2.0045/1M |
qwen/qwen3-vl-235b-a22bQwen3 VL 235B |
— | 128,000 | $0.22155/1M | $0.1055/1M | $2.0045/1M |
qwen/qwen3.5-397b-a17bQwen: Qwen3.5 397B A17B |
— | 262,144 | $0.79125/1M | Not published | $4.7475/1M |
qwen/qwen3.5-9bQwen: Qwen3.5-9B |
IQ 90#113 | 262,144 | $0.1055/1M | Not published | $0.15825/1M |
qwen/qwen3.6-27bQwen: Qwen3.6 27B |
IQ 111#56 | 262,144 | $0.342875/1M | Not published | $3.42875/1M |
runway/gen-4.5Runway Gen-4.5 |
— | 1,000 | selected route | Not published | selected route |
shengshu/vidu-q3ShengShu Vidu Q3 |
— | 2,500 | selected route | Not published | selected route |
z-ai/glm-4.6Z.ai: GLM 4.6 |
— | 198,000 | $0.45365/1M | $0.0844/1M | $1.84625/1M |
z-ai/glm-4.7Z.ai: GLM 4.7 |
IQ 103#74 | 202,752 | $0.58025/1M | $0.11605/1M | $2.79575/1M |
z-ai/glm-4.7-flashZ.ai: GLM 4.7 Flash |
— | 202,752 | $0.0633/1M | $0.01055/1M | $0.422/1M |
z-ai/glm-5Z.ai: GLM 5 |
IQ 106#66 | 198,000 | $1.055/1M | $0.211/1M | $3.376/1M |
z-ai/glm-5-turboZ.ai: GLM 5 Turbo |
— | 202,752 | $1.266/1M | $0.2532/1M | $4.22/1M |
z-ai/glm-5.1Z.ai: GLM 5.1 |
IQ 114#43 | 200,000 | $1.6247/1M | $0.30173/1M | $5.1062/1M |
z-ai/glm-5.2Z.ai: GLM 5.2 |
IQ 120#27 | 1,048,576 | $1.477/1M | $0.2743/1M | $4.642/1M |
z-ai/glm-5v-turboZ.ai: GLM 5V Turbo |
— | 202,752 | $1.5825/1M | $0.3165/1M | $5.275/1M |
Questions
Does Venice have zero data retention?
TrustedRouter does not currently mark Venice as provider-level zero data retention. Use trustedrouter/zdr or provider.min_privacy=zdr to select a different eligible route, and review the linked policy source for changes.
Is Venice end-to-end encrypted?
TrustedRouter does not currently mark Venice as end-to-end encrypted at the provider boundary. The TrustedRouter gateway is still attested, but the selected provider normally receives the request in order to run the model. Use trustedrouter/e2e for the stronger route requirement.
Which Venice models are available through TrustedRouter?
This page currently lists 57 public Venice models, with live pricing, route count, context length, measured performance when available, and links to each model's provider and benchmark pages.