Venice
Explore Venice models on TrustedRouter with current routes, token pricing, policy sources, privacy posture, regional availability, and API support.
Venicevenice
No provider claim| Provider | Venice |
|---|---|
| Provider website | https://venice.ai/ |
| Models | 57 public models |
| Prepaid routes | 57 |
| BYOK routes | 40 |
| Zero data retention | no |
| Confidential compute | no |
| Provider E2EE | no |
| Policy note | Mixed model-specific posture. TrustedRouter cannot independently verify a complete chain from a live Venice endpoint through immutable source and hardware measurements to committed model weights. Venice is therefore not tracked as confidential or E2EE. Exact routes marked Private in Venice's live catalog qualify only as policy-backed ZDR through endpoint-specific records; Anonymized routes do not. Policy source |
Measured performance
46 samplesContinuously sampled across Venice's routed models: p50 TTFT, effective throughput, and success rate. Effective throughput uses provider-reported output tokens over complete request time. Unsupported route and probe-configuration rows are separated from provider downtime. No prompt or output content stored.
| p50 TTFT | 2700 ms |
|---|---|
| Effective throughput | 39 tok/s n=9 |
| Uptime | 93.48% |
| Model | p50 TTFT | p50 TTFB | Effective throughput | Uptime | Config excluded | Availability samples |
|---|---|---|---|---|---|---|
| nvidia/nemotron-3.5-lightning | 827 ms | 826 ms | 50 tok/s n=1 | 100.00% | — | 2 |
| qwen/qwen3-5-35b-a3b | 1064 ms | 1064 ms | — | 100.00% | — | 2 |
| moonshotai/kimi-k3 | 1468 ms | 1468 ms | 39 tok/s n=2 | 100.00% | — | 2 |
| deepseek/deepseek-v4-flash-0731-fast | 1709 ms | 1709 ms | — | 100.00% | — | 1 |
| z-ai/glm-5.2 | 1725 ms | 1725 ms | 20 tok/s n=1 | 100.00% | — | 2 |
| qwen/qwen3-235b-a22b-instruct-2507 | 1777 ms | 1777 ms | — | 100.00% | — | 2 |
| moonshotai/kimi-k2-6 | 1884 ms | 1884 ms | — | 100.00% | — | 2 |
| qwen/qwen3-vl-235b-a22b | 2562 ms | 2561 ms | — | 100.00% | — | 1 |
| qwen/qwen3-next-80b | 2700 ms | 2699 ms | — | 100.00% | — | 1 |
| z-ai/glm-4.7 | 2719 ms | 2719 ms | — | 100.00% | — | 2 |
| qwen/qwen3-235b-a22b-thinking-2507 | 3076 ms | 3075 ms | — | 100.00% | — | 1 |
| google/gemma-4-uncensored | 3085 ms | 3085 ms | — | 100.00% | — | 1 |
| minimax/minimax-m25 | 3513 ms | 3513 ms | — | 100.00% | — | 2 |
| z-ai/glm-5v-turbo | 3521 ms | 3520 ms | — | 100.00% | — | 1 |
| meta-llama/llama-3.2-3b | 3825 ms | 3824 ms | — | 100.00% | — | 1 |
| moonshotai/kimi-k2-7-code | 4001 ms | 4000 ms | — | 100.00% | — | 1 |
| qwen/qwen-3-8-2-4t-a95b | 4218 ms | 4218 ms | — | 100.00% | — | 3 |
| qwen/qwen-3-6-plus | 4370 ms | 4370 ms | — | 100.00% | — | 3 |
| qwen/qwen-3-7-plus | 4545 ms | 4545 ms | — | 100.00% | — | 3 |
| z-ai/glm-4.6 | 4658 ms | 4658 ms | — | 100.00% | — | 1 |
| z-ai/glm-5 | 5333 ms | 5333 ms | — | 100.00% | — | 1 |
| deepseek/deepseek-v4-pro-0423 | 8112 ms | 8111 ms | 28 tok/s n=2 | 100.00% | — | 1 |
| z-ai/glm-5-turbo | 15060 ms | 15059 ms | — | 100.00% | — | 1 |
| qwen/qwen3.6-27b | 889 ms | 888 ms | — | 75.00% | — | 8 |
| deepseek/deepseek-v3.2 | — | — | — | 0.00% | — | 1 |
| deepseek/deepseek-v4-flash | — | — | 50 tok/s n=2 | — | — | 0 |
| deepseek/deepseek-v4-flash-0731 | — | — | 58 tok/s n=1 | — | — | 0 |
Venice performance history · Full provider & model leaderboard.
Models served by Venice.
Each row links to pricing, provider, benchmark, and API pages for the model.
| Model | AI IQ | Context | Endpoints | Prompt | Completion | Routes |
|---|---|---|---|---|---|---|
alibaba/wan-2.7Alibaba Wan 2.7 |
— | 5,000 | 1 | selected route | selected route | prepaid |
bytedance/seedance-2.0ByteDance Seedance 2.0 |
— | 10,000 | 1 | selected route | selected route | prepaid |
bytedance/seedance-2.0-fastByteDance Seedance 2.0 Fast |
— | 10,000 | 1 | selected route | selected route | prepaid |
deepseek/deepseek-v3.2DeepSeek: DeepSeek V3.2 |
IQ 103#64 | 163,840 | 2 | $0.34815/1M | $0.5064/1M | prepaid BYOK |
deepseek/deepseek-v4-flashDeepSeek: DeepSeek V4 Flash 0423 |
IQ 113#41 | 1,048,576 | 2 | $0.14559/1M | $0.290125/1M | prepaid BYOK |
deepseek/deepseek-v4-flash-0731DeepSeek: DeepSeek V4 Flash 0731 |
— | 1,048,576 | 2 | $0.184625/1M | $0.36925/1M | prepaid BYOK |
deepseek/deepseek-v4-flash-0731-fastDeepSeek V4 Flash 0731 Fast |
— | 1,000,000 | 2 | $0.36925/1M | $0.7385/1M | prepaid BYOK |
deepseek/deepseek-v4-proDeepSeek: DeepSeek V4 Pro |
IQ 117#31 | 1,048,576 | 2 | $1.74075/1M | $3.482555/1M | prepaid BYOK |
deepseek/deepseek-v4-pro-0423DeepSeek V4 Pro 0423 |
— | 1,048,576 | 1 | $1.74075/1M | $3.482555/1M | prepaid |
google/gemini-omni-flashGoogle Gemini Omni Flash |
— | 1,048,576 | 1 | selected route | selected route | prepaid |
google/gemma-4-uncensoredGemma 4 Uncensored |
— | 256,000 | 2 | $0.171438/1M | $0.5275/1M | prepaid BYOK |
google/veo-3.1Google Veo 3.1 |
— | 2,500 | 1 | selected route | selected route | prepaid |
google/veo-3.1-fastGoogle Veo 3.1 Fast |
— | 2,500 | 1 | selected route | selected route | prepaid |
kling/o3-proKling Video 3.0 Omni Pro |
— | 3,072 | 1 | selected route | selected route | prepaid |
kling/v3-proKling Video 3.0 Pro |
— | 3,072 | 1 | selected route | selected route | prepaid |
lightricks/ltx-2.3Lightricks LTX 2.3 |
— | 5,000 | 1 | selected route | selected route | prepaid |
lightricks/ltx-2.3-fastLightricks LTX 2.3 Fast |
— | 5,000 | 1 | selected route | selected route | prepaid |
meta-llama/llama-3.2-3bLlama 3.2 3B |
— | 128,000 | 2 | $0.15825/1M | $0.633/1M | prepaid BYOK |
meta-llama/llama-3.3-70bLlama 3.3 70B |
— | 128,000 | 2 | $0.7385/1M | $2.954/1M | prepaid BYOK |
minimax/hailuo-3MiniMax Hailuo 3 (H3) |
— | 7,000 | 1 | selected route | selected route | prepaid |
minimax/minimax-m25MiniMax M2.5 |
— | 198,000 | 2 | $0.28485/1M | $1.00225/1M | prepaid BYOK |
minimax/minimax-m27MiniMax M2.7 |
— | 198,000 | 2 | $0.395625/1M | $1.5825/1M | prepaid BYOK |
minimax/minimax-m3-previewMiniMax M3 Preview |
IQ 114#39 | 524,288 | 2 | $0.3165/1M | $1.266/1M | prepaid BYOK |
moonshotai/kimi-k2-5Kimi K2.5 |
— | 256,000 | 2 | $0.5908/1M | $3.6925/1M | prepaid BYOK |
moonshotai/kimi-k2-6Kimi K2.6 |
— | 256,000 | 2 | $0.79125/1M | $3.6925/1M | prepaid BYOK |
moonshotai/kimi-k2-7-codeKimi K2.7 Code |
— | 256,000 | 2 | $0.79125/1M | $3.6925/1M | prepaid BYOK |
moonshotai/kimi-k3MoonshotAI: Kimi K3 |
IQ 123#16 | 1,048,576 | 2 | $3.95625/1M | $19.78125/1M | prepaid BYOK |
moonshotai/kimi-k3-fast-apiKimi K3 Fast |
— | 1,000,000 | 2 | $4.7475/1M | $23.7375/1M | prepaid BYOK |
nvidia/nemotron-3.5-lightningNVIDIA: Nemotron 3.5 Lightning |
— | 1,000,000 | 2 | $0.1055/1M | $0.26375/1M | prepaid BYOK |
openai/sora-2OpenAI Sora 2 |
— | 2,500 | 1 | selected route | selected route | prepaid |
openai/sora-2-proOpenAI Sora 2 Pro |
— | 2,500 | 1 | selected route | selected route | prepaid |
pixverse/c1PixVerse C1 |
— | 2,500 | 1 | selected route | selected route | prepaid |
qwen/qwen-3-6-plusQwen 3.6 Plus Uncensored |
— | 1,000,000 | 2 | $0.659375/1M | $3.95625/1M | prepaid BYOK |
qwen/qwen-3-7-maxQwen 3.7 Max |
— | 1,000,000 | 2 | $2.8485/1M | $8.49275/1M | prepaid BYOK |
qwen/qwen-3-7-plusQwen 3.7 Plus |
— | 1,000,000 | 2 | $0.5275/1M | $2.11/1M | prepaid BYOK |
qwen/qwen-3-8-2-4t-a95bQwen 3.8 2.4T |
— | 262,144 | 2 | $2.6375/1M | $7.9125/1M | prepaid BYOK |
qwen/qwen-3-8-maxQwen 3.8 Max |
— | 1,000,000 | 2 | $2.6375/1M | $7.9125/1M | prepaid BYOK |
qwen/qwen3-235b-a22b-instruct-2507Qwen3 235B A22B Instruct 2507 |
— | 131,072 | 2 | $0.15825/1M | $0.79125/1M | prepaid BYOK |
qwen/qwen3-235b-a22b-thinking-2507Qwen: Qwen3 235B A22B Thinking 2507 |
— | 262,144 | 2 | $0.47475/1M | $3.6925/1M | prepaid BYOK |
qwen/qwen3-5-35b-a3bQwen 3.5 35B A3B |
— | 256,000 | 2 | $0.329688/1M | $1.31875/1M | prepaid BYOK |
qwen/qwen3-6-35b-a3bQwen 3.6 35B A3B |
— | 256,000 | 2 | $0.1055/1M | $1.055/1M | prepaid BYOK |
qwen/qwen3-coder-480b-a35b-instruct-turboQwen/Qwen3-Coder-480B-A35B-Instruct-Turbo |
— | 262,144 | 2 | $0.36925/1M | $1.5825/1M | prepaid BYOK |
qwen/qwen3-next-80bQwen 3 Next 80b |
— | 256,000 | 2 | $0.36925/1M | $2.0045/1M | prepaid BYOK |
qwen/qwen3-vl-235b-a22bQwen3 VL 235B |
— | 128,000 | 2 | $0.22155/1M | $2.0045/1M | prepaid BYOK |
qwen/qwen3.5-397b-a17bQwen: Qwen3.5 397B A17B |
— | 262,144 | 2 | $0.79125/1M | $4.7475/1M | prepaid BYOK |
qwen/qwen3.5-9bQwen: Qwen3.5-9B |
IQ 93#99 | 262,144 | 2 | $0.1055/1M | $0.15825/1M | prepaid BYOK |
qwen/qwen3.6-27bQwen: Qwen3.6 27B |
IQ 111#47 | 262,144 | 2 | $0.342875/1M | $3.42875/1M | prepaid BYOK |
runway/gen-4.5Runway Gen-4.5 |
— | 1,000 | 1 | selected route | selected route | prepaid |
shengshu/vidu-q3ShengShu Vidu Q3 |
— | 2,500 | 1 | selected route | selected route | prepaid |
z-ai/glm-4.6Z.ai: GLM 4.6 |
— | 204,800 | 2 | $0.45365/1M | $1.84625/1M | prepaid BYOK |
z-ai/glm-4.7Z.ai: GLM 4.7 |
IQ 103#66 | 204,800 | 2 | $0.58025/1M | $2.79575/1M | prepaid BYOK |
z-ai/glm-4.7-flashZ.ai: GLM 4.7 Flash |
— | 202,752 | 2 | $0.0633/1M | $0.422/1M | prepaid BYOK |
z-ai/glm-5Z.ai: GLM 5 |
IQ 105#60 | 204,800 | 2 | $1.055/1M | $3.376/1M | prepaid BYOK |
z-ai/glm-5-turboZ.ai: GLM 5 Turbo |
— | 202,752 | 2 | $1.266/1M | $4.22/1M | prepaid BYOK |
z-ai/glm-5.1Z.ai: GLM 5.1 |
IQ 114#38 | 204,800 | 2 | $1.6247/1M | $5.1062/1M | prepaid BYOK |
z-ai/glm-5.2Z.ai: GLM 5.2 |
IQ 120#19 | 1,048,576 | 2 | $1.477/1M | $4.642/1M | prepaid BYOK |
z-ai/glm-5v-turboZ.ai: GLM 5V Turbo |
— | 202,752 | 2 | $1.5825/1M | $5.275/1M | prepaid BYOK |
Questions
Does Venice have zero data retention?
TrustedRouter does not currently mark Venice as provider-level zero data retention. Use trustedrouter/zdr or provider.min_privacy=zdr to select a different eligible route, and review the linked policy source for changes.
Is Venice end-to-end encrypted?
TrustedRouter does not currently mark Venice as end-to-end encrypted at the provider boundary. The TrustedRouter gateway is still attested, but the selected provider normally receives the request in order to run the model. Use trustedrouter/e2e for the stronger route requirement.
Which Venice models are available through TrustedRouter?
This page currently lists 57 public Venice models, with live pricing, route count, context length, measured performance when available, and links to each model's provider and benchmark pages.