OpenAI compatible API · Attested · Public status

Venice

Explore Venice models on TrustedRouter with current routes, token pricing, policy sources, privacy posture, regional availability, and API support.

Verify gateway
Onebase URL to migrate
100sof models and routes
0prompt or output logs. Always.

Venicevenice

No provider claim

All providers

ProviderVenice
Routing statusActive
Provider websitehttps://venice.ai/
Models57 public models
Credits routes57
Zero data retentionno
Confidential computeno
Provider E2EEno
Policy noteMixed model-specific posture. TrustedRouter cannot independently verify a complete chain from a live Venice endpoint through immutable source and hardware measurements to committed model weights. Venice is therefore not tracked as confidential or E2EE. Exact routes marked Private in Venice's live catalog qualify only as policy-backed ZDR through endpoint-specific records; Anonymized routes do not.
Policy source

Measured performance

47 samples

Continuously sampled across Venice's routed models: p50 TTFT, effective throughput, and success rate. Effective throughput uses provider-reported output tokens over complete request time. Unsupported route and probe-configuration rows are separated from provider downtime. No prompt or output content stored.

p50 TTFT2143 ms
Effective throughput37 tok/s n=9
Uptime91.49%
Modelp50 TTFTEffective throughputUptimeConfig excludedAvailability samples
google/gemma-4-uncensored 817 ms 100.00% 1
minimax/minimax-m3-preview 1016 ms 100.00% 2
qwen/qwen3-vl-235b-a22b 1029 ms 100.00% 2
qwen/qwen3-6-35b-a3b 1083 ms 100.00% 2
qwen/qwen3-235b-a22b-thinking-2507 1143 ms 100.00% 1
qwen/qwen3-5-35b-a3b 1287 ms 100.00% 2
qwen/qwen3-next-80b 1507 ms 100.00% 1
deepseek/deepseek-v4-pro-0423 1582 ms 32 tok/s n=2 100.00% 2
moonshotai/kimi-k3-fast-api 1661 ms 100.00% 1
moonshotai/kimi-k3 1711 ms 37 tok/s n=2 50.00% 4
qwen/qwen3.5-397b-a17b 1741 ms 100.00% 1
moonshotai/kimi-k2-6 1958 ms 100.00% 2
z-ai/glm-5v-turbo 2133 ms 100.00% 1
qwen/qwen3-235b-a22b-instruct-2507 2143 ms 100.00% 1
deepseek/deepseek-v4-flash 2177 ms 19 tok/s n=1 100.00% 2
deepseek/deepseek-v4-flash-0731-fast 2196 ms 100.00% 3
z-ai/glm-4.7 2665 ms 100.00% 3
qwen/qwen-3-7-max 2736 ms 100.00% 1
minimax/minimax-m27 2738 ms 100.00% 1
deepseek/deepseek-v4-pro 2981 ms 26 tok/s n=1 100.00% 2
moonshotai/kimi-k2-7-code 3071 ms 100.00% 1
z-ai/glm-4.6 3800 ms 100.00% 2
qwen/qwen-3-8-2-4t-a95b 3857 ms 100.00% 1
qwen/qwen3-coder-480b-a35b-instruct-turbo 6191 ms 100.00% 2
deepseek/deepseek-v4-flash-0731 6873 ms 52 tok/s n=1 100.00% 1
z-ai/glm-5-turbo 9299 ms 100.00% 1
meta-llama/llama-3.2-3b 9639 ms 100.00% 2
deepseek/deepseek-v3.2 0.00% 1
z-ai/glm-4.7-flash 0.00% 1
z-ai/glm-5.2 56 tok/s n=2 0

Venice performance history · Full provider & model leaderboard.

Provider models

Models served by Venice.

Each row links to pricing, provider, benchmark, and API pages for the model.

Model AI IQ Context Input Cached input Output
alibaba/wan-2.7
Alibaba Wan 2.7
5,000 selected route Not published selected route
bytedance/seedance-2.0
ByteDance Seedance 2.0
10,000 selected route Not published selected route
bytedance/seedance-2.0-fast
ByteDance Seedance 2.0 Fast
10,000 selected route Not published selected route
deepseek/deepseek-v3.2
DeepSeek: DeepSeek V3.2
IQ 103#73 163,840 $0.34815/1M $0.1688/1M $0.5064/1M
deepseek/deepseek-v4-flash
DeepSeek: DeepSeek V4 Flash 0423
IQ 113#46 1,024,000 $0.14559/1M $0.02954/1M $0.290125/1M
deepseek/deepseek-v4-flash-0731
DeepSeek: DeepSeek V4 Flash 0731
1,048,576 $0.184625/1M $0.036925/1M $0.36925/1M
deepseek/deepseek-v4-flash-0731-fast
DeepSeek V4 Flash 0731 Fast
1,000,000 $0.36925/1M $0.092313/1M $0.7385/1M
deepseek/deepseek-v4-pro
DeepSeek: DeepSeek V4 Pro 0423
IQ 117#36 1,024,000 $1.74075/1M $0.34815/1M $3.482555/1M
deepseek/deepseek-v4-pro-0423
DeepSeek V4 Pro 0423
1,024,000 $1.74075/1M $0.34815/1M $3.482555/1M
google/gemini-omni-flash
Google Gemini Omni Flash
1,048,576 selected route Not published selected route
google/gemma-4-uncensored
Gemma 4 Uncensored
256,000 $0.171438/1M Not published $0.5275/1M
google/veo-3.1
Google Veo 3.1
2,500 selected route Not published selected route
google/veo-3.1-fast
Google Veo 3.1 Fast
2,500 selected route Not published selected route
kling/o3-pro
Kling Video 3.0 Omni Pro
3,072 selected route Not published selected route
kling/v3-pro
Kling Video 3.0 Pro
3,072 selected route Not published selected route
lightricks/ltx-2.3
Lightricks LTX 2.3
5,000 selected route Not published selected route
lightricks/ltx-2.3-fast
Lightricks LTX 2.3 Fast
5,000 selected route Not published selected route
meta-llama/llama-3.2-3b
Llama 3.2 3B
128,000 $0.15825/1M Not published $0.633/1M
meta-llama/llama-3.3-70b
Llama 3.3 70B
128,000 $0.7385/1M Not published $2.954/1M
minimax/hailuo-3
MiniMax Hailuo 3 (H3)
7,000 selected route Not published selected route
minimax/minimax-m25
MiniMax M2.5
198,000 $0.28485/1M $0.03165/1M $1.00225/1M
minimax/minimax-m27
MiniMax M2.7
198,000 $0.395625/1M $0.072532/1M $1.5825/1M
minimax/minimax-m3-preview
MiniMax M3 Preview
IQ 114#44 524,288 $0.3165/1M $0.0633/1M $1.266/1M
moonshotai/kimi-k2-5
Kimi K2.5
256,000 $0.5908/1M $0.2321/1M $3.6925/1M
moonshotai/kimi-k2-6
Kimi K2.6
256,000 $0.79125/1M $0.1688/1M $3.6925/1M
moonshotai/kimi-k2-7-code
Kimi K2.7 Code
256,000 $0.79125/1M $0.1688/1M $3.6925/1M
moonshotai/kimi-k3
MoonshotAI: Kimi K3
IQ 123#21 1,048,576 $3.95625/1M $0.395625/1M $19.78125/1M
moonshotai/kimi-k3-fast-api
Kimi K3 Fast
1,000,000 $4.7475/1M $0.47475/1M $23.7375/1M
openai/sora-2
OpenAI Sora 2
2,500 selected route Not published selected route
openai/sora-2-pro
OpenAI Sora 2 Pro
2,500 selected route Not published selected route
pixverse/c1
PixVerse C1
2,500 selected route Not published selected route
qwen/qwen-3-6-plus
Qwen 3.6 Plus Uncensored
1,000,000 $0.659375/1M $0.065938/1M $3.95625/1M
qwen/qwen-3-7-max
Qwen 3.7 Max
1,000,000 $2.8485/1M $0.28485/1M $8.49275/1M
qwen/qwen-3-7-plus
Qwen 3.7 Plus
1,000,000 $0.5275/1M $0.05275/1M $2.11/1M
qwen/qwen-3-8-2-4t-a95b
Qwen 3.8 2.4T
262,144 $2.6375/1M $0.329688/1M $7.9125/1M
qwen/qwen-3-8-27b
Qwen 3.8 27B
262,144 $0.47475/1M Not published $3.376/1M
qwen/qwen-3-8-max
Qwen 3.8 Max
1,000,000 $2.6375/1M $0.329688/1M $7.9125/1M
qwen/qwen3-235b-a22b-instruct-2507
Qwen3 235B A22B Instruct 2507
131,072 $0.15825/1M Not published $0.79125/1M
qwen/qwen3-235b-a22b-thinking-2507
Qwen: Qwen3 235B A22B Thinking 2507
131,072 $0.47475/1M Not published $3.6925/1M
qwen/qwen3-5-35b-a3b
Qwen 3.5 35B A3B
256,000 $0.329688/1M $0.164844/1M $1.31875/1M
qwen/qwen3-6-35b-a3b
Qwen 3.6 35B A3B
256,000 $0.1055/1M Not published $1.055/1M
qwen/qwen3-coder-480b-a35b-instruct-turbo
Qwen/Qwen3-Coder-480B-A35B-Instruct-Turbo
262,144 $0.36925/1M $0.0422/1M $1.5825/1M
qwen/qwen3-next-80b
Qwen 3 Next 80b
256,000 $0.36925/1M Not published $2.0045/1M
qwen/qwen3-vl-235b-a22b
Qwen3 VL 235B
128,000 $0.22155/1M $0.1055/1M $2.0045/1M
qwen/qwen3.5-397b-a17b
Qwen: Qwen3.5 397B A17B
262,144 $0.79125/1M Not published $4.7475/1M
qwen/qwen3.5-9b
Qwen: Qwen3.5-9B
IQ 90#113 262,144 $0.1055/1M Not published $0.15825/1M
qwen/qwen3.6-27b
Qwen: Qwen3.6 27B
IQ 111#56 262,144 $0.342875/1M Not published $3.42875/1M
runway/gen-4.5
Runway Gen-4.5
1,000 selected route Not published selected route
shengshu/vidu-q3
ShengShu Vidu Q3
2,500 selected route Not published selected route
z-ai/glm-4.6
Z.ai: GLM 4.6
198,000 $0.45365/1M $0.0844/1M $1.84625/1M
z-ai/glm-4.7
Z.ai: GLM 4.7
IQ 103#74 202,752 $0.58025/1M $0.11605/1M $2.79575/1M
z-ai/glm-4.7-flash
Z.ai: GLM 4.7 Flash
202,752 $0.0633/1M $0.01055/1M $0.422/1M
z-ai/glm-5
Z.ai: GLM 5
IQ 106#66 198,000 $1.055/1M $0.211/1M $3.376/1M
z-ai/glm-5-turbo
Z.ai: GLM 5 Turbo
202,752 $1.266/1M $0.2532/1M $4.22/1M
z-ai/glm-5.1
Z.ai: GLM 5.1
IQ 114#43 200,000 $1.6247/1M $0.30173/1M $5.1062/1M
z-ai/glm-5.2
Z.ai: GLM 5.2
IQ 120#27 1,048,576 $1.477/1M $0.2743/1M $4.642/1M
z-ai/glm-5v-turbo
Z.ai: GLM 5V Turbo
202,752 $1.5825/1M $0.3165/1M $5.275/1M

Questions

Does Venice have zero data retention?

TrustedRouter does not currently mark Venice as provider-level zero data retention. Use trustedrouter/zdr or provider.min_privacy=zdr to select a different eligible route, and review the linked policy source for changes.

Is Venice end-to-end encrypted?

TrustedRouter does not currently mark Venice as end-to-end encrypted at the provider boundary. The TrustedRouter gateway is still attested, but the selected provider normally receives the request in order to run the model. Use trustedrouter/e2e for the stronger route requirement.

Which Venice models are available through TrustedRouter?

This page currently lists 57 public Venice models, with live pricing, route count, context length, measured performance when available, and links to each model's provider and benchmark pages.

Workspace access

Sign in

Choose a sign in method to access your TrustedRouter workspace.

By signing in you agree to the terms of service and privacy policy.