OpenAI compatible API · Attested · Public status

Venice

Explore Venice models on TrustedRouter with current routes, token pricing, policy sources, privacy posture, regional availability, and API support.

Verify gateway
Onebase URL to migrate
100sof models and routes
0prompt or output logs. Always.

Venicevenice

No provider claim

All providers

ProviderVenice
Provider websitehttps://venice.ai/
Models57 public models
Prepaid routes57
BYOK routes40
Zero data retentionno
Confidential computeno
Provider E2EEno
Policy noteMixed model-specific posture. TrustedRouter cannot independently verify a complete chain from a live Venice endpoint through immutable source and hardware measurements to committed model weights. Venice is therefore not tracked as confidential or E2EE. Exact routes marked Private in Venice's live catalog qualify only as policy-backed ZDR through endpoint-specific records; Anonymized routes do not.
Policy source

Measured performance

46 samples

Continuously sampled across Venice's routed models: p50 TTFT, effective throughput, and success rate. Effective throughput uses provider-reported output tokens over complete request time. Unsupported route and probe-configuration rows are separated from provider downtime. No prompt or output content stored.

p50 TTFT2700 ms
Effective throughput39 tok/s n=9
Uptime93.48%
Modelp50 TTFTp50 TTFBEffective throughputUptimeConfig excludedAvailability samples
nvidia/nemotron-3.5-lightning 827 ms 826 ms 50 tok/s n=1 100.00% 2
qwen/qwen3-5-35b-a3b 1064 ms 1064 ms 100.00% 2
moonshotai/kimi-k3 1468 ms 1468 ms 39 tok/s n=2 100.00% 2
deepseek/deepseek-v4-flash-0731-fast 1709 ms 1709 ms 100.00% 1
z-ai/glm-5.2 1725 ms 1725 ms 20 tok/s n=1 100.00% 2
qwen/qwen3-235b-a22b-instruct-2507 1777 ms 1777 ms 100.00% 2
moonshotai/kimi-k2-6 1884 ms 1884 ms 100.00% 2
qwen/qwen3-vl-235b-a22b 2562 ms 2561 ms 100.00% 1
qwen/qwen3-next-80b 2700 ms 2699 ms 100.00% 1
z-ai/glm-4.7 2719 ms 2719 ms 100.00% 2
qwen/qwen3-235b-a22b-thinking-2507 3076 ms 3075 ms 100.00% 1
google/gemma-4-uncensored 3085 ms 3085 ms 100.00% 1
minimax/minimax-m25 3513 ms 3513 ms 100.00% 2
z-ai/glm-5v-turbo 3521 ms 3520 ms 100.00% 1
meta-llama/llama-3.2-3b 3825 ms 3824 ms 100.00% 1
moonshotai/kimi-k2-7-code 4001 ms 4000 ms 100.00% 1
qwen/qwen-3-8-2-4t-a95b 4218 ms 4218 ms 100.00% 3
qwen/qwen-3-6-plus 4370 ms 4370 ms 100.00% 3
qwen/qwen-3-7-plus 4545 ms 4545 ms 100.00% 3
z-ai/glm-4.6 4658 ms 4658 ms 100.00% 1
z-ai/glm-5 5333 ms 5333 ms 100.00% 1
deepseek/deepseek-v4-pro-0423 8112 ms 8111 ms 28 tok/s n=2 100.00% 1
z-ai/glm-5-turbo 15060 ms 15059 ms 100.00% 1
qwen/qwen3.6-27b 889 ms 888 ms 75.00% 8
deepseek/deepseek-v3.2 0.00% 1
deepseek/deepseek-v4-flash 50 tok/s n=2 0
deepseek/deepseek-v4-flash-0731 58 tok/s n=1 0

Venice performance history · Full provider & model leaderboard.

Provider models

Models served by Venice.

Each row links to pricing, provider, benchmark, and API pages for the model.

Model AI IQ Context Endpoints Prompt Completion Routes
alibaba/wan-2.7
Alibaba Wan 2.7
5,000 1 selected route selected route prepaid
bytedance/seedance-2.0
ByteDance Seedance 2.0
10,000 1 selected route selected route prepaid
bytedance/seedance-2.0-fast
ByteDance Seedance 2.0 Fast
10,000 1 selected route selected route prepaid
deepseek/deepseek-v3.2
DeepSeek: DeepSeek V3.2
IQ 103#64 163,840 2 $0.34815/1M $0.5064/1M prepaid BYOK
deepseek/deepseek-v4-flash
DeepSeek: DeepSeek V4 Flash 0423
IQ 113#41 1,048,576 2 $0.14559/1M $0.290125/1M prepaid BYOK
deepseek/deepseek-v4-flash-0731
DeepSeek: DeepSeek V4 Flash 0731
1,048,576 2 $0.184625/1M $0.36925/1M prepaid BYOK
deepseek/deepseek-v4-flash-0731-fast
DeepSeek V4 Flash 0731 Fast
1,000,000 2 $0.36925/1M $0.7385/1M prepaid BYOK
deepseek/deepseek-v4-pro
DeepSeek: DeepSeek V4 Pro
IQ 117#31 1,048,576 2 $1.74075/1M $3.482555/1M prepaid BYOK
deepseek/deepseek-v4-pro-0423
DeepSeek V4 Pro 0423
1,048,576 1 $1.74075/1M $3.482555/1M prepaid
google/gemini-omni-flash
Google Gemini Omni Flash
1,048,576 1 selected route selected route prepaid
google/gemma-4-uncensored
Gemma 4 Uncensored
256,000 2 $0.171438/1M $0.5275/1M prepaid BYOK
google/veo-3.1
Google Veo 3.1
2,500 1 selected route selected route prepaid
google/veo-3.1-fast
Google Veo 3.1 Fast
2,500 1 selected route selected route prepaid
kling/o3-pro
Kling Video 3.0 Omni Pro
3,072 1 selected route selected route prepaid
kling/v3-pro
Kling Video 3.0 Pro
3,072 1 selected route selected route prepaid
lightricks/ltx-2.3
Lightricks LTX 2.3
5,000 1 selected route selected route prepaid
lightricks/ltx-2.3-fast
Lightricks LTX 2.3 Fast
5,000 1 selected route selected route prepaid
meta-llama/llama-3.2-3b
Llama 3.2 3B
128,000 2 $0.15825/1M $0.633/1M prepaid BYOK
meta-llama/llama-3.3-70b
Llama 3.3 70B
128,000 2 $0.7385/1M $2.954/1M prepaid BYOK
minimax/hailuo-3
MiniMax Hailuo 3 (H3)
7,000 1 selected route selected route prepaid
minimax/minimax-m25
MiniMax M2.5
198,000 2 $0.28485/1M $1.00225/1M prepaid BYOK
minimax/minimax-m27
MiniMax M2.7
198,000 2 $0.395625/1M $1.5825/1M prepaid BYOK
minimax/minimax-m3-preview
MiniMax M3 Preview
IQ 114#39 524,288 2 $0.3165/1M $1.266/1M prepaid BYOK
moonshotai/kimi-k2-5
Kimi K2.5
256,000 2 $0.5908/1M $3.6925/1M prepaid BYOK
moonshotai/kimi-k2-6
Kimi K2.6
256,000 2 $0.79125/1M $3.6925/1M prepaid BYOK
moonshotai/kimi-k2-7-code
Kimi K2.7 Code
256,000 2 $0.79125/1M $3.6925/1M prepaid BYOK
moonshotai/kimi-k3
MoonshotAI: Kimi K3
IQ 123#16 1,048,576 2 $3.95625/1M $19.78125/1M prepaid BYOK
moonshotai/kimi-k3-fast-api
Kimi K3 Fast
1,000,000 2 $4.7475/1M $23.7375/1M prepaid BYOK
nvidia/nemotron-3.5-lightning
NVIDIA: Nemotron 3.5 Lightning
1,000,000 2 $0.1055/1M $0.26375/1M prepaid BYOK
openai/sora-2
OpenAI Sora 2
2,500 1 selected route selected route prepaid
openai/sora-2-pro
OpenAI Sora 2 Pro
2,500 1 selected route selected route prepaid
pixverse/c1
PixVerse C1
2,500 1 selected route selected route prepaid
qwen/qwen-3-6-plus
Qwen 3.6 Plus Uncensored
1,000,000 2 $0.659375/1M $3.95625/1M prepaid BYOK
qwen/qwen-3-7-max
Qwen 3.7 Max
1,000,000 2 $2.8485/1M $8.49275/1M prepaid BYOK
qwen/qwen-3-7-plus
Qwen 3.7 Plus
1,000,000 2 $0.5275/1M $2.11/1M prepaid BYOK
qwen/qwen-3-8-2-4t-a95b
Qwen 3.8 2.4T
262,144 2 $2.6375/1M $7.9125/1M prepaid BYOK
qwen/qwen-3-8-max
Qwen 3.8 Max
1,000,000 2 $2.6375/1M $7.9125/1M prepaid BYOK
qwen/qwen3-235b-a22b-instruct-2507
Qwen3 235B A22B Instruct 2507
131,072 2 $0.15825/1M $0.79125/1M prepaid BYOK
qwen/qwen3-235b-a22b-thinking-2507
Qwen: Qwen3 235B A22B Thinking 2507
262,144 2 $0.47475/1M $3.6925/1M prepaid BYOK
qwen/qwen3-5-35b-a3b
Qwen 3.5 35B A3B
256,000 2 $0.329688/1M $1.31875/1M prepaid BYOK
qwen/qwen3-6-35b-a3b
Qwen 3.6 35B A3B
256,000 2 $0.1055/1M $1.055/1M prepaid BYOK
qwen/qwen3-coder-480b-a35b-instruct-turbo
Qwen/Qwen3-Coder-480B-A35B-Instruct-Turbo
262,144 2 $0.36925/1M $1.5825/1M prepaid BYOK
qwen/qwen3-next-80b
Qwen 3 Next 80b
256,000 2 $0.36925/1M $2.0045/1M prepaid BYOK
qwen/qwen3-vl-235b-a22b
Qwen3 VL 235B
128,000 2 $0.22155/1M $2.0045/1M prepaid BYOK
qwen/qwen3.5-397b-a17b
Qwen: Qwen3.5 397B A17B
262,144 2 $0.79125/1M $4.7475/1M prepaid BYOK
qwen/qwen3.5-9b
Qwen: Qwen3.5-9B
IQ 93#99 262,144 2 $0.1055/1M $0.15825/1M prepaid BYOK
qwen/qwen3.6-27b
Qwen: Qwen3.6 27B
IQ 111#47 262,144 2 $0.342875/1M $3.42875/1M prepaid BYOK
runway/gen-4.5
Runway Gen-4.5
1,000 1 selected route selected route prepaid
shengshu/vidu-q3
ShengShu Vidu Q3
2,500 1 selected route selected route prepaid
z-ai/glm-4.6
Z.ai: GLM 4.6
204,800 2 $0.45365/1M $1.84625/1M prepaid BYOK
z-ai/glm-4.7
Z.ai: GLM 4.7
IQ 103#66 204,800 2 $0.58025/1M $2.79575/1M prepaid BYOK
z-ai/glm-4.7-flash
Z.ai: GLM 4.7 Flash
202,752 2 $0.0633/1M $0.422/1M prepaid BYOK
z-ai/glm-5
Z.ai: GLM 5
IQ 105#60 204,800 2 $1.055/1M $3.376/1M prepaid BYOK
z-ai/glm-5-turbo
Z.ai: GLM 5 Turbo
202,752 2 $1.266/1M $4.22/1M prepaid BYOK
z-ai/glm-5.1
Z.ai: GLM 5.1
IQ 114#38 204,800 2 $1.6247/1M $5.1062/1M prepaid BYOK
z-ai/glm-5.2
Z.ai: GLM 5.2
IQ 120#19 1,048,576 2 $1.477/1M $4.642/1M prepaid BYOK
z-ai/glm-5v-turbo
Z.ai: GLM 5V Turbo
202,752 2 $1.5825/1M $5.275/1M prepaid BYOK

Questions

Does Venice have zero data retention?

TrustedRouter does not currently mark Venice as provider-level zero data retention. Use trustedrouter/zdr or provider.min_privacy=zdr to select a different eligible route, and review the linked policy source for changes.

Is Venice end-to-end encrypted?

TrustedRouter does not currently mark Venice as end-to-end encrypted at the provider boundary. The TrustedRouter gateway is still attested, but the selected provider normally receives the request in order to run the model. Use trustedrouter/e2e for the stronger route requirement.

Which Venice models are available through TrustedRouter?

This page currently lists 57 public Venice models, with live pricing, route count, context length, measured performance when available, and links to each model's provider and benchmark pages.

Workspace access

Sign in

Choose a sign in method to access your TrustedRouter workspace.

By signing in you agree to the terms of service and privacy policy.