OpenAI compatible API · Attested · Public status

Novita AI

Explore Novita AI models on TrustedRouter with current routes, token pricing, policy sources, privacy posture, regional availability, and API support.

Verify gateway
Onebase URL to migrate
100sof models and routes
0prompt or output logs. Always.

Novita AInovita

No provider claim

All providers

ProviderNovita AI
Provider websitehttps://novita.ai/
Models109 public models
Prepaid routes91
BYOK routes109
Zero data retentionnot claimed
Confidential computenot claimed
Provider E2EEnot claimed
Policy noteNo provider-ZDR claim is tracked here. Novita's privacy policy says personal information is not used for model training; customer-content processing is governed by customer agreements.
Policy source

Measured performance

467 samples

Continuously sampled across Novita AI's routed models: p50 TTFT, effective throughput, and success rate. Effective throughput uses provider-reported output tokens over complete request time. Unsupported route and probe-configuration rows are separated from provider downtime. No prompt or output content stored.

p50 TTFT7695 ms
Effective throughput57 tok/s n=14
Uptime33.83%
Modelp50 TTFTp50 TTFBEffective throughputUptimeConfig excludedAvailability samples
microsoft/wizardlm-2-8x22b 7695 ms 4820 ms 25.67% 409
qwen/qwen3-235b-a22b-instruct-2507 985 ms 985 ms 100.00% 1
zai-org/glm-4.5-air 1181 ms 1181 ms 100.00% 1
minimax/minimax-m3 1206 ms 1206 ms 57 tok/s n=2 100.00% 1
qwen/qwen3-omni-30b-a3b-thinking 1230 ms 1229 ms 100.00% 1
deepseek/deepseek-v4-flash-0731 1348 ms 1348 ms 58 tok/s n=1 100.00% 1
moonshotai/kimi-k2-instruct 1394 ms 1393 ms 100.00% 2
qwen/qwen3-235b-a22b-thinking-2507 1403 ms 1403 ms 100.00% 1
sao10k/l31-70b-euryale-v2.2 1802 ms 1802 ms 100.00% 1
openai/gpt-oss-20b 1850 ms 1850 ms 100.00% 1
deepseek/deepseek-r1-distill-llama-70b 1908 ms 1908 ms 100.00% 1
moonshotai/kimi-k2-0905 1923 ms 1923 ms 100.00% 3 provider_error 2
deepseek/deepseek-v3-turbo 2181 ms 2180 ms 100.00% 2
minimax/minimax-m2.5-highspeed 2336 ms 2336 ms 100.00% 1
qwen/qwen3.8-max 2821 ms 2821 ms 100.00% 2
Sao10K/L3-8B-Stheno-v3.2 2859 ms 2859 ms 100.00% 8
qwen/qwen3-next-80b-a3b-instruct 3268 ms 3268 ms 100.00% 1
inclusionai/ring-2.6-1t 3415 ms 3414 ms 100.00% 1
minimaxai/minimax-m1-80k 3650 ms 3650 ms 100.00% 2
qwen/qwen3.6-27b 3714 ms 3713 ms 100.00% 1
moonshotai/kimi-k2.7-code 3806 ms 3806 ms 42 tok/s n=1 100.00% 1
qwen/qwen3.5-397b-a17b 3951 ms 3951 ms 100.00% 1
minimax/minimax-m2.1 4159 ms 4158 ms 100.00% 2
qwen/qwen3-omni-30b-a3b-instruct 4250 ms 4250 ms 100.00% 1
inclusionai/ling-2.6-flash 4267 ms 4267 ms 100.00% 1
google/gemma-4-31b-it 4323 ms 4322 ms 31 tok/s n=1 100.00% 2
qwen/qwen3-vl-235b-a22b-instruct 4430 ms 4430 ms 100.00% 1
deepseek/deepseek-r1-0528 4569 ms 4569 ms 100.00% 1
z-ai/glm-5.2 4707 ms 4707 ms 60 tok/s n=2 100.00% 1
stepfun/step-3.7-flash 8329 ms 8329 ms 100.00% 3
moonshotai/kimi-k2.5 18267 ms 18267 ms 100.00% 1
baidu/ernie-4.5-vl-424b-a47b 100.00% 2
kwaipilot/kat-coder-pro 100.00% 1
mindai/macaron-v1-venti 100.00% 1
qwen/qwen-2.5-72b-instruct 100.00% 1
zai-org/autoglm-phone-9b-multilingual 100.00% 1 provider_error 1
zai-org/glm-4.6 3643 ms 3642 ms 50.00% 2
baichuan/baichuan-m2-32b 0.00% 2
baidu/cobuddy 0.00% 1
deepseek/deepseek-v4-flash 61 tok/s n=2 0
google/gemma-3-12b-it 0.00% 1
inclusionai/ling-3.0-flash 121 tok/s n=1 0
moonshotai/kimi-k2.6 34 tok/s n=1 0
moonshotai/kimi-k3 29 tok/s n=1 0
openai/gpt-oss-120b 20 tok/s n=1 0
qwen/qwen3.7-max 85 tok/s n=1 0

Novita AI performance history · Full provider & model leaderboard.

Provider models

Models served by Novita AI.

Each row links to pricing, provider, benchmark, and API pages for the model.

Model AI IQ Context Endpoints Prompt Completion Routes
Sao10K/L3-8B-Stheno-v3.2
L3 8B Stheno V3.2
8,192 2 $0.05275/1M $0.05275/1M prepaid BYOK
baichuan/baichuan-m2-32b
BaiChuan M2 32B
131,072 2 $0.07385/1M $0.07385/1M prepaid BYOK
baidu/cobuddy
CoBuddy
131,072 2 $0.2954/1M $1.19215/1M prepaid BYOK
baidu/ernie-4.5-21B-a3b
ERNIE 4.5 21B A3B
120,000 2 $0.07385/1M $0.2954/1M prepaid BYOK
baidu/ernie-4.5-21B-a3b-thinking
ERNIE-4.5-21B-A3B-Thinking
131,072 1 $0.07385/1M $0.2954/1M BYOK
baidu/ernie-4.5-300b-a47b-paddle
ERNIE 4.5 300B A47B
123,000 1 $0.2954/1M $1.1605/1M BYOK
baidu/ernie-4.5-vl-28b-a3b-thinking
ERNIE-4.5-VL-28B-A3B-Thinking
131,072 1 $0.41145/1M $0.41145/1M BYOK
baidu/ernie-4.5-vl-424b-a47b
Baidu: ERNIE 4.5 VL 424B A47B
123,000 2 $0.4431/1M $1.31875/1M prepaid BYOK
deepseek/deepseek-ocr
DeepSeek-OCR
8,192 2 $0.03165/1M $0.03165/1M prepaid BYOK
deepseek/deepseek-ocr-2
DeepSeek-OCR 2
8,192 2 $0.03165/1M $0.03165/1M prepaid BYOK
deepseek/deepseek-prover-v2-671b
Deepseek Prover V2 671B
160,000 1 $0.7385/1M $2.6375/1M BYOK
deepseek/deepseek-r1-0528
DeepSeek: R1 0528
163,840 2 $0.7385/1M $2.6375/1M prepaid BYOK
deepseek/deepseek-r1-0528-qwen3-8b
DeepSeek R1 0528 Qwen3 8B
128,000 1 $0.0633/1M $0.09495/1M BYOK
deepseek/deepseek-r1-distill-llama-70b
DeepSeek: R1 Distill Llama 70B
8,192 2 $0.844/1M $0.844/1M prepaid BYOK
deepseek/deepseek-r1-turbo
DeepSeek R1 (Turbo)
64,000 2 $0.7385/1M $2.6375/1M prepaid BYOK
deepseek/deepseek-v3-0324
DeepSeek V3 0324
163,840 2 $0.28485/1M $1.1816/1M prepaid BYOK
deepseek/deepseek-v3-turbo
DeepSeek V3 (Turbo)
64,000 2 $0.422/1M $1.3715/1M prepaid BYOK
deepseek/deepseek-v3.1
DeepSeek V3.1
IQ 96#88 131,072 2 $0.28485/1M $1.055/1M prepaid BYOK
deepseek/deepseek-v3.1-terminus
DeepSeek: DeepSeek V3.1 Terminus
163,840 2 $0.28485/1M $1.055/1M prepaid BYOK
deepseek/deepseek-v3.2
DeepSeek: DeepSeek V3.2
IQ 103#64 163,840 2 $0.283795/1M $0.422/1M prepaid BYOK
deepseek/deepseek-v3.2-exp
DeepSeek: DeepSeek V3.2 Exp
163,840 2 $0.28485/1M $0.43255/1M prepaid BYOK
deepseek/deepseek-v4-flash
DeepSeek: DeepSeek V4 Flash 0423
IQ 113#41 1,048,576 2 $0.1477/1M $0.2954/1M prepaid BYOK
deepseek/deepseek-v4-flash-0731
DeepSeek: DeepSeek V4 Flash 0731
1,048,576 2 $0.1477/1M $0.2954/1M prepaid BYOK
deepseek/deepseek-v4-pro
DeepSeek: DeepSeek V4 Pro
IQ 117#31 1,048,576 2 $1.688/1M $3.376/1M prepaid BYOK
google/gemma-3-12b-it
Google: Gemma 3 12B
131,072 2 $0.05275/1M $0.1055/1M prepaid BYOK
google/gemma-3-27b-it
Google: Gemma 3 27B
262,144 2 $0.125545/1M $0.211/1M prepaid BYOK
google/gemma-4-26b-a4b-it
Google: Gemma 4 26B A4B
IQ 96#89 262,144 2 $0.13715/1M $0.422/1M prepaid BYOK
google/gemma-4-31b-it
Google: Gemma 4 31B
IQ 101#73 262,144 2 $0.1477/1M $0.422/1M prepaid BYOK
gryphe/mythomax-l2-13b
MythoMax 13B
8,192 1 $0.09495/1M $0.09495/1M BYOK
inclusionai/ling-2.6-1t
inclusionAI: Ling-2.6-1T
262,144 2 $0.3165/1M $2.6375/1M prepaid BYOK
inclusionai/ling-2.6-flash
inclusionAI: Ling-2.6-flash
262,144 2 $0.1055/1M $0.3165/1M prepaid BYOK
inclusionai/ling-3.0-flash
Ling-3.0-flash
IQ 92#100 262,144 2 $0.0633/1M $0.1899/1M prepaid BYOK
inclusionai/ring-2.6-1t
inclusionAI: Ring-2.6-1T
IQ 101#75 262,144 2 $0.3165/1M $2.6375/1M prepaid BYOK
kwaipilot/kat-coder-pro
Kat Coder Pro
256,000 2 $0.3165/1M $1.266/1M prepaid BYOK
meta-llama/llama-3-8b-instruct
Llama 3 8B Instruct
8,192 1 $0.0422/1M $0.0422/1M BYOK
meta-llama/llama-3.1-8b-instruct
Meta: Llama 3.1 8B Instruct
131,072 2 $0.0211/1M $0.05275/1M prepaid BYOK
meta-llama/llama-3.2-3b-instruct
Meta: Llama 3.2 3B Instruct
131,072 1 $0.03165/1M $0.05275/1M BYOK
meta-llama/llama-3.3-70b-instruct
Meta: Llama 3.3 70B Instruct
131,072 2 $0.142425/1M $0.422/1M prepaid BYOK
meta-llama/llama-4-maverick-17b-128e-instruct-fp8
Llama 4 Maverick Instruct
1,048,576 2 $0.28485/1M $0.89675/1M prepaid BYOK
meta-llama/llama-4-scout-17b-16e-instruct
Llama 4 Scout Instruct
131,072 2 $0.1899/1M $0.62245/1M prepaid BYOK
microsoft/wizardlm-2-8x22b
WizardLM-2 8x22B
65,535 2 $0.6541/1M $0.6541/1M prepaid BYOK
mindai/macaron-v1-tall
Macaron V1 Tall
262,144 2 $0.47475/1M $2.743/1M prepaid BYOK
mindai/macaron-v1-venti
Macaron V1 Venti
1,048,576 2 $1.5825/1M $4.7475/1M prepaid BYOK
minimax/minimax-m2
MiniMax: MiniMax M2
204,800 2 $0.3165/1M $1.266/1M prepaid BYOK
minimax/minimax-m2.1
MiniMax: MiniMax M2.1
IQ 102#70 204,800 2 $0.3165/1M $1.266/1M prepaid BYOK
minimax/minimax-m2.5
MiniMax: MiniMax M2.5
IQ 105#63 204,800 2 $0.3165/1M $1.266/1M prepaid BYOK
minimax/minimax-m2.5-highspeed
MiniMax M2.5-highspeed
204,800 2 $0.633/1M $2.532/1M prepaid BYOK
minimax/minimax-m2.7
MiniMax: MiniMax M2.7
IQ 109#50 204,800 2 $0.3165/1M $1.266/1M prepaid BYOK
minimax/minimax-m3
MiniMax: MiniMax M3
IQ 114#39 1,048,576 2 $0.3165/1M $1.266/1M prepaid BYOK
minimaxai/minimax-m1-80k
MiniMax M1
1,000,000 2 $0.58025/1M $2.321/1M prepaid BYOK
mistralai/mistral-nemo
Mistral: Mistral Nemo
131,072 2 $0.0422/1M $0.17935/1M prepaid BYOK
moonshotai/kimi-k2-0905
MoonshotAI: Kimi K2 0905
262,144 2 $0.633/1M $2.6375/1M prepaid BYOK
moonshotai/kimi-k2-instruct
Kimi K2 Instruct
IQ 94#96 131,072 2 $0.60135/1M $2.4265/1M prepaid BYOK
moonshotai/kimi-k2-thinking
MoonshotAI: Kimi K2 Thinking
IQ 94#96 262,144 2 $0.633/1M $2.6375/1M prepaid BYOK
moonshotai/kimi-k2.5
MoonshotAI: Kimi K2.5
IQ 111#46 262,144 2 $0.633/1M $3.165/1M prepaid BYOK
moonshotai/kimi-k2.6
MoonshotAI: Kimi K2.6
IQ 119#23 262,144 2 $0.844/1M $3.587/1M prepaid BYOK
moonshotai/kimi-k2.7-code
MoonshotAI: Kimi K2.7 Code
IQ 118#28 262,144 2 $1.00225/1M $4.22/1M prepaid BYOK
moonshotai/kimi-k3
MoonshotAI: Kimi K3
IQ 123#16 1,048,576 2 $3.165/1M $15.825/1M prepaid BYOK
nousresearch/hermes-2-pro-llama-3-8b
Hermes 2 Pro Llama 3 8B
8,192 1 $0.1477/1M $0.1477/1M BYOK
nvidia/nemotron-3-nano-30b-a3b
NVIDIA: Nemotron 3 Nano 30B A3B
262,144 2 $0.05275/1M $0.211/1M prepaid BYOK
openai/gpt-oss-120b
OpenAI: gpt-oss-120b
IQ 105#61 131,072 2 $0.05275/1M $0.26375/1M prepaid BYOK
openai/gpt-oss-20b
OpenAI: gpt-oss-20b
IQ 100#77 131,072 2 $0.0422/1M $0.15825/1M prepaid BYOK
paddlepaddle/paddleocr-vl
PaddleOCR-VL
16,384 1 $0.0211/1M $0.0211/1M BYOK
qwen/qwen-2.5-72b-instruct
Qwen2.5 72B Instruct
32,768 2 $0.4009/1M $0.422/1M prepaid BYOK
qwen/qwen-mt-plus
Qwen MT Plus
16,384 2 $0.26375/1M $0.79125/1M prepaid BYOK
qwen/qwen2.5-7b-instruct
Qwen2.5 7B Instruct
32,000 1 $0.07385/1M $0.07385/1M BYOK
qwen/qwen2.5-vl-72b-instruct
Qwen: Qwen2.5 VL 72B Instruct
128,000 1 $0.844/1M $0.844/1M BYOK
qwen/qwen3-235b-a22b-fp8
Qwen3 235B A22B
40,960 2 $0.211/1M $0.844/1M prepaid BYOK
qwen/qwen3-235b-a22b-instruct-2507
Qwen3 235B A22B Instruct 2507
131,072 2 $0.09495/1M $0.6119/1M prepaid BYOK
qwen/qwen3-235b-a22b-thinking-2507
Qwen: Qwen3 235B A22B Thinking 2507
262,144 2 $0.3165/1M $3.165/1M prepaid BYOK
qwen/qwen3-30b-a3b-fp8
Qwen3 30B A3B
40,960 1 $0.09495/1M $0.47475/1M BYOK
qwen/qwen3-32b-fp8
Qwen3 32B
40,960 1 $0.1055/1M $0.47475/1M BYOK
qwen/qwen3-4b-fp8
Qwen3 4B
128,000 1 $0.03165/1M $0.03165/1M BYOK
qwen/qwen3-8b-fp8
Qwen3 8B
128,000 1 $0.036925/1M $0.14559/1M BYOK
qwen/qwen3-coder-30b-a3b-instruct
Qwen: Qwen3 Coder 30B A3B Instruct
262,144 2 $0.07385/1M $0.28485/1M prepaid BYOK
qwen/qwen3-coder-480b-a35b-instruct
Qwen3 Coder 480B A35B Instruct
262,144 2 $0.4009/1M $1.63525/1M prepaid BYOK
qwen/qwen3-coder-next
Qwen: Qwen3 Coder Next
262,144 2 $0.211/1M $1.5825/1M prepaid BYOK
qwen/qwen3-max
Qwen3 Max
262,144 2 $2.22605/1M $8.91475/1M prepaid BYOK
qwen/qwen3-next-80b-a3b-instruct
Qwen: Qwen3 Next 80B A3B Instruct
262,144 2 $0.15825/1M $1.5825/1M prepaid BYOK
qwen/qwen3-omni-30b-a3b-instruct
Qwen3 Omni 30B A3B Instruct
65,536 2 $0.26375/1M $1.02335/1M prepaid BYOK
qwen/qwen3-omni-30b-a3b-thinking
Qwen3 Omni 30B A3B Thinking
65,536 2 $0.26375/1M $1.02335/1M prepaid BYOK
qwen/qwen3-vl-235b-a22b-instruct
Qwen: Qwen3 VL 235B A22B Instruct
262,144 2 $0.3165/1M $1.5825/1M prepaid BYOK
qwen/qwen3-vl-235b-a22b-thinking
Qwen: Qwen3 VL 235B A22B Thinking
131,072 2 $1.0339/1M $4.16725/1M prepaid BYOK
qwen/qwen3-vl-30b-a3b-instruct
Qwen: Qwen3 VL 30B A3B Instruct
262,144 2 $0.211/1M $0.7385/1M prepaid BYOK
qwen/qwen3.5-122b-a10b
Qwen: Qwen3.5-122B-A10B
262,144 2 $0.422/1M $3.376/1M prepaid BYOK
qwen/qwen3.5-27b
Qwen: Qwen3.5-27B
262,144 2 $0.3165/1M $2.532/1M prepaid BYOK
qwen/qwen3.5-35b-a3b
Qwen: Qwen3.5-35B-A3B
262,144 2 $0.26375/1M $2.11/1M prepaid BYOK
qwen/qwen3.5-397b-a17b
Qwen: Qwen3.5 397B A17B
262,144 2 $0.633/1M $3.798/1M prepaid BYOK
qwen/qwen3.6-27b
Qwen: Qwen3.6 27B
IQ 111#47 262,144 2 $0.633/1M $3.798/1M prepaid BYOK
qwen/qwen3.6-35b-a3b
Qwen: Qwen3.6 35B A3B
IQ 100#78 262,144 2 $0.26164/1M $1.566675/1M prepaid BYOK
qwen/qwen3.7-max
Qwen3.7-Max
IQ 118#29 1,000,000 2 $1.31875/1M $3.95625/1M prepaid BYOK
qwen/qwen3.8-max
Qwen3.8 Max
IQ 119#26 1,000,000 2 $2.11/1M $6.33/1M prepaid BYOK
sao10k/l3-70b-euryale-v2.1
L3 70B Euryale V2.1
8,192 1 $1.5614/1M $1.5614/1M BYOK
sao10k/l3-8b-lunaris
Sao10k L3 8B Lunaris
8,192 2 $0.05275/1M $0.05275/1M prepaid BYOK
sao10k/l31-70b-euryale-v2.2
L31 70B Euryale V2.2
8,192 2 $1.5614/1M $1.5614/1M prepaid BYOK
stepfun/step-3.7-flash
StepFun: Step 3.7 Flash
IQ 101#76 262,144 2 $0.211/1M $1.21325/1M prepaid BYOK
tencent/hy3
Tencent: Hy3
IQ 103#67 262,144 2 $0.1477/1M $0.6119/1M prepaid BYOK
xiaomimimo/mimo-v2-flash
XiaomiMiMo/MiMo-V2-Flash
IQ 105#62 262,144 1 $0.1055/1M $0.3165/1M BYOK
xiaomimimo/mimo-v2.5-pro
XiaomiMiMo/MiMo-V2.5-Pro
IQ 115#36 1,048,576 2 $2.11/1M $6.33/1M prepaid BYOK
z-ai/glm-5.2
Z.ai: GLM 5.2
IQ 120#19 1,048,576 2 $1.477/1M $4.642/1M prepaid BYOK
zai-org/autoglm-phone-9b-multilingual
AutoGLM-Phone-9B-Multilingual
65,536 2 $0.036925/1M $0.14559/1M prepaid BYOK
zai-org/glm-4.5-air
zai-org/glm-4.5-air
131,072 2 $0.13715/1M $0.89675/1M prepaid BYOK
zai-org/glm-4.5v
GLM 4.5V
65,536 2 $0.633/1M $1.899/1M prepaid BYOK
zai-org/glm-4.6
GLM 4.6
204,800 2 $0.58025/1M $2.321/1M prepaid BYOK
zai-org/glm-4.6v
GLM 4.6V
131,072 2 $0.3165/1M $0.9495/1M prepaid BYOK
zai-org/glm-4.7
GLM-4.7
IQ 103#66 204,800 2 $0.633/1M $2.321/1M prepaid BYOK
zai-org/glm-4.7-flash
GLM-4.7-Flash
200,000 2 $0.07385/1M $0.422/1M prepaid BYOK
zai-org/glm-5
GLM-5
IQ 105#60 202,800 2 $1.055/1M $3.376/1M prepaid BYOK
zai-org/glm-5.1
GLM-5.1
IQ 114#38 204,800 2 $1.4559/1M $4.642/1M prepaid BYOK

Questions

Does Novita AI have zero data retention?

TrustedRouter does not currently mark Novita AI as provider-level zero data retention. Use trustedrouter/zdr or provider.min_privacy=zdr to select a different eligible route, and review the linked policy source for changes.

Is Novita AI end-to-end encrypted?

TrustedRouter does not currently mark Novita AI as end-to-end encrypted at the provider boundary. The TrustedRouter gateway is still attested, but the selected provider normally receives the request in order to run the model. Use trustedrouter/e2e for the stronger route requirement.

Which Novita AI models are available through TrustedRouter?

This page currently lists 109 public Novita AI models, with live pricing, route count, context length, measured performance when available, and links to each model's provider and benchmark pages.

Workspace access

Sign in

Choose a sign in method to access your TrustedRouter workspace.

By signing in you agree to the terms of service and privacy policy.