OpenAI compatible API · Attested · Public status

Telnyx

Explore Telnyx models on TrustedRouter with current routes, token pricing, policy sources, privacy posture, regional availability, and API support.

Verify gateway
Onebase URL to migrate
100sof models and routes
0prompt or output logs. Always.

Telnyxtelnyx

No provider claim

All providers

Telnyx Inference serves OpenAI-compatible chat completions on Telnyx-owned GPUs. Marketplace routes include Telnyx-hosted models only, not its third-party BYOK passthrough catalog.

These privacy labels describe Telnyx, the upstream model provider. ZDR is a retention policy; verified confidential inference additionally requires attested provider compute and end-to-end encryption.

ProviderTelnyx
Legal entityTelnyx LLC; Illinois, USA limited liability company
HeadquartersAustin, Texas, USA
Registered and operating address600 Congress Ave, Floor 14, Austin, TX 78701, USA
DUNS966115342
EIN27-0273220
CEODavid Casem
Serving regionsUSA, EU, Australia and UAE; availability varies by model. Region availability alone is not a residency guarantee.
APIOpenAI-compatible chat completions, streaming, tools, structured output and reasoning content on supported models.
CatalogAuthenticated native OpenAI catalog; not yet a Provider Reliability Contract v2 declaration.
Routing statusActive
Provider websitehttps://telnyx.com/products/inference
Models17 public models
Credits routes17
Zero data retentionnot claimed
Verified confidential inferenceNot verified
Policy noteNo provider-ZDR or confidential-compute claim is tracked here. Telnyx's privacy policy is linked for users who need to review inference data handling.
Policy source
Documentation and policy sources

Measured performance

22 samples

Continuously sampled across Telnyx's routed models: p50 TTFT, effective throughput, and success rate. Effective throughput uses provider-reported output tokens over complete request time. Unsupported route and probe-configuration rows are separated from provider downtime. No prompt or output content stored.

p50 TTFT1523 ms
Effective throughput96 tok/s n=2
Uptime100.00%
Modelp50 TTFTEffective throughputUptimeConfig excludedAvailability samples
moonshotai/kimi-k3 861 ms 100.00% 2
minimax/minimax-m3 1170 ms 132 tok/s n=1 100.00% 4
z-ai/glm-5.3-flash 1327 ms 100.00% 3
qwen/qwen3.8-27b 1523 ms 100.00% 2
meta-llama/llama-3.1-70b-instruct 1612 ms 100.00% 1
meta-llama/llama-3.3-70b-instruct 1645 ms 100.00% 1
z-ai/glm-5.1 1717 ms 100.00% 2
deepseek/deepseek-v4-flash-0731 1837 ms 100.00% 1
moonshotai/kimi-k2.6 1838 ms 100.00% 1
moonshotai/kimi-k2.5 2020 ms 100.00% 2
qwen/qwen3-235b-a22b 2668 ms 100.00% 1
google/gemma-2b-it 3363 ms 100.00% 1
meta-llama/llama-3.1-8b-instruct 3759 ms 100.00% 1
z-ai/glm-5.2 60 tok/s n=1 0

Telnyx performance history · Full provider & model leaderboard.

Provider models

Models served by Telnyx.

Each row links to pricing, provider, benchmark, and API pages for the model.

Model AI IQ Context Input Cached input Output
deepseek/deepseek-v4-flash-0731
DeepSeek: DeepSeek V4 Flash 0731
1,048,576 $0.13715/1M $0.03165/1M $0.2743/1M
deepseek/deepseek-v4.1-flash
DeepSeek: DeepSeek V4.1 Flash
IQ 116#38 1,048,576 $0.3165/1M $0.01/1M $1.266/1M
google/gemma-2b-it
gemma-2b-it
8,192 $0.211/1M $0.211/1M $0.211/1M
meta-llama/llama-3.1-70b-instruct
Meta: Llama 3.1 70B Instruct
131,072 $0.633/1M $0.633/1M $0.633/1M
meta-llama/llama-3.1-8b-instruct
Meta: Llama 3.1 8B Instruct
131,072 $0.211/1M $0.211/1M $0.211/1M
meta-llama/llama-3.3-70b-instruct
Meta: Llama 3.3 70B Instruct
131,072 $0.633/1M $0.633/1M $0.633/1M
minimax/minimax-m2.7
MiniMax: MiniMax M2.7
IQ 107#72 196,608 $0.22155/1M $0.03165/1M $1.266/1M
minimax/minimax-m3
MiniMax: MiniMax M3
IQ 115#45 524,288 $0.28485/1M $0.0844/1M $1.1605/1M
moonshotai/kimi-k2.5
MoonshotAI: Kimi K2.5
IQ 111#57 262,144 $0.633/1M $0.422/1M $3.165/1M
moonshotai/kimi-k2.6
MoonshotAI: Kimi K2.6
IQ 118#35 262,144 $0.701575/1M $0.0844/1M $4.22/1M
moonshotai/kimi-k3
MoonshotAI: Kimi K3
IQ 121#27 1,048,576 $2.8485/1M $0.28485/1M $14.2425/1M
qwen/qwen3-235b-a22b
Qwen3-235B-A22B
32,768 $0.633/1M $0.422/1M $2.11/1M
qwen/qwen3.8-27b
Qwen: Qwen3.8 27B
IQ 110#60 262,144 $0.422/1M $0.05275/1M $3.165/1M
z-ai/glm-5.1
Z.ai: GLM 5.1
IQ 113#49 202,752 $1.0339/1M $0.13715/1M $4.642/1M
z-ai/glm-5.2
Z.ai: GLM 5.2
IQ 120#28 1,048,576 $1.055/1M $0.211/1M $4.22/1M
z-ai/glm-5.3
Z.ai: GLM 5.3
IQ 123#18 1,048,576 $1.31875/1M $0.2532/1M $4.22/1M
z-ai/glm-5.3-flash
Z.ai: GLM 5.3 Flash
IQ 116#40 1,048,576 $0.142425/1M $0.028485/1M $0.47475/1M

Questions

Does Telnyx have zero data retention?

TrustedRouter does not currently mark Telnyx as provider-level zero data retention. Use trustedrouter/zdr or provider.min_privacy=zdr to select a different eligible route, and review the linked policy source for changes.

Is Telnyx end-to-end encrypted?

TrustedRouter does not currently mark Telnyx as end-to-end encrypted at the provider boundary. The TrustedRouter gateway is still attested, but the selected provider normally receives the request in order to run the model. Use trustedrouter/e2e for the stronger route requirement.

Which Telnyx models are available through TrustedRouter?

This page currently lists 17 public Telnyx models, with live pricing, route count, context length, measured performance when available, and links to each model's provider and benchmark pages.

Workspace access

Sign in

Choose a sign in method to access your TrustedRouter workspace.

By signing in you agree to the terms of service and privacy policy.