OpenAI compatible API · Attested · Public status
Venice
Venice models on TrustedRouter with prices, routes, policy notes, and source links.
Onebase URL to migrate
100sof models and routes
0prompt or output logs. Always.
venice
No provider claim
| Provider | Venice |
|---|---|
| Models | 28 public models |
| Prepaid routes | 28 |
| BYOK routes | 12 |
| Zero data retention | no |
| Confidential compute | no |
| Provider E2EE | no |
| Policy note | Mixed model-specific posture. TrustedRouter cannot independently verify a complete chain from a live Venice endpoint through immutable source and hardware measurements to committed model weights. Venice is therefore not tracked as confidential or E2EE. Exact routes marked Private in Venice's live catalog qualify only as policy-backed ZDR through endpoint-specific records; Anonymized routes do not. Policy source |
Measured performance
209 samplesContinuously sampled across Venice's routed models: p50 TTFT, effective throughput, and success rate. Effective throughput uses provider-reported output tokens over complete request time. Unsupported route and probe-configuration rows are separated from provider downtime. No prompt or output content stored.
| p50 TTFT | 2799 ms |
|---|---|
| Effective throughput | — |
| Uptime | 97.13% |
| Model | p50 TTFT | p50 TTFB | Effective throughput | Uptime | Config excluded | Availability samples |
|---|---|---|---|---|---|---|
| qwen/qwen3.5-9b | 2550 ms | 2550 ms | — | 100.00% | — | 17 |
| z-ai/glm-5.2 | 2578 ms | 2578 ms | — | 100.00% | — | 19 |
| qwen/qwen3.6-27b | 2646 ms | 2646 ms | — | 100.00% | — | 15 |
| qwen/qwen3-235b-a22b-thinking-2507 | 2700 ms | 2700 ms | — | 100.00% | — | 14 |
| z-ai/glm-4.7 | 2799 ms | 2799 ms | — | 100.00% | — | 26 |
| z-ai/glm-4.6 | 2992 ms | 2992 ms | — | 100.00% | — | 11 |
| z-ai/glm-5.1 | 3083 ms | 3082 ms | — | 100.00% | — | 17 |
| qwen/qwen3.5-397b-a17b | 3233 ms | 3233 ms | — | 100.00% | — | 16 |
| z-ai/glm-5 | 3396 ms | 3396 ms | — | 100.00% | — | 16 |
| z-ai/glm-5v-turbo | 3980 ms | 3980 ms | — | 100.00% | — | 18 |
| z-ai/glm-5-turbo | 3734 ms | 3734 ms | — | 86.67% | — | 15 |
| z-ai/glm-4.7-flash | 2694 ms | 2694 ms | — | 84.00% | — | 25 |
Venice performance history · Full provider & model leaderboard.
Provider models
Models served by Venice.
Each row links to pricing, provider, benchmark, and API pages for the model.
| Model | AI IQ | Context | Endpoints | Prompt | Completion | Routes |
|---|---|---|---|---|---|---|
alibaba/wan-2.7Alibaba Wan 2.7 |
— | 5,000 | 1 | selected route | selected route | prepaid |
bytedance/seedance-2.0ByteDance Seedance 2.0 |
— | 10,000 | 1 | selected route | selected route | prepaid |
bytedance/seedance-2.0-fastByteDance Seedance 2.0 Fast |
— | 10,000 | 1 | selected route | selected route | prepaid |
google/gemini-omni-flashGoogle Gemini Omni Flash |
— | 1,048,576 | 1 | selected route | selected route | prepaid |
google/veo-3.1Google Veo 3.1 |
— | 2,500 | 1 | selected route | selected route | prepaid |
google/veo-3.1-fastGoogle Veo 3.1 Fast |
— | 2,500 | 1 | selected route | selected route | prepaid |
kling/o3-proKling Video 3.0 Omni Pro |
— | 3,072 | 1 | selected route | selected route | prepaid |
kling/v3-proKling Video 3.0 Pro |
— | 3,072 | 1 | selected route | selected route | prepaid |
lightricks/ltx-2.3Lightricks LTX 2.3 |
— | 5,000 | 1 | selected route | selected route | prepaid |
lightricks/ltx-2.3-fastLightricks LTX 2.3 Fast |
— | 5,000 | 1 | selected route | selected route | prepaid |
minimax/hailuo-3MiniMax Hailuo 3 (H3) |
— | 7,000 | 1 | selected route | selected route | prepaid |
openai/sora-2OpenAI Sora 2 |
— | 2,500 | 1 | selected route | selected route | prepaid |
openai/sora-2-proOpenAI Sora 2 Pro |
— | 2,500 | 1 | selected route | selected route | prepaid |
pixverse/c1PixVerse C1 |
— | 2,500 | 1 | selected route | selected route | prepaid |
qwen/qwen3-235b-a22b-thinking-2507Qwen: Qwen3 235B A22B Thinking 2507 |
— | 262,144 | 2 | $0.4725/1M | $3.675/1M | prepaid BYOK |
qwen/qwen3.5-397b-a17bQwen: Qwen3.5 397B A17B |
— | 262,144 | 2 | $0.7875/1M | $4.725/1M | prepaid BYOK |
qwen/qwen3.5-9bQwen: Qwen3.5-9B |
IQ 93#90 | 262,144 | 2 | $0.105/1M | $0.1575/1M | prepaid BYOK |
qwen/qwen3.6-27bQwen: Qwen3.6 27B |
IQ 111#40 | 262,144 | 2 | $0.3465/1M | $3.4125/1M | prepaid BYOK |
runway/gen-4.5Runway Gen-4.5 |
— | 1,000 | 1 | selected route | selected route | prepaid |
shengshu/vidu-q3ShengShu Vidu Q3 |
— | 2,500 | 1 | selected route | selected route | prepaid |
z-ai/glm-4.6Z.ai: GLM 4.6 |
— | 204,800 | 2 | $0.4515/1M | $1.8375/1M | prepaid BYOK |
z-ai/glm-4.7Z.ai: GLM 4.7 |
IQ 103#59 | 204,800 | 2 | $0.5775/1M | $2.7825/1M | prepaid BYOK |
z-ai/glm-4.7-flashZ.ai: GLM 4.7 Flash |
— | 202,752 | 2 | $0.063/1M | $0.42/1M | prepaid BYOK |
z-ai/glm-5Z.ai: GLM 5 |
IQ 105#52 | 204,800 | 2 | $1.05/1M | $3.36/1M | prepaid BYOK |
z-ai/glm-5-turboZ.ai: GLM 5 Turbo |
— | 202,752 | 2 | $1.26/1M | $4.2/1M | prepaid BYOK |
z-ai/glm-5.1Z.ai: GLM 5.1 |
IQ 114#32 | 204,800 | 2 | $1.617/1M | $5.082/1M | prepaid BYOK |
z-ai/glm-5.2Z.ai: GLM 5.2 |
IQ 120#16 | 1,048,576 | 2 | $1.47/1M | $4.62/1M | prepaid BYOK |
z-ai/glm-5v-turboZ.ai: GLM 5V Turbo |
— | 202,752 | 2 | $1.575/1M | $5.25/1M | prepaid BYOK |