OpenAI compatible API · Attested · Public status
Crusoe
Crusoe models on TrustedRouter with prices, routes, policy notes, and source links.
Onebase URL to migrate
100sof models and routes
0prompt or output logs. Always.
crusoe
No provider claim
| Provider | Crusoe |
|---|---|
| Models | 14 public models |
| Prepaid routes | 14 |
| BYOK routes | 14 |
| Zero data retention | not claimed |
| Confidential compute | not claimed |
| Provider E2EE | not claimed |
| Policy note | No provider-ZDR claim is tracked here. Crusoe's Managed Inference docs and pricing/catalog pages are linked for model and API data-handling review. Policy source |
Measured performance
209 samplesContinuously sampled across Crusoe's routed models: p50 TTFT, effective throughput, and success rate. Effective throughput uses provider-reported output tokens over complete request time. Unsupported route and probe-configuration rows are separated from provider downtime. No prompt or output content stored.
| p50 TTFT | 2627 ms |
|---|---|
| Effective throughput | — |
| Uptime | 97.13% |
| Model | p50 TTFT | p50 TTFB | Effective throughput | Uptime | Config excluded | Availability samples |
|---|---|---|---|---|---|---|
| deepseek/deepseek-v4-flash | 2385 ms | 2385 ms | — | 100.00% | — | 19 |
| yutori/n1.5 | 2437 ms | 2437 ms | — | 100.00% | — | 20 |
| meta-llama/llama-3.3-70b-instruct | 2439 ms | 2439 ms | — | 100.00% | — | 16 |
| nvidia/nemotron-3-super-120b-a12b | 2532 ms | 2532 ms | — | 100.00% | — | 13 |
| qwen/qwen3-235b-a22b-2507 | 2620 ms | 2620 ms | — | 100.00% | — | 9 |
| nvidia/nemotron-3-nano-30b-a3b | 2627 ms | 2627 ms | — | 100.00% | — | 13 |
| nvidia/nemotron-3-nano-omni-reasoning-30b-a3b | 2665 ms | 2665 ms | — | 100.00% | — | 15 |
| openai/gpt-oss-120b | 2714 ms | 2714 ms | — | 100.00% | — | 14 |
| google/gemma-4-31b-it | 3057 ms | 3057 ms | — | 100.00% | — | 7 |
| z-ai/glm-5.1 | 3693 ms | 3693 ms | — | 100.00% | — | 23 |
| moonshotai/kimi-k2.6 | 2773 ms | 2773 ms | — | 94.12% | — | 17 |
| z-ai/glm-5.2 | 3288 ms | 3288 ms | — | 93.75% | — | 16 |
| deepseek/deepseek-v3-0324 | 2585 ms | 2585 ms | — | 92.86% | — | 14 |
| deepseek/deepseek-v4-pro | 2403 ms | 2403 ms | — | 76.92% | — | 13 |
Crusoe performance history · Full provider & model leaderboard.
Provider models
Models served by Crusoe.
Each row links to pricing, provider, benchmark, and API pages for the model.
| Model | AI IQ | Context | Endpoints | Prompt | Completion | Routes |
|---|---|---|---|---|---|---|
deepseek/deepseek-v3-0324DeepSeek V3 0324 |
— | 163,840 | 2 | $0.525/1M | $1.575/1M | prepaid BYOK |
deepseek/deepseek-v4-flashDeepSeek: DeepSeek V4 Flash |
IQ 108#46 | 1,048,576 | 2 | $0.147/1M | $0.294/1M | prepaid BYOK |
deepseek/deepseek-v4-proDeepSeek: DeepSeek V4 Pro |
IQ 115#30 | 1,048,576 | 2 | $1.827/1M | $3.654/1M | prepaid BYOK |
google/gemma-4-31b-itGoogle: Gemma 4 31B |
IQ 101#66 | 262,144 | 2 | $0.147/1M | $0.42/1M | prepaid BYOK |
meta-llama/llama-3.3-70b-instructMeta: Llama 3.3 70B Instruct |
— | 131,072 | 2 | $0.2625/1M | $0.7875/1M | prepaid BYOK |
moonshotai/kimi-k2.6MoonshotAI: Kimi K2.6 |
IQ 119#19 | 262,144 | 2 | $0.735/1M | $3.675/1M | prepaid BYOK |
nvidia/nemotron-3-nano-30b-a3bNVIDIA: Nemotron 3 Nano 30B A3B |
— | 262,144 | 2 | $0.0525/1M | $0.21/1M | prepaid BYOK |
nvidia/nemotron-3-nano-omni-reasoning-30b-a3bnvidia/Nemotron-3-Nano-Omni-Reasoning-30B-A3B |
— | 262,144 | 2 | $0.315/1M | $1.9215/1M | prepaid BYOK |
nvidia/nemotron-3-super-120b-a12bNVIDIA: Nemotron 3 Super |
— | 1,000,000 | 2 | $0.315/1M | $2.52/1M | prepaid BYOK |
openai/gpt-oss-120bOpenAI: gpt-oss-120b |
IQ 105#53 | 131,072 | 2 | $0.0525/1M | $0.2625/1M | prepaid BYOK |
qwen/qwen3-235b-a22b-2507Qwen: Qwen3 235B A22B Instruct 2507 |
— | 262,144 | 2 | $0.231/1M | $0.84/1M | prepaid BYOK |
yutori/n1.5yutori/n1.5 |
— | 128,000 | 2 | $1.575/1M | $5.25/1M | prepaid BYOK |
z-ai/glm-5.1Z.ai: GLM 5.1 |
IQ 114#32 | 204,800 | 2 | $1.26/1M | $4.62/1M | prepaid BYOK |
z-ai/glm-5.2Z.ai: GLM 5.2 |
IQ 120#16 | 1,048,576 | 2 | $1.47/1M | $4.62/1M | prepaid BYOK |