OpenAI compatible API · Attested · Public status
DigitalOcean Gradient AI
DigitalOcean Gradient AI models on TrustedRouter with prices, routes, policy notes, and source links.
Onebase URL to migrate
100sof models and routes
0prompt or output logs. Always.
digitalocean
No provider claim
| Provider | DigitalOcean Gradient AI |
|---|---|
| Models | 22 public models |
| Prepaid routes | 22 |
| BYOK routes | 22 |
| Zero data retention | not claimed |
| Confidential compute | not claimed |
| Provider E2EE | not claimed |
| Policy note | No provider-ZDR claim is tracked here. DigitalOcean's Gradient AI model and pricing documentation is linked for data-handling review. Policy source |
Measured performance
66 samplesContinuously sampled across DigitalOcean Gradient AI's routed models: p50 TTFT, effective throughput, and success rate. Effective throughput uses provider-reported output tokens over complete request time. Unsupported route and probe-configuration rows are separated from provider downtime. No prompt or output content stored.
| p50 TTFT | 1554 ms |
|---|---|
| Effective throughput | 28 tok/s n=7 |
| Uptime | 98.48% |
| Model | p50 TTFT | p50 TTFB | Effective throughput | Uptime | Config excluded | Availability samples |
|---|---|---|---|---|---|---|
| minimax/minimax-m2.5 | 651 ms | 651 ms | — | 100.00% | — | 1 |
| nvidia/nemotron-3-super-120b | 898 ms | 898 ms | — | 100.00% | — | 2 |
| nvidia/nemotron-3-nano-omni | 962 ms | 961 ms | — | 100.00% | — | 4 |
| qwen/qwen3.5-397b-a17b | 965 ms | 965 ms | — | 100.00% | — | 2 |
| mistralai/ministral-3-14b-instruct | 1112 ms | 1112 ms | — | 100.00% | — | 2 |
| z-ai/glm-5 | 1341 ms | 1341 ms | — | 100.00% | — | 3 |
| qwen/qwen3-coder-flash | 1399 ms | 1399 ms | — | 100.00% | — | 6 |
| meta-llama/llama-4-maverick | 1429 ms | 1429 ms | — | 100.00% | — | 4 |
| moonshotai/kimi-k2.6 | 1474 ms | 1474 ms | 29 tok/s n=1 | 100.00% | — | 6 |
| z-ai/glm-5.1 | 1554 ms | 1553 ms | — | 100.00% | — | 3 |
| moonshotai/kimi-k2.5 | 1572 ms | 1572 ms | — | 100.00% | — | 5 |
| deepseek/deepseek-v3.2 | 1596 ms | 1596 ms | — | 100.00% | — | 3 |
| nvidia/nemotron-3-ultra-550b | 1637 ms | 1637 ms | — | 100.00% | — | 3 |
| meta-llama/llama-3.3-70b-instruct | 1689 ms | 1689 ms | — | 100.00% | — | 2 |
| nvidia/nemotron-nano-12b-v2-vl | 1689 ms | 1688 ms | — | 100.00% | — | 7 |
| deepseek/deepseek-r1-distill-llama-70b | 3010 ms | 3010 ms | — | 100.00% | — | 2 |
| deepseek/deepseek-v4-flash | 3423 ms | 3423 ms | 20 tok/s n=1 | 100.00% | — | 3 |
| google/gemma-4-31b-it | 4371 ms | 4371 ms | — | 100.00% | — | 1 |
| deepseek/deepseek-v4-pro | 4645 ms | 4645 ms | 26 tok/s n=2 | 100.00% | — | 5 |
| z-ai/glm-5.2 | 5351 ms | 5351 ms | 28 tok/s n=2 | 50.00% | — | 2 |
| xiaomi/mimo-v2.5-pro | — | — | 64 tok/s n=1 | — | — | 0 |
DigitalOcean Gradient AI performance history · Full provider & model leaderboard.
Provider models
Models served by DigitalOcean Gradient AI.
Each row links to pricing, provider, benchmark, and API pages for the model.
| Model | AI IQ | Context | Endpoints | Prompt | Completion | Routes |
|---|---|---|---|---|---|---|
deepseek/deepseek-r1-distill-llama-70bDeepSeek: R1 Distill Llama 70B |
— | 8,192 | 2 | $1.0395/1M | $1.0395/1M | prepaid BYOK |
deepseek/deepseek-v3.2DeepSeek: DeepSeek V3.2 |
IQ 103#56 | 163,840 | 2 | $0.44625/1M | $1.428/1M | prepaid BYOK |
deepseek/deepseek-v4-flashDeepSeek: DeepSeek V4 Flash |
IQ 108#46 | 1,048,576 | 2 | $0.1176/1M | $0.2352/1M | prepaid BYOK |
deepseek/deepseek-v4-proDeepSeek: DeepSeek V4 Pro |
IQ 115#30 | 1,048,576 | 2 | $1.4616/1M | $2.9232/1M | prepaid BYOK |
google/gemma-4-31b-itGoogle: Gemma 4 31B |
IQ 101#65 | 262,144 | 2 | $0.189/1M | $0.525/1M | prepaid BYOK |
meta-llama/llama-3.3-70b-instructMeta: Llama 3.3 70B Instruct |
— | 131,072 | 2 | $0.6825/1M | $0.6825/1M | prepaid BYOK |
meta-llama/llama-4-maverickMeta: Llama 4 Maverick |
IQ 90#90 | 1,048,576 | 2 | $0.2625/1M | $0.9135/1M | prepaid BYOK |
minimax/minimax-m2.5MiniMax: MiniMax M2.5 |
IQ 105#54 | 204,800 | 2 | $0.23625/1M | $0.945/1M | prepaid BYOK |
mistralai/ministral-3-14b-instructMinistral 3 14B Instruct |
— | 262,144 | 2 | $0.21/1M | $0.21/1M | prepaid BYOK |
moonshotai/kimi-k2.5MoonshotAI: Kimi K2.5 |
IQ 111#39 | 262,144 | 2 | $0.39375/1M | $2.12625/1M | prepaid BYOK |
moonshotai/kimi-k2.6MoonshotAI: Kimi K2.6 |
IQ 119#19 | 262,144 | 2 | $0.798/1M | $3.36/1M | prepaid BYOK |
nvidia/nemotron-3-nano-omniNemotron Nano 3 Omni |
— | 65,536 | 2 | $0.525/1M | $0.945/1M | prepaid BYOK |
nvidia/nemotron-3-super-120bNemotron-3-Super-120B |
— | 1,000,000 | 2 | $0.2205/1M | $0.47775/1M | prepaid BYOK |
nvidia/nemotron-3-ultra-550bNemotron 3 Ultra |
— | 131,072 | 2 | $0.945/1M | $1.785/1M | prepaid BYOK |
nvidia/nemotron-nano-12b-v2-vlNemotron Nano 12B v2 VL |
— | 128,000 | 2 | $0.21/1M | $0.63/1M | prepaid BYOK |
qwen/qwen3-32bQwen: Qwen3 32B |
— | 131,072 | 2 | $0.2625/1M | $0.5775/1M | prepaid BYOK |
qwen/qwen3-coder-flashQwen3 Coder Flash |
— | 262,144 | 2 | $0.4725/1M | $1.785/1M | prepaid BYOK |
qwen/qwen3.5-397b-a17bQwen: Qwen3.5 397B A17B |
— | 262,144 | 2 | $0.40425/1M | $2.5725/1M | prepaid BYOK |
xiaomi/mimo-v2.5-proXiaomi: MiMo-V2.5-Pro |
IQ 116#27 | 1,050,000 | 2 | $0.63/1M | $3.15/1M | prepaid BYOK |
z-ai/glm-5Z.ai: GLM 5 |
IQ 105#51 | 204,800 | 2 | $0.7875/1M | $2.52/1M | prepaid BYOK |
z-ai/glm-5.1Z.ai: GLM 5.1 |
IQ 114#32 | 204,800 | 2 | $1.02375/1M | $4.515/1M | prepaid BYOK |
z-ai/glm-5.2Z.ai: GLM 5.2 |
IQ 120#16 | 1,048,576 | 2 | $1.1025/1M | $4.62/1M | prepaid BYOK |