OpenAI compatible API · Attested · Public status

DigitalOcean Gradient AI

DigitalOcean Gradient AI models on TrustedRouter with prices, routes, policy notes, and source links.

Verify gateway
Onebase URL to migrate
100sof models and routes
0prompt or output logs. Always.

digitalocean

No provider claim

All providers

ProviderDigitalOcean Gradient AI
Models22 public models
Prepaid routes22
BYOK routes22
Zero data retentionnot claimed
Confidential computenot claimed
Provider E2EEnot claimed
Policy noteNo provider-ZDR claim is tracked here. DigitalOcean's Gradient AI model and pricing documentation is linked for data-handling review.
Policy source

Measured performance

66 samples

Continuously sampled across DigitalOcean Gradient AI's routed models: p50 TTFT, effective throughput, and success rate. Effective throughput uses provider-reported output tokens over complete request time. Unsupported route and probe-configuration rows are separated from provider downtime. No prompt or output content stored.

p50 TTFT1554 ms
Effective throughput28 tok/s n=7
Uptime98.48%
Modelp50 TTFTp50 TTFBEffective throughputUptimeConfig excludedAvailability samples
minimax/minimax-m2.5 651 ms 651 ms 100.00% 1
nvidia/nemotron-3-super-120b 898 ms 898 ms 100.00% 2
nvidia/nemotron-3-nano-omni 962 ms 961 ms 100.00% 4
qwen/qwen3.5-397b-a17b 965 ms 965 ms 100.00% 2
mistralai/ministral-3-14b-instruct 1112 ms 1112 ms 100.00% 2
z-ai/glm-5 1341 ms 1341 ms 100.00% 3
qwen/qwen3-coder-flash 1399 ms 1399 ms 100.00% 6
meta-llama/llama-4-maverick 1429 ms 1429 ms 100.00% 4
moonshotai/kimi-k2.6 1474 ms 1474 ms 29 tok/s n=1 100.00% 6
z-ai/glm-5.1 1554 ms 1553 ms 100.00% 3
moonshotai/kimi-k2.5 1572 ms 1572 ms 100.00% 5
deepseek/deepseek-v3.2 1596 ms 1596 ms 100.00% 3
nvidia/nemotron-3-ultra-550b 1637 ms 1637 ms 100.00% 3
meta-llama/llama-3.3-70b-instruct 1689 ms 1689 ms 100.00% 2
nvidia/nemotron-nano-12b-v2-vl 1689 ms 1688 ms 100.00% 7
deepseek/deepseek-r1-distill-llama-70b 3010 ms 3010 ms 100.00% 2
deepseek/deepseek-v4-flash 3423 ms 3423 ms 20 tok/s n=1 100.00% 3
google/gemma-4-31b-it 4371 ms 4371 ms 100.00% 1
deepseek/deepseek-v4-pro 4645 ms 4645 ms 26 tok/s n=2 100.00% 5
z-ai/glm-5.2 5351 ms 5351 ms 28 tok/s n=2 50.00% 2
xiaomi/mimo-v2.5-pro 64 tok/s n=1 0

DigitalOcean Gradient AI performance history · Full provider & model leaderboard.

Provider models

Models served by DigitalOcean Gradient AI.

Each row links to pricing, provider, benchmark, and API pages for the model.

Model AI IQ Context Endpoints Prompt Completion Routes
deepseek/deepseek-r1-distill-llama-70b
DeepSeek: R1 Distill Llama 70B
8,192 2 $1.0395/1M $1.0395/1M prepaid BYOK
deepseek/deepseek-v3.2
DeepSeek: DeepSeek V3.2
IQ 103#56 163,840 2 $0.44625/1M $1.428/1M prepaid BYOK
deepseek/deepseek-v4-flash
DeepSeek: DeepSeek V4 Flash
IQ 108#46 1,048,576 2 $0.1176/1M $0.2352/1M prepaid BYOK
deepseek/deepseek-v4-pro
DeepSeek: DeepSeek V4 Pro
IQ 115#30 1,048,576 2 $1.4616/1M $2.9232/1M prepaid BYOK
google/gemma-4-31b-it
Google: Gemma 4 31B
IQ 101#65 262,144 2 $0.189/1M $0.525/1M prepaid BYOK
meta-llama/llama-3.3-70b-instruct
Meta: Llama 3.3 70B Instruct
131,072 2 $0.6825/1M $0.6825/1M prepaid BYOK
meta-llama/llama-4-maverick
Meta: Llama 4 Maverick
IQ 90#90 1,048,576 2 $0.2625/1M $0.9135/1M prepaid BYOK
minimax/minimax-m2.5
MiniMax: MiniMax M2.5
IQ 105#54 204,800 2 $0.23625/1M $0.945/1M prepaid BYOK
mistralai/ministral-3-14b-instruct
Ministral 3 14B Instruct
262,144 2 $0.21/1M $0.21/1M prepaid BYOK
moonshotai/kimi-k2.5
MoonshotAI: Kimi K2.5
IQ 111#39 262,144 2 $0.39375/1M $2.12625/1M prepaid BYOK
moonshotai/kimi-k2.6
MoonshotAI: Kimi K2.6
IQ 119#19 262,144 2 $0.798/1M $3.36/1M prepaid BYOK
nvidia/nemotron-3-nano-omni
Nemotron Nano 3 Omni
65,536 2 $0.525/1M $0.945/1M prepaid BYOK
nvidia/nemotron-3-super-120b
Nemotron-3-Super-120B
1,000,000 2 $0.2205/1M $0.47775/1M prepaid BYOK
nvidia/nemotron-3-ultra-550b
Nemotron 3 Ultra
131,072 2 $0.945/1M $1.785/1M prepaid BYOK
nvidia/nemotron-nano-12b-v2-vl
Nemotron Nano 12B v2 VL
128,000 2 $0.21/1M $0.63/1M prepaid BYOK
qwen/qwen3-32b
Qwen: Qwen3 32B
131,072 2 $0.2625/1M $0.5775/1M prepaid BYOK
qwen/qwen3-coder-flash
Qwen3 Coder Flash
262,144 2 $0.4725/1M $1.785/1M prepaid BYOK
qwen/qwen3.5-397b-a17b
Qwen: Qwen3.5 397B A17B
262,144 2 $0.40425/1M $2.5725/1M prepaid BYOK
xiaomi/mimo-v2.5-pro
Xiaomi: MiMo-V2.5-Pro
IQ 116#27 1,050,000 2 $0.63/1M $3.15/1M prepaid BYOK
z-ai/glm-5
Z.ai: GLM 5
IQ 105#51 204,800 2 $0.7875/1M $2.52/1M prepaid BYOK
z-ai/glm-5.1
Z.ai: GLM 5.1
IQ 114#32 204,800 2 $1.02375/1M $4.515/1M prepaid BYOK
z-ai/glm-5.2
Z.ai: GLM 5.2
IQ 120#16 1,048,576 2 $1.1025/1M $4.62/1M prepaid BYOK
Workspace access

Sign in

Choose a sign in method. New email and OAuth accounts include $0.10 in starter credit; wallet-only accounts start at $0.

By signing in you agree to the terms of service and privacy policy.