OpenAI compatible API · Attested · Public status
DigitalOcean Gradient AI
Explore DigitalOcean Gradient AI models on TrustedRouter with current routes, token pricing, policy sources, privacy posture, regional availability, and API support.
Onebase URL to migrate
100sof models and routes
0prompt or output logs. Always.
DigitalOcean Gradient AIdigitalocean
No provider claim| Provider | DigitalOcean Gradient AI |
|---|---|
| Provider website | https://www.digitalocean.com/products/gradient-ai-platform |
| Models | 22 public models |
| Prepaid routes | 22 |
| BYOK routes | 22 |
| Zero data retention | not claimed |
| Confidential compute | not claimed |
| Provider E2EE | not claimed |
| Policy note | No provider-ZDR claim is tracked here. DigitalOcean's Gradient AI model and pricing documentation is linked for data-handling review. Policy source |
Measured performance
61 samplesContinuously sampled across DigitalOcean Gradient AI's routed models: p50 TTFT, effective throughput, and success rate. Effective throughput uses provider-reported output tokens over complete request time. Unsupported route and probe-configuration rows are separated from provider downtime. No prompt or output content stored.
| p50 TTFT | 1709 ms |
|---|---|
| Effective throughput | 39 tok/s n=6 |
| Uptime | 95.08% |
| Model | p50 TTFT | p50 TTFB | Effective throughput | Uptime | Config excluded | Availability samples |
|---|---|---|---|---|---|---|
| nvidia/nemotron-3-nano-omni | 727 ms | 727 ms | — | 100.00% | — | 1 |
| nvidia/nemotron-3-ultra-550b | 981 ms | 980 ms | — | 100.00% | — | 1 |
| moonshotai/kimi-k2.5 | 1197 ms | 1197 ms | — | 100.00% | — | 3 |
| mistralai/ministral-3-14b-instruct | 1210 ms | 1210 ms | — | 100.00% | — | 2 |
| meta-llama/llama-3.3-70b-instruct | 1216 ms | 1215 ms | — | 100.00% | — | 2 |
| meta-llama/llama-4-maverick | 1388 ms | 1388 ms | — | 100.00% | — | 3 |
| qwen/qwen3-coder-flash | 1395 ms | 1395 ms | — | 100.00% | — | 2 |
| qwen/qwen3.5-397b-a17b | 1682 ms | 1681 ms | — | 100.00% | — | 6 |
| z-ai/glm-5 | 1709 ms | 1709 ms | — | 100.00% | — | 8 |
| nvidia/nemotron-nano-12b-v2-vl | 2243 ms | 2243 ms | — | 100.00% | — | 1 |
| qwen/qwen3-32b | 2353 ms | 2353 ms | — | 100.00% | — | 5 |
| z-ai/glm-5.2 | 2558 ms | 2557 ms | 69 tok/s n=1 | 100.00% | — | 5 |
| deepseek/deepseek-v3.2 | 2712 ms | 2712 ms | — | 100.00% | — | 1 |
| google/gemma-4-31b-it | 2915 ms | 2915 ms | — | 100.00% | — | 3 |
| xiaomi/mimo-v2.5-pro | 3269 ms | 3269 ms | — | 100.00% | — | 1 |
| deepseek/deepseek-v4-pro | 3312 ms | 3311 ms | 60 tok/s n=1 | 100.00% | — | 3 |
| deepseek/deepseek-v4-flash | 4449 ms | 4449 ms | 14 tok/s n=2 | 100.00% | — | 2 |
| z-ai/glm-5.1 | 4954 ms | 4954 ms | — | 100.00% | — | 1 |
| moonshotai/kimi-k2.6 | 1414 ms | 1414 ms | 39 tok/s n=2 | 75.00% | — | 4 |
| nvidia/nemotron-3-super-120b | 4043 ms | 4043 ms | — | 71.43% | — | 7 |
DigitalOcean Gradient AI performance history · Full provider & model leaderboard.
Provider models
Models served by DigitalOcean Gradient AI.
Each row links to pricing, provider, benchmark, and API pages for the model.
| Model | AI IQ | Context | Endpoints | Prompt | Completion | Routes |
|---|---|---|---|---|---|---|
deepseek/deepseek-r1-distill-llama-70bDeepSeek: R1 Distill Llama 70B |
— | 8,192 | 2 | $1.0395/1M | $1.0395/1M | prepaid BYOK |
deepseek/deepseek-v3.2DeepSeek: DeepSeek V3.2 |
IQ 103#57 | 163,840 | 2 | $0.44625/1M | $1.428/1M | prepaid BYOK |
deepseek/deepseek-v4-flashDeepSeek: DeepSeek V4 Flash 0423 |
IQ 108#46 | 1,048,576 | 2 | $0.1176/1M | $0.2352/1M | prepaid BYOK |
deepseek/deepseek-v4-proDeepSeek: DeepSeek V4 Pro |
IQ 115#30 | 1,048,576 | 2 | $1.4616/1M | $2.9232/1M | prepaid BYOK |
google/gemma-4-31b-itGoogle: Gemma 4 31B |
IQ 101#66 | 262,144 | 2 | $0.189/1M | $0.525/1M | prepaid BYOK |
meta-llama/llama-3.3-70b-instructMeta: Llama 3.3 70B Instruct |
— | 131,072 | 2 | $0.6825/1M | $0.6825/1M | prepaid BYOK |
meta-llama/llama-4-maverickMeta: Llama 4 Maverick |
IQ 90#91 | 1,048,576 | 2 | $0.2625/1M | $0.9135/1M | prepaid BYOK |
minimax/minimax-m2.5MiniMax: MiniMax M2.5 |
IQ 105#56 | 204,800 | 2 | $0.23625/1M | $0.945/1M | prepaid BYOK |
mistralai/ministral-3-14b-instructMinistral 3 14B Instruct |
— | 262,144 | 2 | $0.21/1M | $0.21/1M | prepaid BYOK |
moonshotai/kimi-k2.5MoonshotAI: Kimi K2.5 |
IQ 111#39 | 262,144 | 2 | $0.39375/1M | $2.12625/1M | prepaid BYOK |
moonshotai/kimi-k2.6MoonshotAI: Kimi K2.6 |
IQ 119#19 | 262,144 | 2 | $0.798/1M | $3.36/1M | prepaid BYOK |
nvidia/nemotron-3-nano-omniNemotron Nano 3 Omni |
— | 65,536 | 2 | $0.525/1M | $0.945/1M | prepaid BYOK |
nvidia/nemotron-3-super-120bNemotron-3-Super-120B |
— | 1,000,000 | 2 | $0.2205/1M | $0.47775/1M | prepaid BYOK |
nvidia/nemotron-3-ultra-550bNemotron 3 Ultra |
— | 131,072 | 2 | $0.945/1M | $1.785/1M | prepaid BYOK |
nvidia/nemotron-nano-12b-v2-vlNemotron Nano 12B v2 VL |
— | 128,000 | 2 | $0.21/1M | $0.63/1M | prepaid BYOK |
qwen/qwen3-32bQwen: Qwen3 32B |
— | 131,072 | 2 | $0.2625/1M | $0.5775/1M | prepaid BYOK |
qwen/qwen3-coder-flashQwen3 Coder Flash |
— | 262,144 | 2 | $0.4725/1M | $1.785/1M | prepaid BYOK |
qwen/qwen3.5-397b-a17bQwen: Qwen3.5 397B A17B |
— | 262,144 | 2 | $0.40425/1M | $2.5725/1M | prepaid BYOK |
xiaomi/mimo-v2.5-proXiaomi: MiMo-V2.5-Pro |
IQ 116#27 | 1,050,000 | 2 | $0.63/1M | $3.15/1M | prepaid BYOK |
z-ai/glm-5Z.ai: GLM 5 |
IQ 105#52 | 204,800 | 2 | $0.7875/1M | $2.52/1M | prepaid BYOK |
z-ai/glm-5.1Z.ai: GLM 5.1 |
IQ 114#32 | 204,800 | 2 | $1.02375/1M | $4.515/1M | prepaid BYOK |
z-ai/glm-5.2Z.ai: GLM 5.2 |
IQ 120#16 | 1,048,576 | 2 | $1.1025/1M | $4.62/1M | prepaid BYOK |