Providers

DigitalOcean

The developer cloud's Gradient AI platform — open-weight models served alongside its core infrastructure.

Models on DigitalOcean

Open in catalog

25 of 25

DeepSeek V3.2Chatdeepseek-v3.2164K contextWeights Routeemergency$0.56 / $0.84per M input / output
DeepSeek V4 FlashChatdeepseek-v4-flash1M contextWeights Routefallback$0.14 / $0.28per M input / output
DeepSeek V4 Flash 0731Chatdeepseek-v4-flash-07311M contextWeights Routeemergency$0.22 / $0.66per M input / output
DeepSeek V4 ProChatdeepseek-v4-pro1M contextWeights Routeemergency$1.74 / $3.48per M input / output
DeepSeek V4 Pro 0813Chatdeepseek-v4-pro-08131M contextRouteemergency$1.32 / $3.96per M input / output
Gemma 4 31BChatgemma-4-31b262K contextRouteemergency$0.18 / $0.50per M input / output
GLM 5.1Chatglm-5.1164K contextRouteprimary$1.40 / $4.40per M input / output
GLM 5.2Chatglm-5.21M contextWeights Routeemergency$1.40 / $4.40per M input / output
GLM-5Chatglm-5128K contextWeights Routeprimary$1.00 / $3.20per M input / output
GPT-OSS 120BChatgpt-oss-120b131K contextWeights Routeemergency$0.35 / $0.75per M input / output
GPT-OSS 20BChatgpt-oss-20b131K contextWeights Routeemergency$0.07 / $0.45per M input / output
Kimi K2.5Chatkimi-k2.5262K contextRouteemergency$0.60 / $3.00per M input / output
Kimi K2.6Chatkimi-k2.6262K contextRouteemergency$0.95 / $4.00per M input / output
Kimi K3Chatkimi-k31M contextRouteemergency$3.00 / $15.00per M input / output
Llama 4 MaverickChatllama-4-maverick1M contextWeights Routefallback$0.25 / $0.87per M input / output
MiniMax M2.5Chatminimax-m2.5197K contextRouteemergency$0.30 / $1.20per M input / output
Ministral 3 14BChatministral-3-14b262K contextRoutefallback$0.20 / $0.20per M input / output
Nemotron 3 Nano OmniChatnemotron-3-nano-omni66K contextRouteprimary$0.50 / $0.90per M input / output
Nemotron 3 SuperChatnemotron-3-super262K contextRoutefallback$0.30 / $0.65per M input / output
Nemotron 3 UltraChatnemotron-3-ultra262K contextRouteemergency$0.90 / $2.40per M input / output
Nemotron Nano 12B v2 VLChatnemotron-nano-12b-v2-vl128K contextRouteprimary$0.20 / $0.60per M input / output
Qwen 3.5 397B A17BChatqwen-3.5-397b-a17b262K contextRoutefallback$0.55 / $3.50per M input / output
Qwen 3.8 MaxChatqwen-3.8-max262K contextRouteemergency$2.00 / $6.00per M input / output
Stable Diffusion 3.5 LargeImagestable-diffusion-3.5-largeRouteprimary$0.08/image
Wan 2.2 T2V A14BVideowan-2.2-t2v-a14bRouteprimary$0.6/video

Performance

Request success
100%
Time to first token
15.0s
Generation speed
64.9tok/s

Response latency estimates

Time to first token

Generation speed · tok/s

Request success

Text-model timing and speed, plus API request success. Daily latency percentiles are estimates.

Direct routingCall DigitalOcean explicitly
digitalocean· routing
# pin every request to DigitalOcean
curl https://www.ninjachat.ai/api/v1/chat/completions \
  -H "Authorization: Bearer $NINJACHAT_API_KEY" \
  -H "Content-Type: application/json" \
  -d '{
    "model": "glm-5.1",
    "messages": [{ "role": "user", "content": "Hello!" }],
    "routing": { "providers": { "only": ["digitalocean"] } }
  }'