DeepSeek's open reasoning flagship with conformant GMI tool use and a lower-cost DeepInfra text fallback.
| Provider | Role | Context | Price · in / out | p50 | p95 | Uptime |
|---|---|---|---|---|---|---|
| ddeepinfraNo trainingNo training on your data | Primary | 1M | $2.09 / $4.18/MToksame price, any rail | 684ms | 684ms | 100% |
| bbasetenNo trainingNo training on your data. Zero-retention available | Fallback | 1M | 731ms | 821ms |
Collecting — charts appear after two days.
curl https://www.ninjachat.ai/api/v1/chat/completions \
-H "Authorization: Bearer $NINJACHAT_API_KEY" \
-H "Content-Type: application/json" \
-d '{
"model": "deepseek-v4-pro",
"messages": [{ "role": "user", "content": "Hello!" }],
"stream": true
}'DeepSeek's current V4 Pro release with million-token context and three fully probed neutral-host rails.
DeepSeek V4 Pro costs $2.09/M input tokens and $4.18/M output tokens.
DeepSeek V4 Pro accepts text and returns text.
POST /api/v1/chat/completions with `"model": "deepseek-v4-pro"`.
DeepSeek R1 0528, DeepSeek V3, DeepSeek V3.2, DeepSeek V4 Flash, DeepSeek V4 Flash 0731, DeepSeek V4 Pro 0813.
| 100% |
| ddigitaloceanNo trainingNo training on your data. Zero-retention available | Emergency | 1M | 648ms | 648ms | 100% |
routing.strategyproviders.onlyproviders.excludeproviders.ordermax_cost_usddata_policydocs →DeepSeek's widely used V3.2 model on two independently operated neutral hosts.
DeepSeek's current V4 Flash release with million-token context and two fully probed neutral-host rails.
The most popular OpenRouter model this week, served directly through Fireworks.