DeepSeek's open reasoning flagship with conformant GMI tool use and a lower-cost DeepInfra text fallback.
| Provider | Role | Context | Price · in / out | p50 | p95 | Uptime |
|---|---|---|---|---|---|---|
| ddeepinfraNo trainingNo training on your data | Primary | 1M | $0.24 / $0.33/MToksame price, any rail | 700ms | 700ms | 100% |
| ddigitaloceanNo trainingNo training on your data. Zero-retention available | Fallback | 1M | 1.4s | 4.2s |
Collecting — charts appear after two days.
curl https://www.ninjachat.ai/api/v1/chat/completions \
-H "Authorization: Bearer $NINJACHAT_API_KEY" \
-H "Content-Type: application/json" \
-d '{
"model": "deepseek-v4-flash",
"messages": [{ "role": "user", "content": "Hello!" }],
"stream": true
}'DeepSeek's current V4 Flash release with million-token context and two fully probed neutral-host rails.
DeepSeek V4 Flash costs $0.24/M input tokens and $0.33/M output tokens.
DeepSeek V4 Flash accepts text and returns text.
POST /api/v1/chat/completions with `"model": "deepseek-v4-flash"`.
DeepSeek R1 0528, DeepSeek V3, DeepSeek V3.2, DeepSeek V4 Flash 0731, DeepSeek V4 Pro, DeepSeek V4 Pro 0813.
| 100% |
routing.strategyproviders.onlyproviders.excludeproviders.ordermax_cost_usddata_policydocs →DeepSeek's widely used V3.2 model on two independently operated neutral hosts.
The most popular OpenRouter model this week, served directly through Fireworks.
DeepSeek's current V4 Pro release with million-token context and three fully probed neutral-host rails.