Qwen's sparse 35B/3B-active multimodal model for efficient coding and agents.
| Provider | Role | Context | Price · in / out | p50 | p95 | Uptime |
|---|---|---|---|---|---|---|
| ddeepinfraNo trainingNo training on your data | Primary | 262K | $1.98 / $5.94/MToksame price, any rail | 1.7s | 1.7s | 100% |
| Fallback |
Collecting — charts appear after two days.
curl https://www.ninjachat.ai/api/v1/chat/completions \
-H "Authorization: Bearer $NINJACHAT_API_KEY" \
-H "Content-Type: application/json" \
-d '{
"model": "qwen-3.8-max",
"messages": [{ "role": "user", "content": "Hello!" }],
"stream": true
}'Qwen's 2.4T-parameter sparse-MoE frontier model for autonomous long-horizon work.
Qwen 3.8 Max costs $1.98/M input tokens and $5.94/M output tokens.
Qwen 3.8 Max accepts text and returns text.
POST /api/v1/chat/completions with `"model": "qwen-3.8-max"`.
Qwen 3.6 35B A3B, Qwen 3.7 Plus, Qwen 3.8 2.4T A95B, Qwen 3.8 27B, Qwen3 Next 80B A3B, QwQ 32B.
| 262K |
| — |
| — |
| — |
| ggmicloudNo trainingNo training on your data | Emergency | 262K | — | — | — |
routing.strategyproviders.onlyproviders.excludeproviders.ordermax_cost_usddata_policydocs →Qwen's open-weight 2.4T-parameter sparse-MoE flagship with 95B active parameters.
Qwen's efficient 27B model on independently operated DeepInfra and GMI rails.
Qwen's efficient 80B/3B-active instruct model on two independent direct rails.