Qwen's dense 27B multimodal model for capable coding, reasoning, and visual tool workflows.
Qwen's 122B/10B-active multimodal MoE, balancing stronger quality with efficient inference.
| Provider | Context | Price · in / out |
|---|---|---|
| ddeepinfraNo trainingNo training on your data | 262K | $0.44 / $2.88/MTok |
routing.strategyproviders.onlyproviders.excludeproviders.ordermax_cost_usddata_policydocs →Collecting — charts appear after two days.
curl https://www.ninjachat.ai/api/v1/chat/completions \
-H "Authorization: Bearer $NINJACHAT_API_KEY" \
-H "Content-Type: application/json" \
-d '{
"model": "qwen-3.5-122b-a10b",
"messages": [{ "role": "user", "content": "Hello!" }],
"stream": true
}'Qwen's 122B/10B-active multimodal MoE, balancing stronger quality with efficient inference.
Qwen 3.5 122B A10B costs $0.44/M input tokens and $2.88/M output tokens.
Qwen 3.5 122B A10B accepts text and image and returns text.
POST /api/v1/chat/completions with `"model": "qwen-3.5-122b-a10b"`.
Qwen 3.5 27B, Qwen 3.5 397B A17B, Qwen 3.6 27B, Qwen 3.6 35B A3B, Qwen 3.7 Plus, Qwen 3.8 2.4T A95B, Qwen 3.8 27B, Qwen 3.8 Max, Qwen3 Next 80B A3B, QwQ 32B.
Qwen's compact multimodal reasoning model on a cost-efficient DeepInfra rail with ultra-fast Groq failover.
Qwen's sparse 35B/3B-active multimodal model for efficient coding and agents.
Qwen's fast multimodal flagship for agent loops, coding, and tool use.