Providers

Makora

Verified

A high-speed inference platform optimized for low per-user latency on frontier open-weight models.

Models on Makora

Open in catalog

6 of 6

DeepSeek V4 Flash 0731Chatdeepseek-v4-flash-07311M contextWeights Routeemergency$0.22 / $0.66per M input / output
Gemma 4 26B A4BChatgemma-4-26b-a4b256K contextRouteemergency$0.13 / $0.40per M input / output
GLM 5.3Chatglm-5.31M contextRouteemergency$1.40 / $4.40per M input / output
GLM 5.3 FlashChatglm-5.3-flash1M contextRouteemergency$0.15 / $0.50per M input / output
Kimi K3Chatkimi-k31M contextRouteemergency$3.00 / $15.00per M input / output
Qwen 3.8 Flash NextChatqwen-3.8-flash-next262K contextRouteprimary$0.15 / $0.47per M input / output
Direct routingCall Makora explicitly
makora· routing
# pin every request to Makora
curl https://www.ninjachat.ai/api/v1/chat/completions \
  -H "Authorization: Bearer $NINJACHAT_API_KEY" \
  -H "Content-Type: application/json" \
  -d '{
    "model": "qwen-3.8-flash-next",
    "messages": [{ "role": "user", "content": "Hello!" }],
    "routing": { "providers": { "only": ["makora"] } }
  }'