NinjaChat
ModelsAPIMCPPricingDocs
Sign inCreate free key
ModelsAPIMCPPricingDocs
Providers

Novita AI

VerifiedNo trainingZero retention

A cost-efficient GPU cloud serving a wide catalog of open-weight text and media models.

Start buildingView requests
ModelsPerformanceAPIRequest logs

Models on Novita AI

Open in catalog

21 of 21

Qwen 3.8 FlashChatqwen-3.8-flashReleased Aug 26, 20261M context$0.15 / $0.47per M input / output
Qwen3 Coder NextChatqwen3-coder-next262K context$0.20 / $1.50per M input / output
Step 3.7 FlashChatstep-3.7-flashReleased May 28, 2026256K context$0.20 / $1.15per M input / output
Qwen3 VL 30B A3B InstructChatqwen3-vl-30b-a3b262K contextWeights $0.20 / $0.70per M input / output
DeepSeek V3.1Chatdeepseek-v3.1164K context$0.25 / $0.95per M input / output
DeepSeek V3.2Chatdeepseek-v3.2164K contextWeights $0.56 / $0.84per M input / output
DeepSeek V4 Flash 0731Chatdeepseek-v4-flash-07311M contextWeights $0.22 / $0.66per M input / output
DeepSeek V4 Flash Vision ExpChatdeepseek-v4-flash-vision-exp1M context$0.44 / $1.32per M input / output
DeepSeek V4 Pro 0813Chatdeepseek-v4-pro-08131M context$1.32 / $3.96per M input / output
Gemma 4 26B A4BChatgemma-4-26b-a4b256K context$0.13 / $0.40per M input / output
Gemma 4 31BChatgemma-4-31b262K context$0.18 / $0.50per M input / output
GLM 5.2Chatglm-5.2Released June 20261M contextWeights $1.40 / $4.40per M input / output
GLM-5Chatglm-5Released February 2026128K contextWeights $1.00 / $3.20per M input / output
GPT-OSS 120BChatgpt-oss-120bReleased Aug 5, 2025131K contextWeights $0.35 / $0.75per M input / output
Kimi K3Chatkimi-k3Released Jul 16, 20261M context$3.00 / $15.00per M input / output
Ling 3.0 FlashChatling-3.0-flash131K context$0.06 / $0.18per M input / output
Ling 3.0 Flash FinChatling-3.0-flash-fin262K context$0.06 / $0.18per M input / output
MiniMax M2.7Chatminimax-m2.7197K contextWeights $0.30 / $1.20per M input / output
MiniMax M3Chatminimax-m3Released May 31, 2026512K context$0.30 / $1.20per M input / output
Nemotron 3 Nano 30B A3BChatnemotron-3-nano262K contextWeights $0.05 / $0.20per M input / output
Qwen3 VL 235B A22B InstructChatqwen3-vl-235b-a22b262K context$0.20 / $0.88per M input / output

Traffic & reliability

Token volume
240k
Reliability
100%
Inspect 21 requestsFilter model catalog
Direct routingCall Novita AI explicitly
novita· routing
# pin every request to Novita AI
curl https://www.ninjachat.ai/api/v1/chat/completions \
  -H "Authorization: Bearer $NINJACHAT_API_KEY" \
  -H "Content-Type: application/json" \
  -d '{
    "model": "qwen-3.8-flash",
    "messages": [{ "role": "user", "content": "Hello!" }],
    "routing": { "providers": { "only": ["novita"] } }
  }'
NinjaChat

Every AI. One app.

Download on the App Store

Product

  • Dashboard
  • ninja for iMessage
  • Pricing
  • Free AI Tools
  • Affiliate Program
  • iOS App

Developers

  • API
  • API Models
  • Router
  • MCP / Agents
  • Search API
  • API Pricing
  • Migration Guides
  • API Docs

Models

  • Model Council
  • Seed 1.8
  • Gemini 2.5 Flash
  • Gemini 2.5 Pro
  • Gemini 3 Flash Preview
  • View All Models

Company

  • Blog
  • Uncensored AI
  • Community
  • Careers
  • Support
  • Privacy Policy
  • Terms of Service

Copyright © 2026 NinjaChat AI. Product of Bloon All Rights Reserved.