NVIDIA's 550B/55B-active frontier reasoning model for demanding agent workloads.
| Provider | Role | Context | Price · in / out |
|---|---|---|---|
| Primary | 262K | $0.20 / $0.35/MToksame price, any rail | |
| ddeepinfraNo trainingNo training on your data | Fallback | 262K |
routing.strategyCollecting — charts appear after two days.
curl https://www.ninjachat.ai/api/v1/chat/completions \
-H "Authorization: Bearer $NINJACHAT_API_KEY" \
-H "Content-Type: application/json" \
-d '{
"model": "nemotron-3.5-lightning",
"messages": [{ "role": "user", "content": "Hello!" }],
"stream": true
}'NVIDIA's 3B-active hybrid Mamba-Transformer reasoning model at five cents per input MTok.
Nemotron 3.5 Lightning costs $0.20/M input tokens and $0.35/M output tokens.
Nemotron 3.5 Lightning accepts text and returns text.
POST /api/v1/chat/completions with `"model": "nemotron-3.5-lightning"`.
Nemotron 3 Ultra.
providers.onlyproviders.excludeproviders.ordermax_cost_usddata_policyAnthropic's fastest Claude variant — lightning-quick responses.
Anthropic's most intelligent model — use when quality is paramount.
Anthropic's Opus 5 for complex agentic coding and enterprise knowledge work.