Mistral's largest Ministral 3 model, pairing strong small-model quality with symmetric low token pricing.
Mistral's low-latency coding model for chat generation, function calling, and repository-scale context.
| Provider | Context | Price · in / out |
|---|---|---|
| mmistralNo trainingNo training on your data. Zero-retention available | 256K | $0.45 / $1.08/MTok |
routing.strategyproviders.onlyproviders.excludeproviders.ordermax_cost_usddata_policydocs →Collecting — charts appear after two days.
curl https://www.ninjachat.ai/api/v1/chat/completions \
-H "Authorization: Bearer $NINJACHAT_API_KEY" \
-H "Content-Type: application/json" \
-d '{
"model": "codestral-2508",
"messages": [{ "role": "user", "content": "Hello!" }],
"stream": true
}'Mistral's low-latency coding model for chat generation, function calling, and repository-scale context.
Codestral 2508 costs $0.45/M input tokens and $1.08/M output tokens.
Codestral 2508 accepts text and returns text.
POST /api/v1/chat/completions with `"model": "codestral-2508"`.
Ministral 3 14B, Ministral 3 3B, Ministral 3 8B.
Anthropic's highest-capability generally available model for long-running agents.
Anthropic's fastest Claude variant — lightning-quick responses.
Anthropic's most intelligent model — use when quality is paramount.