Providers

Cerebras

Wafer-scale AI silicon — some of the fastest open-model token throughput available anywhere.

Models on Cerebras

Open in catalog

1 of 1

GPT-OSS 120BChatgpt-oss-120b131K contextWeights Routefallback$0.35 / $0.75per M input / output
Direct routingCall Cerebras explicitly
cerebras· routing
# pin every request to Cerebras
curl https://www.ninjachat.ai/api/v1/chat/completions \
  -H "Authorization: Bearer $NINJACHAT_API_KEY" \
  -H "Content-Type: application/json" \
  -d '{
    "model": "gpt-oss-120b",
    "messages": [{ "role": "user", "content": "Hello!" }],
    "routing": { "providers": { "only": ["cerebras"] } }
  }'