Gemini's high-efficiency flash variant.
Google's Gemini 3.5 Flash served through two funded neutral-host rails with full multimodal routing.
| Provider | Role | Context | Price · in / out | p50 | p95 | Uptime |
|---|---|---|---|---|---|---|
| ggmicloudNo trainingNo training on your data | Primary | 1M | $1.68 / $10.08/MToksame price, any rail | 1.9s | 2.2s | 100% |
| ddeepinfraNo trainingNo training on your data | Fallback | 1M | 3.0s | 3.0s | 100% |
Collecting — charts appear after two days.
curl https://www.ninjachat.ai/api/v1/chat/completions \
-H "Authorization: Bearer $NINJACHAT_API_KEY" \
-H "Content-Type: application/json" \
-d '{
"model": "gemini-3.5-flash",
"messages": [{ "role": "user", "content": "Hello!" }],
"stream": true
}'Google's Gemini 3.5 Flash served through two funded neutral-host rails with full multimodal routing.
Gemini 3.5 Flash costs $1.68/M input tokens and $10.08/M output tokens.
Gemini 3.5 Flash accepts text and image and returns text.
POST /api/v1/chat/completions with `"model": "gemini-3.5-flash"`.
Gemini 2.5 Flash, Gemini 2.5 Pro, Gemini 3 Flash, Gemini 3 Pro, Gemini 3.1 Flash Lite, Gemini 3.1 Pro, Gemini 3.5 Flash-Lite, Gemini 3.6 Flash, Gemini 3.7 Flash, Nano Banana, Nano Banana 2, Nano Banana Pro.
routing.strategyproviders.onlyproviders.excludeproviders.ordermax_cost_usddata_policydocs →Google's ultra-fast model with a 1M token context window.
Gemini 3 Pro — Google's latest standard-tier powerhouse.
Google's low-cost million-token Gemini Flash Lite model with two funded multimodal rails.
Google's most capable standard model — excellent creative output.