Gemini's high-efficiency flash variant.
Google's fastest, lowest-cost GA Gemini model for high-volume agentic and multimodal workloads.
| Provider | Role | Context | Price · in / out | p50 | p95 | Uptime |
|---|---|---|---|---|---|---|
| Primary | 1M | $0.45 / $2.80/MToksame price, any rail | 764ms | 959ms | 100% | |
| ggmicloudNo trainingNo training on your data | Fallback | 1M | 1.7s | 2.8s |
Collecting — charts appear after two days.
curl https://www.ninjachat.ai/api/v1/chat/completions \
-H "Authorization: Bearer $NINJACHAT_API_KEY" \
-H "Content-Type: application/json" \
-d '{
"model": "gemini-3.5-flash-lite",
"messages": [{ "role": "user", "content": "Hello!" }],
"stream": true
}'Google's fastest, lowest-cost GA Gemini model for high-volume agentic and multimodal workloads.
Gemini 3.5 Flash-Lite costs $0.45/M input tokens and $2.80/M output tokens.
Gemini 3.5 Flash-Lite accepts text and image and returns text.
POST /api/v1/chat/completions with `"model": "gemini-3.5-flash-lite"`.
Gemini 2.5 Flash, Gemini 2.5 Pro, Gemini 3 Flash, Gemini 3 Pro, Gemini 3.1 Pro, Gemini 3.6 Flash, Gemini 3.7 Flash, Nano Banana, Nano Banana 2, Nano Banana Pro.
| 100% |
routing.strategyproviders.onlyproviders.excludeproviders.ordermax_cost_usddata_policydocs →Google's ultra-fast model with a 1M token context window.
Gemini 3 Pro — Google's latest standard-tier powerhouse.
Google's most capable standard model — excellent creative output.