Google's newest Flash-class model: a million-token context, native image and video input, and reasoning tuned for agentic coding at Flash speed.
Text-model timing and speed, plus API request success. Daily latency percentiles are estimates. Your request logs
curl https://www.ninjachat.ai/api/v1/chat/completions \
-H "Authorization: Bearer $NINJACHAT_API_KEY" \
-H "Content-Type: application/json" \
-d '{
"model": "gemini-3.7-flash",
"messages": [{ "role": "user", "content": "Hello!" }],
"stream": true
}'Published USD rates. Actual cost depends on input, output, cache usage and request options.
Unedited answers, latency and price for every model, measured through the NinjaChat API in September 2026.