Gemini 3.1 Flash Lite

DeepInfra

Google's low-cost million-token Gemini 3.1 Flash Lite for high-volume multimodal work.

Modalities
IntextimageOuttext
Input / output
$0.25 / $1.50USD per 1M tokens
Context
1M
Providers
Live
Released
May 7, 2026

Playground

Preparing playground

Providers

fallback

API

POST/api/v1/chat/completionsOpenAI-compatible
gemini-3.1-flash-lite
curl https://www.ninjachat.ai/api/v1/chat/completions \
  -H "Authorization: Bearer $NINJACHAT_API_KEY" \
  -H "Content-Type: application/json" \
  -d '{
    "model": "gemini-3.1-flash-lite",
    "messages": [{ "role": "user", "content": "Hello!" }],
    "stream": true
  }'
messagestemperaturemax_completion_tokenstop_pstopfrequency_penaltypresence_penaltyseedstreamuserroutingtoolstool_choiceimage_url content partsreasoningreasoning_effort

Pricing

Input
$0.25/M tokens
Cached input
$0.025/M tokens
Output
$1.50/M tokens

Published USD rates. Actual cost depends on input, output, cache usage and request options.