Googlegemini-3.1-flash-lite

google/gemini-3.1-flash-lite

fast
Open in Playground

Google gemini-3.1-flash-lite: the lightest, lowest-latency Gemini tier for high-volume, cost-sensitive workloads. Text and image input.

Pricing

USD per 1M tokens, billed per token. Cached input applies where the provider caches prompts.

Input
$0.250
Output
$1.50
Cached input
$0.025

Specifications

Context length
—
Max output tokens
—
Input
text, image
Output
text
Endpoints
gemini, openai
Released
—
Alias accepted
gemini-3.1-flash-lite

Example request

curl https://api.nxioai.com/api/v1/chat/completions \
  -H "Authorization: Bearer $NXIO_API_KEY" \
  -H "Content-Type: application/json" \
  -d '{"model":"google/gemini-3.1-flash-lite","messages":[{"role":"user","content":"Hello"}]}'

Price history

Every price change opens a new version. Past requests are billed with the version in effect at the time.

EffectiveUntilInputOutputCached inputSource
Oct 6, 2026current$0.250$1.50$0.025sync