Googlegemini-3.8-flash

google/gemini-3.8-flash

fastreasoning
Open in Playground

Google gemini-3.8-flash: the fast, price-efficient Gemini Flash tier with long context and optional thinking. Text and image input.

Pricing

USD per 1M tokens, billed per token. Cached input applies where the provider caches prompts.

Input
$0.750
Output
$3.75
Cached input
$0.075

Specifications

Context length
1M
Max output tokens
—
Input
text, image
Output
text
Endpoints
gemini, openai
Released
Sep 2, 2026
Alias accepted
gemini-3.8-flash

Example request

curl https://api.nxioai.com/api/v1/chat/completions \
  -H "Authorization: Bearer $NXIO_API_KEY" \
  -H "Content-Type: application/json" \
  -d '{"model":"google/gemini-3.8-flash","messages":[{"role":"user","content":"Hello"}]}'

Price history

Every price change opens a new version. Past requests are billed with the version in effect at the time.

EffectiveUntilInputOutputCached inputSource
Oct 6, 2026current$0.750$3.75$0.075sync