Googlegemini-2.5-flash
google/gemini-2.5-flash
fastreasoning
Google's fast, price-efficient Gemini 2.5 workhorse with a 1M-token context window, up to 65K output tokens and optional thinking. Accepts text and images. Released 2025-06.
Pricing
USD per 1M tokens, billed per token. Cached input applies where the provider caches prompts.
- Input
- $0.300
- Output
- $2.50
- Cached input
- $0.030
Specifications
- Context length
- 1M
- Max output tokens
- 65,536
- Input
- text, image
- Output
- text
- Endpoints
- gemini, openai
- Released
- —
- Alias accepted
- gemini-2.5-flash
Example request
curl https://api.nxioai.com/api/v1/chat/completions \
-H "Authorization: Bearer $NXIO_API_KEY" \
-H "Content-Type: application/json" \
-d '{"model":"google/gemini-2.5-flash","messages":[{"role":"user","content":"Hello"}]}'Price history
Every price change opens a new version. Past requests are billed with the version in effect at the time.
| Effective | Until | Input | Output | Cached input | Source |
|---|---|---|---|---|---|
| Oct 6, 2026 | current | $0.300 | $2.50 | $0.030 | sync |