Qwen (Alibaba)qwen3-vl-flash

qwen/qwen3-vl-flash

vision
Open in Playground

Alibaba qwen3-vl-flash: Qwen vision-language model for image understanding and multimodal chat.

Pricing

USD per 1M tokens, billed per token. Cached input applies where the provider caches prompts.

Input
$0.022
Output
$0.216
Cached input
$0.0022

Specifications

Context length
262K
Max output tokens
—
Input
text, image
Output
text
Endpoints
openai, anthropic
Released
—
Alias accepted
qwen3-vl-flash

Example request

curl https://api.nxioai.com/api/v1/chat/completions \
  -H "Authorization: Bearer $NXIO_API_KEY" \
  -H "Content-Type: application/json" \
  -d '{"model":"qwen/qwen3-vl-flash","messages":[{"role":"user","content":"Hello"}]}'

Price history

Every price change opens a new version. Past requests are billed with the version in effect at the time.

EffectiveUntilInputOutputCached inputSource
Oct 6, 2026current$0.022$0.216$0.0022sync