Z.ai (Zhipu)GLM-5.3-Flash

z-ai/glm-5.3-flash

featuredtextcodingfast
Open in Playground

Z.ai GLM-5.3-Flash: fast, low-cost GLM tier for high-throughput chat and coding.

Pricing

USD per 1M tokens, billed per token. Cached input applies where the provider caches prompts.

Input
$0.118
Output
$0.413
Cached input
$0.034

Specifications

Context length
—
Max output tokens
—
Input
text
Output
text
Endpoints
openai, anthropic
Released
—
Alias accepted
GLM-5.3-Flash

Example request

curl https://api.nxioai.com/api/v1/chat/completions \
  -H "Authorization: Bearer $NXIO_API_KEY" \
  -H "Content-Type: application/json" \
  -d '{"model":"z-ai/glm-5.3-flash","messages":[{"role":"user","content":"Hello"}]}'

Price history

Every price change opens a new version. Past requests are billed with the version in effect at the time.

EffectiveUntilInputOutputCached inputSource
Oct 6, 2026current$0.118$0.413$0.034sync