Qwen (Alibaba)Qwen3-Max-Thinking
qwen/qwen3-max-thinking
reasoning
Alibaba Qwen3-Max-Thinking: Qwen reasoning variant with extended thinking (reasoning tokens billed as output).
Pricing
USD per 1M tokens, billed per token. Cached input applies where the provider caches prompts.
- Input
- $0.362
- Output
- $1.45
- Cached input
- $0.362
Specifications
- Context length
- 252K
- Max output tokens
- —
- Input
- text
- Output
- text
- Endpoints
- openai
- Released
- Jan 23, 2026
- Alias accepted
- Qwen3-Max-Thinking
Example request
curl https://api.nxioai.com/api/v1/chat/completions \
-H "Authorization: Bearer $NXIO_API_KEY" \
-H "Content-Type: application/json" \
-d '{"model":"qwen/qwen3-max-thinking","messages":[{"role":"user","content":"Hello"}]}'Price history
Every price change opens a new version. Past requests are billed with the version in effect at the time.
| Effective | Until | Input | Output | Cached input | Source |
|---|---|---|---|---|---|
| Oct 6, 2026 | current | $0.362 | $1.45 | $0.362 | sync |