Alibaba Cloud (Qwen) · Flagship
Qwen3.8-Max API pricing and specs
The first open-weight Qwen Max model, with thinking and non-thinking modes and up to 262K tokens of reasoning.
$1.65
$0.21
$4.95
Explicit cache reads cost $0.137 per million and cache creation $2.063. The Singapore region charges $2 input and $6 output. Batch pricing is offered only in the Beijing region.
Spec sheet
- API model string
qwen3.8-max- Provider
- Alibaba Cloud (Qwen)
- Context
- 1M tokens
- Max output
- 131K tokens
- Released
- August 3, 2026
- Batch API
- No batch discount listed
- Documented agent features
- Tool callingStructured outputBuilt-in web searchAdjustable reasoningImage input
What a task costs
Four reference workloads priced at this model's list rates, with caching applied after the first call. Your numbers will differ; the calculator lets you change every assumption.
| Workload | Per task | Per 1,000 tasks |
|---|---|---|
| Support bot | $0.039 | $39 |
| Coding agent | $1.63 | $1,629 |
| Browser agent | $1.15 | $1,153 |
| Research agent | $1.04 | $1,037 |
Other Alibaba Cloud (Qwen) models
Similar cost elsewhere
Models from other providers whose cost on the 30-step coding workload is closest to this one.