Qwen3.8 Max
alibaba · alibaba/qwen3.8-max
2.4-trillion-parameter MoE flagship for coding, professional work, multimodal understanding, and long-horizon agentic workflows
≈667k Chinese characters of context · Reads images · Step-by-step reasoning · Calls tools / agents · Structured JSON output
Read a 200,000-Chinese-character book and write a 2,000-Chinese-character summary: about $0.62
An order-of-magnitude estimate at 1.5 tokens per Chinese character and the vendor's own rate — not a quote. Full method: the glossary on the AI Models page.
Specs
| Context | 1,000,000(≈667k Chinese characters) |
|---|---|
| Max output | 131,072 |
| Structured output | Yes |
| Open weights | No |
| Released | 2026-08-03 |
Source: models.dev snapshot 2026-08-29
Who serves it, at what price
| Provider | In /M | Out /M | Cache read /M |
|---|---|---|---|
| Alibabafirst-party | $2 | $6 | $0.25 |
| Abacus | $2 | $6 | — |
| AIHubMix | $1.69 | $5.07 | $0.169 |
| Alibaba (China) | $1.77744 | $5.33231 | $0.22218 |
| Alibaba Token Plan | Billed by plan | Billed by plan | $0 |
| Alibaba Token Plan (China) | Billed by plan | Billed by plan | $0 |
| DigitalOcean | $2 | $6 | $0.2 |
| Charm Hyper | $2 | $6 | $0.25 |
| DevPass (LLM Gateway) | $1.815 | $5.4461 | $0.21 |
| NanoGPT | $2 | $6 | $0.25 |
| OpenCode Go | $2 | $6 | $0.25 |
| Requesty | $2 | $6 | $0.25 |
| Vancine | $1.6 | $4.8 | $0.2 |
Rows tagged “first-party” are the model vendor's own pricing; untagged rows are resale channels and may differ.
When to choose it
This model is on our watch list but has no written assessment yet — we do not write what we cannot support. The specs and prices above still stand.
Leaderboard
Not yet on any leaderboard we track. That does not mean the model performs poorly — it means we have not taken this snapshot for it yet, and we do not publish a score we have not measured.