GLM-5.3-Flash
zhipuai · zhipuai/glm-5.3-flash
Native multimodal GLM model for efficient coding and long-horizon agent tasks
≈667k Chinese characters of context · Reads images · Step-by-step reasoning · Calls tools / agents · Structured JSON output
Read a 200,000-Chinese-character book and write a 2,000-Chinese-character summary: about $0.02
An order-of-magnitude estimate at 1.5 tokens per Chinese character and the vendor's own rate — not a quote. Full method: the glossary on the AI Models page.
Specs
| Context | 1,000,000(≈667k Chinese characters) |
|---|---|
| Max output | 131,072 |
| Structured output | Yes |
| Open weights | No |
| Released | 2026-08-26 |
Source: models.dev snapshot 2026-08-29
Who serves it, at what price
| Provider | In /M | Out /M | Cache read /M |
|---|---|---|---|
| Zhipu AIfirst-party | $0.075 | $0.25 | $0.015 |
| Cortecs | $0.201 | $0.5 | $0.05 |
| CrofAI | $0.07 | $0.22 | $0.01 |
| DigitalOcean | $0.15 | $0.5 | $0.03 |
| DevPass (LLM Gateway) | $0.13 | $0.4 | $0.024 |
| Ollama Cloud | — | — | — |
| OpenCode Go | $0.075 | $0.25 | $0.015 |
| Requesty | $0.15 | $0.5 | $0.03 |
| Vancine | $0.06 | $0.2 | $0.012 |
| Vivgrid | $0.15 | $0.5 | $0.04 |
| Z.AI | $0.075 | $0.25 | $0.015 |
| Z.AI Coding Plan | Billed by plan | Billed by plan | $0 |
| Zhipu AI Coding Plan | Billed by plan | Billed by plan | $0 |
Rows tagged “first-party” are the model vendor's own pricing; untagged rows are resale channels and may differ.
When to choose it
We have not assessed this model. It comes from the full models.dev snapshot — the specs are the vendor's published figures, and a price is the vendor's own only on rows tagged first-party; the rest are resale channels. The fit is simply not something anyone here has written.
Leaderboard
Not yet on any leaderboard we track. That does not mean the model performs poorly — it means we have not taken this snapshot for it yet, and we do not publish a score we have not measured.