GLM 5.3 Flash
GLM-5.3-Flash is a native multimodal model from Z.ai. It is suited for efficient coding and long-horizon agent tasks. Its hybrid sparse and linear attention architecture maintains accurate long-context behavior while...
What published route metadata supports
Only source-backed fields are filled
LLMPrice does not infer a release date, knowledge cutoff, or open-weight status from a model name, Hugging Face link, or pricing route.
| Endpoint / route | Model ID | Input | Cached | Output | Context | Source |
|---|---|---|---|---|---|---|
| OpenRouterBatch | z-ai/glm-5.3-flash:batch | $0.075 | $0.015 | $0.25 | 1.05M | OpenRouter API ↗ |
| OpenRouterStandard | z-ai/glm-5.3-flash | $0.15 | $0.05 | $0.5 | 1.31M | OpenRouter API ↗ |
Hugging Face references
Presence of a Hugging Face ID is not treated as proof that model weights are open.
4 verified price changes since Aug 26, 2026
1 route availability change recorded. No earlier prices are inferred.
- Sep 22, 2026OpenRouter · Standard
Input $0.09 → $0.15 · Cached input $0.018 → $0.05 · Output $0.3 → $0.5
- Sep 16, 2026OpenRouter · Standard
Input $0.15 → $0.09 · Cached input $0.03 → $0.018 · Output $0.5 → $0.3
- Sep 11, 2026OpenRouter · Standard
Input $0.075 → $0.15 · Cached input $0.015 → $0.03 · Output $0.25 → $0.5
- Sep 9, 2026OpenRouter · Batch
Input $0.15 → $0.075 · Cached input $0.03 → $0.015 · Output $0.5 → $0.25
- Aug 28, 2026OpenRouter · Batch
New pricing route added