GLM-5.3-Flash
by engy · family glm-flash · listed Aug 2026RTVW
Native multimodal GLM model for efficient coding and long-horizon agent tasks
Current rates USD per 1M tokens
Input$0.135
Output$0.45
Cache read$0.027
Cache write—
Against the market
On the record
- Model id
glm-5.3-flash- Context window
- 262K tokens
- Max output
- 33K tokens
- Input modalities
- text, image
- Output modalities
- text
- Reasoning
- yes
- Tool calls
- yes
- Open weights
- yes
- Knowledge cutoff
- —
- Released
- 2026-08-26
- Last updated
- 2026-08-26
- Provider docs
- engy.ai/pricing
Back of the envelope
≈
Also listed at
- Zhipu AI Coding Planfree in · free out
- Volcengine Ark Coding Planfree in · free out
- Z.AI Coding Planfree in · free out
- SCNet Token Planfree in · free out
- Nvidiafree in · free out
- Pareto Inference$0.03 in · $0.10 out
- CrofAI$0.07 in · $0.22 out
- 302.AI$0.075 in · $0.25 out
- OrcaRouter$0.075 in · $0.25 out
- Merge Gateway$0.075 in · $0.25 out
- TokenGo$0.075 in · $0.025 out
- LLM Gateway$0.088 in · $0.25 out
- DevPass (LLM Gateway)$0.088 in · $0.25 out
- NanoGPT$0.10 in · $0.30 out
- Vultr$0.10 in · $0.35 out