GLM-4.7-Flash
by Hugging Face · family glm · listed Aug 2025RTW
Efficient GLM model for fast reasoning, coding, and agent workflows
Current rates USD per 1M tokens
Inputfree
Outputfree
Cache read—
Cache write—
Against the market
On the record
- Model id
zai-org/GLM-4.7-Flash- Context window
- 200K tokens
- Max output
- 128K tokens
- Input modalities
- text
- Output modalities
- text
- Reasoning
- yes
- Tool calls
- yes
- Open weights
- yes
- Knowledge cutoff
- Apr 2025
- Released
- 2025-08-08
- Last updated
- 2025-08-08
- Provider docs
- huggingface.co/docs/inference-providers
Back of the envelope
≈
Also listed at
- Z.AIfree in · free out
- Zhipu AIfree in · free out
- CrofAI$0.04 in · $0.30 out
- Deep Infra$0.06 in · $0.40 out
- OpenRouter$0.06 in · $0.40 out
- LLM Gateway$0.06 in · $0.40 out
- Eden AI$0.06 in · $0.40 out
- DevPass (LLM Gateway)$0.06 in · $0.40 out
- Kilo Gateway$0.06 in · $0.40 out
- Cloudflare Workers AI$0.0605 in · $0.40 out
- Cloudflare AI Gateway$0.0605 in · $0.40 out
- NanoGPT$0.07 in · $0.40 out
- Jiekou.AI$0.07 in · $0.40 out
- Merge Gateway$0.07 in · $0.40 out
- Vercel AI Gateway$0.07 in · $0.40 out