by NanoGPT · family glm-flash · listed Jul 2026RTVW
GLM 5.3 Flash Uncensored is an uncensored fine-tune of the efficient 320B mixture-of-experts reasoning model, built for unrestricted chat, creative writing, coding, agentic work, tool use, and long-context tasks.
Current rates USD per 1M tokens
Input$0.20
Output$0.80
Cache read$0.07
Cache write—
Against the market
Input price vs. maker flagships ($/1M tokens)Output price vs. maker flagships ($/1M tokens)