GLM-4.6V scales its context window to 128k tokens in training, and achieves SoTA performance in visual understanding among models of similar parameter scales. Integrates native Function Calling capabilities, bridging 'visual perception' and 'executable action' for multimodal agents. Direct via Z-AI (Zhipu).
Current rates USD per 1M tokens
Input$0.60
Output$0.90
Cache read$0.30
Cache write—
Against the market
Input price vs. maker flagships ($/1M tokens)Output price vs. maker flagships ($/1M tokens)