Fast GLM vision model for screenshots, documents, and multimodal agent tasks
Fast GLM vision model for screenshots, documents, and multimodal agent tasks. Supports a context window of 200,000 tokens. Supports extended reasoning. Supports tool calling. Input priced at $5 per million tokens. Weights are not publicly released; access is via the provider API.
Zhipu AI
api
paid
No benchmark results have been added yet.