All Models
GLM 4.6V Original
GLM-4.6V scales its context window to 128k tokens in training, and achieves SoTA performance in visual understanding among models of similar parameter scales. Integrates native Function Calling capabilities, bridging 'visual perception' and 'executable action' for multimodal agents. Direct via Z-AI (Zhipu).
Available Providers (1)
| Provider | Model ID | Input Cost | Output Cost | Context | Max Output | Docs |
|---|---|---|---|---|---|---|
z-ai/glm-4.6v-original | $0.60/MTok | $0.90/MTok | 128K | 24K |
Capabilities
Reasoning
Tool Calling
Attachments
Open Weights
Structured Output