GLM 4.6V
GLM-4.6V is a large multimodal model designed for high-fidelity visual understanding and long-context reasoning across images, documents, and mixed media. It supports up to 128K tokens, processes complex page layouts and charts directly as visual inputs, and integrates native multimodal function calling to connect perception with downstream tool execution. The model also enables interleaved image-text generation and UI reconstruction workflows, including screenshot-to-HTML synthesis and iterative visual editing.
z-ai/glm-4.6v
- Context
- 131k tokens
- Completion cap
- 32,768
- Tools
- Yes
- JSON
- Yes
- Released
- 2025-12-08
Call it from Geek Hub
Same OpenAI SDK. Change the base URL and the model id.
curl https://api.geekhub.mx/v1/chat/completions \
-H "Authorization: Bearer $GEEKHUB_API_KEY" \
-H "Content-Type: application/json" \
-d '{
"model": "z-ai/glm-4.6v",
"messages": [{"role": "user", "content": "Hola"}]
}'Get an API keyFrequently asked questions
- What is GLM 4.6V?
- GLM-4.6V is a large multimodal model designed for high-fidelity visual understanding and long-context reasoning across images, documents, and mixed media. It supports up to 128K tokens, processes complex page layouts and charts directly as visual inputs, and integrates native multimodal function calling to connect perception with downstream tool execution. The model also enables interleaved image-text generation and UI reconstruction workflows, including screenshot-to-HTML synthesis and iterative visual editing. GLM 4.6V runs on the Geek Hub API (OpenAI-compatible). Model id: z-ai/glm-4.6v.
- Is GLM 4.6V free?
- No. Input is $0.327 per 1M tokens and output is $0.9811 per 1M tokens on Geek Hub (markup included).
- What is the context length of GLM 4.6V?
- GLM 4.6V has a 131k tokens context window. It supports up to 32,768 completion tokens.
- Does GLM 4.6V support tool calling and structured outputs?
- GLM 4.6V accepts tools and tool_choice for function calling. It also supports structured outputs via a JSON schema in response_format.
- When was GLM 4.6V released?
- GLM 4.6V was released on 2025-12-08.
More models from Zhipu
- GLM 4.5GLM-4.5 is our latest flagship foundation model, purpose-built for agent-based applications. It leverages a Mixture-of-Experts (MoE) architecture and supports a context length of u
- GLM 4.5 AirGLM-4.5-Air is the lightweight variant of our latest flagship model family, also purpose-built for agent-centric applications. Like GLM-4.5, it adopts the Mixture-of-Experts (MoE)
- GLM 4.5VGLM-4.5V is a vision-language foundation model for multimodal agent applications. Built on a Mixture-of-Experts (MoE) architecture with 106B parameters and 12B activated parameters
- GLM 4.6Compared with GLM-4.5, this generation brings several key improvements: Longer context window: The context window has been expanded from 128K to 200K tokens, enabling the model to
- GLM 4.7GLM-4.7 is Z.ai’s latest flagship model, featuring upgrades in two key areas: enhanced programming capabilities and more stable multi-step reasoning/execution. It demonstrates sign
- GLM 4.7 FlashAs a 30B-class SOTA model, GLM-4.7-Flash offers a new option that balances performance and efficiency. It is further optimized for agentic coding use cases, strengthening coding ca