GLM 4.7 Flash

As a 30B-class SOTA model, GLM-4.7-Flash offers a new option that balances performance and efficiency. It is further optimized for agentic coding use cases, strengthening coding capabilities, long-horizon task planning, and tool collaboration, and has achieved leading performance among open-source models of the same size on several current public benchmark leaderboards.

ChatZhipu203k tokens$0.0654 / $0.436 · 1M

z-ai/glm-4.7-flash

Contexto
203k tokens
Máx. completion
16,384
Tools
JSON
Lanzamiento
2026-01-19

Llámalo desde Geek Hub

El mismo SDK de OpenAI. Cambia el base URL y el id del modelo.

curl https://api.geekhub.mx/v1/chat/completions \
  -H "Authorization: Bearer $GEEKHUB_API_KEY" \
  -H "Content-Type: application/json" \
  -d '{
    "model": "z-ai/glm-4.7-flash",
    "messages": [{"role": "user", "content": "Hola"}]
  }'
Consigue tu API key

Preguntas frecuentes

¿Qué es GLM 4.7 Flash?
As a 30B-class SOTA model, GLM-4.7-Flash offers a new option that balances performance and efficiency. It is further optimized for agentic coding use cases, strengthening coding capabilities, long-horizon task planning, and tool collaboration, and has achieved leading performance among open-source models of the same size on several current public benchmark leaderboards. GLM 4.7 Flash corre en el API de Geek Hub (compatible con OpenAI). Id: z-ai/glm-4.7-flash.
¿GLM 4.7 Flash es gratis?
No. El input cuesta $0.0654 / 1M tokens y el output $0.436 / 1M tokens en Geek Hub (markup incluido).
¿Cuál es el contexto de GLM 4.7 Flash?
GLM 4.7 Flash tiene una ventana de 203k tokens. Soporta hasta 16,384 tokens de completion.
¿GLM 4.7 Flash soporta tool calling y structured outputs?
GLM 4.7 Flash acepta tools y tool_choice para function calling. También soporta structured outputs con un JSON schema en response_format.
¿Cuándo se lanzó GLM 4.7 Flash?
GLM 4.7 Flash se lanzó el 19 de enero de 2026.

Más modelos de Zhipu