GLM 5.3 Flash (batch)

GLM-5.3-Flash is a native multimodal model from Z.ai. It is suited for efficient coding and long-horizon agent tasks. Its hybrid sparse and linear attention architecture maintains accurate long-context behavior while reducing compute overhead.

ChatZhipu1M tokens$0.1635 / $0.545 · 1M

z-ai/glm-5.3-flash:batch

Contexto
1M tokens
Máx. completion
943,717
Tools
JSON
Lanzamiento
2026-08-26

Llámalo desde Geek Hub

El mismo SDK de OpenAI. Cambia el base URL y el id del modelo.

curl https://api.geekhub.mx/v1/chat/completions \
  -H "Authorization: Bearer $GEEKHUB_API_KEY" \
  -H "Content-Type: application/json" \
  -d '{
    "model": "z-ai/glm-5.3-flash:batch",
    "messages": [{"role": "user", "content": "Hola"}]
  }'
Consigue tu API key

Preguntas frecuentes

¿Qué es GLM 5.3 Flash (batch)?
GLM-5.3-Flash is a native multimodal model from Z.ai. It is suited for efficient coding and long-horizon agent tasks. Its hybrid sparse and linear attention architecture maintains accurate long-context behavior while reducing compute overhead. GLM 5.3 Flash (batch) corre en el API de Geek Hub (compatible con OpenAI). Id: z-ai/glm-5.3-flash:batch.
¿GLM 5.3 Flash (batch) es gratis?
No. El input cuesta $0.1635 / 1M tokens y el output $0.545 / 1M tokens en Geek Hub (markup incluido).
¿Cuál es el contexto de GLM 5.3 Flash (batch)?
GLM 5.3 Flash (batch) tiene una ventana de 1M tokens. Soporta hasta 943,717 tokens de completion.
¿GLM 5.3 Flash (batch) soporta tool calling y structured outputs?
GLM 5.3 Flash (batch) acepta tools y tool_choice para function calling. También soporta structured outputs con un JSON schema en response_format.
¿Cuándo se lanzó GLM 5.3 Flash (batch)?
GLM 5.3 Flash (batch) se lanzó el 26 de agosto de 2026.

Más modelos de Zhipu