z-ai logo
z-ai

Z.ai: GLM 5.2 (free)

Descripción de la fuente (inglés)

GLM 5.2 is a large-scale reasoning model from Z.ai. It supports text input and output with a 1M-token context window, and is suited for long-horizon agent workflows, project-level software engineering,...

Resumen

Especificaciones del modelo

Contexto
32.768 tokens
Salida máxima
29.491 tokens
Arquitectura
text->text
Tokenizador
Other
Corte de conocimiento
No indicado
Moderado
No
OPENROUTER

Precios completos

Tarifas sincronizadas desde OpenRouter, por millón de tokens.

Entrada
Gratis
por millón de tokens
Salida
Gratis
por millón de tokens
API

Configuración del modelo

Parámetros predeterminados

temperature
1
top_p
0.95

Razonamiento

Razonamiento obligatorio
No
Parámetros predeterminados
high
Capacidades y modalidades
xhigh, high
API

Inicio rápido

Configura OPENROUTER_API_KEY localmente. Python requiere requests; JavaScript se ejecuta en Node.js. Mantén la clave en el servidor.

Documentación de la API

curl --fail-with-body https://openrouter.ai/api/v1/chat/completions \
  -H "Authorization: Bearer $OPENROUTER_API_KEY" \
  -H "Content-Type: application/json" \
  -d '{"model":"z-ai/glm-5.2:free","messages":[{"role":"user","content":"Hello!"}]}'
Capacidades y modalidades

EntradaSalida

Entrada
text
Salida
text
Razonamiento

· xhigh · high

API

Parámetros API compatibles

frequency_penaltyinclude_reasoningmax_tokensmin_ppresence_penaltyreasoningreasoning_effortrepetition_penaltyseedstoptemperaturetop_ktop_p
Proveedores disponibles

1 Proveedor

Verificado: 16 de septiembre de 2026

Proveedores en vivo en OpenRouter

Disponibilidad, latencia, rendimiento y enrutamiento cambian continuamente. Consulta la fuente para datos actuales.

OpenRouter

Decart

fp4
Contexto
33K
Salida máxima
29K
Entrada
Gratis
Salida
Gratis
Lectura de caché
No indicado
Escritura de caché
No indicado
z-ai

4 modelos

Todos los modelos
z-ai logoz-ai

Z.ai: GLM 5.3 FlashX

GLM-5.3-FlashX is the high-speed variant of Z.ai's GLM-5.3-Flash, a native multimodal model delivering inference speeds of up to 200 tokens/s. Built on the same hybrid sparse and linear attention architecture...

Contexto
1.0M
Entrada
text · image · video
Salida
text
Entrada: $0.37Salida: $1.25por millón de tokens
Ver detalles del modelo
z-ai logoz-ai

Z.ai: GLM 5.3 Flash

GLM-5.3-Flash is a native multimodal model from Z.ai. It is suited for efficient coding and long-horizon agent tasks. Its hybrid sparse and linear attention architecture maintains accurate long-context behavior while...

Contexto
1.3M
Entrada
text · image · video
Salida
text
Entrada: $0.15Salida: $0.5por millón de tokens
Ver detalles del modelo
z-ai logoz-ai

Z.ai: GLM 5.3 Flash (batch)

GLM-5.3-Flash is a native multimodal model from Z.ai. It is suited for efficient coding and long-horizon agent tasks. Its hybrid sparse and linear attention architecture maintains accurate long-context behavior while...

Contexto
1.0M
Entrada
text · image · video
Salida
text
Entrada: $0.06Salida: $0.2por millón de tokens
Ver detalles del modelo
z-ai logoz-ai

Z.ai: GLM 5.3

GLM-5.3 is a large-scale reasoning model from Z.ai, built for complex software engineering and long-horizon agent tasks. It supports text input and output with a 1M-token context window, and improves...

Contexto
1.3M
Entrada
text
Salida
text
Entrada: $0.5614Salida: $1.7644por millón de tokens
Ver detalles del modelo