z-ai logo
z-ai

Z.ai: GLM 5.2 (free)

GLM 5.2 is a large-scale reasoning model from Z.ai. This model is free to use. 256,000 token context window, maximum output of 256,000 tokens. Higher uptime with 27 providers. Includes independent benchmarks from Artificial Analysis.

Panoramica

Specifiche del modello

Contesto
256.000 tokens
Output massimo
230.400 tokens
Architettura
text->text
Tokenizer
Other
Limite di conoscenza
Non indicato
Moderato
No
OPENROUTER

Prezzi completi

Tariffe sincronizzate da OpenRouter, per milione di token.

Input
Gratuito
/M tokens
Output
Gratuito
/M tokens
API

Configurazione del modello

Parametri predefiniti

temperature
1
top_p
0.95

Ragionamento

Moderato
No
Parametri predefiniti
high
Capacità e modalità
xhigh, high
API

Avvio rapido

Chiama questo modello tramite l’API compatibile OpenAI di OpenRouter.

curl https://openrouter.ai/api/v1/chat/completions \
  -H "Authorization: Bearer $OPENROUTER_API_KEY" \
  -H "Content-Type: application/json" \
  -d '{"model":"z-ai/glm-5.2:free","messages":[{"role":"user","content":"Hello!"}]}'
Capacità e modalità

InputOutput

Input
text
Output
text
Ragionamento

· xhigh · high

API

Parametri API supportati

frequency_penaltyinclude_reasoningmax_tokensmin_ppresence_penaltyreasoningreasoning_effortrepetition_penaltyresponse_formatseedstopstructured_outputstemperaturetool_choicetoolstop_ktop_p
Provider disponibili

33 provider

Provider live su OpenRouter

Disponibilità, latenza, throughput e routing cambiano continuamente. Consulta la fonte per i dati correnti.

OpenRouter

DeepInfra

fp4
Contesto
1.0M
Output massimo
164K
Input
$0.487
Output
$1.56
Lettura cache
$0.091
Scrittura cache
Non indicato

Sail Research

fp8
Contesto
1.0M
Output massimo
131K
Input
$0.500
Output
$3.15
Lettura cache
$0.115
Scrittura cache
Non indicato

Ambient

fp8
Contesto
203K
Output massimo
182K
Input
$0.600
Output
$2.00
Lettura cache
$0.150
Scrittura cache
Non indicato

Decart

fp4
Contesto
1.0M
Output massimo
944K
Input
$0.684
Output
$2.28
Lettura cache
$0.114
Scrittura cache
Non indicato

DigitalOcean

unknown
Contesto
262K
Output massimo
236K
Input
$0.700
Output
$2.20
Lettura cache
$0.105
Scrittura cache
Non indicato

Inceptron

fp4
Contesto
1.0M
Output massimo
944K
Input
$0.710
Output
$2.35
Lettura cache
$0.120
Scrittura cache
Non indicato

Makora

fp4
Contesto
980K
Output massimo
128K
Input
$0.720
Output
$2.38
Lettura cache
$0.120
Scrittura cache
Non indicato

StreamLake

fp8
Contesto
1.0M
Output massimo
128K
Input
$0.735
Output
$2.31
Lettura cache
$0.137
Scrittura cache
Non indicato

Novita

fp8
Contesto
1.0M
Output massimo
131K
Input
$0.742
Output
$2.33
Lettura cache
$0.138
Scrittura cache
Non indicato

CoreWeave

fp4
Contesto
1.0M
Output massimo
944K
Input
$0.760
Output
$2.42
Lettura cache
$0.140
Scrittura cache
Non indicato

Alibaba

fp8
Contesto
1.0M
Output massimo
131K
Input
$0.966
Output
$3.04
Lettura cache
$0.193
Scrittura cache
Non indicato

GMICloud

fp8
Contesto
1.0M
Output massimo
944K
Input
$1.05
Output
$3.30
Lettura cache
$0.195
Scrittura cache
Non indicato

Reka

fp8
Contesto
262K
Output massimo
131K
Input
$1.10
Output
$3.75
Lettura cache
$0.200
Scrittura cache
Non indicato

SiliconFlow

fp8
Contesto
1.0M
Output massimo
262K
Input
$1.19
Output
$3.74
Lettura cache
$0.221
Scrittura cache
Non indicato

Phala

fp8
Contesto
1.0M
Output massimo
131K
Input
$1.26
Output
$3.00
Lettura cache
$0.220
Scrittura cache
Non indicato

AtlasCloud

fp8
Contesto
1.0M
Output massimo
131K
Input
$1.26
Output
$3.96
Lettura cache
$0.234
Scrittura cache
Non indicato

Baidu

fp8
Contesto
1.0M
Output massimo
131K
Input
$1.40
Output
$4.40
Lettura cache
$0.260
Scrittura cache
Non indicato

BaseTen

fp8
Contesto
1.0M
Output massimo
262K
Input
$1.40
Output
$4.40
Lettura cache
$0.140
Scrittura cache
Non indicato

Mistral

unknown
Contesto
1.0M
Output massimo
128K
Input
$1.40
Output
$4.40
Lettura cache
$0.140
Scrittura cache
Non indicato

Mistral

unknown
Contesto
1.0M
Output massimo
128K
Input
$1.40
Output
$4.40
Lettura cache
$0.140
Scrittura cache
Non indicato

Fireworks

unknown
Contesto
1.0M
Output massimo
944K
Input
$1.40
Output
$4.40
Lettura cache
$0.140
Scrittura cache
Non indicato

Cloudflare

unknown
Contesto
262K
Output massimo
236K
Input
$1.40
Output
$4.40
Lettura cache
$0.260
Scrittura cache
Non indicato

Z.AI

fp8
Contesto
1.0M
Output massimo
131K
Input
$1.40
Output
$4.40
Lettura cache
$0.260
Scrittura cache
Non indicato

Parasail

fp4
Contesto
262K
Output massimo
236K
Input
$1.40
Output
$4.40
Lettura cache
$0.260
Scrittura cache
Non indicato

Together

unknown
Contesto
512K
Output massimo
461K
Input
$1.40
Output
$4.40
Lettura cache
$0.260
Scrittura cache
Non indicato

Crusoe

fp8
Contesto
1.0M
Output massimo
944K
Input
$1.40
Output
$4.40
Lettura cache
$0.260
Scrittura cache
Non indicato

Venice

fp8
Contesto
1M
Output massimo
131K
Input
$1.40
Output
$4.40
Lettura cache
$0.260
Scrittura cache
Non indicato

Friendli

unknown
Contesto
1.0M
Output massimo
944K
Input
$1.40
Output
$4.40
Lettura cache
$0.260
Scrittura cache
Non indicato

Mistral

unknown
Contesto
1.0M
Output massimo
128K
Input
$1.54
Output
$4.84
Lettura cache
$0.154
Scrittura cache
Non indicato

BaseTen

fp8
Contesto
1.0M
Output massimo
262K
Input
$2.10
Output
$6.60
Lettura cache
$0.210
Scrittura cache
Non indicato

Fireworks

unknown
Contesto
1.0M
Output massimo
944K
Input
$2.10
Output
$6.60
Lettura cache
$0.210
Scrittura cache
Non indicato

Fireworks

unknown
Contesto
1.0M
Output massimo
944K
Input
$2.10
Output
$6.60
Lettura cache
$0.210
Scrittura cache
Non indicato

Alibaba

fp8
Contesto
1.0M
Output massimo
131K
Input
$2.31
Output
$7.26
Lettura cache
$0.462
Scrittura cache
Non indicato
z-ai

modelli

Tutti i modelli
z-ai logoz-ai

Z.ai: GLM 5.3 Flash (batch)

GLM-5.3-Flash is a native multimodal model from Z.ai. It is suited for efficient coding and long-horizon agent tasks. Its hybrid sparse and linear attention architecture maintains accurate long-context behavior while...

Contesto
1.0M
Input
text · image · video
Output
text
Input: $0.150Output: $0.500per milione di token
Vedi i dettagli del modello
z-ai logoz-ai

Z.ai: GLM 5.2

GLM 5.2 is a large-scale reasoning model from Z.ai. It supports text input and output with a 1M-token context window, and is suited for long-horizon agent workflows, project-level software engineering,...

Contesto
1.0M
Input
text
Output
text
Input: $1.19Output: $3.74per milione di token
Vedi i dettagli del modello
z-ai logoz-ai

Z.ai: GLM 5.1

GLM-5.1 delivers a major leap in coding capability, with particularly significant gains in handling long-horizon tasks. Unlike previous models built around minute-level interactions, GLM-5.1 can work independently and continuously on...

Contesto
205K
Input
text
Output
text
Input: $0.966Output: $3.04per milione di token
Vedi i dettagli del modello
z-ai logoz-ai

Z.ai: GLM 5

GLM-5 is Z.ai’s flagship open-source foundation model engineered for complex systems design and long-horizon agent workflows. Built for expert developers, it delivers production-grade performance on large-scale programming tasks, rivaling leading...

Contesto
205K
Input
text
Output
text
Input: $0.600Output: $1.92per milione di token
Vedi i dettagli del modello