z-ai logo
z-ai

Z.ai: GLM 5.2 (free)

GLM 5.2 is a large-scale reasoning model from Z.ai. This model is free to use. 256,000 token context window, maximum output of 256,000 tokens. Higher uptime with 27 providers. Includes independent benchmarks from Artificial Analysis.

模型概览

模型规格

上下文
256,000 tokens
最大输出
230,400 tokens
架构
text->text
分词器
Other
知识截止时间
暂未提供
内容审核
OPENROUTER

完整价格

同步自 OpenRouter。Token 价格均按每百万 Token 展示。

输入
免费
/M tokens
输出
免费
/M tokens
API

模型配置

默认参数

temperature
1
top_p
0.95

推理能力

内容审核
默认参数
high
能力与模态
xhigh, high
API

快速调用

通过 OpenRouter 的 OpenAI 兼容 API 调用这个确切的模型 ID。

curl https://openrouter.ai/api/v1/chat/completions \
  -H "Authorization: Bearer $OPENROUTER_API_KEY" \
  -H "Content-Type: application/json" \
  -d '{"model":"z-ai/glm-5.2:free","messages":[{"role":"user","content":"Hello!"}]}'
能力与模态

输入输出

输入
text
输出
text
推理能力

· xhigh · high

API

支持的 API 参数

frequency_penaltyinclude_reasoningmax_tokensmin_ppresence_penaltyreasoningreasoning_effortrepetition_penaltyresponse_formatseedstopstructured_outputstemperaturetool_choicetoolstop_ktop_p
可用提供方

33 个提供方

在 OpenRouter 查看实时 Provider

Provider 可用性、延迟、吞吐量和路由会持续变化,请前往来源页面查看实时运行数据。

OpenRouter

DeepInfra

fp4
上下文
1.0M
最大输出
164K
输入
$0.487
输出
$1.56
缓存读取
$0.091
缓存写入
暂未提供

Sail Research

fp8
上下文
1.0M
最大输出
131K
输入
$0.500
输出
$3.15
缓存读取
$0.115
缓存写入
暂未提供

Ambient

fp8
上下文
203K
最大输出
182K
输入
$0.600
输出
$2.00
缓存读取
$0.150
缓存写入
暂未提供

Decart

fp4
上下文
1.0M
最大输出
944K
输入
$0.684
输出
$2.28
缓存读取
$0.114
缓存写入
暂未提供

DigitalOcean

unknown
上下文
262K
最大输出
236K
输入
$0.700
输出
$2.20
缓存读取
$0.105
缓存写入
暂未提供

Inceptron

fp4
上下文
1.0M
最大输出
944K
输入
$0.710
输出
$2.35
缓存读取
$0.120
缓存写入
暂未提供

Makora

fp4
上下文
980K
最大输出
128K
输入
$0.720
输出
$2.38
缓存读取
$0.120
缓存写入
暂未提供

StreamLake

fp8
上下文
1.0M
最大输出
128K
输入
$0.735
输出
$2.31
缓存读取
$0.137
缓存写入
暂未提供

Novita

fp8
上下文
1.0M
最大输出
131K
输入
$0.742
输出
$2.33
缓存读取
$0.138
缓存写入
暂未提供

CoreWeave

fp4
上下文
1.0M
最大输出
944K
输入
$0.760
输出
$2.42
缓存读取
$0.140
缓存写入
暂未提供

Alibaba

fp8
上下文
1.0M
最大输出
131K
输入
$0.966
输出
$3.04
缓存读取
$0.193
缓存写入
暂未提供

GMICloud

fp8
上下文
1.0M
最大输出
944K
输入
$1.05
输出
$3.30
缓存读取
$0.195
缓存写入
暂未提供

Reka

fp8
上下文
262K
最大输出
131K
输入
$1.10
输出
$3.75
缓存读取
$0.200
缓存写入
暂未提供

SiliconFlow

fp8
上下文
1.0M
最大输出
262K
输入
$1.19
输出
$3.74
缓存读取
$0.221
缓存写入
暂未提供

Phala

fp8
上下文
1.0M
最大输出
131K
输入
$1.26
输出
$3.00
缓存读取
$0.220
缓存写入
暂未提供

AtlasCloud

fp8
上下文
1.0M
最大输出
131K
输入
$1.26
输出
$3.96
缓存读取
$0.234
缓存写入
暂未提供

Baidu

fp8
上下文
1.0M
最大输出
131K
输入
$1.40
输出
$4.40
缓存读取
$0.260
缓存写入
暂未提供

BaseTen

fp8
上下文
1.0M
最大输出
262K
输入
$1.40
输出
$4.40
缓存读取
$0.140
缓存写入
暂未提供

Mistral

unknown
上下文
1.0M
最大输出
128K
输入
$1.40
输出
$4.40
缓存读取
$0.140
缓存写入
暂未提供

Mistral

unknown
上下文
1.0M
最大输出
128K
输入
$1.40
输出
$4.40
缓存读取
$0.140
缓存写入
暂未提供

Fireworks

unknown
上下文
1.0M
最大输出
944K
输入
$1.40
输出
$4.40
缓存读取
$0.140
缓存写入
暂未提供

Cloudflare

unknown
上下文
262K
最大输出
236K
输入
$1.40
输出
$4.40
缓存读取
$0.260
缓存写入
暂未提供

Z.AI

fp8
上下文
1.0M
最大输出
131K
输入
$1.40
输出
$4.40
缓存读取
$0.260
缓存写入
暂未提供

Parasail

fp4
上下文
262K
最大输出
236K
输入
$1.40
输出
$4.40
缓存读取
$0.260
缓存写入
暂未提供

Together

unknown
上下文
512K
最大输出
461K
输入
$1.40
输出
$4.40
缓存读取
$0.260
缓存写入
暂未提供

Crusoe

fp8
上下文
1.0M
最大输出
944K
输入
$1.40
输出
$4.40
缓存读取
$0.260
缓存写入
暂未提供

Venice

fp8
上下文
1M
最大输出
131K
输入
$1.40
输出
$4.40
缓存读取
$0.260
缓存写入
暂未提供

Friendli

unknown
上下文
1.0M
最大输出
944K
输入
$1.40
输出
$4.40
缓存读取
$0.260
缓存写入
暂未提供

Mistral

unknown
上下文
1.0M
最大输出
128K
输入
$1.54
输出
$4.84
缓存读取
$0.154
缓存写入
暂未提供

BaseTen

fp8
上下文
1.0M
最大输出
262K
输入
$2.10
输出
$6.60
缓存读取
$0.210
缓存写入
暂未提供

Fireworks

unknown
上下文
1.0M
最大输出
944K
输入
$2.10
输出
$6.60
缓存读取
$0.210
缓存写入
暂未提供

Fireworks

unknown
上下文
1.0M
最大输出
944K
输入
$2.10
输出
$6.60
缓存读取
$0.210
缓存写入
暂未提供

Alibaba

fp8
上下文
1.0M
最大输出
131K
输入
$2.31
输出
$7.26
缓存读取
$0.462
缓存写入
暂未提供
z-ai

个模型

全部模型
z-ai logoz-ai

Z.ai: GLM 5.3 Flash (batch)

GLM-5.3-Flash is a native multimodal model from Z.ai. It is suited for efficient coding and long-horizon agent tasks. Its hybrid sparse and linear attention architecture maintains accurate long-context behavior while...

上下文
1.0M
输入
text · image · video
输出
text
输入: $0.150输出: $0.500每百万 Token
查看模型详情
z-ai logoz-ai

Z.ai: GLM 5.2

GLM 5.2 is a large-scale reasoning model from Z.ai. It supports text input and output with a 1M-token context window, and is suited for long-horizon agent workflows, project-level software engineering,...

上下文
1.0M
输入
text
输出
text
输入: $1.19输出: $3.74每百万 Token
查看模型详情
z-ai logoz-ai

Z.ai: GLM 5.1

GLM-5.1 delivers a major leap in coding capability, with particularly significant gains in handling long-horizon tasks. Unlike previous models built around minute-level interactions, GLM-5.1 can work independently and continuously on...

上下文
205K
输入
text
输出
text
输入: $0.966输出: $3.04每百万 Token
查看模型详情
z-ai logoz-ai

Z.ai: GLM 5

GLM-5 is Z.ai’s flagship open-source foundation model engineered for complex systems design and long-horizon agent workflows. Built for expert developers, it delivers production-grade performance on large-scale programming tasks, rivaling leading...

上下文
205K
输入
text
输出
text
输入: $0.600输出: $1.92每百万 Token
查看模型详情