z-ai logo
z-ai

Z.ai: GLM 5.1

GLM-5.1 delivers a major leap in coding capability, with particularly significant gains in handling long-horizon tasks. $0.91 per million input tokens, $2.86 per million output tokens. 204,800 token context window. Higher uptime with 17 providers. Includes independent benchmarks from Artificial Analysis.

Présentation

Caractéristiques du modèle

Contexte
204 800 tokens
Sortie maximale
128 000 tokens
Architecture
text->text
Tokenizer
Other
Limite des connaissances
Non indiqué
Modéré
Non
OPENROUTER

Tarification complète

Tarifs synchronisés depuis OpenRouter, par million de tokens.

Entrée
$0.91
/M tokens
Sortie
$2.86
/M tokens
Lecture du cache
$0.169
/M tokens
API

Configuration du modèle

Paramètres par défaut

temperature
1
top_p
0.95

Raisonnement

Modéré
Non
Paramètres par défaut
Oui
API

Démarrage rapide

Appelez ce modèle via l’API compatible OpenAI d’OpenRouter.

curl https://openrouter.ai/api/v1/chat/completions \
  -H "Authorization: Bearer $OPENROUTER_API_KEY" \
  -H "Content-Type: application/json" \
  -d '{"model":"z-ai/glm-5.1","messages":[{"role":"user","content":"Hello!"}]}'
Capacités et modalités

EntréeSortie

Entrée
text
Sortie
text
Raisonnement

Oui

API

Paramètres API pris en charge

frequency_penaltyinclude_reasoninglogit_biaslogprobsmax_tokensmin_ppresence_penaltyreasoningrepetition_penaltyresponse_formatseedstopstructured_outputstemperaturetool_choicetoolstop_ktop_logprobstop_p
Fournisseurs disponibles

17 fournisseurs

Fournisseurs en direct sur OpenRouter

Disponibilité, latence, débit et routage évoluent en continu. Consultez la source pour les données en temps réel.

OpenRouter

GMICloud

fp8
Contexte
203K
Sortie maximale
182K
Entrée
$0.910
Sortie
$2.86
Lecture du cache
$0.169
Écriture du cache
Non indiqué

StreamLake

fp8
Contexte
200K
Sortie maximale
128K
Entrée
$0.966
Sortie
$3.04
Lecture du cache
$0.179
Écriture du cache
Non indiqué

Chutes

fp8
Contexte
203K
Sortie maximale
66K
Entrée
$0.980
Sortie
$3.08
Lecture du cache
$0.098
Écriture du cache
Non indiqué

DeepInfra

fp4
Contexte
203K
Sortie maximale
66K
Entrée
$1.05
Sortie
$3.50
Lecture du cache
$0.205
Écriture du cache
Non indiqué

SiliconFlow

fp8
Contexte
205K
Sortie maximale
131K
Entrée
$1.19
Sortie
$3.74
Lecture du cache
$0.600
Écriture du cache
Non indiqué

Crusoe

fp8
Contexte
203K
Sortie maximale
182K
Entrée
$1.20
Sortie
$4.40
Lecture du cache
$0.250
Écriture du cache
Non indiqué

Phala

unknown
Contexte
203K
Sortie maximale
128K
Entrée
$1.21
Sortie
$4.20
Lecture du cache
$0.600
Écriture du cache
Non indiqué

AtlasCloud

fp8
Contexte
203K
Sortie maximale
182K
Entrée
$1.26
Sortie
$3.96
Lecture du cache
$0.234
Écriture du cache
Non indiqué

DigitalOcean

unknown
Contexte
164K
Sortie maximale
147K
Entrée
$1.30
Sortie
$4.30
Lecture du cache
$0.260
Écriture du cache
Non indiqué

Alibaba

fp8
Contexte
203K
Sortie maximale
131K
Entrée
$1.33
Sortie
$4.18
Lecture du cache
$0.247
Écriture du cache
Non indiqué

Novita

fp8
Contexte
205K
Sortie maximale
131K
Entrée
$1.38
Sortie
$4.40
Lecture du cache
$0.260
Écriture du cache
Non indiqué

Baidu

fp8
Contexte
203K
Sortie maximale
131K
Entrée
$1.40
Sortie
$4.40
Lecture du cache
$0.260
Écriture du cache
Non indiqué

Nebius

fp8
Contexte
203K
Sortie maximale
182K
Entrée
$1.40
Sortie
$4.40
Lecture du cache
Non indiqué
Écriture du cache
Non indiqué

Z.AI

fp8
Contexte
203K
Sortie maximale
131K
Entrée
$1.40
Sortie
$4.40
Lecture du cache
$0.260
Écriture du cache
Non indiqué

Parasail

fp8
Contexte
203K
Sortie maximale
131K
Entrée
$1.40
Sortie
$4.40
Lecture du cache
$0.260
Écriture du cache
Non indiqué

Friendli

unknown
Contexte
203K
Sortie maximale
182K
Entrée
$1.40
Sortie
$4.40
Lecture du cache
$0.260
Écriture du cache
Non indiqué

Venice

fp8
Contexte
200K
Sortie maximale
80K
Entrée
$1.54
Sortie
$4.84
Lecture du cache
$0.286
Écriture du cache
Non indiqué
z-ai

modèles

Tous les modèles
z-ai logoz-ai

Z.ai: GLM 5.3 Flash (batch)

GLM-5.3-Flash is a native multimodal model from Z.ai. It is suited for efficient coding and long-horizon agent tasks. Its hybrid sparse and linear attention architecture maintains accurate long-context behavior while...

Contexte
1.0M
Entrée
text · image · video
Sortie
text
Entrée: $0.150Sortie: $0.500par million de tokens
Voir la fiche du modèle
z-ai logoz-ai

Z.ai: GLM 5.2

GLM 5.2 is a large-scale reasoning model from Z.ai. It supports text input and output with a 1M-token context window, and is suited for long-horizon agent workflows, project-level software engineering,...

Contexte
1.0M
Entrée
text
Sortie
text
Entrée: $1.19Sortie: $3.74par million de tokens
Voir la fiche du modèle
z-ai logoz-ai

Z.ai: GLM 5.2 (free)

GLM 5.2 is a large-scale reasoning model from Z.ai. It supports text input and output with a 1M-token context window, and is suited for long-horizon agent workflows, project-level software engineering,...

Contexte
256K
Entrée
text
Sortie
text
Entrée: GratuitSortie: Gratuitpar million de tokens
Voir la fiche du modèle
z-ai logoz-ai

Z.ai: GLM 5

GLM-5 is Z.ai’s flagship open-source foundation model engineered for complex systems design and long-horizon agent workflows. Built for expert developers, it delivers production-grade performance on large-scale programming tasks, rivaling leading...

Contexte
205K
Entrée
text
Sortie
text
Entrée: $0.600Sortie: $1.92par million de tokens
Voir la fiche du modèle