z-ai logo
z-ai

Z.ai: GLM 4.7

Description de la source (anglais)

GLM-4.7 is Z.ai’s latest flagship model, featuring upgrades in two key areas: enhanced programming capabilities and more stable multi-step reasoning/execution. It demonstrates significant improvements in executing complex agent tasks while...

Présentation

Caractéristiques du modèle

Contexte
204 800 tokens
Sortie maximale
131 072 tokens
Architecture
text->text
Tokenizer
Other
Limite des connaissances
Non indiqué
Modéré
Non
OPENROUTER

Tarification complète

Tarifs synchronisés depuis OpenRouter, par million de tokens.

Entrée
$0.4
par million de tokens
Sortie
$1.75
par million de tokens
Lecture du cache
$0.08
par million de tokens
API

Configuration du modèle

Paramètres par défaut

temperature
1
top_p
0.95

Raisonnement

Raisonnement obligatoire
Non
Paramètres par défaut
Oui
API

Démarrage rapide

Définissez OPENROUTER_API_KEY localement. Python nécessite requests ; JavaScript s’exécute dans Node.js. Gardez la clé côté serveur.

Documentation de l’API

curl --fail-with-body https://openrouter.ai/api/v1/chat/completions \
  -H "Authorization: Bearer $OPENROUTER_API_KEY" \
  -H "Content-Type: application/json" \
  -d '{"model":"z-ai/glm-4.7","messages":[{"role":"user","content":"Hello!"}]}'
Capacités et modalités

EntréeSortie

Entrée
text
Sortie
text
Raisonnement

Oui

API

Paramètres API pris en charge

frequency_penaltyinclude_reasoninglogit_biasmax_tokensmin_ppresence_penaltyreasoningrepetition_penaltyresponse_formatseedstopstructured_outputstemperaturetool_choicetoolstop_atop_ktop_p
Fournisseurs disponibles

7 fournisseurs

Vérifié: 7 septembre 2026

Fournisseurs en direct sur OpenRouter

Disponibilité, latence, débit et routage évoluent en continu. Consultez la source pour les données en temps réel.

OpenRouter

DeepInfra

fp4
Contexte
203K
Sortie maximale
131K
Entrée
$0.4
Sortie
$1.75
Lecture du cache
$0.08
Écriture du cache
Non indiqué

AtlasCloud

fp8
Contexte
203K
Sortie maximale
182K
Entrée
$0.52
Sortie
$1.85
Lecture du cache
$0.12
Écriture du cache
Non indiqué

Novita

fp8
Contexte
205K
Sortie maximale
131K
Entrée
$0.54
Sortie
$1.98
Lecture du cache
$0.099
Écriture du cache
Non indiqué

Venice

fp4
Contexte
198K
Sortie maximale
16K
Entrée
$0.55
Sortie
$2.65
Lecture du cache
$0.11
Écriture du cache
Non indiqué

Google

Non indiqué
Contexte
200K
Sortie maximale
128K
Entrée
$0.6
Sortie
$2.2
Lecture du cache
Non indiqué
Écriture du cache
Non indiqué

Z.AI

fp4
Contexte
203K
Sortie maximale
131K
Entrée
$0.6
Sortie
$2.2
Lecture du cache
$0.11
Écriture du cache
Non indiqué

Mancer 2

fp4
Contexte
131K
Sortie maximale
118K
Entrée
$0.7
Sortie
$2.5
Lecture du cache
Non indiqué
Écriture du cache
Non indiqué
z-ai

4 modèles

Tous les modèles
z-ai logoz-ai

Z.ai: GLM 5.3 FlashX

GLM-5.3-FlashX is the high-speed variant of Z.ai's GLM-5.3-Flash, a native multimodal model delivering inference speeds of up to 200 tokens/s. Built on the same hybrid sparse and linear attention architecture...

Contexte
1.0M
Entrée
text · image · video
Sortie
text
Entrée: $0.37Sortie: $1.25par million de tokens
Voir la fiche du modèle
z-ai logoz-ai

Z.ai: GLM 5.3 Flash

GLM-5.3-Flash is a native multimodal model from Z.ai. It is suited for efficient coding and long-horizon agent tasks. Its hybrid sparse and linear attention architecture maintains accurate long-context behavior while...

Contexte
1.3M
Entrée
text · image · video
Sortie
text
Entrée: $0.15Sortie: $0.5par million de tokens
Voir la fiche du modèle
z-ai logoz-ai

Z.ai: GLM 5.3 Flash (batch)

GLM-5.3-Flash is a native multimodal model from Z.ai. It is suited for efficient coding and long-horizon agent tasks. Its hybrid sparse and linear attention architecture maintains accurate long-context behavior while...

Contexte
1.0M
Entrée
text · image · video
Sortie
text
Entrée: $0.06Sortie: $0.2par million de tokens
Voir la fiche du modèle
z-ai logoz-ai

Z.ai: GLM 5.3

GLM-5.3 is a large-scale reasoning model from Z.ai, built for complex software engineering and long-horizon agent tasks. It supports text input and output with a 1M-token context window, and improves...

Contexte
1.3M
Entrée
text
Sortie
text
Entrée: $0.5614Sortie: $1.7644par million de tokens
Voir la fiche du modèle