qwen logo
qwen

Qwen: Qwen3.5 397B A17B

The Qwen3.5 series 397B-A17B native vision-language model is built on a hybrid architecture that integrates a linear attention mechanism with a sparse mixture-of-experts model, achieving higher inference efficiency. $0.39 per million input tokens, $2.34 per million output tokens. 262,144 token context window, maximum output of 65,536 tokens. Higher uptime with 10 providers. Includes independent benchmarks from Artificial Analysis.

Présentation

Caractéristiques du modèle

Contexte
262 144 tokens
Sortie maximale
65 536 tokens
Architecture
text+image+video->text
Tokenizer
Qwen3
Limite des connaissances
Non indiqué
Modéré
Non
OPENROUTER

Tarification complète

Tarifs synchronisés depuis OpenRouter, par million de tokens.

Entrée
$0.39
/M tokens
Sortie
$2.34
/M tokens
API

Configuration du modèle

Paramètres par défaut

temperature
0.6
top_p
0.95
top_k
20

Raisonnement

Modéré
Non
Paramètres par défaut
Non
API

Démarrage rapide

Appelez ce modèle via l’API compatible OpenAI d’OpenRouter.

curl https://openrouter.ai/api/v1/chat/completions \
  -H "Authorization: Bearer $OPENROUTER_API_KEY" \
  -H "Content-Type: application/json" \
  -d '{"model":"qwen/qwen3.5-397b-a17b","messages":[{"role":"user","content":"Hello!"}]}'
Capacités et modalités

EntréeSortie

Entrée
textimagevideo
Sortie
text
Raisonnement

Non

API

Paramètres API pris en charge

frequency_penaltyinclude_reasoninglogit_biaslogprobsmax_tokensmin_ppresence_penaltyreasoningrepetition_penaltyresponse_formatseedstopstructured_outputstemperaturetool_choicetoolstop_ktop_logprobstop_p
Fournisseurs disponibles

10 fournisseurs

Fournisseurs en direct sur OpenRouter

Disponibilité, latence, débit et routage évoluent en continu. Consultez la source pour les données en temps réel.

OpenRouter

Alibaba

unknown
Contexte
262K
Sortie maximale
66K
Entrée
$0.390
Sortie
$2.34
Lecture du cache
Non indiqué
Écriture du cache
Non indiqué

DeepInfra

fp8
Contexte
262K
Sortie maximale
82K
Entrée
$0.450
Sortie
$3.00
Lecture du cache
$0.220
Écriture du cache
Non indiqué

Parasail

fp8
Contexte
262K
Sortie maximale
236K
Entrée
$0.500
Sortie
$3.60
Lecture du cache
$0.300
Écriture du cache
Non indiqué

DigitalOcean

unknown
Contexte
131K
Sortie maximale
118K
Entrée
$0.550
Sortie
$3.50
Lecture du cache
$0.111
Écriture du cache
Non indiqué

AtlasCloud

fp8
Contexte
262K
Sortie maximale
66K
Entrée
$0.550
Sortie
$3.50
Lecture du cache
$0.550
Écriture du cache
Non indiqué

Phala

unknown
Contexte
262K
Sortie maximale
236K
Entrée
$0.550
Sortie
$3.50
Lecture du cache
$0.225
Écriture du cache
Non indiqué

Novita

unknown
Contexte
262K
Sortie maximale
66K
Entrée
$0.600
Sortie
$3.60
Lecture du cache
Non indiqué
Écriture du cache
Non indiqué

GMICloud

fp8
Contexte
262K
Sortie maximale
236K
Entrée
$0.600
Sortie
$3.60
Lecture du cache
Non indiqué
Écriture du cache
Non indiqué

StreamLake

unknown
Contexte
256K
Sortie maximale
64K
Entrée
$0.600
Sortie
$3.60
Lecture du cache
$0.120
Écriture du cache
Non indiqué

Venice

unknown
Contexte
128K
Sortie maximale
33K
Entrée
$0.750
Sortie
$4.50
Lecture du cache
Non indiqué
Écriture du cache
Non indiqué
qwen

modèles

Tous les modèles
qwen logoqwen

Qwen: Qwen3.8 Flash

Qwen3.8 Flash is a multimodal reasoning model from Alibaba. It is suited for coding assistance, agentic workflows, visual understanding, document and codebase analysis, desktop interaction, chart analysis, and long-video analysis.

Contexte
1M
Entrée
text · image · video
Sortie
text
Entrée: $0.150Sortie: $0.470par million de tokens
Voir la fiche du modèle
qwen logoqwen

Qwen: Qwen3.8 27B

Qwen3.8 27B is an open-weight dense vision-language model from Qwen. It is suited for coding, professional workflows, research, multimodal interaction, and long-running agent tasks, with flexible thinking that can be...

Contexte
1M
Entrée
text · image · video
Sortie
text
Entrée: $0.425Sortie: $2.55par million de tokens
Voir la fiche du modèle
qwen logoqwen

Qwen3 Reranker 8B

Qwen3 Reranker 8B is a text reranking model from Alibaba Cloud built on the Qwen3 architecture. It evaluates query-document pairs to produce relevance scores for use in retrieval and RAG...

Contexte
41K
Entrée
text
Sortie
rerank
Entrée: GratuitSortie: Gratuitpar million de tokens
Voir la fiche du modèle
qwen logoqwen

Qwen: Qwen3 ASR 1.7B

Qwen3 ASR 1.7B is an automatic speech recognition model from Qwen. It supports multilingual language identification and transcription across 30 languages and 22 Chinese dialects, with streaming and offline inference...

Contexte
Non indiqué
Entrée
audio
Sortie
transcription
Entrée: $7.50Sortie: Gratuitpar million de tokens
Voir la fiche du modèle