qwen logo
qwen

Qwen: Qwen3 Next 80B A3B Instruct

Beschreibung der Quelle (Englisch)

Qwen3-Next-80B-A3B-Instruct is an instruction-tuned chat model in the Qwen3-Next series optimized for fast, stable responses without “thinking” traces. It targets complex tasks across reasoning, code generation, knowledge QA, and multilingual...

Überblick

Modellspezifikationen

Kontext
262.144 tokens
Maximale Ausgabe
16.384 tokens
Architektur
text->text
Tokenizer
Qwen3
Wissensstand
2025-09-30
Moderiert
Nein
OPENROUTER

Vollständige Preise

Mit OpenRouter synchronisierte Preise; Tokenpreise gelten pro eine Million Token.

Eingabe
$0.09
pro 1 Mio. Token
Ausgabe
$1.1
pro 1 Mio. Token
API

Schnellstart

Setzen Sie OPENROUTER_API_KEY lokal. Python benötigt requests; JavaScript läuft in Node.js. Bewahren Sie den Schlüssel auf dem Server auf.

API-Dokumentation

curl --fail-with-body https://openrouter.ai/api/v1/chat/completions \
  -H "Authorization: Bearer $OPENROUTER_API_KEY" \
  -H "Content-Type: application/json" \
  -d '{"model":"qwen/qwen3-next-80b-a3b-instruct","messages":[{"role":"user","content":"Hello!"}]}'
Fähigkeiten und Modalitäten

EingabeAusgabe

Eingabe
text
Ausgabe
text
API

Unterstützte API-Parameter

frequency_penaltylogit_biaslogprobsmax_tokensmin_ppresence_penaltyrepetition_penaltyresponse_formatseedstopstructured_outputstemperaturetool_choicetoolstop_ktop_logprobstop_p
Verfügbare Anbieter

5 Anbieter

Geprüft: 9. September 2026

Live-Anbieter bei OpenRouter

Verfügbarkeit, Latenz, Durchsatz und Routing ändern sich laufend. Aktuelle Betriebsdaten stehen auf der Quellseite.

OpenRouter

DeepInfra

fp8
Kontext
262K
Maximale Ausgabe
16K
Eingabe
$0.09
Ausgabe
$1.1
Cache-Lesen
Nicht angegeben
Cache-Schreiben
Nicht angegeben

Alibaba

Nicht angegeben
Kontext
131K
Maximale Ausgabe
33K
Eingabe
$0.0975
Ausgabe
$0.78
Cache-Lesen
Nicht angegeben
Cache-Schreiben
Nicht angegeben

Parasail

fp8
Kontext
262K
Maximale Ausgabe
236K
Eingabe
$0.1
Ausgabe
$1.1
Cache-Lesen
$0.07
Cache-Schreiben
Nicht angegeben

Google

Nicht angegeben
Kontext
262K
Maximale Ausgabe
236K
Eingabe
$0.15
Ausgabe
$1.2
Cache-Lesen
Nicht angegeben
Cache-Schreiben
Nicht angegeben

Novita

bf16
Kontext
131K
Maximale Ausgabe
33K
Eingabe
$0.15
Ausgabe
$1.5
Cache-Lesen
Nicht angegeben
Cache-Schreiben
Nicht angegeben
qwen

4 Modelle

Alle Modelle
qwen logoqwen

Qwen: Qwen3.8 Omni Flash

Qwen3.8 Omni Flash is an omni-modal reasoning model from Alibaba, the first Qwen model built around agentic capabilities with native audio-video understanding. It is suited for audio-video analysis and summarization,...

Kontext
1M
Eingabe
text · image · audio · video
Ausgabe
text
Eingabe: $0.15Ausgabe: $0.47pro 1 Mio. Token
Modelldetails ansehen
qwen logoqwen

Qwen: Qwen3.8 Max (0902)

Qwen3.8 Max 0902 is an updated snapshot of Qwen3.8 Max from Alibaba's Qwen team. It is a 2.4-trillion-parameter mixture-of-experts model that accepts text, image, and video input and returns text,...

Kontext
1M
Eingabe
text · image · video
Ausgabe
text
Eingabe: $2Ausgabe: $6pro 1 Mio. Token
Modelldetails ansehen
qwen logoqwen

Qwen: Qwen3.8 Flash

Qwen3.8 Flash is a multimodal reasoning model from Alibaba. It is suited for coding assistance, agentic workflows, visual understanding, document and codebase analysis, desktop interaction, chart analysis, and long-video analysis.

Kontext
1M
Eingabe
text · image · video
Ausgabe
text
Eingabe: $0.15Ausgabe: $0.47pro 1 Mio. Token
Modelldetails ansehen
qwen logoqwen

Qwen: Qwen3.8 27B

Qwen3.8 27B is an open-weight dense vision-language model from Qwen. It is suited for coding, professional workflows, research, multimodal interaction, and long-running agent tasks, with flexible thinking that can be...

Kontext
1M
Eingabe
text · image · video
Ausgabe
text
Eingabe: $0.42Ausgabe: $3pro 1 Mio. Token
Modelldetails ansehen