qwen logo
qwen

Qwen: Qwen3 Next 80B A3B Instruct

Descrizione della fonte (inglese)

Qwen3-Next-80B-A3B-Instruct is an instruction-tuned chat model in the Qwen3-Next series optimized for fast, stable responses without “thinking” traces. It targets complex tasks across reasoning, code generation, knowledge QA, and multilingual...

Panoramica

Specifiche del modello

Contesto
262.144 tokens
Output massimo
16.384 tokens
Architettura
text->text
Tokenizer
Qwen3
Limite di conoscenza
2025-09-30
Moderato
No
OPENROUTER

Prezzi completi

Tariffe sincronizzate da OpenRouter, per milione di token.

Input
$0.09
per milione di token
Output
$1.1
per milione di token
API

Avvio rapido

Imposta OPENROUTER_API_KEY localmente. Python richiede requests; JavaScript viene eseguito in Node.js. Conserva la chiave sul server.

Documentazione API

curl --fail-with-body https://openrouter.ai/api/v1/chat/completions \
  -H "Authorization: Bearer $OPENROUTER_API_KEY" \
  -H "Content-Type: application/json" \
  -d '{"model":"qwen/qwen3-next-80b-a3b-instruct","messages":[{"role":"user","content":"Hello!"}]}'
Capacità e modalità

InputOutput

Input
text
Output
text
API

Parametri API supportati

frequency_penaltylogit_biaslogprobsmax_tokensmin_ppresence_penaltyrepetition_penaltyresponse_formatseedstopstructured_outputstemperaturetool_choicetoolstop_ktop_logprobstop_p
Provider disponibili

5 provider

Verificato: 9 settembre 2026

Provider live su OpenRouter

Disponibilità, latenza, throughput e routing cambiano continuamente. Consulta la fonte per i dati correnti.

OpenRouter

DeepInfra

fp8
Contesto
262K
Output massimo
16K
Input
$0.09
Output
$1.1
Lettura cache
Non indicato
Scrittura cache
Non indicato

Alibaba

Non indicato
Contesto
131K
Output massimo
33K
Input
$0.0975
Output
$0.78
Lettura cache
Non indicato
Scrittura cache
Non indicato

Parasail

fp8
Contesto
262K
Output massimo
236K
Input
$0.1
Output
$1.1
Lettura cache
$0.07
Scrittura cache
Non indicato

Google

Non indicato
Contesto
262K
Output massimo
236K
Input
$0.15
Output
$1.2
Lettura cache
Non indicato
Scrittura cache
Non indicato

Novita

bf16
Contesto
131K
Output massimo
33K
Input
$0.15
Output
$1.5
Lettura cache
Non indicato
Scrittura cache
Non indicato
qwen

4 modelli

Tutti i modelli
qwen logoqwen

Qwen: Qwen3.8 Omni Flash

Qwen3.8 Omni Flash is an omni-modal reasoning model from Alibaba, the first Qwen model built around agentic capabilities with native audio-video understanding. It is suited for audio-video analysis and summarization,...

Contesto
1M
Input
text · image · audio · video
Output
text
Input: $0.15Output: $0.47per milione di token
Vedi i dettagli del modello
qwen logoqwen

Qwen: Qwen3.8 Max (0902)

Qwen3.8 Max 0902 is an updated snapshot of Qwen3.8 Max from Alibaba's Qwen team. It is a 2.4-trillion-parameter mixture-of-experts model that accepts text, image, and video input and returns text,...

Contesto
1M
Input
text · image · video
Output
text
Input: $2Output: $6per milione di token
Vedi i dettagli del modello
qwen logoqwen

Qwen: Qwen3.8 Flash

Qwen3.8 Flash is a multimodal reasoning model from Alibaba. It is suited for coding assistance, agentic workflows, visual understanding, document and codebase analysis, desktop interaction, chart analysis, and long-video analysis.

Contesto
1M
Input
text · image · video
Output
text
Input: $0.15Output: $0.47per milione di token
Vedi i dettagli del modello
qwen logoqwen

Qwen: Qwen3.8 27B

Qwen3.8 27B is an open-weight dense vision-language model from Qwen. It is suited for coding, professional workflows, research, multimodal interaction, and long-running agent tasks, with flexible thinking that can be...

Contesto
1M
Input
text · image · video
Output
text
Input: $0.42Output: $3per milione di token
Vedi i dettagli del modello