qwen logo
qwen

Qwen: Qwen3.5-35B-A3B

The Qwen3.5 Series 35B-A3B is a native vision-language model designed with a hybrid architecture that integrates linear attention mechanisms and a sparse mixture-of-experts model, achieving higher inference efficiency. $0.08 per million input tokens, $0.75 per million output tokens. 262,144 token context window, maximum output of 16,384 tokens. Higher uptime with 8 providers. Includes independent benchmarks from Artificial Analysis.

Überblick

Modellspezifikationen

Kontext
262.144 tokens
Maximale Ausgabe
235.929 tokens
Architektur
text+image+video->text
Tokenizer
Qwen3
Wissensstand
Nicht angegeben
Moderiert
Nein
OPENROUTER

Vollständige Preise

Mit OpenRouter synchronisierte Preise; Tokenpreise gelten pro eine Million Token.

Eingabe
$0.08
/M tokens
Ausgabe
$0.75
/M tokens
API

Modellkonfiguration

Standardparameter

temperature
1
top_p
0.95
top_k
20

Schlussfolgern

Moderiert
Nein
Standardparameter
Nein
API

Schnellstart

Dieses Modell über die OpenAI-kompatible API von OpenRouter aufrufen.

curl https://openrouter.ai/api/v1/chat/completions \
  -H "Authorization: Bearer $OPENROUTER_API_KEY" \
  -H "Content-Type: application/json" \
  -d '{"model":"qwen/qwen3.5-35b-a3b","messages":[{"role":"user","content":"Hello!"}]}'
Fähigkeiten und Modalitäten

EingabeAusgabe

Eingabe
textimagevideo
Ausgabe
text
Schlussfolgern

Nein

API

Unterstützte API-Parameter

frequency_penaltyinclude_reasoninglogit_biaslogprobsmax_tokensmin_ppresence_penaltyreasoningrepetition_penaltyresponse_formatseedstopstructured_outputstemperaturetool_choicetoolstop_ktop_logprobstop_p
Verfügbare Anbieter

8 Anbieter

Live-Anbieter bei OpenRouter

Verfügbarkeit, Latenz, Durchsatz und Routing ändern sich laufend. Aktuelle Betriebsdaten stehen auf der Quellseite.

OpenRouter

Darkbloom

fp4
Kontext
262K
Maximale Ausgabe
16K
Eingabe
$0.080
Ausgabe
$0.750
Cache-Lesen
Nicht angegeben
Cache-Schreiben
Nicht angegeben

DeepInfra

fp8
Kontext
262K
Maximale Ausgabe
82K
Eingabe
$0.140
Ausgabe
$1.00
Cache-Lesen
$0.050
Cache-Schreiben
Nicht angegeben

Parasail

fp8
Kontext
262K
Maximale Ausgabe
236K
Eingabe
$0.150
Ausgabe
$1.00
Cache-Lesen
$0.050
Cache-Schreiben
Nicht angegeben

Alibaba

unknown
Kontext
262K
Maximale Ausgabe
66K
Eingabe
$0.163
Ausgabe
$1.30
Cache-Lesen
Nicht angegeben
Cache-Schreiben
Nicht angegeben

AtlasCloud

fp8
Kontext
262K
Maximale Ausgabe
66K
Eingabe
$0.225
Ausgabe
$1.80
Cache-Lesen
$0.225
Cache-Schreiben
Nicht angegeben

SiliconFlow

fp8
Kontext
262K
Maximale Ausgabe
236K
Eingabe
$0.240
Ausgabe
$1.80
Cache-Lesen
Nicht angegeben
Cache-Schreiben
Nicht angegeben

CoreWeave

fp8
Kontext
262K
Maximale Ausgabe
236K
Eingabe
$0.250
Ausgabe
$1.25
Cache-Lesen
$0.250
Cache-Schreiben
Nicht angegeben

Venice

unknown
Kontext
256K
Maximale Ausgabe
16K
Eingabe
$0.313
Ausgabe
$1.25
Cache-Lesen
$0.156
Cache-Schreiben
Nicht angegeben
qwen

Modelle

Alle Modelle
qwen logoqwen

Qwen: Qwen3.8 Flash

Qwen3.8 Flash is a multimodal reasoning model from Alibaba. It is suited for coding assistance, agentic workflows, visual understanding, document and codebase analysis, desktop interaction, chart analysis, and long-video analysis.

Kontext
1M
Eingabe
text · image · video
Ausgabe
text
Eingabe: $0.150Ausgabe: $0.470pro 1 Mio. Token
Modelldetails ansehen
qwen logoqwen

Qwen: Qwen3.8 27B

Qwen3.8 27B is an open-weight dense vision-language model from Qwen. It is suited for coding, professional workflows, research, multimodal interaction, and long-running agent tasks, with flexible thinking that can be...

Kontext
1M
Eingabe
text · image · video
Ausgabe
text
Eingabe: $0.425Ausgabe: $2.55pro 1 Mio. Token
Modelldetails ansehen
qwen logoqwen

Qwen3 Reranker 8B

Qwen3 Reranker 8B is a text reranking model from Alibaba Cloud built on the Qwen3 architecture. It evaluates query-document pairs to produce relevance scores for use in retrieval and RAG...

Kontext
41K
Eingabe
text
Ausgabe
rerank
Eingabe: KostenlosAusgabe: Kostenlospro 1 Mio. Token
Modelldetails ansehen
qwen logoqwen

Qwen: Qwen3 ASR 1.7B

Qwen3 ASR 1.7B is an automatic speech recognition model from Qwen. It supports multilingual language identification and transcription across 30 languages and 22 Chinese dialects, with streaming and offline inference...

Kontext
Nicht angegeben
Eingabe
audio
Ausgabe
transcription
Eingabe: $7.50Ausgabe: Kostenlospro 1 Mio. Token
Modelldetails ansehen