qwen logo
qwen

Qwen: Qwen3.5-35B-A3B

The Qwen3.5 Series 35B-A3B is a native vision-language model designed with a hybrid architecture that integrates linear attention mechanisms and a sparse mixture-of-experts model, achieving higher inference efficiency. $0.08 per million input tokens, $0.75 per million output tokens. 262,144 token context window, maximum output of 16,384 tokens. Higher uptime with 8 providers. Includes independent benchmarks from Artificial Analysis.

概要

モデル仕様

コンテキスト
262,144 tokens
最大出力
235,929 tokens
アーキテクチャ
text+image+video->text
トークナイザー
Qwen3
知識カットオフ
情報なし
モデレーション
いいえ
OPENROUTER

料金の詳細

OpenRouterと同期した料金です。トークン料金は100万トークン単位です。

入力
$0.08
/M tokens
出力
$0.75
/M tokens
API

モデル設定

既定パラメータ

temperature
1
top_p
0.95
top_k
20

推論

モデレーション
いいえ
既定パラメータ
いいえ
API

クイックスタート

OpenRouterのOpenAI互換APIでこのモデルIDを呼び出します。

curl https://openrouter.ai/api/v1/chat/completions \
  -H "Authorization: Bearer $OPENROUTER_API_KEY" \
  -H "Content-Type: application/json" \
  -d '{"model":"qwen/qwen3.5-35b-a3b","messages":[{"role":"user","content":"Hello!"}]}'
機能とモダリティ

入力出力

入力
textimagevideo
出力
text
推論

いいえ

API

対応APIパラメータ

frequency_penaltyinclude_reasoninglogit_biaslogprobsmax_tokensmin_ppresence_penaltyreasoningrepetition_penaltyresponse_formatseedstopstructured_outputstemperaturetool_choicetoolstop_ktop_logprobstop_p
利用可能なプロバイダー

8 プロバイダー

OpenRouterでプロバイダーを見る

可用性、遅延、スループット、ルーティングは常時変化します。最新情報は出典ページで確認してください。

OpenRouter

Darkbloom

fp4
コンテキスト
262K
最大出力
16K
入力
$0.080
出力
$0.750
キャッシュ読込
情報なし
キャッシュ書込
情報なし

DeepInfra

fp8
コンテキスト
262K
最大出力
82K
入力
$0.140
出力
$1.00
キャッシュ読込
$0.050
キャッシュ書込
情報なし

Parasail

fp8
コンテキスト
262K
最大出力
236K
入力
$0.150
出力
$1.00
キャッシュ読込
$0.050
キャッシュ書込
情報なし

Alibaba

unknown
コンテキスト
262K
最大出力
66K
入力
$0.163
出力
$1.30
キャッシュ読込
情報なし
キャッシュ書込
情報なし

AtlasCloud

fp8
コンテキスト
262K
最大出力
66K
入力
$0.225
出力
$1.80
キャッシュ読込
$0.225
キャッシュ書込
情報なし

SiliconFlow

fp8
コンテキスト
262K
最大出力
236K
入力
$0.240
出力
$1.80
キャッシュ読込
情報なし
キャッシュ書込
情報なし

CoreWeave

fp8
コンテキスト
262K
最大出力
236K
入力
$0.250
出力
$1.25
キャッシュ読込
$0.250
キャッシュ書込
情報なし

Venice

unknown
コンテキスト
256K
最大出力
16K
入力
$0.313
出力
$1.25
キャッシュ読込
$0.156
キャッシュ書込
情報なし
qwen logoqwen

Qwen: Qwen3.8 Flash

Qwen3.8 Flash is a multimodal reasoning model from Alibaba. It is suited for coding assistance, agentic workflows, visual understanding, document and codebase analysis, desktop interaction, chart analysis, and long-video analysis.

コンテキスト
1M
入力
text · image · video
出力
text
入力: $0.150出力: $0.470100万トークンあたり
モデル詳細を見る
qwen logoqwen

Qwen: Qwen3.8 27B

Qwen3.8 27B is an open-weight dense vision-language model from Qwen. It is suited for coding, professional workflows, research, multimodal interaction, and long-running agent tasks, with flexible thinking that can be...

コンテキスト
1M
入力
text · image · video
出力
text
入力: $0.425出力: $2.55100万トークンあたり
モデル詳細を見る
qwen logoqwen

Qwen3 Reranker 8B

Qwen3 Reranker 8B is a text reranking model from Alibaba Cloud built on the Qwen3 architecture. It evaluates query-document pairs to produce relevance scores for use in retrieval and RAG...

コンテキスト
41K
入力
text
出力
rerank
入力: 無料出力: 無料100万トークンあたり
モデル詳細を見る
qwen logoqwen

Qwen: Qwen3 ASR 1.7B

Qwen3 ASR 1.7B is an automatic speech recognition model from Qwen. It supports multilingual language identification and transcription across 30 languages and 22 Chinese dialects, with streaming and offline inference...

コンテキスト
情報なし
入力
audio
出力
transcription
入力: $7.50出力: 無料100万トークンあたり
モデル詳細を見る