qwen logo
qwen

Qwen: Qwen3 235B A22B Instruct 2507

Qwen3-235B-A22B-Instruct-2507 is a multilingual, instruction-tuned mixture-of-experts language model based on the Qwen3-235B architecture, with 22B active parameters per forward pass. $0.0875 per million input tokens, $0.35 per million output tokens. 262,144 token context window. Higher uptime with 10 providers. Includes independent benchmarks from Artificial Analysis.

模型概览

模型规格

上下文
262,144 tokens
最大输出
235,929 tokens
架构
text->text
分词器
Qwen3
知识截止时间
2025-06-30
内容审核
OPENROUTER

完整价格

同步自 OpenRouter。Token 价格均按每百万 Token 展示。

输入
$0.0875
/M tokens
输出
$0.35
/M tokens
缓存读取
$0.0175
/M tokens
API

快速调用

通过 OpenRouter 的 OpenAI 兼容 API 调用这个确切的模型 ID。

curl https://openrouter.ai/api/v1/chat/completions \
  -H "Authorization: Bearer $OPENROUTER_API_KEY" \
  -H "Content-Type: application/json" \
  -d '{"model":"qwen/qwen3-235b-a22b-2507","messages":[{"role":"user","content":"Hello!"}]}'
能力与模态

输入输出

输入
text
输出
text
API

支持的 API 参数

frequency_penaltylogit_biaslogprobsmax_tokensmin_ppresence_penaltyrepetition_penaltyresponse_formatseedstopstructured_outputstemperaturetool_choicetoolstop_ktop_logprobstop_p
可用提供方

11 个提供方

在 OpenRouter 查看实时 Provider

Provider 可用性、延迟、吞吐量和路由会持续变化,请前往来源页面查看实时运行数据。

OpenRouter

GMICloud

fp8
上下文
262K
最大输出
236K
输入
$0.087
输出
$0.350
缓存读取
$0.018
缓存写入
暂未提供

DeepInfra

fp8
上下文
262K
最大输出
16K
输入
$0.090
输出
$0.550
缓存读取
暂未提供
缓存写入
暂未提供

Novita

fp8
上下文
131K
最大输出
16K
输入
$0.090
输出
$0.580
缓存读取
暂未提供
缓存写入
暂未提供

Parasail

fp8
上下文
131K
最大输出
118K
输入
$0.140
输出
$0.800
缓存读取
$0.050
缓存写入
暂未提供

Alibaba

unknown
上下文
131K
最大输出
33K
输入
$0.150
输出
$0.598
缓存读取
暂未提供
缓存写入
暂未提供

Venice

fp8
上下文
128K
最大输出
16K
输入
$0.150
输出
$0.750
缓存读取
暂未提供
缓存写入
暂未提供

Nebius

fp8
上下文
262K
最大输出
236K
输入
$0.200
输出
$0.600
缓存读取
暂未提供
缓存写入
暂未提供

AtlasCloud

fp8
上下文
131K
最大输出
118K
输入
$0.200
输出
$0.880
缓存读取
$0.200
缓存写入
暂未提供

StreamLake

unknown
上下文
128K
最大输出
32K
输入
$0.210
输出
$0.840
缓存读取
暂未提供
缓存写入
暂未提供

Google

unknown
上下文
262K
最大输出
16K
输入
$0.220
输出
$0.880
缓存读取
暂未提供
缓存写入
暂未提供

Google

unknown
上下文
262K
最大输出
16K
输入
$0.250
输出
$1.00
缓存读取
暂未提供
缓存写入
暂未提供
qwen

个模型

全部模型
qwen logoqwen

Qwen: Qwen3.8 Flash

Qwen3.8 Flash is a multimodal reasoning model from Alibaba. It is suited for coding assistance, agentic workflows, visual understanding, document and codebase analysis, desktop interaction, chart analysis, and long-video analysis.

上下文
1M
输入
text · image · video
输出
text
输入: $0.150输出: $0.470每百万 Token
查看模型详情
qwen logoqwen

Qwen: Qwen3.8 27B

Qwen3.8 27B is an open-weight dense vision-language model from Qwen. It is suited for coding, professional workflows, research, multimodal interaction, and long-running agent tasks, with flexible thinking that can be...

上下文
1M
输入
text · image · video
输出
text
输入: $0.425输出: $2.55每百万 Token
查看模型详情
qwen logoqwen

Qwen3 Reranker 8B

Qwen3 Reranker 8B is a text reranking model from Alibaba Cloud built on the Qwen3 architecture. It evaluates query-document pairs to produce relevance scores for use in retrieval and RAG...

上下文
41K
输入
text
输出
rerank
输入: 免费输出: 免费每百万 Token
查看模型详情
qwen logoqwen

Qwen: Qwen3 ASR 1.7B

Qwen3 ASR 1.7B is an automatic speech recognition model from Qwen. It supports multilingual language identification and transcription across 30 languages and 22 Chinese dialects, with streaming and offline inference...

上下文
暂未提供
输入
audio
输出
transcription
输入: $7.50输出: 免费每百万 Token
查看模型详情