google logo
google

Google: Chirp 3

来源介绍(英文)

Chirp 3 is Google's latest multilingual speech-to-text model. It offers enhanced transcription accuracy across 24 GA languages and 77+ preview languages, with support for automatic language detection, automatic punctuation, and...

模型概览

模型规格

上下文
暂未提供
最大输出
暂未提供
架构
audio->transcription
分词器
Other
知识截止时间
暂未提供
内容审核
OPENROUTER

完整价格

同步自 OpenRouter。Token 价格均按每百万 Token 展示。

音频时长
$0.000267
每秒
API

快速调用

在本地设置 OPENROUTER_API_KEY。Python 需安装 requests;JavaScript 在 Node.js 中运行。密钥应保存在服务端。

API 文档

curl --fail-with-body https://openrouter.ai/api/v1/audio/transcriptions \
  -H "Authorization: Bearer $OPENROUTER_API_KEY" \
  -F 'model=google/chirp-3' \
  -F 'file=@audio.wav'
能力与模态

输入输出

输入
audio
输出
transcription
API

支持的 API 参数

可用提供方

1 提供方

核验日期: 2026年9月19日

在 OpenRouter 查看实时 Provider

Provider 可用性、延迟、吞吐量和路由会持续变化,请前往来源页面查看实时运行数据。

OpenRouter

Google

暂未提供
上下文
暂未提供
最大输出
暂未提供
google

4 个模型

全部模型
google logogoogle

Google: Gemini 3.8 Flash

Gemini 3.8 Flash is Google's most intelligent Flash model with significant gains from 3.7 Flash across software engineering, agentic tasks, and multi-step reasoning.

上下文
1.0M
输入
text · image · video · file · audio
输出
text
输入: $0.75输出: $3.75每百万 Token
查看模型详情
google logogoogle

Google: Gemini 3.8 Flash (batch)

Gemini 3.8 Flash is Google's most intelligent Flash model with significant gains from 3.7 Flash across software engineering, agentic tasks, and multi-step reasoning.

上下文
1.0M
输入
text · image · video · file · audio
输出
text
输入: $0.375输出: $1.875每百万 Token
查看模型详情
google logogoogle

Google: Gemini 3.7 Flash

Gemini 3.7 Flash is a multimodal model from Google for fast agentic workflows, coding, and complex multi-step reasoning. It is designed for tasks that require responsive performance and reliable multi-step...

上下文
1.0M
输入
text · image · video · file · audio
输出
text
输入: $0.75输出: $3.75每百万 Token
查看模型详情
google logogoogle

Google: Gemini 3.7 Flash (batch)

Gemini 3.7 Flash is a multimodal model from Google for fast agentic workflows, coding, and complex multi-step reasoning. It is designed for tasks that require responsive performance and reliable multi-step...

上下文
1.0M
输入
text · image · video · file · audio
输出
text
输入: $0.375输出: $1.875每百万 Token
查看模型详情