google logo
google

Google: Gemini 3.1 Flash TTS Preview

Source description (English)

Gemini 3.1 Flash TTS Preview is a text-to-speech model from Google, and a substantial generational step up from Gemini 2.5 Flash TTS. It takes text input and produces audio output...

Overview

Model specifications

Context
32,768 tokens
Maximum output
16,384 tokens
Architecture
text->speech
Tokenizer
Gemini
Knowledge cutoff
Not provided
Moderated
No
OPENROUTER

Complete pricing

Synchronized OpenRouter rates. Token prices are shown per one million tokens.

Input
$1
per million text tokens
Output
$20
per million audio tokens
API

Model configuration

Supported voices

ZephyrPuckCharonKoreFenrirLedaOrusAoedeCallirrhoeAutonoeEnceladusIapetusUmbrielAlgiebaDespinaErinomeAlgenibRasalgethiLaomedeiaAchernarAlnilamSchedarGacruxPulcherrimaAchirdZubenelgenubiVindemiatrixSadachbiaSadaltagerSulafat
API

Quick start

Set OPENROUTER_API_KEY locally. Python requires requests; JavaScript runs in Node.js. Keep the key on the server.

Choose a supported voice. Replace YOUR_VOICE_ID if no voice is listed.

API documentation

curl --fail-with-body https://openrouter.ai/api/v1/audio/speech \
  -H "Authorization: Bearer $OPENROUTER_API_KEY" \
  -H "Content-Type: application/json" \
  -d '{"model":"google/gemini-3.1-flash-tts-preview","input":"Hello!","voice":"Zephyr","response_format":"mp3"}' \
  --output speech.mp3
Capabilities

InputOutput

Input
text
Output
speech
API

Supported API parameters

Available providers

1 Provider

Checked: September 16, 2026

Live providers on OpenRouter

Provider availability, latency, throughput and routing can change continuously. Open the source page for current operational data.

OpenRouter

Google

Not provided
Context
33K
Maximum output
16K
google

4 models

All models