google logo
google

Google: Gemini 3.8 Flash TTS

Description de la source (anglais)

Gemini 3.8 Flash TTS is a text-to-speech model from Google and the successor to Gemini 3.1 Flash TTS Preview. It is the creative tier of the 3.8 TTS family, suited...

Présentation

Caractéristiques du modèle

Contexte
8 192 tokens
Sortie maximale
7 372 tokens
Architecture
text->speech
Tokenizer
Gemini
Limite des connaissances
Non indiqué
Modéré
Non
OPENROUTER

Tarification complète

Tarifs synchronisés depuis OpenRouter, par million de tokens.

Entrée
$0.5
par million de tokens de texte
Sortie
$9
par million de tokens audio
API

Configuration du modèle

Voix prises en charge

ZephyrPuckCharonKoreFenrirLedaOrusAoedeCallirrhoeAutonoeEnceladusIapetusUmbrielAlgiebaDespinaErinomeAlgenibRasalgethiLaomedeiaAchernarAlnilamSchedarGacruxPulcherrimaAchirdZubenelgenubiVindemiatrixSadachbiaSadaltagerSulafat
API

Démarrage rapide

Définissez OPENROUTER_API_KEY localement. Python nécessite requests ; JavaScript s’exécute dans Node.js. Gardez la clé côté serveur.

Choisissez une voix compatible. Remplacez YOUR_VOICE_ID si aucune voix n’est indiquée.

Documentation de l’API

curl --fail-with-body https://openrouter.ai/api/v1/audio/speech \
  -H "Authorization: Bearer $OPENROUTER_API_KEY" \
  -H "Content-Type: application/json" \
  -d '{"model":"google/gemini-3.8-flash-tts","input":"Hello!","voice":"Zephyr","response_format":"mp3"}' \
  --output speech.mp3
Capacités et modalités

Entrée → Sortie

Entrée
text
Sortie
speech
API

Paramètres API pris en charge

Fournisseurs disponibles

1 Fournisseur

Vérifié: 24 septembre 2026

Fournisseurs en direct sur OpenRouter

Disponibilité, latence, débit et routage évoluent en continu. Consultez la source pour les données en temps réel.

OpenRouter

Google AI Studio

Non indiqué
Contexte
8K
Sortie maximale
7K
google

4 modèles

Tous les modèles
google logogoogle

Google: Gemini 3.5 Transcribe

Gemini 3.5 Transcribe is a speech-to-text model from Google. It is suited for synchronous transcription that needs word-level timestamps or speaker diarization, with support for up to eight speakers. Audio...

Contexte
98K
Entrée
audio
Sortie
transcription
Entrée: $2 par million de tokensSortie: $12 par million de tokens
Voir la fiche du modèle →
google logogoogle

Google: Gemini 3.8 Flash Lite TTS

Gemini 3.8 Flash Lite TTS is a text-to-speech model from Google and the fast, high-throughput member of the 3.8 TTS family alongside Gemini 3.8 Flash TTS. It is suited for...

Contexte
8K
Entrée
text
Sortie
speech
Entrée: $0.5 par million de tokens de texteSortie: $6 par million de tokens audio
Voir la fiche du modèle →
google logogoogle

Google: Gemini 3.8 Flash

Gemini 3.8 Flash is Google's most intelligent Flash model with significant gains from 3.7 Flash across software engineering, agentic tasks, and multi-step reasoning.

Contexte
1.0M
Entrée
text · image · video · file · audio
Sortie
text
Entrée: $0.75Sortie: $3.75par million de tokens
Voir la fiche du modèle →
google logogoogle

Google: Gemini 3.8 Flash (batch)

Gemini 3.8 Flash is Google's most intelligent Flash model with significant gains from 3.7 Flash across software engineering, agentic tasks, and multi-step reasoning.

Contexte
1.0M
Entrée
text · image · video · file · audio
Sortie
text
Entrée: $0.375Sortie: $1.875par million de tokens
Voir la fiche du modèle →