microsoft logo
microsoft

Microsoft: MAI-Image-2.5

Microsoft's MAI-Image-2.5 is a high-quality image generation model available via Azure AI Foundry. Priced from $5 per M tokens. 4,096 token context window, maximum output of 1,024 tokens. Includes independent benchmarks from Artificial Analysis.

Présentation

Caractéristiques du modèle

Contexte
4 096 tokens
Sortie maximale
1 024 tokens
Architecture
text+image->image
Tokenizer
Other
Limite des connaissances
Non indiqué
Modéré
Non
OPENROUTER

Tarification complète

Tarifs synchronisés depuis OpenRouter, par million de tokens.

Entrée
$5
/M tokens
Image Input
$8
/M tokens
Image
$47
/M tokens
API

Démarrage rapide

Appelez ce modèle via l’API compatible OpenAI d’OpenRouter.

curl https://openrouter.ai/api/v1/chat/completions \
  -H "Authorization: Bearer $OPENROUTER_API_KEY" \
  -H "Content-Type: application/json" \
  -d '{"model":"microsoft/mai-image-2.5","messages":[{"role":"user","content":"Hello!"}]}'
Capacités et modalités

EntréeSortie

Entrée
textimage
Sortie
image
API

Paramètres API pris en charge

max_completion_tokensmax_tokenstemperature
Fournisseurs disponibles

1 Fournisseur

Fournisseurs en direct sur OpenRouter

Disponibilité, latence, débit et routage évoluent en continu. Consultez la source pour les données en temps réel.

OpenRouter

Azure

unknown
Contexte
4K
Sortie maximale
1K
Entrée
$5.00
Sortie
Gratuit
Lecture du cache
Non indiqué
Écriture du cache
Non indiqué
microsoft

modèles

Tous les modèles
microsoft logomicrosoft

Microsoft: MAI-Image-2.5 Pro

Microsoft's MAI-Image-2.5 is a high-quality image generation model available via Azure AI Foundry. It produces photorealistic and artistic images from text prompts with support for various aspect ratios.

Contexte
4K
Entrée
text · image
Sortie
image
Entrée: $5.00Sortie: Gratuitpar million de tokens
Voir la fiche du modèle
microsoft logomicrosoft

Microsoft: MAI-Voice-2-Flash

MAI-Voice-2-Flash is a low-latency text-to-speech model from Microsoft for voice agents, assistants, call centers, accessibility, narration, and other interactive applications. It generates expressive 24 kHz mono speech across 15 languages...

Contexte
Non indiqué
Entrée
text
Sortie
speech
Entrée: $15.00Sortie: Gratuitpar million de tokens
Voir la fiche du modèle
microsoft logomicrosoft

Microsoft: MAI-Voice-2

MAI-Voice-2 is an expressive text-to-speech model from Microsoft. It is suited for conversational assistants, media narration, accessibility, education, and other long-form voice applications. It supports 15 languages across 18 locales,...

Contexte
Non indiqué
Entrée
text
Sortie
speech
Entrée: $22.00Sortie: Gratuitpar million de tokens
Voir la fiche du modèle
microsoft logomicrosoft

Microsoft: MAI-Transcribe 1.5

MAI-Transcribe 1.5 is a multilingual speech-to-text model from Microsoft AI. It is suited for captions, call transcription, subtitling, accessibility, and other voice-enabled applications, with reliable transcription across 43 languages, diverse...

Contexte
Non indiqué
Entrée
audio
Sortie
transcription
Entrée: $360000.00Sortie: Gratuitpar million de tokens
Voir la fiche du modèle