meta-llama logo
meta-llama

Meta: Llama 4 Scout

Llama 4 Scout 17B Instruct (16E) is a mixture-of-experts (MoE) language model developed by Meta, activating 17 billion parameters out of a total of 109B. $0.10 per million input tokens, $0.30 per million output tokens. 1,310,720 token context window, maximum output of 16,384 tokens. Higher uptime with 4 providers. Includes independent benchmarks from Artificial Analysis.

Présentation

Caractéristiques du modèle

Contexte
1 310 720 tokens
Sortie maximale
8 192 tokens
Architecture
text+image->text
Tokenizer
Llama4
Limite des connaissances
2024-08-31
Modéré
Non
OPENROUTER

Tarification complète

Tarifs synchronisés depuis OpenRouter, par million de tokens.

Entrée
$0.1
/M tokens
Sortie
$0.3
/M tokens
API

Démarrage rapide

Appelez ce modèle via l’API compatible OpenAI d’OpenRouter.

curl https://openrouter.ai/api/v1/chat/completions \
  -H "Authorization: Bearer $OPENROUTER_API_KEY" \
  -H "Content-Type: application/json" \
  -d '{"model":"meta-llama/llama-4-scout","messages":[{"role":"user","content":"Hello!"}]}'
Capacités et modalités

EntréeSortie

Entrée
textimage
Sortie
text
API

Paramètres API pris en charge

frequency_penaltylogit_biasmax_tokensmin_ppresence_penaltyrepetition_penaltyresponse_formatseedstopstructured_outputstemperaturetool_choicetoolstop_ktop_p
Fournisseurs disponibles

4 fournisseurs

Fournisseurs en direct sur OpenRouter

Disponibilité, latence, débit et routage évoluent en continu. Consultez la source pour les données en temps réel.

OpenRouter

DeepInfra

fp8
Contexte
328K
Sortie maximale
16K
Entrée
$0.100
Sortie
$0.300
Lecture du cache
Non indiqué
Écriture du cache
Non indiqué

Groq

unknown
Contexte
131K
Sortie maximale
8K
Entrée
$0.110
Sortie
$0.340
Lecture du cache
$0.055
Écriture du cache
Non indiqué

Novita

bf16
Contexte
131K
Sortie maximale
118K
Entrée
$0.180
Sortie
$0.590
Lecture du cache
Non indiqué
Écriture du cache
Non indiqué

Google

unknown
Contexte
1.3M
Sortie maximale
8K
Entrée
$0.250
Sortie
$0.700
Lecture du cache
Non indiqué
Écriture du cache
Non indiqué
meta-llama

modèles

Tous les modèles