meta-llama logo
meta-llama

Meta: Llama 3.1 8B Instruct

Meta's latest class of model (Llama 3.1) launched with a variety of sizes & flavors. $0.02 per million input tokens, $0.04 per million output tokens. 131,072 token context window, maximum output of 16,384 tokens. Higher uptime with 5 providers. Includes independent benchmarks from Artificial Analysis.

概要

モデル仕様

コンテキスト
131,072 tokens
最大出力
117,964 tokens
アーキテクチャ
text->text
トークナイザー
Llama3
知識カットオフ
2023-12-31
モデレーション
いいえ
OPENROUTER

料金の詳細

OpenRouterと同期した料金です。トークン料金は100万トークン単位です。

入力
$0.02
/M tokens
出力
$0.04
/M tokens
API

クイックスタート

OpenRouterのOpenAI互換APIでこのモデルIDを呼び出します。

curl https://openrouter.ai/api/v1/chat/completions \
  -H "Authorization: Bearer $OPENROUTER_API_KEY" \
  -H "Content-Type: application/json" \
  -d '{"model":"meta-llama/llama-3.1-8b-instruct","messages":[{"role":"user","content":"Hello!"}]}'
機能とモダリティ

入力出力

入力
text
出力
text
API

対応APIパラメータ

frequency_penaltylogit_biaslogprobsmax_tokensmin_ppresence_penaltyrepetition_penaltyresponse_formatseedstopstructured_outputstemperaturetool_choicetoolstop_ktop_logprobstop_p
利用可能なプロバイダー

5 プロバイダー

OpenRouterでプロバイダーを見る

可用性、遅延、スループット、ルーティングは常時変化します。最新情報は出典ページで確認してください。

OpenRouter

DeepInfra

fp8
コンテキスト
131K
最大出力
16K
入力
$0.020
出力
$0.040
キャッシュ読込
情報なし
キャッシュ書込
情報なし

Novita

fp8
コンテキスト
16K
最大出力
15K
入力
$0.020
出力
$0.050
キャッシュ読込
情報なし
キャッシュ書込
情報なし

Groq

unknown
コンテキスト
131K
最大出力
118K
入力
$0.050
出力
$0.080
キャッシュ読込
$0.025
キャッシュ書込
情報なし

Cloudflare

fp8
コンテキスト
32K
最大出力
29K
入力
$0.152
出力
$0.287
キャッシュ読込
情報なし
キャッシュ書込
情報なし

CoreWeave

bf16
コンテキスト
131K
最大出力
118K
入力
$0.220
出力
$0.220
キャッシュ読込
$0.220
キャッシュ書込
情報なし
meta-llama

モデル

すべてのモデル
meta-llama logometa-llama

Meta: Llama Guard 4 12B

Llama Guard 4 is a Llama 4 Scout-derived multimodal pretrained model, fine-tuned for content safety classification. Similar to previous versions, it can be used to classify content in both LLM...

コンテキスト
164K
入力
image · text
出力
text
入力: $0.180出力: $0.180100万トークンあたり
モデル詳細を見る
meta-llama logometa-llama

Meta: Llama 4 Maverick

Llama 4 Maverick 17B Instruct (128E) is a high-capacity multimodal language model from Meta, built on a mixture-of-experts (MoE) architecture with 128 experts and 17 billion active parameters per forward...

コンテキスト
1.0M
入力
text · image
出力
text
入力: $0.200出力: $0.696100万トークンあたり
モデル詳細を見る
meta-llama logometa-llama

Meta: Llama 4 Scout

Llama 4 Scout 17B Instruct (16E) is a mixture-of-experts (MoE) language model developed by Meta, activating 17 billion parameters out of a total of 109B. It supports native multimodal input...

コンテキスト
1.3M
入力
text · image
出力
text
入力: $0.110出力: $0.340100万トークンあたり
モデル詳細を見る
meta-llama logometa-llama

Meta: Llama 3.3 70B Instruct

The Meta Llama 3.3 multilingual large language model (LLM) is a pretrained and instruction tuned generative model in 70B (text in/text out). The Llama 3.3 instruction tuned text only model...

コンテキスト
131K
入力
text
出力
text
入力: $0.710出力: $0.710100万トークンあたり
モデル詳細を見る