meta-llama logo
meta-llama

Meta: Llama 3.3 70B Instruct

The Meta Llama 3.3 multilingual large language model (LLM) is a pretrained and instruction tuned generative model in 70B (text in/text out). $0.10 per million input tokens, $0.32 per million output tokens. 131,072 token context window, maximum output of 16,384 tokens. Higher uptime with 12 providers. Includes independent benchmarks from Artificial Analysis.

概要

モデル仕様

コンテキスト
131,072 tokens
最大出力
115,200 tokens
アーキテクチャ
text->text
トークナイザー
Llama3
知識カットオフ
2023-12-31
モデレーション
いいえ
OPENROUTER

料金の詳細

OpenRouterと同期した料金です。トークン料金は100万トークン単位です。

入力
$0.1
/M tokens
出力
$0.32
/M tokens
API

クイックスタート

OpenRouterのOpenAI互換APIでこのモデルIDを呼び出します。

curl https://openrouter.ai/api/v1/chat/completions \
  -H "Authorization: Bearer $OPENROUTER_API_KEY" \
  -H "Content-Type: application/json" \
  -d '{"model":"meta-llama/llama-3.3-70b-instruct","messages":[{"role":"user","content":"Hello!"}]}'
機能とモダリティ

入力出力

入力
text
出力
text
API

対応APIパラメータ

frequency_penaltylogit_biaslogprobsmax_tokensmin_ppresence_penaltyrepetition_penaltyresponse_formatseedstopstructured_outputstemperaturetool_choicetoolstop_ktop_logprobstop_p
利用可能なプロバイダー

13 プロバイダー

OpenRouterでプロバイダーを見る

可用性、遅延、スループット、ルーティングは常時変化します。最新情報は出典ページで確認してください。

OpenRouter

DeepInfra

fp8
コンテキスト
131K
最大出力
16K
入力
$0.100
出力
$0.320
キャッシュ読込
情報なし
キャッシュ書込
情報なし

Nebius

fp8
コンテキスト
131K
最大出力
118K
入力
$0.130
出力
$0.400
キャッシュ読込
情報なし
キャッシュ書込
情報なし

Novita

bf16
コンテキスト
12K
最大出力
11K
入力
$0.135
出力
$0.400
キャッシュ読込
情報なし
キャッシュ書込
情報なし

AkashML

fp8
コンテキスト
131K
最大出力
128K
入力
$0.200
出力
$0.520
キャッシュ読込
$0.100
キャッシュ書込
情報なし

Parasail

fp8
コンテキスト
131K
最大出力
16K
入力
$0.220
出力
$0.500
キャッシュ読込
$0.110
キャッシュ書込
情報なし

Crusoe

bf16
コンテキスト
131K
最大出力
118K
入力
$0.250
出力
$0.750
キャッシュ読込
$0.130
キャッシュ書込
情報なし

Cloudflare

fp8
コンテキスト
24K
最大出力
22K
入力
$0.293
出力
$2.25
キャッシュ読込
情報なし
キャッシュ書込
情報なし

SambaNova

unknown
コンテキスト
131K
最大出力
3K
入力
$0.450
出力
$0.900
キャッシュ読込
情報なし
キャッシュ書込
情報なし

Groq

unknown
コンテキスト
131K
最大出力
33K
入力
$0.590
出力
$0.790
キャッシュ読込
$0.295
キャッシュ書込
情報なし

CoreWeave

fp16
コンテキスト
128K
最大出力
115K
入力
$0.710
出力
$0.710
キャッシュ読込
$0.710
キャッシュ書込
情報なし

Google

unknown
コンテキスト
128K
最大出力
8K
入力
$0.720
出力
$0.720
キャッシュ読込
情報なし
キャッシュ書込
情報なし

Google

unknown
コンテキスト
128K
最大出力
115K
入力
$0.720
出力
$0.720
キャッシュ読込
情報なし
キャッシュ書込
情報なし

Together

unknown
コンテキスト
131K
最大出力
2K
入力
$1.04
出力
$1.04
キャッシュ読込
情報なし
キャッシュ書込
情報なし
meta-llama

モデル

すべてのモデル
meta-llama logometa-llama

Meta: Llama Guard 4 12B

Llama Guard 4 is a Llama 4 Scout-derived multimodal pretrained model, fine-tuned for content safety classification. Similar to previous versions, it can be used to classify content in both LLM...

コンテキスト
164K
入力
image · text
出力
text
入力: $0.180出力: $0.180100万トークンあたり
モデル詳細を見る
meta-llama logometa-llama

Meta: Llama 4 Maverick

Llama 4 Maverick 17B Instruct (128E) is a high-capacity multimodal language model from Meta, built on a mixture-of-experts (MoE) architecture with 128 experts and 17 billion active parameters per forward...

コンテキスト
1.0M
入力
text · image
出力
text
入力: $0.200出力: $0.696100万トークンあたり
モデル詳細を見る
meta-llama logometa-llama

Meta: Llama 4 Scout

Llama 4 Scout 17B Instruct (16E) is a mixture-of-experts (MoE) language model developed by Meta, activating 17 billion parameters out of a total of 109B. It supports native multimodal input...

コンテキスト
1.3M
入力
text · image
出力
text
入力: $0.110出力: $0.340100万トークンあたり
モデル詳細を見る
meta-llama logometa-llama

Meta: Llama 3.2 1B Instruct

Llama 3.2 1B is a 1-billion-parameter language model focused on efficiently performing natural language tasks, such as summarization, dialogue, and multilingual text analysis. Its smaller size allows it to operate...

コンテキスト
60K
入力
text
出力
text
入力: $0.027出力: $0.201100万トークンあたり
モデル詳細を見る