nvidia logo
nvidia

NVIDIA: Nemotron 3 Ultra (free)

NVIDIA Nemotron 3 Ultra is an open frontier-reasoning and orchestration model from NVIDIA, with 55B active parameters out of 550B total (MoE). This model is free to use. 1,000,000 token context window, maximum output of 65,536 tokens. Higher uptime with 5 providers. Includes independent benchmarks from Artificial Analysis.

概要

モデル仕様

コンテキスト
1,000,000 tokens
最大出力
65,536 tokens
アーキテクチャ
text->text
トークナイザー
Other
知識カットオフ
情報なし
モデレーション
いいえ
OPENROUTER

料金の詳細

OpenRouterと同期した料金です。トークン料金は100万トークン単位です。

入力
無料
/M tokens
出力
無料
/M tokens
API

モデル設定

既定パラメータ

temperature
1
top_p
0.95

推論

モデレーション
いいえ
既定パラメータ
high
機能とモダリティ
high, medium
API

クイックスタート

OpenRouterのOpenAI互換APIでこのモデルIDを呼び出します。

curl https://openrouter.ai/api/v1/chat/completions \
  -H "Authorization: Bearer $OPENROUTER_API_KEY" \
  -H "Content-Type: application/json" \
  -d '{"model":"nvidia/nemotron-3-ultra-550b-a55b:free","messages":[{"role":"user","content":"Hello!"}]}'
機能とモダリティ

入力出力

入力
text
出力
text
推論

はい · high · medium

API

対応APIパラメータ

include_reasoningmax_tokensreasoningreasoning_effortseedtemperaturetool_choicetoolstop_p
利用可能なプロバイダー

3 プロバイダー

OpenRouterでプロバイダーを見る

可用性、遅延、スループット、ルーティングは常時変化します。最新情報は出典ページで確認してください。

OpenRouter

DeepInfra

fp4
コンテキスト
262K
最大出力
16K
入力
$0.500
出力
$2.20
キャッシュ読込
$0.100
キャッシュ書込
情報なし

BaseTen

fp4
コンテキスト
203K
最大出力
183K
入力
$0.600
出力
$2.40
キャッシュ読込
$0.120
キャッシュ書込
情報なし

Venice

fp8
コンテキスト
256K
最大出力
33K
入力
$0.625
出力
$3.13
キャッシュ読込
$0.188
キャッシュ書込
情報なし
nvidia

モデル

すべてのモデル
nvidia logonvidia

NVIDIA: Nemotron 3.5 ASR Streaming Multilingual 0.6B

Nemotron 3.5 ASR Streaming Multilingual 0.6B is a speech recognition model from NVIDIA. Its prompt-conditioned, cache-aware FastConformer-RNNT design targets low-latency transcription across more than 40 languages for real-time captioning, voice...

コンテキスト
情報なし
入力
audio
出力
transcription
入力: $3.33出力: 無料100万トークンあたり
モデル詳細を見る
nvidia logonvidia

NVIDIA: Nemotron 3.5 Lightning

NVIDIA Nemotron 3.5 Lightning is an open mixture-of-experts model from NVIDIA, with 3B active parameters out of 30B total. It is suited for high-throughput agentic workloads and specialized tasks that...

コンテキスト
262K
入力
text
出力
text
入力: $0.080出力: $0.200100万トークンあたり
モデル詳細を見る
nvidia logonvidia

NVIDIA: Nemotron 3.5 Lightning (free)

NVIDIA Nemotron 3.5 Lightning is an open mixture-of-experts model from NVIDIA, with 3B active parameters out of 30B total. It is suited for high-throughput agentic workloads and specialized tasks that...

コンテキスト
1M
入力
text
出力
text
入力: 無料出力: 無料100万トークンあたり
モデル詳細を見る
nvidia logonvidia

NVIDIA: Nemotron 3 Embed 1B (free)

NVIDIA Nemotron 3 Embed 1B is an open text embedding model from NVIDIA, optimized for high-throughput, low-latency retrieval. It is suited for enterprise search, RAG, code retrieval, and agentic retrieval...

コンテキスト
33K
入力
text
出力
embeddings
入力: 無料出力: 無料100万トークンあたり
モデル詳細を見る