nvidia logo
nvidia

NVIDIA: Llama Nemotron Rerank VL 1B V2 (free)

出典の説明(英語)

Llama Nemotron Rerank VL 1B V2 is a 1.7B multimodal reranking model from NVIDIA. It evaluates the relevance of document images and text against user queries, designed for vision RAG...

概要

モデル仕様

コンテキスト
10,240 tokens
最大出力
9,216 tokens
アーキテクチャ
text+image->rerank
トークナイザー
Other
知識カットオフ
情報なし
モデレーション
いいえ
OPENROUTER

料金の詳細

OpenRouterと同期した料金です。トークン料金は100万トークン単位です。

入力
無料
100万トークンあたり
出力
無料
100万トークンあたり
API

クイックスタート

ローカルにOPENROUTER_API_KEYを設定してください。Pythonにはrequestsが必要です。JavaScriptはNode.jsで実行し、キーはサーバー側で管理してください。

APIドキュメント

curl --fail-with-body https://openrouter.ai/api/v1/rerank \
  -H "Authorization: Bearer $OPENROUTER_API_KEY" \
  -H "Content-Type: application/json" \
  -d '{"model":"nvidia/llama-nemotron-rerank-vl-1b-v2:free","query":"What is the capital of France?","documents":["Paris is the capital of France.","Berlin is the capital of Germany."],"top_n":1}'
機能とモダリティ

入力出力

入力
textimage
出力
rerank
API

対応APIパラメータ

利用可能なプロバイダー

1 プロバイダー

確認日: 2026年9月16日

OpenRouterでプロバイダーを見る

可用性、遅延、スループット、ルーティングは常時変化します。最新情報は出典ページで確認してください。

OpenRouter

Nvidia

情報なし
コンテキスト
10K
最大出力
9K
nvidia

4 モデル

すべてのモデル
nvidia logonvidia

NVIDIA: Nemotron 3.5 ASR Streaming Multilingual 0.6B

Nemotron 3.5 ASR Streaming Multilingual 0.6B is a speech recognition model from NVIDIA. Its prompt-conditioned, cache-aware FastConformer-RNNT design targets low-latency transcription across more than 40 languages for real-time captioning, voice...

コンテキスト
情報なし
入力
audio
出力
transcription
音声の長さ: $0.000003 1秒あたり
モデル詳細を見る
nvidia logonvidia

NVIDIA: Nemotron 3.5 Lightning

NVIDIA Nemotron 3.5 Lightning is an open mixture-of-experts model from NVIDIA, with 3B active parameters out of 30B total. It is suited for high-throughput agentic workloads and specialized tasks that...

コンテキスト
262K
入力
text
出力
text
入力: $0.08出力: $0.2100万トークンあたり
モデル詳細を見る
nvidia logonvidia

NVIDIA: Nemotron 3.5 Lightning (free)

NVIDIA Nemotron 3.5 Lightning is an open mixture-of-experts model from NVIDIA, with 3B active parameters out of 30B total. It is suited for high-throughput agentic workloads and specialized tasks that...

コンテキスト
1M
入力
text
出力
text
入力: 無料出力: 無料100万トークンあたり
モデル詳細を見る
nvidia logonvidia

NVIDIA: Nemotron 3 Embed 1B (free)

NVIDIA Nemotron 3 Embed 1B is an open text embedding model from NVIDIA, optimized for high-throughput, low-latency retrieval. It is suited for enterprise search, RAG, code retrieval, and agentic retrieval...

コンテキスト
33K
入力
text
出力
embeddings
入力: 無料100万トークンあたり
モデル詳細を見る