nvidia logo
nvidia

NVIDIA: Llama Nemotron Rerank VL 1B V2 (free)

Beschreibung der Quelle (Englisch)

Llama Nemotron Rerank VL 1B V2 is a 1.7B multimodal reranking model from NVIDIA. It evaluates the relevance of document images and text against user queries, designed for vision RAG...

Überblick

Modellspezifikationen

Kontext
10.240 tokens
Maximale Ausgabe
9.216 tokens
Architektur
text+image->rerank
Tokenizer
Other
Wissensstand
Nicht angegeben
Moderiert
Nein
OPENROUTER

Vollständige Preise

Mit OpenRouter synchronisierte Preise; Tokenpreise gelten pro eine Million Token.

Eingabe
Kostenlos
pro 1 Mio. Token
Ausgabe
Kostenlos
pro 1 Mio. Token
API

Schnellstart

Setzen Sie OPENROUTER_API_KEY lokal. Python benötigt requests; JavaScript läuft in Node.js. Bewahren Sie den Schlüssel auf dem Server auf.

API-Dokumentation

curl --fail-with-body https://openrouter.ai/api/v1/rerank \
  -H "Authorization: Bearer $OPENROUTER_API_KEY" \
  -H "Content-Type: application/json" \
  -d '{"model":"nvidia/llama-nemotron-rerank-vl-1b-v2:free","query":"What is the capital of France?","documents":["Paris is the capital of France.","Berlin is the capital of Germany."],"top_n":1}'
Fähigkeiten und Modalitäten

EingabeAusgabe

Eingabe
textimage
Ausgabe
rerank
API

Unterstützte API-Parameter

Verfügbare Anbieter

1 Anbieter

Geprüft: 16. September 2026

Live-Anbieter bei OpenRouter

Verfügbarkeit, Latenz, Durchsatz und Routing ändern sich laufend. Aktuelle Betriebsdaten stehen auf der Quellseite.

OpenRouter

Nvidia

Nicht angegeben
Kontext
10K
Maximale Ausgabe
9K
nvidia

4 Modelle

Alle Modelle
nvidia logonvidia

NVIDIA: Nemotron 3.5 ASR Streaming Multilingual 0.6B

Nemotron 3.5 ASR Streaming Multilingual 0.6B is a speech recognition model from NVIDIA. Its prompt-conditioned, cache-aware FastConformer-RNNT design targets low-latency transcription across more than 40 languages for real-time captioning, voice...

Kontext
Nicht angegeben
Eingabe
audio
Ausgabe
transcription
Audiodauer: $0.000003 pro Sekunde
Modelldetails ansehen
nvidia logonvidia

NVIDIA: Nemotron 3.5 Lightning

NVIDIA Nemotron 3.5 Lightning is an open mixture-of-experts model from NVIDIA, with 3B active parameters out of 30B total. It is suited for high-throughput agentic workloads and specialized tasks that...

Kontext
262K
Eingabe
text
Ausgabe
text
Eingabe: $0.08Ausgabe: $0.2pro 1 Mio. Token
Modelldetails ansehen
nvidia logonvidia

NVIDIA: Nemotron 3.5 Lightning (free)

NVIDIA Nemotron 3.5 Lightning is an open mixture-of-experts model from NVIDIA, with 3B active parameters out of 30B total. It is suited for high-throughput agentic workloads and specialized tasks that...

Kontext
1M
Eingabe
text
Ausgabe
text
Eingabe: KostenlosAusgabe: Kostenlospro 1 Mio. Token
Modelldetails ansehen
nvidia logonvidia

NVIDIA: Nemotron 3 Embed 1B (free)

NVIDIA Nemotron 3 Embed 1B is an open text embedding model from NVIDIA, optimized for high-throughput, low-latency retrieval. It is suited for enterprise search, RAG, code retrieval, and agentic retrieval...

Kontext
33K
Eingabe
text
Ausgabe
embeddings
Eingabe: Kostenlospro 1 Mio. Token
Modelldetails ansehen