Modellspezifikationen
- Kontext
- 10.240 tokens
- Maximale Ausgabe
- 9.216 tokens
- Architektur
- text+image->rerank
- Tokenizer
- Other
- Wissensstand
- Nicht angegeben
- Moderiert
- Nein
Vollständige Preise
Mit OpenRouter synchronisierte Preise; Tokenpreise gelten pro eine Million Token.
- Eingabe
- Kostenlos pro 1 Mio. Token
- Ausgabe
- Kostenlos pro 1 Mio. Token
Schnellstart
Setzen Sie OPENROUTER_API_KEY lokal. Python benötigt requests; JavaScript läuft in Node.js. Bewahren Sie den Schlüssel auf dem Server auf.
curl --fail-with-body https://openrouter.ai/api/v1/rerank \
-H "Authorization: Bearer $OPENROUTER_API_KEY" \
-H "Content-Type: application/json" \
-d '{"model":"nvidia/llama-nemotron-rerank-vl-1b-v2:free","query":"What is the capital of France?","documents":["Paris is the capital of France.","Berlin is the capital of Germany."],"top_n":1}'Eingabe → Ausgabe
Unterstützte API-Parameter
1 Anbieter
Geprüft: 16. September 2026
Verfügbarkeit, Latenz, Durchsatz und Routing ändern sich laufend. Aktuelle Betriebsdaten stehen auf der Quellseite.
Nvidia
Nicht angegeben- Kontext
- 10K
- Maximale Ausgabe
- 9K