nvidia logo
nvidia

NVIDIA: Nemotron 3 Embed 1B (free)

Beschreibung der Quelle (Englisch)

NVIDIA Nemotron 3 Embed 1B is an open text embedding model from NVIDIA, optimized for high-throughput, low-latency retrieval. It is suited for enterprise search, RAG, code retrieval, and agentic retrieval...

Überblick

Modellspezifikationen

Kontext
32.768 tokens
Maximale Ausgabe
29.491 tokens
Architektur
text->embeddings
Tokenizer
Other
Wissensstand
Nicht angegeben
Moderiert
Nein
OPENROUTER

Vollständige Preise

Mit OpenRouter synchronisierte Preise; Tokenpreise gelten pro eine Million Token.

Eingabe
Kostenlos
pro 1 Mio. Token
API

Schnellstart

Setzen Sie OPENROUTER_API_KEY lokal. Python benötigt requests; JavaScript läuft in Node.js. Bewahren Sie den Schlüssel auf dem Server auf.

API-Dokumentation

curl --fail-with-body https://openrouter.ai/api/v1/embeddings \
  -H "Authorization: Bearer $OPENROUTER_API_KEY" \
  -H "Content-Type: application/json" \
  -d '{"model":"nvidia/nemotron-3-embed-1b:free","input":"Your text here"}'
Fähigkeiten und Modalitäten

EingabeAusgabe

Eingabe
text
Ausgabe
embeddings
API

Unterstützte API-Parameter

Verfügbare Anbieter

1 Anbieter

Geprüft: 16. September 2026

Live-Anbieter bei OpenRouter

Verfügbarkeit, Latenz, Durchsatz und Routing ändern sich laufend. Aktuelle Betriebsdaten stehen auf der Quellseite.

OpenRouter

Nvidia

Nicht angegeben
Kontext
33K
Maximale Ausgabe
29K
Eingabe
Kostenlos
Ausgabe
Kostenlos
Cache-Lesen
Nicht angegeben
Cache-Schreiben
Nicht angegeben
nvidia

4 Modelle

Alle Modelle
nvidia logonvidia

NVIDIA: Nemotron 3.5 ASR Streaming Multilingual 0.6B

Nemotron 3.5 ASR Streaming Multilingual 0.6B is a speech recognition model from NVIDIA. Its prompt-conditioned, cache-aware FastConformer-RNNT design targets low-latency transcription across more than 40 languages for real-time captioning, voice...

Kontext
Nicht angegeben
Eingabe
audio
Ausgabe
transcription
Audiodauer: $0.000003 pro Sekunde
Modelldetails ansehen
nvidia logonvidia

NVIDIA: Nemotron 3.5 Lightning

NVIDIA Nemotron 3.5 Lightning is an open mixture-of-experts model from NVIDIA, with 3B active parameters out of 30B total. It is suited for high-throughput agentic workloads and specialized tasks that...

Kontext
262K
Eingabe
text
Ausgabe
text
Eingabe: $0.08Ausgabe: $0.2pro 1 Mio. Token
Modelldetails ansehen
nvidia logonvidia

NVIDIA: Nemotron 3.5 Lightning (free)

NVIDIA Nemotron 3.5 Lightning is an open mixture-of-experts model from NVIDIA, with 3B active parameters out of 30B total. It is suited for high-throughput agentic workloads and specialized tasks that...

Kontext
1M
Eingabe
text
Ausgabe
text
Eingabe: KostenlosAusgabe: Kostenlospro 1 Mio. Token
Modelldetails ansehen
nvidia logonvidia

NVIDIA: Llama Nemotron Rerank VL 1B V2 (free)

Llama Nemotron Rerank VL 1B V2 is a 1.7B multimodal reranking model from NVIDIA. It evaluates the relevance of document images and text against user queries, designed for vision RAG...

Kontext
10K
Eingabe
text · image
Ausgabe
rerank
Eingabe: Kostenlos pro 1 Mio. TokenAusgabe: Kostenlos pro 1 Mio. Token
Modelldetails ansehen