nvidia logo
nvidia

NVIDIA: Llama Nemotron Rerank VL 1B V2 (free)

Descrizione della fonte (inglese)

Llama Nemotron Rerank VL 1B V2 is a 1.7B multimodal reranking model from NVIDIA. It evaluates the relevance of document images and text against user queries, designed for vision RAG...

Panoramica

Specifiche del modello

Contesto
10.240 tokens
Output massimo
9216 tokens
Architettura
text+image->rerank
Tokenizer
Other
Limite di conoscenza
Non indicato
Moderato
No
OPENROUTER

Prezzi completi

Tariffe sincronizzate da OpenRouter, per milione di token.

Input
Gratuito
per milione di token
Output
Gratuito
per milione di token
API

Avvio rapido

Imposta OPENROUTER_API_KEY localmente. Python richiede requests; JavaScript viene eseguito in Node.js. Conserva la chiave sul server.

Documentazione API

curl --fail-with-body https://openrouter.ai/api/v1/rerank \
  -H "Authorization: Bearer $OPENROUTER_API_KEY" \
  -H "Content-Type: application/json" \
  -d '{"model":"nvidia/llama-nemotron-rerank-vl-1b-v2:free","query":"What is the capital of France?","documents":["Paris is the capital of France.","Berlin is the capital of Germany."],"top_n":1}'
Capacità e modalità

InputOutput

Input
textimage
Output
rerank
API

Parametri API supportati

Provider disponibili

1 Provider

Verificato: 16 settembre 2026

Provider live su OpenRouter

Disponibilità, latenza, throughput e routing cambiano continuamente. Consulta la fonte per i dati correnti.

OpenRouter

Nvidia

Non indicato
Contesto
10K
Output massimo
9K
nvidia

4 modelli

Tutti i modelli
nvidia logonvidia

NVIDIA: Nemotron 3.5 ASR Streaming Multilingual 0.6B

Nemotron 3.5 ASR Streaming Multilingual 0.6B is a speech recognition model from NVIDIA. Its prompt-conditioned, cache-aware FastConformer-RNNT design targets low-latency transcription across more than 40 languages for real-time captioning, voice...

Contesto
Non indicato
Input
audio
Output
transcription
Durata audio: $0.000003 al secondo
Vedi i dettagli del modello
nvidia logonvidia

NVIDIA: Nemotron 3.5 Lightning

NVIDIA Nemotron 3.5 Lightning is an open mixture-of-experts model from NVIDIA, with 3B active parameters out of 30B total. It is suited for high-throughput agentic workloads and specialized tasks that...

Contesto
262K
Input
text
Output
text
Input: $0.08Output: $0.2per milione di token
Vedi i dettagli del modello
nvidia logonvidia

NVIDIA: Nemotron 3.5 Lightning (free)

NVIDIA Nemotron 3.5 Lightning is an open mixture-of-experts model from NVIDIA, with 3B active parameters out of 30B total. It is suited for high-throughput agentic workloads and specialized tasks that...

Contesto
1M
Input
text
Output
text
Input: GratuitoOutput: Gratuitoper milione di token
Vedi i dettagli del modello
nvidia logonvidia

NVIDIA: Nemotron 3 Embed 1B (free)

NVIDIA Nemotron 3 Embed 1B is an open text embedding model from NVIDIA, optimized for high-throughput, low-latency retrieval. It is suited for enterprise search, RAG, code retrieval, and agentic retrieval...

Contesto
33K
Input
text
Output
embeddings
Input: Gratuitoper milione di token
Vedi i dettagli del modello