nvidia logo
nvidia

NVIDIA: Nemotron 3.5 Lightning (free)

Descripción de la fuente (inglés)

NVIDIA Nemotron 3.5 Lightning is an open mixture-of-experts model from NVIDIA, with 3B active parameters out of 30B total. It is suited for high-throughput agentic workloads and specialized tasks that...

Resumen

Especificaciones del modelo

Contexto
1.000.000 tokens
Salida máxima
65.536 tokens
Arquitectura
text->text
Tokenizador
Other
Corte de conocimiento
No indicado
Moderado
No
OPENROUTER

Precios completos

Tarifas sincronizadas desde OpenRouter, por millón de tokens.

Entrada
Gratis
por millón de tokens
Salida
Gratis
por millón de tokens
API

Configuración del modelo

Razonamiento

Razonamiento obligatorio
No
Parámetros predeterminados
No
API

Inicio rápido

Configura OPENROUTER_API_KEY localmente. Python requiere requests; JavaScript se ejecuta en Node.js. Mantén la clave en el servidor.

Documentación de la API

curl --fail-with-body https://openrouter.ai/api/v1/chat/completions \
  -H "Authorization: Bearer $OPENROUTER_API_KEY" \
  -H "Content-Type: application/json" \
  -d '{"model":"nvidia/nemotron-3.5-lightning:free","messages":[{"role":"user","content":"Hello!"}]}'
Capacidades y modalidades

EntradaSalida

Entrada
text
Salida
text
Razonamiento

No

API

Parámetros API compatibles

include_reasoningmax_tokensreasoningseedtemperaturetool_choicetoolstop_p
Proveedores disponibles

1 Proveedor

Verificado: 7 de septiembre de 2026

Proveedores en vivo en OpenRouter

Disponibilidad, latencia, rendimiento y enrutamiento cambian continuamente. Consulta la fuente para datos actuales.

OpenRouter

Nvidia

nvfp4
Contexto
1M
Salida máxima
66K
Entrada
Gratis
Salida
Gratis
Lectura de caché
No indicado
Escritura de caché
No indicado
nvidia

4 modelos

Todos los modelos
nvidia logonvidia

NVIDIA: Nemotron 3.5 ASR Streaming Multilingual 0.6B

Nemotron 3.5 ASR Streaming Multilingual 0.6B is a speech recognition model from NVIDIA. Its prompt-conditioned, cache-aware FastConformer-RNNT design targets low-latency transcription across more than 40 languages for real-time captioning, voice...

Contexto
No indicado
Entrada
audio
Salida
transcription
Duración del audio: $0.000003 por segundo
Ver detalles del modelo
nvidia logonvidia

NVIDIA: Nemotron 3.5 Lightning

NVIDIA Nemotron 3.5 Lightning is an open mixture-of-experts model from NVIDIA, with 3B active parameters out of 30B total. It is suited for high-throughput agentic workloads and specialized tasks that...

Contexto
262K
Entrada
text
Salida
text
Entrada: $0.08Salida: $0.2por millón de tokens
Ver detalles del modelo
nvidia logonvidia

NVIDIA: Nemotron 3 Embed 1B (free)

NVIDIA Nemotron 3 Embed 1B is an open text embedding model from NVIDIA, optimized for high-throughput, low-latency retrieval. It is suited for enterprise search, RAG, code retrieval, and agentic retrieval...

Contexto
33K
Entrada
text
Salida
embeddings
Entrada: Gratispor millón de tokens
Ver detalles del modelo
nvidia logonvidia

NVIDIA: Llama Nemotron Rerank VL 1B V2 (free)

Llama Nemotron Rerank VL 1B V2 is a 1.7B multimodal reranking model from NVIDIA. It evaluates the relevance of document images and text against user queries, designed for vision RAG...

Contexto
10K
Entrada
text · image
Salida
rerank
Entrada: Gratis por millón de tokensSalida: Gratis por millón de tokens
Ver detalles del modelo