N
nvidia

NVIDIA: Nemotron 3.5 Lightning (free)

NVIDIA Nemotron 3.5 Lightning is an open mixture-of-experts model from NVIDIA, with 3B active parameters out of 30B total. It is suited for high-throughput agentic workloads and specialized tasks that...

Présentation

Caractéristiques du modèle

Context
1 000 000 tokens
Sortie maximale
65 536 tokens
Architecture
text->text
Tokenizer
Other
Limite des connaissances
Not provided
Modéré
Non
Capacités et modalités

InputOutput

Input
text
Output
text
Raisonnement

Non

API

Paramètres API pris en charge

include_reasoningmax_tokensreasoningseedtemperaturetool_choicetoolstop_p
Fournisseurs disponibles

2 fournisseurs

DeepInfra

bf16
Context
262K
Sortie maximale
131K
Input
$0.080
Output
$0.200
Lecture du cache
$0.040
Écriture du cache
Not provided

CoreWeave

bf16
Context
262K
Sortie maximale
236K
Input
$0.100
Output
$0.250
Lecture du cache
$0.050
Écriture du cache
Not provided
nvidia

models

Tous les modèles