nvidia logo
nvidia

NVIDIA: Nemotron 3 Nano Omni (free)

NVIDIA Nemotron™ 3 Nano Omni is a 30B-A3B open multimodal model designed to function as a perception and context sub-agent in enterprise agent systems. This model is free to use. 256,000 token context window, maximum output of 65,536 tokens. Includes independent benchmarks from Artificial Analysis.

模型概览

模型规格

上下文
256,000 tokens
最大输出
65,536 tokens
架构
text+image+audio+video->text
分词器
Other
知识截止时间
暂未提供
内容审核
OPENROUTER

完整价格

同步自 OpenRouter。Token 价格均按每百万 Token 展示。

输入
免费
/M tokens
输出
免费
/M tokens
API

模型配置

默认参数

temperature
0.6
top_p
0.95

推理能力

内容审核
默认参数
API

快速调用

通过 OpenRouter 的 OpenAI 兼容 API 调用这个确切的模型 ID。

curl https://openrouter.ai/api/v1/chat/completions \
  -H "Authorization: Bearer $OPENROUTER_API_KEY" \
  -H "Content-Type: application/json" \
  -d '{"model":"nvidia/nemotron-3-nano-omni-30b-a3b-reasoning:free","messages":[{"role":"user","content":"Hello!"}]}'
能力与模态

输入输出

输入
textaudioimagevideo
输出
text
推理能力

API

支持的 API 参数

include_reasoningmax_tokensreasoningseedtemperaturetool_choicetoolstop_p
可用提供方

0 个提供方

在 OpenRouter 查看实时 Provider

Provider 可用性、延迟、吞吐量和路由会持续变化,请前往来源页面查看实时运行数据。

OpenRouter

最近一次同步没有提供端点详情。

nvidia

个模型

全部模型
nvidia logonvidia

NVIDIA: Nemotron 3.5 ASR Streaming Multilingual 0.6B

Nemotron 3.5 ASR Streaming Multilingual 0.6B is a speech recognition model from NVIDIA. Its prompt-conditioned, cache-aware FastConformer-RNNT design targets low-latency transcription across more than 40 languages for real-time captioning, voice...

上下文
暂未提供
输入
audio
输出
transcription
输入: $3.33输出: 免费每百万 Token
查看模型详情
nvidia logonvidia

NVIDIA: Nemotron 3.5 Lightning

NVIDIA Nemotron 3.5 Lightning is an open mixture-of-experts model from NVIDIA, with 3B active parameters out of 30B total. It is suited for high-throughput agentic workloads and specialized tasks that...

上下文
262K
输入
text
输出
text
输入: $0.080输出: $0.200每百万 Token
查看模型详情
nvidia logonvidia

NVIDIA: Nemotron 3.5 Lightning (free)

NVIDIA Nemotron 3.5 Lightning is an open mixture-of-experts model from NVIDIA, with 3B active parameters out of 30B total. It is suited for high-throughput agentic workloads and specialized tasks that...

上下文
1M
输入
text
输出
text
输入: 免费输出: 免费每百万 Token
查看模型详情
nvidia logonvidia

NVIDIA: Nemotron 3 Embed 1B (free)

NVIDIA Nemotron 3 Embed 1B is an open text embedding model from NVIDIA, optimized for high-throughput, low-latency retrieval. It is suited for enterprise search, RAG, code retrieval, and agentic retrieval...

上下文
33K
输入
text
输出
embeddings
输入: 免费输出: 免费每百万 Token
查看模型详情