nvidia logo
nvidia

NVIDIA: Nemotron 3 Nano 30B A3B

출처 설명 (영어)

NVIDIA Nemotron 3 Nano 30B A3B is a small language MoE model with highest compute efficiency and accuracy for developers to build specialized agentic AI systems. The model is fully...

개요

모델 사양

컨텍스트
262,144 tokens
최대 출력
235,929 tokens
아키텍처
text->text
토크나이저
Other
지식 기준일
정보 없음
검토됨
아니요
OPENROUTER

전체 가격

OpenRouter에서 동기화한 요금이며 토큰 가격은 100만 토큰 기준입니다.

입력
$0.05
백만 토큰당
출력
$0.2
백만 토큰당
캐시 읽기
$0.03
백만 토큰당
API

모델 구성

추론

추론 필수 여부
아니요
기본 매개변수
아니요
API

빠른 시작

로컬에 OPENROUTER_API_KEY를 설정하세요. Python에는 requests가 필요하며 JavaScript는 Node.js에서 실행됩니다. 키는 서버에 보관하세요.

API 문서

curl --fail-with-body https://openrouter.ai/api/v1/chat/completions \
  -H "Authorization: Bearer $OPENROUTER_API_KEY" \
  -H "Content-Type: application/json" \
  -d '{"model":"nvidia/nemotron-3-nano-30b-a3b","messages":[{"role":"user","content":"Hello!"}]}'
기능 및 모달리티

입력출력

입력
text
출력
text
추론

아니요

API

지원 API 매개변수

frequency_penaltyinclude_reasoninglogit_biaslogprobsmax_tokensmin_ppresence_penaltyreasoningrepetition_penaltyresponse_formatseedstopstructured_outputstemperaturetool_choicetoolstop_ktop_logprobstop_p
사용 가능한 제공업체

4 개 제공업체

확인일: 2026년 9월 22일

OpenRouter 실시간 제공업체

가용성, 지연 시간, 처리량과 라우팅은 계속 바뀝니다. 최신 운영 데이터는 출처 페이지에서 확인하세요.

OpenRouter

Crusoe

fp8
컨텍스트
262K
최대 출력
236K
입력
$0.05
출력
$0.2
캐시 읽기
$0.03
캐시 쓰기
정보 없음

Novita

fp4
컨텍스트
262K
최대 출력
33K
입력
$0.05
출력
$0.2
캐시 읽기
정보 없음
캐시 쓰기
정보 없음

DeepInfra

fp4
컨텍스트
262K
최대 출력
228K
입력
$0.05
출력
$0.2
캐시 읽기
$0.025
캐시 쓰기
정보 없음

Nebius

fp8
컨텍스트
262K
최대 출력
236K
입력
$0.06
출력
$0.24
캐시 읽기
정보 없음
캐시 쓰기
정보 없음
nvidia

4 개 모델

모든 모델
nvidia logonvidia

NVIDIA: Nemotron 3.5 ASR Streaming Multilingual 0.6B

Nemotron 3.5 ASR Streaming Multilingual 0.6B is a speech recognition model from NVIDIA. Its prompt-conditioned, cache-aware FastConformer-RNNT design targets low-latency transcription across more than 40 languages for real-time captioning, voice...

컨텍스트
정보 없음
입력
audio
출력
transcription
오디오 길이: $0.000003 초당
모델 상세 보기
nvidia logonvidia

NVIDIA: Nemotron 3.5 Lightning

NVIDIA Nemotron 3.5 Lightning is an open mixture-of-experts model from NVIDIA, with 3B active parameters out of 30B total. It is suited for high-throughput agentic workloads and specialized tasks that...

컨텍스트
262K
입력
text
출력
text
입력: $0.08출력: $0.2백만 토큰당
모델 상세 보기
nvidia logonvidia

NVIDIA: Nemotron 3.5 Lightning (free)

NVIDIA Nemotron 3.5 Lightning is an open mixture-of-experts model from NVIDIA, with 3B active parameters out of 30B total. It is suited for high-throughput agentic workloads and specialized tasks that...

컨텍스트
1M
입력
text
출력
text
입력: 무료출력: 무료백만 토큰당
모델 상세 보기
nvidia logonvidia

NVIDIA: Nemotron 3 Embed 1B (free)

NVIDIA Nemotron 3 Embed 1B is an open text embedding model from NVIDIA, optimized for high-throughput, low-latency retrieval. It is suited for enterprise search, RAG, code retrieval, and agentic retrieval...

컨텍스트
33K
입력
text
출력
embeddings
입력: 무료백만 토큰당
모델 상세 보기