nvidia logo
nvidia

NVIDIA: Nemotron 3 Super (free)

출처 설명 (영어)

NVIDIA Nemotron 3 Super is a 120B-parameter open hybrid MoE model, activating just 12B parameters for maximum compute efficiency and accuracy in complex multi-agent applications. Built on a hybrid Mamba-Transformer...

개요

모델 사양

컨텍스트
262,144 tokens
최대 출력
235,929 tokens
아키텍처
text->text
토크나이저
Other
지식 기준일
정보 없음
검토됨
아니요
OPENROUTER

전체 가격

OpenRouter에서 동기화한 요금이며 토큰 가격은 100만 토큰 기준입니다.

입력
무료
백만 토큰당
출력
무료
백만 토큰당
API

모델 구성

기본 매개변수

temperature
1
top_p
0.95

추론

추론 필수 여부
아니요
기본 매개변수
medium
기능 및 모달리티
medium, low
API

빠른 시작

로컬에 OPENROUTER_API_KEY를 설정하세요. Python에는 requests가 필요하며 JavaScript는 Node.js에서 실행됩니다. 키는 서버에 보관하세요.

API 문서

curl --fail-with-body https://openrouter.ai/api/v1/chat/completions \
  -H "Authorization: Bearer $OPENROUTER_API_KEY" \
  -H "Content-Type: application/json" \
  -d '{"model":"nvidia/nemotron-3-super-120b-a12b:free","messages":[{"role":"user","content":"Hello!"}]}'
기능 및 모달리티

입력출력

입력
text
출력
text
추론

· medium · low

API

지원 API 매개변수

include_reasoningmax_tokensreasoningreasoning_effortresponse_formatseedstructured_outputstemperaturetool_choicetoolstop_p
사용 가능한 제공업체

1 제공업체

확인일: 2026년 9월 7일

OpenRouter 실시간 제공업체

가용성, 지연 시간, 처리량과 라우팅은 계속 바뀝니다. 최신 운영 데이터는 출처 페이지에서 확인하세요.

OpenRouter

Nvidia

정보 없음
컨텍스트
262K
최대 출력
236K
입력
무료
출력
무료
캐시 읽기
정보 없음
캐시 쓰기
정보 없음
nvidia

4 개 모델

모든 모델
nvidia logonvidia

NVIDIA: Nemotron 3.5 ASR Streaming Multilingual 0.6B

Nemotron 3.5 ASR Streaming Multilingual 0.6B is a speech recognition model from NVIDIA. Its prompt-conditioned, cache-aware FastConformer-RNNT design targets low-latency transcription across more than 40 languages for real-time captioning, voice...

컨텍스트
정보 없음
입력
audio
출력
transcription
오디오 길이: $0.000003 초당
모델 상세 보기
nvidia logonvidia

NVIDIA: Nemotron 3.5 Lightning

NVIDIA Nemotron 3.5 Lightning is an open mixture-of-experts model from NVIDIA, with 3B active parameters out of 30B total. It is suited for high-throughput agentic workloads and specialized tasks that...

컨텍스트
262K
입력
text
출력
text
입력: $0.08출력: $0.2백만 토큰당
모델 상세 보기
nvidia logonvidia

NVIDIA: Nemotron 3.5 Lightning (free)

NVIDIA Nemotron 3.5 Lightning is an open mixture-of-experts model from NVIDIA, with 3B active parameters out of 30B total. It is suited for high-throughput agentic workloads and specialized tasks that...

컨텍스트
1M
입력
text
출력
text
입력: 무료출력: 무료백만 토큰당
모델 상세 보기
nvidia logonvidia

NVIDIA: Nemotron 3 Embed 1B (free)

NVIDIA Nemotron 3 Embed 1B is an open text embedding model from NVIDIA, optimized for high-throughput, low-latency retrieval. It is suited for enterprise search, RAG, code retrieval, and agentic retrieval...

컨텍스트
33K
입력
text
출력
embeddings
입력: 무료백만 토큰당
모델 상세 보기