모델 사양
- Context
- 32,768 tokens
- 최대 출력
- 29,491 tokens
- 아키텍처
- text->embeddings
- 토크나이저
- Other
- 지식 기준일
- Not provided
- 검토됨
- 아니요
Input → Output
지원 API 매개변수
max_tokensseedtemperaturetop_p0 개 제공업체
최근 동기화에서 엔드포인트 세부 정보가 제공되지 않았습니다.
NVIDIA Nemotron 3 Embed 1B is an open text embedding model from NVIDIA, optimized for high-throughput, low-latency retrieval. It is suited for enterprise search, RAG, code retrieval, and agentic retrieval...
max_tokensseedtemperaturetop_p최근 동기화에서 엔드포인트 세부 정보가 제공되지 않았습니다.
Nemotron 3.5 ASR Streaming Multilingual 0.6B is a speech recognition model from NVIDIA. Its prompt-conditioned, cache-aware FastConformer-RNNT design targets low-latency transcription across more than 40 languages for real-time captioning, voice...
NVIDIA Nemotron 3.5 Lightning is an open mixture-of-experts model from NVIDIA, with 3B active parameters out of 30B total. It is suited for high-throughput agentic workloads and specialized tasks that...
NVIDIA Nemotron 3.5 Lightning is an open mixture-of-experts model from NVIDIA, with 3B active parameters out of 30B total. It is suited for high-throughput agentic workloads and specialized tasks that...