nvidia logo
nvidia

NVIDIA: Llama Nemotron Rerank VL 1B V2 (free)

출처 설명 (영어)

Llama Nemotron Rerank VL 1B V2 is a 1.7B multimodal reranking model from NVIDIA. It evaluates the relevance of document images and text against user queries, designed for vision RAG...

개요

모델 사양

컨텍스트
10,240 tokens
최대 출력
9,216 tokens
아키텍처
text+image->rerank
토크나이저
Other
지식 기준일
정보 없음
검토됨
아니요
OPENROUTER

전체 가격

OpenRouter에서 동기화한 요금이며 토큰 가격은 100만 토큰 기준입니다.

입력
무료
백만 토큰당
출력
무료
백만 토큰당
API

빠른 시작

로컬에 OPENROUTER_API_KEY를 설정하세요. Python에는 requests가 필요하며 JavaScript는 Node.js에서 실행됩니다. 키는 서버에 보관하세요.

API 문서

curl --fail-with-body https://openrouter.ai/api/v1/rerank \
  -H "Authorization: Bearer $OPENROUTER_API_KEY" \
  -H "Content-Type: application/json" \
  -d '{"model":"nvidia/llama-nemotron-rerank-vl-1b-v2:free","query":"What is the capital of France?","documents":["Paris is the capital of France.","Berlin is the capital of Germany."],"top_n":1}'
기능 및 모달리티

입력출력

입력
textimage
출력
rerank
API

지원 API 매개변수

사용 가능한 제공업체

1 제공업체

확인일: 2026년 9월 16일

OpenRouter 실시간 제공업체

가용성, 지연 시간, 처리량과 라우팅은 계속 바뀝니다. 최신 운영 데이터는 출처 페이지에서 확인하세요.

OpenRouter

Nvidia

정보 없음
컨텍스트
10K
최대 출력
9K
nvidia

4 개 모델

모든 모델
nvidia logonvidia

NVIDIA: Nemotron 3.5 ASR Streaming Multilingual 0.6B

Nemotron 3.5 ASR Streaming Multilingual 0.6B is a speech recognition model from NVIDIA. Its prompt-conditioned, cache-aware FastConformer-RNNT design targets low-latency transcription across more than 40 languages for real-time captioning, voice...

컨텍스트
정보 없음
입력
audio
출력
transcription
오디오 길이: $0.000003 초당
모델 상세 보기
nvidia logonvidia

NVIDIA: Nemotron 3.5 Lightning

NVIDIA Nemotron 3.5 Lightning is an open mixture-of-experts model from NVIDIA, with 3B active parameters out of 30B total. It is suited for high-throughput agentic workloads and specialized tasks that...

컨텍스트
262K
입력
text
출력
text
입력: $0.08출력: $0.2백만 토큰당
모델 상세 보기
nvidia logonvidia

NVIDIA: Nemotron 3.5 Lightning (free)

NVIDIA Nemotron 3.5 Lightning is an open mixture-of-experts model from NVIDIA, with 3B active parameters out of 30B total. It is suited for high-throughput agentic workloads and specialized tasks that...

컨텍스트
1M
입력
text
출력
text
입력: 무료출력: 무료백만 토큰당
모델 상세 보기
nvidia logonvidia

NVIDIA: Nemotron 3 Embed 1B (free)

NVIDIA Nemotron 3 Embed 1B is an open text embedding model from NVIDIA, optimized for high-throughput, low-latency retrieval. It is suited for enterprise search, RAG, code retrieval, and agentic retrieval...

컨텍스트
33K
입력
text
출력
embeddings
입력: 무료백만 토큰당
모델 상세 보기