openai logo
openai

OpenAI: gpt-oss-20b (batch)

Descrizione della fonte (inglese)

gpt-oss-20b is an open-weight 21B parameter model released by OpenAI under the Apache 2.0 license. It uses a Mixture-of-Experts (MoE) architecture with 3.6B active parameters per forward pass, optimized for...

Panoramica

Specifiche del modello

Contesto
131.072 tokens
Output massimo
117.964 tokens
Architettura
text->text
Tokenizer
GPT
Limite di conoscenza
2024-06-30
Moderato
No
OPENROUTER

Prezzi completi

Tariffe sincronizzate da OpenRouter, per milione di token.

Input
$0.024
per milione di token
Output
$0.112
per milione di token
API

Configurazione del modello

Ragionamento

Ragionamento obbligatorio
Parametri predefiniti
medium
Capacità e modalità
high, medium, low
API

Avvio rapido

Imposta OPENROUTER_API_KEY localmente. Python richiede requests; JavaScript viene eseguito in Node.js. Conserva la chiave sul server.

Questa richiesta invia un batch asincrono con l’ID del modello base. L’ID dell’attività conferma l’invio, non il completamento.

Documentazione API

curl --fail-with-body https://openrouter.ai/api/beta/batches \
  -H "Authorization: Bearer $OPENROUTER_API_KEY" \
  -H "Content-Type: application/json" \
  -d '{"endpoint":"/v1/chat/completions","model":"openai/gpt-oss-20b","requests":[{"custom_id":"request-1","body":{"messages":[{"role":"user","content":"Hello!"}]}}]}'

# Replace JOB_ID with the returned id to check progress.
curl --fail-with-body "https://openrouter.ai/api/beta/batches/JOB_ID" \
  -H "Authorization: Bearer $OPENROUTER_API_KEY"
Capacità e modalità

InputOutput

Input
text
Output
text
Ragionamento

No · high · medium · low

API

Parametri API supportati

frequency_penaltyinclude_reasoninglogit_biasmax_tokensmin_ppresence_penaltyreasoningreasoning_effortrepetition_penaltyresponse_formatseedstopstructured_outputstemperaturetool_choicetoolstop_ktop_p
Provider disponibili

1 Provider

Verificato: 23 settembre 2026

Provider live su OpenRouter

Disponibilità, latenza, throughput e routing cambiano continuamente. Consulta la fonte per i dati correnti.

OpenRouter

DeepInfra

bf16
Contesto
131K
Output massimo
118K
Input
$0.024
Output
$0.112
Lettura cache
Non indicato
Scrittura cache
Non indicato
openai

4 modelli

Tutti i modelli
openai logoopenai

OpenAI: GPT-6 Luna Pro

GPT-6 Luna Pro is the same underlying model as GPT-6 Luna, served with reasoning.mode set to pro for higher-quality responses on complex tasks. Learn more in OpenAI's docs: https://developers.openai.com/api/docs/guides/reasoning#reasoning-mode

Contesto
1.1M
Input
file · image · text
Output
text
Input: $0.1Output: $0.5per milione di token
Vedi i dettagli del modello
openai logoopenai

OpenAI: GPT-6 Luna Pro (batch)

GPT-6 Luna Pro is the same underlying model as GPT-6 Luna, served with reasoning.mode set to pro for higher-quality responses on complex tasks. Learn more in OpenAI's docs: https://developers.openai.com/api/docs/guides/reasoning#reasoning-mode

Contesto
1.1M
Input
file · image · text
Output
text
Input: $0.05Output: $0.25per milione di token
Vedi i dettagli del modello
openai logoopenai

OpenAI: GPT-6 Luna

GPT-6 Luna is the fast, cost-efficient model in OpenAI's GPT-6 series, positioned below GPT-6 Sol. It is suited for high-volume and latency-sensitive workloads such as chat, classification, and lightweight agentic...

Contesto
1.1M
Input
file · image · text
Output
text
Input: $0.1Output: $0.5per milione di token
Vedi i dettagli del modello
openai logoopenai

OpenAI: GPT-6 Luna (batch)

GPT-6 Luna is the fast, cost-efficient model in OpenAI's GPT-6 series, positioned below GPT-6 Sol. It is suited for high-volume and latency-sensitive workloads such as chat, classification, and lightweight agentic...

Contesto
1.1M
Input
file · image · text
Output
text
Input: $0.05Output: $0.25per milione di token
Vedi i dettagli del modello