openai/gpt-5.4-mini
The efficient member of the GPT-5 family, tuned for high-throughput, latency-sensitive workloads. Text + image input, 400K-token context, strong reasoning/coding/tool use at the lowest GPT-5 per-token rate.
Fonte: launch_0023 · Verificado em 2026-06-01
Fast $0.15 / $0.9 /M
Leitura de cache $0.0075/M · Gravação de cache —/M
| Faixa | Entrada | Leitura de cache | Gravação de cache | Saída |
|---|---|---|---|---|
| Standard | $0.075 | $0.0075 | — | $0.45 |
| Priority | $0.15 | $0.015 | — | $0.9 |
Você pode usar qualquer ID em chamadas de API.
openai/gpt-5.4-miniSubstitua o placeholder ONEHOP_KEY pela sua API key. Criar →
from openai import OpenAI
client = OpenAI(
base_url="https://api.onehop.ai/v1",
api_key="<ONEHOP_KEY>",
)
completion = client.chat.completions.create(
model="openai/gpt-5.4-mini",
messages=[{"role": "user", "content": "What is the meaning of life?"}],
)
print(completion.choices[0].message.content)OpenAI's flagship GPT-5.6 for the hardest reasoning, coding, and agentic work.
$0.25/M↓ · $1.5/M↑
Balanced GPT-5.6 quality and cost for production workloads.
$0.125/M↓ · $0.75/M↑
The most efficient GPT-5.6 for high-volume and latency-sensitive work.
$0.05/M↓ · $0.3/M↑
OpenAI GPT-5.5 — flagship, 1M context, top reasoning & agentic coding.
$0.5/M↓ · $3/M↑
OpenAI GPT-5.4 — strong general model, 1M context, lower cost than 5.5.
$0.25/M↓ · $1.5/M↑
3.577 requisições nos últimos 30 dias