openai/gpt-5.4-mini
The efficient member of the GPT-5 family, tuned for high-throughput, latency-sensitive workloads. Text + image input, 400K-token context, strong reasoning/coding/tool use at the lowest GPT-5 per-token rate.
Quelle: launch_0023 · Verifiziert am 2026-06-01
Fast $0.15 / $0.9 /M
Cache-Lesezugriff $0.0075/M · Cache-Schreibzugriff —/M
| Stufe | Eingabe | Cache-Lesezugriff | Cache-Schreibzugriff | Ausgabe |
|---|---|---|---|---|
| Standard | $0.075 | $0.0075 | — | $0.45 |
| Priority | $0.15 | $0.015 | — | $0.9 |
Zum API-Aufruf kann jede der IDs verwendet werden.
openai/gpt-5.4-miniErsetzen Sie den Platzhalter ONEHOP_KEY durch Ihren API key. Erstellen →
from openai import OpenAI
client = OpenAI(
base_url="https://api.onehop.ai/v1",
api_key="<ONEHOP_KEY>",
)
completion = client.chat.completions.create(
model="openai/gpt-5.4-mini",
messages=[{"role": "user", "content": "What is the meaning of life?"}],
)
print(completion.choices[0].message.content)OpenAI's flagship GPT-5.6 for the hardest reasoning, coding, and agentic work.
$0.25/M↓ · $1.5/M↑
Balanced GPT-5.6 quality and cost for production workloads.
$0.125/M↓ · $0.75/M↑
The most efficient GPT-5.6 for high-volume and latency-sensitive work.
$0.05/M↓ · $0.3/M↑
OpenAI GPT-5.5 — flagship, 1M context, top reasoning & agentic coding.
$0.5/M↓ · $3/M↑
OpenAI GPT-5.4 — strong general model, 1M context, lower cost than 5.5.
$0.25/M↓ · $1.5/M↑
4.046 Anfragen in den letzten 30 Tagen