GPT-5.4 MiniAPI
OpenAI GPT-5.4 Mini — fast, cost-efficient reasoning.
Input price
$0.225
per 1M tokens
Output price
$1.35
per 1M tokens
Prompt caching
Pass cache_control on a stable prefix and cache hits bill at a fraction of input. Cache writes bill at the plain input rate.
Auto-derived from input × 0.2
Specs
- Model ID
gpt-5-4-mini- Endpoint
POST /v1/chat/completions- Category
- Text
- Provider
- OpenAI
- Capabilities
- functionstreamingthinking
Drop-in OpenAI compatibility
Point your existing OpenAI SDK at api.kunavo.com/v1 and swap the model id. No streaming changes, no SDK changes.
View docsGPT API pricing guideTry it
Set KUNAVO_API_KEY from /app/keys, then run any of the snippets below.
curl https://api.kunavo.com/v1/chat/completions \
-H "Authorization: Bearer $KUNAVO_API_KEY" \
-H "Content-Type: application/json" \
-d '{
"model": "gpt-5-4-mini",
"messages": [
{"role": "user", "content": "Hello, GPT-5.4 Mini"}
],
"stream": false
}'Kunavo vs calling OpenAI directly
Same GPT-5.4 Mini weights, same responses. What a gateway changes is the account, the SDK and the bill — here is the honest side-by-side.
| Kunavo | OpenAI direct | |
|---|---|---|
| Price (per 1M tokens) | $0.225 / $1.35 (−70%) | $0.75 / $4.50 |
| Account | One Kunavo account, key in 2 minutes, $5 minimum top-up | A OpenAI account plus its own billing setup |
| SDK | Keep the OpenAI SDK — change base_url, set model to "gpt-5-4-mini" | OpenAI's SDK, or their OpenAI-compatible layer where offered |
| Same key also reaches | Claude, Gemini, GPT, image, video and audio models | OpenAI's own catalog only |
| Billing | One Stripe balance, pay-as-you-go, never expires; failed requests unbilled | A separate invoice per provider |
| Rate limits | Shared gateway capacity, no contractual per-account limit | OpenAI's own tier limits, raised by usage history or contract |
Go direct to OpenAI when you need a contractual rate limit, an enterprise SLA, or a provider-only feature on day one. Use Kunavo when you want one key, one bill and a lower per-token rate across every provider.
FAQ
How much does GPT-5.4 Mini cost?
On Kunavo, GPT-5.4 Mini is $0.225 per 1M input tokens and $1.35 per 1M output tokens — about 70% under OpenAI's official price. Failed requests are never billed, and it's pay-as-you-go — you only pay for successful calls.
Can I call GPT-5.4 Mini with the OpenAI SDK?
Yes. Set base_url to https://api.kunavo.com/v1, pass your Kunavo key, and set model to "gpt-5-4-mini". Requests and responses are OpenAI-compatible.
What endpoint does GPT-5.4 Mini use?
GPT-5.4 Mini is called via POST /v1/chat/completions on Kunavo's OpenAI-compatible API.
What is GPT-5.4 Mini good for?
GPT-5.4 Mini is a text model from OpenAI, with support for function, streaming, thinking.
What is different from calling OpenAI directly?
The model is the same, at about 70% under OpenAI's list price. What changes is around it: one key and one Stripe balance instead of a OpenAI account with its own billing, the OpenAI SDK with base_url swapped instead of a provider-specific SDK, and the same key also reaching Claude, Gemini, GPT, image, video and audio models. The tradeoff is that gateway capacity is shared — if you need a contractual rate limit or an enterprise SLA on GPT-5.4 Mini specifically, go direct to OpenAI.
Is GPT-5.4 Mini cheaper on Kunavo?
Yes — GPT-5.4 Mini is about 70% under OpenAI's official list price on Kunavo, with no subscription and no minimum monthly spend. You pay per token from a $5 top-up, failed requests are never billed, and the balance never expires.