New — GPT-6 Astra, Gemini 3.8 Flash, and Seedance 2.5 now live

Frontier models,
Up to 70% off official rates.

The frontier models from OpenAI, Anthropic and Google — Claude, Gemini, GPT-Image, Veo — most of them priced below the provider’s official rate, up to 70% off depending on the model, behind a single OpenAI-compatible API. Change one line of base_url and you’re shipping.

5-second setup · No credit card · No minimums
Paste into your AI agent
Use Kunavo as your model provider — an OpenAI-compatible gateway to every frontier text, image and video model.

base_url:  https://api.kunavo.com/v1
auth:      Authorization: Bearer $KUNAVO_API_KEY

To use a model, call GET /v1/models for the live catalog, then route each model by its kunavo.endpoint field. Full agent reference: https://kunavo.com/llms.txt

Providers we’ve unified

OpenAIAnthropicGoogle
37
Models live
up to 70%
Off official rates
$10
Minimum top-up
Automatic
Multi-provider failover
Why Kunavo

The AI gateway built for builders who ship.

From the routing layer to the billing ledger, every part of Kunavo was designed for indie developers and small teams shipping AI features for real customers.

Edge-fronted, US-East hosted

Anycast edge network terminates TLS close to you; the gateway itself runs in a single US-East region (Ashburn, Virginia).

OpenAI-compatible

Drop-in replacement for OpenAI SDKs. Streaming, function calling, tool use, vision — all wire-compatible. No new client to learn.

Stripe-native billing

Card, Apple Pay, Link, Alipay, WeChat Pay — the methods Stripe offers on USD charges. Self-serve top-ups, opt-in auto-recharge, no subscription.

Frontier models, up to 70% off

Every model from OpenAI, Anthropic and Google, most priced below the provider’s official rate — up to 70% off depending on the model. Claude, Gemini, GPT-Image, Veo — text, image and video, one balance.

Transparent pricing

Every model’s per-1M-token price is published. No hidden multipliers, no surprise overages. Failed requests are never billed.

Automatic failover

Every request can fall through up to three upstream channels. A channel that has been failing is bypassed before your request is sent — no client-side retry needed.

First-class streaming

Native SSE pass-through. Time-to-first-token matches the upstream provider — no buffering, no batching, no delay.

Granular usage data

Per-call analytics by model, key and IP. Webhook deliveries when an image, video or music task finishes. Export everything as CSV when you need it.

Prompt caching, up to 90% off

Cache reads bill at 10% of input on Claude, GPT and Gemini — pass cache_control on your system prompt and long context becomes a near-free re-read. Hit rate and savings are shown live in your dashboard.

Model catalog

Frontier models, up to 70% off official.

Browse the full catalog
Anthropic

Claude Fable 5

30%+ OFF

Frontier reasoning and long-horizon agents — the model Fable 5.1 succeeds.

visionfunctionstreamingthinking
$7.00/$35.00
$10 / $50per 1M tokens
Anthropic

Claude Opus 5

60%+ OFFnew

Near-flagship Opus reasoning at half the price of Fable 5 — vision and agentic coding.

visionfunctionstreamingthinking
$2.00/$10.00
$5 / $25per 1M tokens
Anthropic

Claude Opus 4.8

50%+ OFF

Anthropic Opus 4.8 — stronger agentic coding and honesty.

visionfunctionstreamingthinking
$2.50/$12.50
$5 / $25per 1M tokens
Anthropic

Claude Opus 4.7

60%+ OFF

Anthropic Opus 4.7 — previous-generation Opus; reasoning and vision.

visionfunctionstreamingthinking
$2.00/$10.00
$5 / $25per 1M tokens
Anthropic

Claude Opus 4.6

60%+ OFF

Anthropic Opus 4.6 — deep reasoning, exceptional agentic ability.

visionfunctionstreamingthinking
$2.00/$10.00
$5 / $25per 1M tokens
Anthropic

Claude Sonnet 5

Near-Opus coding and agentic quality at Sonnet cost.

visionfunctionstreamingthinking
$2.00/$10.00
$2 / $10per 1M tokens
Anthropic

Claude Sonnet 4.6

60%+ OFFhot

Balanced speed/quality — the everyday production workhorse, elite coding.

visionfunctionstreamingthinking
$1.20/$6.00
$3 / $15per 1M tokens
Anthropic

Claude Haiku 4.5

60%+ OFF

Anthropic Haiku 4.5 — fast and cost-efficient.

visionfunctionstreaming
$0.40/$2.00
$1 / $5per 1M tokens
Google

Gemini 3.8 Flash

65%+ OFFnew

Google's newest Flash — long-horizon software engineering and agentic execution.

visionfunctionstreamingthinking
$0.525/$2.625
$1.5 / $7.5per 1M tokens
Google

Gemini 3.7 Flash

65%+ OFFnew

Google's newest Flash — stronger coding and agentic execution at half the 3.6 official rate.

visionfunctionstreamingthinking
$0.525/$2.625
$1.5 / $7.5per 1M tokens
Google

Gemini 3.6 Flash

30%+ OFFnew

Gemini 3.6 Flash — thinking-by-default at Flash latency, with native audio input.

visionfunctionstreamingthinking
$1.05/$5.25
$1.5 / $7.5per 1M tokens
Google

Gemini 3.1 Pro

65%+ OFF

Gemini 3.1 Pro — Google's flagship for coding, agents, and cross-modal analysis.

visionfunctionstreamingthinking
$0.70/$4.20
$2 / $12per 1M tokens
Google

Gemini 2.5 Flash

70%+ OFF

Previous-gen Gemini Flash — extreme value.

visionfunctionstreaminglong-context
$0.09/$0.75
$0.3 / $2.5per 1M tokens
OpenAI

GPT-6 Astra

60%+ OFFnew

OpenAI's newest flagship — frontier reasoning and long-horizon agentic coding.

functionstreamingthinkinglong-context
$4.00/$20.00
$10 / $50per 1M tokens
OpenAI

GPT-5.6 Sol

60%+ OFF

OpenAI's newest flagship — top-tier reasoning and agentic coding.

functionstreamingthinkinglong-context
$2.00/$12.00
$5 / $30per 1M tokens
OpenAI

GPT-5.6 Terra

65%+ OFF

GPT-5.6 mid tier — the everyday workhorse of the 5.6 family.

functionstreamingthinkinglong-context
$0.70/$4.20
$2 / $12per 1M tokens
OpenAI

GPT-5.6 Luna

65%+ OFF

GPT-5.6 fast tier — high-volume, low-latency execution at mini-class cost.

functionstreaminglong-context
$0.07/$0.42
$0.2 / $1.2per 1M tokens
OpenAI

GPT-5.5

60%+ OFF

OpenAI GPT-5.5 — flagship reasoning and agentic tool use.

functionstreamingthinkinglong-context
$2.00/$12.00
$5 / $30per 1M tokens
For AI agents

Point your agent at llms.txt
It uses every model itself.

Hand one instruction to Claude Code, Cursor, Cline — or any OpenAI-compatible agent. It reads the live model catalog from Kunavo and drives text, image and video models on its own. No SDK, no glue code.

  • OpenAI-wire compatible — agents need no custom integration
  • GET /v1/models is the live catalog — never hardcode model names
  • One key for every modality: text, image, video, audio
Paste into your AI agent
Use Kunavo as your model provider — an OpenAI-compatible gateway to every frontier text, image and video model.

base_url:  https://api.kunavo.com/v1
auth:      Authorization: Bearer $KUNAVO_API_KEY

To use a model, call GET /v1/models for the live catalog, then route each model by its kunavo.endpoint field. Full agent reference: https://kunavo.com/llms.txt
Top up & save

The more you pre-pay, the more you save.

Pre-paid wallet. $10 starts you up. No subscription, no minimum, balance never expires.

Starter

Just exploring

$10
  • Access to every model and endpoint
  • Per-call usage analytics
  • Email support
  • No subscription, no expiry — top up from $10
Sign up free
Most popular

Builder

Limited · +$10

Shipping a product

$100
  • $100 deposit = $110 credit
  • 10% bonus, limited time
  • Priority email support
  • Everything every account gets
Top up $100

Scale

Limited · +$250

Running production traffic

$1000
  • $1000 deposit = $1250 credit
  • 25% bonus, limited time
  • Priority email support
  • Everything every account gets
Top up $1000

Enterprise

Limited · +$2000

High-volume scale

$5000
  • $5000 deposit = $7000 credit
  • 40% bonus, limited time
  • Priority email support
  • Everything every account gets
Top up $5000

Included with every account

  • Unlimited API keys
  • Per-key monthly spend limits
  • Per-key IP allowlists
  • Auto-recharge from a saved card — any whole-dollar amount from $10, opt-in
  • Webhooks for image, video and music tasks
  • Usage API and CSV export
  • OpenAI-compatible and Anthropic Messages endpoints
Guides

Start with the popular guides.

Browse all guides
FAQ

Everything you’re
wondering about.

Didn’t answer your question? Email us at contact@kunavo.com — we reply within 24 hours.

  • Kunavo is purpose-built for indie developers and small teams shipping production AI features. Three real differences: (1) we cover text, image, video and music on one balance; (2) Stripe-native checkout, Alipay, Apple Pay, WeChat Pay all included — no off-platform invoices; (3) full transparency on routing — we never silently swap your model to a cheaper one.

  • Most models are priced below the provider’s official list price — up to 70% off depending on the model; a few sit at list, and each model page says so. Bigger top-ups add a bonus on top. You also save operationally: one account, one balance, one SDK, no commitment minimums. The per-1M-token price for every model is published on /pricing — easy to compare against the upstream listing anytime.

  • Yes. We implement the full set of OpenAI endpoints: /v1/chat/completions, /v1/embeddings, /v1/images/generations, /v1/models and /v1/video/generations. Streaming, function calling, vision and tool use all behave identically. Projects using the OpenAI SDK migrate by changing base_url — that’s it.

  • No. Kunavo is a pre-paid wallet. Top-ups stay in your account forever — no subscriptions, no monthly minimums, no expiration. Account closure refunds remaining balance to your original payment method.

  • Never. 4xx and 5xx responses are not billed. Streaming responses that disconnect mid-flight are billed only for the tokens actually delivered. Every charge is visible per-call in the usage dashboard, exportable as CSV for accounting.

  • Cards (Visa, Mastercard, Amex, JCB, UnionPay), Apple Pay, Link, Cash App Pay, Klarna, Amazon Pay, Alipay and WeChat Pay. Local rails follow the buyer's country — Pix in Brazil, Bancontact in Belgium, BLIK in Poland, EPS in Austria, MB WAY in Portugal — with Stripe converting the USD price to local currency at checkout. SEPA Direct Debit, BACS and BECS are not among the methods we accept; a card issued in those countries works normally. Auto-recharge is opt-in and card-only: save a card once and the wallet refills itself by an amount you choose, from $10.

  • Kunavo runs in one US-East region (Ashburn, Virginia) behind an anycast edge that terminates TLS near you. Accounts, billing and usage data live in a single primary database with daily encrypted volume snapshots.

Three minutes to your first call.

One OpenAI-compatible API for Claude, Gemini, GPT-Image, Veo and Suno — $10 minimum top-up, pay only for what you call.