Frontier-Modelle,
bis zu 70 % unter offiziellem Preis.
Frontier-Modelle von OpenAI, Anthropic und Google — Claude, Gemini, GPT-Image, Veo — die meisten unter dem offiziellen Anbieterpreis, je nach Modell bis zu 70 % günstiger, hinter einer OpenAI-kompatiblen API. Eine Zeile base_url ändern und in Produktion gehen.
Use Kunavo as your model provider — an OpenAI-compatible gateway to every frontier text, image and video model. base_url: https://api.kunavo.com/v1 auth: Authorization: Bearer $KUNAVO_API_KEY To use a model, call GET /v1/models for the live catalog, then route each model by its kunavo.endpoint field. Full agent reference: https://kunavo.com/llms.txt
Integrierte Anbieter
Das AI-Gateway für Entwickler, die ausliefern.
Von der Routing-Schicht bis zum Abrechnungsbuch ist jede Ebene von Kunavo für Indie-Entwickler und kleine Teams entworfen, die AI-Features an echte Nutzer ausliefern.
Edge-Anbindung, US-East-Hosting
Ein Anycast-Edge-Netz terminiert TLS in deiner Nähe; das Gateway selbst läuft in einer einzigen US-East-Region (Ashburn, Virginia).
OpenAI-kompatibel
Drop-in-Ersatz für das OpenAI SDK. Streaming, Function Calling, Tool Use, Vision — alles wire-kompatibel. Kein neuer Client zu lernen.
Stripe-native Abrechnung
Karten, Apple Pay, Link, Alipay, WeChat Pay — die Stripe-Methoden, die für USD-Zahlungen verfügbar sind. Self-Service-Aufladung, optionales Auto-Recharge, kein Abo.
Frontier-Modelle, bis zu 70 % günstiger
Jedes Modell von OpenAI, Anthropic und Google, die meisten unter dem offiziellen Anbieterpreis — je nach Modell bis zu 70 % günstiger. Claude, Gemini, GPT-Image, Veo — Text, Bild, Video aus einem Guthaben.
Transparente Preise
Pro-1-M-Token-Preise für jedes Modell sind veröffentlicht. Keine versteckten Multiplikatoren, keine Überraschungen, keine Abrechnung fehlgeschlagener Anfragen.
Automatisches Failover
Jede Anfrage kann über bis zu drei Upstream-Kanäle durchfallen. Ein Kanal, der zuletzt gescheitert ist, wird umgangen, bevor deine Anfrage überhaupt rausgeht — kein Retry im Client nötig.
Erstklassiges Streaming
Native SSE-Passthrough-Implementierung. Time-to-first-token ist identisch mit dem Upstream — kein Puffern, kein Batching, keine Latenz.
Granulare Nutzungsdaten
Call-by-Call-Analytics nach Modell, Key und IP. Webhook-Zustellung, sobald ein Bild-, Video- oder Musik-Task fertig ist. CSV-Export jederzeit verfügbar.
Prompt-Caching, bis zu 90 % günstiger
Cache-Reads werden bei Claude, GPT und Gemini mit 10 % des Input-Tarifs abgerechnet — ein cache_control in deinem System-Prompt verwandelt lange Kontexte in nahezu kostenlose Re-Reads. Hit-Rate und Ersparnis live im Dashboard.
What to build with Kunavo.
- Customer Support
AI customer support
The fastest-ROI AI deployment in any B2C SaaS — automate ticket triage, draft 80% of responses, and escalate the rest cleanly. Production code, real cost numbers, and the compliance pitfalls that catch teams off-guard.
Explore - Knowledge Base
RAG chatbot API
Most internal knowledge bases are dead documentation — nobody finds anything. A Claude-backed RAG chatbot turns them into a real assistant that cites sources and refuses when it doesn't know. Here's the production pattern.
Explore - Trust & Safety
AI content moderation
Modern moderation isn't just regex — it's nuance: sarcasm, dog whistles, brand-context misuse, image+text combinations. LLMs do this far better than rule-based systems, at a price that scales.
Explore - Developer Tools
AI code assistant
Cursor, Aider, Cline, Continue.dev — they're all powered by the same handful of frontier LLMs. If you're building a coding tool (or a co-pilot inside your own dev product), here's the architecture and the cost reality.
Explore - Data Processing
AI data extraction
The boring, valuable use case. Invoices, receipts, contracts, leads, resumes — anywhere you'd previously have built a parser, an LLM with JSON-mode does it in 30 lines, more accurately, and you can ship in a day instead of a quarter.
Explore
Frontier-Modelle, bis zu 70 % unter offiziellem Preis.
Claude Fable 5
Frontier reasoning and long-horizon agents — the model Fable 5.1 succeeds.
Claude Opus 5
Near-flagship Opus reasoning at half the price of Fable 5 — vision and agentic coding.
Claude Opus 4.8
Anthropic Opus 4.8 — stronger agentic coding and honesty.
Claude Opus 4.7
Anthropic Opus 4.7 — previous-generation Opus; reasoning and vision.
Claude Opus 4.6
Anthropic Opus 4.6 — deep reasoning, exceptional agentic ability.
Claude Sonnet 5
Near-Opus coding and agentic quality at Sonnet cost.
Claude Sonnet 4.6
Balanced speed/quality — the everyday production workhorse, elite coding.
Claude Haiku 4.5
Anthropic Haiku 4.5 — fast and cost-efficient.
Gemini 3.8 Flash
Google's newest Flash — long-horizon software engineering and agentic execution.
Gemini 3.7 Flash
Google's newest Flash — stronger coding and agentic execution at half the 3.6 official rate.
Gemini 3.6 Flash
Gemini 3.6 Flash — thinking-by-default at Flash latency, with native audio input.
Gemini 3.1 Pro
Gemini 3.1 Pro — Google's flagship for coding, agents, and cross-modal analysis.
Gemini 2.5 Flash
Previous-gen Gemini Flash — extreme value.
GPT-6 Astra
OpenAI's newest flagship — frontier reasoning and long-horizon agentic coding.
GPT-5.6 Sol
OpenAI's newest flagship — top-tier reasoning and agentic coding.
GPT-5.6 Terra
GPT-5.6 mid tier — the everyday workhorse of the 5.6 family.
GPT-5.6 Luna
GPT-5.6 fast tier — high-volume, low-latency execution at mini-class cost.
GPT-5.5
OpenAI GPT-5.5 — flagship reasoning and agentic tool use.
Richte deinen Agenten llms.txt
— läuft autonom.
Gib Claude Code, Cursor, Cline — oder jedem OpenAI-kompatiblen Agenten — eine einzige Anweisung. Er lädt den Live-Modellkatalog von Kunavo und steuert Text-, Bild- und Video-Modelle autonom. Kein SDK, kein Glue-Code nötig.
- OpenAI-wire-kompatibel — Agenten brauchen keine Custom-Integration
- GET /v1/models ist der Live-Katalog — niemals Modellnamen hardcoden
- Ein Key, alle Modalitäten: Text, Bild, Video, Audio
Use Kunavo as your model provider — an OpenAI-compatible gateway to every frontier text, image and video model. base_url: https://api.kunavo.com/v1 auth: Authorization: Bearer $KUNAVO_API_KEY To use a model, call GET /v1/models for the live catalog, then route each model by its kunavo.endpoint field. Full agent reference: https://kunavo.com/llms.txt
Je mehr du im Voraus auflädst, desto mehr sparst du.
Prepaid-Wallet. Ab $10. Keine Abos, kein Mindestumsatz, Guthaben verfällt nie.
Starter
Erste Tests
- Zugang zu allen Modellen und Endpunkten
- Call-by-Call-Analytics
- E-Mail-Support
- Kein Abo, kein Verfall — ab $10 aufladen
Builder
Limitiert · +$10Du baust ein Produkt
- $110 Guthaben bei $100 Top-Up
- 10 % Bonus, zeitlich begrenzt
- Priorisierter E-Mail-Support
- Alles, was jeder Account bekommt
Scale
Limitiert · +$250Produktiver Traffic
- $1.250 Guthaben bei $1.000 Top-Up
- 25 % Bonus, zeitlich begrenzt
- Priorisierter E-Mail-Support
- Alles, was jeder Account bekommt
Enterprise
Limitiert · +$2000Großer Maßstab
- $7.000 Guthaben bei $5.000 Top-Up
- 40 % Bonus, zeitlich begrenzt
- Priorisierter E-Mail-Support
- Alles, was jeder Account bekommt
In jedem Account enthalten
- Unbegrenzte API-Keys
- Monatliches Ausgabenlimit pro Key
- IP-Allowlist pro Key
- Auto-Recharge von einer hinterlegten Karte — beliebiger Betrag ab $10, opt-in
- Webhooks für Bild-, Video- und Musik-Tasks
- Usage-API und CSV-Export
- OpenAI-kompatible und Anthropic-Messages-Endpunkte
Start with the popular guides.
- API-Key
Gemini API Key erstellen
Schritt für Schritt in Google AI Studio — oder ein Kunavo-Key ohne Google-Cloud-Projekt.
Read - Preise
Gemini API Preise 2026
Flash & Pro pro 1M Tokens, rund 70% unter Googles Liste, mit nachrechenbaren Beispielen.
Read - Video
Sora API Anleitung
KI-Video am OpenAI-artigen Endpoint — heute live über Google Veo 3.1 ab $0,18 pro Clip.
Read - Setup
Gemini API key
Get a Gemini key and call it with the OpenAI SDK — or use one Kunavo key for everything.
Read - Setup
Claude API key
Get a Claude key and call Claude via the OpenAI SDK or the native Messages API.
Read - Pricing
Claude API pricing 2026
Per-model rates ~60% under Anthropic, worked cost examples, and prompt-caching savings.
Read - Pricing
Gemini API pricing 2026
Gemini 3.8 / 3.7 Flash and 2.5 rates per 1M tokens, set against Google's intro pricing, with worked examples.
Read - Pricing
GPT API pricing 2026
GPT-5.5, 5.6 Sol, Terra and Luna rates ~60% under OpenAI's list, with worked examples.
Read - Concept
LLM gateway
One API for every model — routing, fallback, billing, observability and security, explained.
Read - API
OpenAI-compatible API
Point any OpenAI SDK at hosted Claude, Gemini and GPT by changing only base_url.
Read - Tools
LLM cost calculator
Estimate per-request and monthly API cost for Claude, Gemini and GPT — with savings vs list.
Read - Video
Text-to-video API
Generate video from a prompt on an OpenAI-style endpoint — live today on Google Veo 3.1.
Read - Video
Veo 3.1 API
Google Veo 3.1 text-to-video with native audio, ~50–70% under Google's list.
Read - Audio
Suno API
Generate full songs from a prompt via /v1/audio/music — metered per request, no subscription.
Read - Image
Nano Banana API
Google's image models through an OpenAI-compatible endpoint, ~25–50% under Google's list.
Read - Image
GPT-Image-2 API
OpenAI's image model on the OpenAI-compatible images endpoint at about half of list.
Read - Compare
OpenRouter alternative
Kunavo vs OpenRouter — text breadth vs multimodal, pricing, API keys and payments.
Read - Compare
OpenRouter alternatives 2026
Seven honest alternatives by motivation — cheaper, multimodal, BYO-key or open source.
Read - Models
Gemini 3 API
Gemini 3.8 and 3.7 Flash are live — what they cost against Google's introductory rate, and how to call them.
Read - Setup
CC Switch setup
Add Kunavo to CC Switch for Claude Code and Codex — which API format to pick, and why.
Read
Recent deep dives.
- 实战·5 min
在中国稳定调用 Claude / GPT / Gemini — Kunavo 中国友好路由实测
从北京/上海/深圳调用 Kunavo:无需代理即可直连。支付宝/微信支付/双币卡都可用。健壮重试代码 + 何时真的需要代理。
- 教學·6 min
用 Veo 3 為台灣品牌做短影片廣告(Sora 同端點待上線)— 5 分鐘完整教學
從文字 prompt 到 9:16 直式 Reels / TikTok / IG 短影片,全流程教學。圖生影片用既有產品照片動起來。5 種台灣品牌實際應用、繁中文字渲染注意事項、商業可用品質的進階技巧。
- 実装ガイド·8 min
日本語 RAG チャットボットを Claude で構築 — 5,000 文書のナレッジベースを 30 行で
社内ドキュメント 5,000 件を Claude Sonnet 4.6 で検索可能にする RAG 完全実装。1 クエリ約 0.9 円(prompt caching 適用後)。埋め込みは Kunavo では提供しておらず、その工程のみ OpenAI に直接課金されます。日本語特有のトークン消費・ハルシネーション対策・本番投入チェックリスト含む。
Alles, was du dich
fragst.
Keine Antwort auf deine Frage? Schreib an contact@kunavo.com — wir antworten innerhalb von 24 Stunden.
Kunavo ist speziell für Indie-Entwickler und kleine Teams gebaut, die produktive AI-Features ausliefern. Drei echte Unterschiede: (1) Wir decken Text, Bild, Video und Musik über ein Guthaben ab; (2) Stripe-native Checkouts, Alipay, Apple Pay, WeChat Pay alles dabei — keine Off-Plattform-Rechnungen; (3) Volle Routing-Transparenz — wir tauschen dein Modell niemals heimlich gegen ein günstigeres aus.
Die meisten Modelle liegen unter dem offiziellen Listenpreis des Anbieters — je nach Modell bis zu 70 % günstiger; einige liegen genau auf Listenpreis, und die jeweilige Modellseite sagt das. Größere Top-Ups bringen zusätzlich einen Bonus. Operativ sparst du außerdem: ein Konto, ein Guthaben, ein SDK, keine Mindestabnahme. Der Pro-1-M-Token-Preis jedes Modells steht auf /pricing — jederzeit mit dem Upstream-Listenpreis vergleichbar.
Ja. Wir implementieren das komplette OpenAI-Endpoint-Set: /v1/chat/completions, /v1/embeddings, /v1/images/generations, /v1/models und /v1/video/generations. Streaming, Function Calling, Vision und Tool Use verhalten sich identisch. Projekte, die das OpenAI SDK nutzen, migrieren durch Ändern der base_url — das war's.
Nein. Kunavo ist ein Prepaid-Wallet. Top-Ups bleiben für immer auf deinem Konto — keine Abos, keine Mindestumsätze pro Monat, kein Verfall. Bei Konto-Schließung erstatten wir Restguthaben auf das ursprüngliche Zahlungsmittel.
Niemals. 4xx- und 5xx-Antworten werden nicht abgerechnet. Streaming-Antworten, die mittendrin abbrechen, werden nur für die tatsächlich gelieferten Tokens belastet. Jede Belastung ist call-by-call im Usage-Dashboard sichtbar und als CSV exportierbar.
Karten (Visa, Mastercard, Amex, JCB, UnionPay), Apple Pay, Link, Cash App Pay, Klarna, Amazon Pay, Alipay und WeChat Pay. Lokale Zahlungswege richten sich nach dem Land des Käufers — Pix in Brasilien, Bancontact in Belgien, BLIK in Polen, EPS in Österreich, MB WAY in Portugal —, wobei Stripe den USD-Preis im Checkout in die Landeswährung umrechnet. SEPA-Lastschrift, BACS und BECS gehören nicht zu den von uns akzeptierten Methoden; eine in diesen Ländern ausgestellte Karte funktioniert normal. Auto-Recharge ist opt-in und nur mit Karte möglich: einmal eine Karte hinterlegen, und das Guthaben füllt sich um einen Betrag Ihrer Wahl ab $10 selbst auf.
Kunavo läuft in einer einzigen US-East-Region (Ashburn, Virginia) hinter einem Anycast-Edge, das TLS in deiner Nähe terminiert. Accounts, Abrechnung und Nutzungsdaten liegen in einer einzigen primären Datenbank mit täglichen verschlüsselten Volume-Snapshots.
Drei Minuten bis zum ersten Aufruf.
Eine OpenAI-kompatible API für Claude, Gemini, GPT-Image, Veo und Suno — $10 Mindestaufladung, du zahlst nur für das, was du aufrufst.