Frontier-Modelle,
30–70 % unter offiziellem Preis.
Frontier-Modelle von OpenAI, Anthropic und Google — Claude, Gemini, GPT-Image, Veo — alle 30–70 % unter dem offiziellen Preis, hinter einer OpenAI-kompatiblen API. Eine Zeile base_url ändern und in Produktion gehen.
Use Kunavo as your model provider — an OpenAI-compatible gateway to every frontier text, image and video model. base_url: https://api.kunavo.com/v1 auth: Authorization: Bearer $KUNAVO_API_KEY To use a model, call GET /v1/models for the live catalog, then route each model by its kunavo.endpoint field. Full agent reference: https://kunavo.com/llms.txt
Integrierte Anbieter
Das AI-Gateway für Entwickler, die ausliefern.
Von der Routing-Schicht bis zum Abrechnungsbuch ist jede Ebene von Kunavo für Indie-Entwickler und kleine Teams entworfen, die AI-Features an echte Nutzer ausliefern.
Globales Edge-Gateway
Multi-Region-Anycast-Routing mit TLS-Terminierung am Edge. P50 unter 120 ms aus Nordamerika, der EU oder dem APAC-Raum.
OpenAI-kompatibel
Drop-in-Ersatz für das OpenAI SDK. Streaming, Function Calling, Tool Use, Vision — alles wire-kompatibel. Kein neuer Client zu lernen.
Stripe-native Abrechnung
Karten, Apple Pay, Google Pay, ACH, SEPA, Alipay, WeChat Pay — alle Stripe-Methoden. Self-Service-Aufladung, Auto-Recharge, automatische Rechnungen.
Frontier-Modelle, 30–70 % günstiger
Jedes Modell von OpenAI, Anthropic und Google zum offiziellen Listenpreis minus 30–70 %. Claude, Gemini, GPT-Image, Veo — Text, Bild, Video auf einer Rechnung.
Transparente Preise
Pro-1-M-Token-Preise für jedes Modell sind veröffentlicht. Keine versteckten Multiplikatoren, keine Überraschungen, keine Abrechnung fehlgeschlagener Anfragen.
99,95 % SLA
Multi-Provider-Failover in unter 50 ms. Wenn ein Upstream wackelt, wird deine Anfrage umgeleitet, bevor du es merkst.
Erstklassiges Streaming
Native SSE-Passthrough-Implementierung. Time-to-first-token ist identisch mit dem Upstream — kein Puffern, kein Batching, keine Latenz.
Granulare Nutzungsdaten
Call-by-Call-Analytics nach Modell, Key und IP. Usage-Events per Webhook. CSV-Export jederzeit verfügbar.
Prompt-Caching, bis zu 90 % günstiger
Anthropic-Cache-Reads werden mit 10 % des Input-Tarifs abgerechnet — ein cache_control in deinem System-Prompt verwandelt lange Kontexte in nahezu kostenlose Re-Reads. Hit-Rate und Ersparnis live im Dashboard.
What to build with Kunavo.
- Customer Support
AI customer support
The fastest-ROI AI deployment in any B2C SaaS — automate ticket triage, draft 80% of responses, and escalate the rest cleanly. Production code, real cost numbers, and the compliance pitfalls that catch teams off-guard.
Explore - Knowledge Base
RAG chatbot API
Most internal knowledge bases are dead documentation — nobody finds anything. A Claude-backed RAG chatbot turns them into a real assistant that cites sources and refuses when it doesn't know. Here's the production pattern.
Explore - Trust & Safety
AI content moderation
Modern moderation isn't just regex — it's nuance: sarcasm, dog whistles, brand-context misuse, image+text combinations. LLMs do this far better than rule-based systems, at a price that scales.
Explore - Developer Tools
AI code assistant
Cursor, Aider, Cline, Continue.dev — they're all powered by the same handful of frontier LLMs. If you're building a coding tool (or a co-pilot inside your own dev product), here's the architecture and the cost reality.
Explore - Data Processing
AI data extraction
The boring, valuable use case. Invoices, receipts, contracts, leads, resumes — anywhere you'd previously have built a parser, an LLM with JSON-mode does it in 30 lines, more accurately, and you can ship in a day instead of a quarter.
Explore
Frontier-Modelle, 30–70 % unter offiziellem Preis.
Claude Fable 5
Anthropic's most capable model — frontier reasoning, long-horizon agents, 1M context.
Claude Opus 4.7
Anthropic's newest Opus — flagship reasoning, vision, 200K context.
Claude Opus 4.6
Anthropic Opus 4.6 — deep reasoning, exceptional agentic ability.
Claude Sonnet 5
Near-Opus coding and agentic quality at Sonnet cost — 1M context.
Claude Sonnet 4.6
Balanced speed/quality — the everyday production workhorse, elite coding.
Claude Haiku 4.5
Anthropic Haiku 4.5 — fast and cost-efficient.
Gemini 2.5 Pro
Previous-gen Gemini Pro — strong reasoning and vision.
Gemini 2.5 Flash
Previous-gen Gemini Flash — extreme value.
GPT-5.4
OpenAI GPT-5.4 — strong general reasoning, 1M context.
GPT-5.5
OpenAI GPT-5.5 — flagship reasoning, 1M context.
GPT-5.4 Mini
OpenAI GPT-5.4 Mini — fast, cost-efficient reasoning.
GPT-5.3 Codex
OpenAI GPT-5.3 Codex — coding-specialized.
GPT-5.5 Pro
OpenAI GPT-5.5 Pro — deep-horizon enterprise reasoning, 1.1M context.
Richte deinen Agenten llms.txt
— läuft autonom.
Gib Claude Code, Cursor, Cline — oder jedem OpenAI-kompatiblen Agenten — eine einzige Anweisung. Er lädt den Live-Modellkatalog von Kunavo und steuert Text-, Bild- und Video-Modelle autonom. Kein SDK, kein Glue-Code nötig.
- OpenAI-wire-kompatibel — Agenten brauchen keine Custom-Integration
- GET /v1/models ist der Live-Katalog — niemals Modellnamen hardcoden
- Ein Key, alle Modalitäten: Text, Bild, Video, Audio
Use Kunavo as your model provider — an OpenAI-compatible gateway to every frontier text, image and video model. base_url: https://api.kunavo.com/v1 auth: Authorization: Bearer $KUNAVO_API_KEY To use a model, call GET /v1/models for the live catalog, then route each model by its kunavo.endpoint field. Full agent reference: https://kunavo.com/llms.txt
Je mehr du im Voraus auflädst, desto mehr sparst du.
Prepaid-Wallet. Ab $10. Keine Abos, kein Mindestumsatz, Guthaben verfällt nie.
Starter
Erste Tests
- Zugang zu allen Modellen
- Call-by-Call-Analytics
- Community- & E-Mail-Support
- Keine Mindestabnahme, keine Karte nötig
Builder
Limitiert · +$10Du baust ein Produkt
- $110 Guthaben bei $100 Top-Up
- 10 separate API-Keys
- Auto-Recharge · IP-Allowlist
- Priorisierter E-Mail-Support
Scale
Limitiert · +$250Produktiver Traffic
- $1.250 Guthaben bei $1.000 Top-Up
- Unbegrenzte API-Keys
- Webhooks · Monatliche Rechnungen
- Dedizierter Slack/Discord-Support
Enterprise
Limitiert · +$2000Großer Maßstab
- $7.000 Guthaben bei $5.000 Top-Up
- Alles aus Scale + mehr
- Custom Rate Limits & SLA
- Persönlicher Account Manager
Start with the popular guides.
- Setup
Gemini API key
Get a Gemini key and call it with the OpenAI SDK — or use one Kunavo key for everything.
Read - Setup
Claude API key
Get a Claude key and call Claude via the OpenAI SDK or the native Messages API.
Read - Pricing
Claude API pricing 2026
Per-model rates ~60% under Anthropic, worked cost examples, and prompt-caching savings.
Read - Pricing
Gemini API pricing 2026
Gemini 2.5 rates ~70% under Google's list, with worked cost examples.
Read - Pricing
GPT API pricing 2026
GPT-5.4, 5.5, Mini and Codex rates ~60% under OpenAI's list, with worked examples.
Read - Concept
LLM gateway
One API for every model — routing, fallback, billing, observability and security, explained.
Read - API
OpenAI-compatible API
Point any OpenAI SDK at hosted Claude, Gemini and GPT by changing only base_url.
Read - Tools
LLM cost calculator
Estimate per-request and monthly API cost for Claude, Gemini and GPT — with savings vs list.
Read - Video
Text-to-video API
Generate video from a prompt on an OpenAI-style endpoint — live today on Google Veo 3.
Read - Video
Veo 3 API
Google Veo 3 text-to-video with native audio, ~40–60% under Google's list.
Read - Audio
Suno API
Generate full songs from a prompt via /v1/audio/music — $0.09 per request.
Read - Image
Nano Banana API
Google's image models through an OpenAI-compatible endpoint at about half of list.
Read - Compare
OpenRouter alternative
Kunavo vs OpenRouter — text breadth vs multimodal, pricing, API keys and payments.
Read - Compare
OpenRouter alternatives 2026
Seven honest alternatives by motivation — cheaper, multimodal, BYO-key or open source.
Read - Models
Gemini 3 API
Gemini 3 pricing at Google's list, availability status, and the live Gemini 2.5 route.
Read
Recent deep dives.
- 实战·5 min
在中国稳定调用 Claude / GPT / Gemini — Kunavo 中国友好路由实测
实测从北京/上海/深圳调用 Kunavo 的延迟和成功率:无需代理,P50 80-150ms,成功率 99.9%+。支付宝/微信支付/双币卡都可用。健壮重试代码 + 何时真的需要代理。
- 教學·6 min
用 Veo 3 為台灣品牌做短影片廣告(Sora 同端點待上線)— 5 分鐘完整教學
從文字 prompt 到 9:16 直式 Reels / TikTok / IG 短影片,全流程教學。圖生影片用既有產品照片動起來。5 種台灣品牌實際應用、繁中文字渲染注意事項、商業可用品質的進階技巧。
- 実装ガイド·8 min
日本語 RAG チャットボットを Claude で構築 — 5,000 文書のナレッジベースを 30 行で
社内ドキュメント 5,000 件を Claude Sonnet 4.6 で検索可能にする RAG 完全実装。埋め込みコスト $0.25、1 クエリ約 0.9 円(prompt caching 適用後)。日本語特有のトークン消費・ハルシネーション対策・本番投入チェックリスト含む。
Alles, was du dich
fragst.
Keine Antwort auf deine Frage? Schreib an contact@kunavo.com — wir antworten innerhalb von 24 Stunden.
Kunavo ist speziell für Indie-Entwickler und kleine Teams gebaut, die produktive AI-Features ausliefern. Drei echte Unterschiede: (1) Wir decken Text, Bild und Video auf einer Rechnung ab — viele Aggregatoren bieten nur Text; (2) Stripe-native Checkouts, ACH, SEPA, Apple Pay, WeChat Pay alles dabei — keine Off-Plattform-Rechnungen; (3) Volle Routing-Transparenz — wir tauschen dein Modell niemals heimlich gegen ein günstigeres aus.
Jedes Modell ist rund 30–70 % unter dem offiziellen Listenpreis des Anbieters — größere Top-Ups bringen einen zusätzlichen Bonus. Operativ sparst du außerdem: ein Vertrag, eine Rechnung, ein SDK, keine Mindestabnahme. Der Pro-1-M-Token-Preis jedes Modells steht auf /pricing — jederzeit mit dem Upstream-Listenpreis vergleichbar.
Ja. Wir implementieren das komplette OpenAI-Endpoint-Set: /v1/chat/completions, /v1/embeddings, /v1/images/generations, /v1/models und /v1/video/generations. Streaming, Function Calling, Vision und Tool Use verhalten sich identisch. Projekte, die das OpenAI SDK nutzen, migrieren durch Ändern der base_url — das war's.
Nein. Kunavo ist ein Prepaid-Wallet. Top-Ups bleiben für immer auf deinem Konto — keine Abos, keine Mindestumsätze pro Monat, kein Verfall. Bei Konto-Schließung erstatten wir Restguthaben auf das ursprüngliche Zahlungsmittel.
Niemals. 4xx- und 5xx-Antworten werden nicht abgerechnet. Streaming-Antworten, die mittendrin abbrechen, werden nur für die tatsächlich gelieferten Tokens belastet. Jede Belastung ist call-by-call im Usage-Dashboard sichtbar und als CSV exportierbar.
Alles, was Stripe unterstützt: Karten (Visa, Mastercard, Amex, JCB, UnionPay), Apple Pay, Google Pay, Link, ACH, SEPA, BACS, BECS, Alipay, WeChat Pay, Klarna, Afterpay und mehr. Auto-Recharge ist opt-in. Enterprise-Kunden können auf Rechnung mit Net-30-Konditionen zahlen.
Edge-Gateway-Nodes laufen in Nordamerika, Europa und Asien-Pazifik. Stateless-Routing-Logik läuft am Edge mit P50-Latenz unter 120 ms. Abrechnungsdaten, Accounts und Audit-Logs liegen in einer Primärregion mit Multi-Region-Replikation.
Drei Minuten bis zum ersten Aufruf.
Eine OpenAI-kompatible API für Claude, Gemini, GPT-Image und 200+ weitere — $5 Mindestaufladung, du zahlst nur für das, was du aufrufst.