New — Gemini 3, Claude Opus 4.7 and Veo 3 now live

Frontier models,
Up to 70% off official rates.

The frontier models from OpenAI, Anthropic and Google — Claude, Gemini, GPT-Image, Veo — every one priced 30–70% below the provider’s official rate, behind a single OpenAI-compatible API. Change one line of base_url and you’re shipping.

5-second setup · No credit card · No minimums
Paste into your AI agent
Use Kunavo as your model provider — an OpenAI-compatible gateway to every frontier text, image and video model.

base_url:  https://api.kunavo.com/v1
auth:      Authorization: Bearer $KUNAVO_API_KEY

To use a model, call GET /v1/models for the live catalog, then route each model by its kunavo.endpoint field. Full agent reference: https://kunavo.com/llms.txt

Providers we’ve unified

OpenAIAnthropicGoogle
30–70%
Off official rates
3,200+
Active developers
240M+
Monthly API calls
99.95%
Uptime SLA
<120ms
P50 latency
41
Models live
Why Kunavo

The AI gateway built for builders who ship.

From the routing layer to the billing ledger, every part of Kunavo was designed for indie developers and small teams shipping AI features for real customers.

Global edge gateway

Multi-region Anycast routing with TLS termination at the edge. P50 under 120ms from anywhere — North America, EU, APAC.

OpenAI-compatible

Drop-in replacement for OpenAI SDKs. Streaming, function calling, tool use, vision — all wire-compatible. No new client to learn.

Stripe-native billing

Card, Apple Pay, Link, ACH, Alipay, WeChat Pay — the methods Stripe offers on USD charges. Self-serve top-ups, auto-recharge, invoices.

Frontier models, up to 70% off

Every model from OpenAI, Anthropic and Google, priced 30–70% below the provider’s official rate. Claude, Gemini, GPT-Image, Veo — text, image and video, one bill.

Transparent pricing

Every model’s per-1M-token price is published. No hidden multipliers, no surprise overages. Failed requests are never billed.

99.95% SLA

Multi-provider failover happens in under 50ms. If one upstream wobbles, your request is rerouted before you notice.

First-class streaming

Native SSE pass-through. Time-to-first-token matches the upstream provider — no buffering, no batching, no delay.

Granular usage data

Per-call analytics by model, key and IP. Webhook deliveries for usage events. Export everything as CSV when you need it.

Prompt caching, up to 90% off

Anthropic cache reads bill at 10% of input — pass cache_control on your system prompt and long context becomes a near-free re-read. Hit rate and savings are shown live in your dashboard.

Model catalog

Frontier models, up to 70% off official.

Browse the full catalog
Anthropic

Claude Fable 5.1

30%+ OFFnew

Anthropic's most capable model — frontier reasoning and long-horizon agentic work.

visionfunctionstreamingthinking
$7.00/$35.00
$10 / $50per 1M tokens
Anthropic

Claude Fable 5

30%+ OFF

Frontier reasoning and long-horizon agents — the model Fable 5.1 succeeds.

visionfunctionstreamingthinking
$7.00/$35.00
$10 / $50per 1M tokens
Anthropic

Claude Opus 5

60%+ OFFnew

Near-flagship Opus reasoning at half the price of Fable 5 — vision and agentic coding.

visionfunctionstreamingthinking
$2.00/$10.00
$5 / $25per 1M tokens
Anthropic

Claude Opus 5 Fast

30%+ OFFnew

Opus 5 tuned for latency — same frontier reasoning, faster output.

visionfunctionstreamingthinking
$7.00/$35.00
$10 / $50per 1M tokens
Anthropic

Claude Opus 4.8

50%+ OFFnew

Anthropic Opus 4.8 — stronger agentic coding and honesty.

visionfunctionstreamingthinking
$2.50/$12.50
$5 / $25per 1M tokens
Anthropic

Claude Opus 4.7

60%+ OFF

Anthropic Opus 4.7 — flagship reasoning and vision.

visionfunctionstreamingthinking
$2.00/$10.00
$5 / $25per 1M tokens
Anthropic

Claude Opus 4.6

60%+ OFF

Anthropic Opus 4.6 — deep reasoning, exceptional agentic ability.

visionfunctionstreamingthinking
$2.00/$10.00
$5 / $25per 1M tokens
Anthropic

Claude Sonnet 5

new

Near-Opus coding and agentic quality at Sonnet cost.

visionfunctionstreamingthinking
$2.00/$10.00
$2 / $10per 1M tokens
Anthropic

Claude Sonnet 4.6

60%+ OFFhot

Balanced speed/quality — the everyday production workhorse, elite coding.

visionfunctionstreamingthinking
$1.20/$6.00
$3 / $15per 1M tokens
Anthropic

Claude Haiku 4.5

60%+ OFF

Anthropic Haiku 4.5 — fast and cost-efficient.

visionfunctionstreaming
$0.40/$2.00
$1 / $5per 1M tokens
Google

Gemini 3.8 Flash

65%+ OFFnew

Google's newest Flash — long-horizon software engineering and agentic execution.

visionfunctionstreamingthinking
$0.525/$2.625
$1.5 / $7.5per 1M tokens
Google

Gemini 3.7 Flash

65%+ OFFnew

Google's newest Flash — stronger coding and agentic execution at half the 3.6 official rate.

visionfunctionstreamingthinking
$0.525/$2.625
$1.5 / $7.5per 1M tokens
Google

Gemini 3.6 Flash

30%+ OFF

Gemini 3.6 Flash — thinking-by-default at Flash latency, with native audio input.

visionfunctionstreamingthinking
$1.05/$5.25
$1.5 / $7.5per 1M tokens
Google

Gemini 3.1 Pro

65%+ OFF

Gemini 3.1 Pro — Google's flagship for coding, agents, and cross-modal analysis.

visionfunctionstreamingthinking
$0.70/$4.20
$2 / $12per 1M tokens
Google

Gemini 2.5 Flash

70%+ OFF

Previous-gen Gemini Flash — extreme value.

visionfunctionstreaminglong-context
$0.09/$0.75
$0.3 / $2.5per 1M tokens
OpenAI

GPT-6 Astra

60%+ OFFnew

OpenAI's newest flagship — frontier reasoning and long-horizon agentic coding.

functionstreamingthinkinglong-context
$4.00/$20.00
$10 / $50per 1M tokens
OpenAI

GPT-5.6 Sol

60%+ OFFnew

OpenAI's newest flagship — top-tier reasoning and agentic coding.

functionstreamingthinkinglong-context
$2.00/$12.00
$5 / $30per 1M tokens
OpenAI

GPT-5.6 Terra

65%+ OFFnew

GPT-5.6 mid tier — the everyday workhorse of the 5.6 family.

functionstreamingthinkinglong-context
$0.70/$4.20
$2 / $12per 1M tokens
OpenAI

GPT-5.6 Luna

65%+ OFFnew

GPT-5.6 fast tier — high-volume, low-latency execution at mini-class cost.

functionstreaminglong-context
$0.07/$0.42
$0.2 / $1.2per 1M tokens
OpenAI

GPT-5.4

60%+ OFFnew

OpenAI GPT-5.4 — strong general reasoning and tool use.

functionstreamingthinkinglong-context
$1.00/$6.00
$2.5 / $15per 1M tokens
OpenAI

GPT-5.5

60%+ OFFnew

OpenAI GPT-5.5 — flagship reasoning and agentic tool use.

functionstreamingthinkinglong-context
$2.00/$12.00
$5 / $30per 1M tokens
OpenAI

GPT-5.4 Mini

70%+ OFF

OpenAI GPT-5.4 Mini — fast, cost-efficient reasoning.

functionstreamingthinkinglong-context
$0.225/$1.35
$0.75 / $4.5per 1M tokens
OpenAI

GPT-5.3 Codex

60%+ OFF

OpenAI GPT-5.3 Codex — coding-specialized.

functionstreamingthinkinglong-context
$0.70/$5.60
$1.75 / $14per 1M tokens
OpenAI

GPT-5.5 Pro

60%+ OFFnew

OpenAI GPT-5.5 Pro — deep-horizon enterprise reasoning.

functionstreamingthinkinglong-context
$12.00/$72.00
$30 / $180per 1M tokens
For AI agents

Point your agent at llms.txt
It uses every model itself.

Hand one instruction to Claude Code, Cursor, Cline — or any OpenAI-compatible agent. It reads the live model catalog from Kunavo and drives text, image and video models on its own. No SDK, no glue code.

  • OpenAI-wire compatible — agents need no custom integration
  • GET /v1/models is the live catalog — never hardcode model names
  • One key for every modality: text, image, video, audio
Paste into your AI agent
Use Kunavo as your model provider — an OpenAI-compatible gateway to every frontier text, image and video model.

base_url:  https://api.kunavo.com/v1
auth:      Authorization: Bearer $KUNAVO_API_KEY

To use a model, call GET /v1/models for the live catalog, then route each model by its kunavo.endpoint field. Full agent reference: https://kunavo.com/llms.txt
Top up & save

The more you pre-pay, the more you save.

Pre-paid wallet. $10 starts you up. No subscription, no minimum, balance never expires.

Starter

Just exploring

$10
  • Access to every model
  • Per-call usage analytics
  • Community & email support
  • No minimum, no credit card
Sign up free
Most popular

Builder

Limited · +$10

Shipping a product

$100
  • $100 deposit = $110 credit
  • 10 isolated API keys
  • Auto-recharge · IP allowlist
  • Priority email support
Top up $100

Scale

Limited · +$250

Running production traffic

$1000
  • $1000 deposit = $1250 credit
  • Unlimited API keys
  • Webhooks · monthly invoices
  • Dedicated Slack/Discord support
Top up $1000

Enterprise

Limited · +$2000

High-volume scale

$5000
  • $5000 deposit = $7000 credit
  • Everything in Scale
  • Custom rate limits & SLA
  • Dedicated account manager
Top up $5000
Guides

Start with the popular guides.

Browse all guides
FAQ

Everything you’re
wondering about.

Didn’t answer your question? Email us at contact@kunavo.com — we reply within 24 hours.

  • Kunavo is purpose-built for indie developers and small teams shipping production AI features. Three real differences: (1) we cover text, image and video under one bill — many aggregators are text-only; (2) Stripe-native checkout, ACH, Alipay, Apple Pay, WeChat Pay all included — no off-platform invoices; (3) full transparency on routing — we never silently swap your model to a cheaper one.

  • Every model is priced at roughly 30–70% below the provider's official list price — and bigger top-ups add a further bonus on top. You also save operationally: one contract, one invoice, one SDK, no commitment minimums. The per-1M-token price for every model is published on /pricing — easy to compare against the upstream listing anytime.

  • Yes. We implement the full set of OpenAI endpoints: /v1/chat/completions, /v1/embeddings, /v1/images/generations, /v1/models and /v1/video/generations. Streaming, function calling, vision and tool use all behave identically. Projects using the OpenAI SDK migrate by changing base_url — that’s it.

  • No. Kunavo is a pre-paid wallet. Top-ups stay in your account forever — no subscriptions, no monthly minimums, no expiration. Account closure refunds remaining balance to your original payment method.

  • Never. 4xx and 5xx responses are not billed. Streaming responses that disconnect mid-flight are billed only for the tokens actually delivered. Every charge is visible per-call in the usage dashboard, exportable as CSV for accounting.

  • Cards (Visa, Mastercard, Amex, JCB, UnionPay), Apple Pay, Link, Cash App Pay, Klarna, Amazon Pay, ACH bank transfer, Alipay and WeChat Pay. Local rails follow the buyer's country — Pix in Brazil, Bancontact in Belgium, BLIK in Poland, EPS in Austria, MB WAY in Portugal — with Stripe converting the USD price to local currency at checkout. SEPA Direct Debit, BACS and BECS are not among the methods we accept; a card issued in those countries works normally. Auto-recharge is opt-in. Enterprise customers can pay by invoice with Net 30 terms.

  • Edge gateway nodes are deployed across North America, Europe and Asia-Pacific. Stateless routing logic runs at the edge for sub-120ms P50 latency. Billing data, accounts and audit logs are stored in a primary region with multi-region replication.

Three minutes to your first call.

One OpenAI-compatible API for Claude, Gemini, GPT-Image, Veo and Suno — $10 minimum top-up, pay only for what you call.