Back to catalog
OpenAI·Textnew65%+ off vs official

GPT-5.6 LunaAPI

GPT-5.6 fast tier — high-volume, low-latency execution at mini-class cost.

Input price

$0.07

per 1M tokens

Output price

$0.42

per 1M tokens

Prompt caching

Pass cache_control on a stable prefix and cache hits bill at a fraction of input. Cache writes bill at the plain input rate.

Cache read
$0.014

Auto-derived from input × 0.2

Specs

Model ID
gpt-5-6-luna
Endpoint
POST /v1/chat/completions
Category
Text
Provider
OpenAI
Capabilities
functionstreaminglong-context

Drop-in OpenAI compatibility

Point your existing OpenAI SDK at api.kunavo.com/v1 and swap the model id. No streaming changes, no SDK changes.

View docsGPT API pricing guide

Try it

Set KUNAVO_API_KEY from /app/keys, then run any of the snippets below.

curl
curl https://api.kunavo.com/v1/chat/completions \
  -H "Authorization: Bearer $KUNAVO_API_KEY" \
  -H "Content-Type: application/json" \
  -d '{
    "model": "gpt-5-6-luna",
    "messages": [
      {"role": "user", "content": "Hello, GPT-5.6 Luna"}
    ],
    "stream": false
  }'

Kunavo vs calling OpenAI directly

Same GPT-5.6 Luna weights, same responses. What a gateway changes is the account, the SDK and the bill — here is the honest side-by-side.

KunavoOpenAI direct
Price (per 1M tokens)$0.07 / $0.42 (−65%)$0.20 / $1.20
AccountOne Kunavo account, key in 2 minutes, $10 minimum top-upA OpenAI account plus its own billing setup
SDKKeep the OpenAI SDK — change base_url, set model to "gpt-5-6-luna"OpenAI's SDK, or their OpenAI-compatible layer where offered
Same key also reachesClaude, Gemini, GPT, image, video and audio modelsOpenAI's own catalog only
BillingOne Stripe balance, pay-as-you-go, never expires; failed requests unbilledA separate invoice per provider
Rate limitsShared gateway capacity, no contractual per-account limitOpenAI's own tier limits, raised by usage history or contract

Go direct to OpenAI when you need a contractual rate limit, an enterprise SLA, or a provider-only feature on day one. Use Kunavo when you want one key, one bill and a lower per-token rate across every provider.

FAQ

How much does GPT-5.6 Luna cost?

On Kunavo, GPT-5.6 Luna is $0.07 per 1M input tokens and $0.42 per 1M output tokens — about 65% under OpenAI's official price. Failed requests are never billed, and it's pay-as-you-go — you only pay for successful calls.

Can I call GPT-5.6 Luna with the OpenAI SDK?

Yes. Set base_url to https://api.kunavo.com/v1, pass your Kunavo key, and set model to "gpt-5-6-luna". Requests and responses are OpenAI-compatible.

What endpoint does GPT-5.6 Luna use?

GPT-5.6 Luna is called via POST /v1/chat/completions on Kunavo's OpenAI-compatible API.

What is GPT-5.6 Luna good for?

GPT-5.6 Luna is a text model from OpenAI, with support for function, streaming, long-context.

What is different from calling OpenAI directly?

The model is the same, at about 65% under OpenAI's list price. What changes is around it: one key and one Stripe balance instead of a OpenAI account with its own billing, the OpenAI SDK with base_url swapped instead of a provider-specific SDK, and the same key also reaching Claude, Gemini, GPT, image, video and audio models. The tradeoff is that gateway capacity is shared — if you need a contractual rate limit or an enterprise SLA on GPT-5.6 Luna specifically, go direct to OpenAI.

Is GPT-5.6 Luna cheaper on Kunavo?

Yes — GPT-5.6 Luna is about 65% under OpenAI's official list price on Kunavo, with no subscription and no minimum monthly spend. You pay per token from a $10 top-up, failed requests are never billed, and the balance never expires.