The ChatGPT API is the OpenAI API: it's billed per token and has no monthly fee. As of October 5, 2026, OpenAI's latest flagship, GPT-6 Astra, is officially priced at $10.00 input / $50.00 output per 1M tokens; on Kunavo it's $4.00 / $20.00. The everyday workhorse GPT-5.6 Terra costs $0.70 / $4.20 (officially $2.00 / $12.00); the new GPT-6 Sol and GPT-6 Luna, released on September 22, cost $0.80 / $4.00 and $0.04 / $0.20, respectively. Output costs more than input, and reasoning tokens for reasoning models are also billed at the output rate.
The point that causes the most confusion: ChatGPT subscriptions and the OpenAI API are billed separately. A ChatGPT Plus subscription doesn't include an API key or API credits. To call the API from code, you need to apply for a separate key and pay separately per token. This page covers only the latter—calling the API from code.
Rates verified on October 5, 2026. OpenAI's official API prices are listed at openai.com/api/pricing, and ChatGPT plan prices at chatgpt.com/pricing (plan prices change often, so this page doesn't repeat them). Kunavo's figures are drawn from this site's model catalog and match the amounts actually billed by the API.
ChatGPT subscription vs. OpenAI API—two separate bills
| ChatGPT subscription | OpenAI API | |
|---|---|---|
| What you pay for | Monthly plan fee (Free, Go, Plus, Pro, etc.) | Per token, no monthly fee |
| Where it's used | Chat on chatgpt.com and in the app | Call from your code, product, or script |
| API key | Not included | Yes |
| Usage limits | The plan's usage limits | Rate limits and your own budget |
| Months without usage | Monthly fee continues | $0 |
When is a subscription a better deal? Every $20 buys about 1,600 turns on GPT-5.6 Sol or about 4,600 turns on GPT-5.6 Terra (2,000 input / 700 output tokens per turn). If you mainly chat heavily in ChatGPT and use more than that per month while staying within the plan's limits, a monthly plan is a better deal. If you're calling from code or your usage varies widely, paying per token costs less.
GPT API pricing table—per 1M tokens
| Model | Official OpenAI price (input / output) | Kunavo (input / output) | Price difference | Best for |
|---|---|---|---|---|
| GPT-6 Astra | $10.00 / $50.00 | $4.00 / $20.00 | Save approximately 60% | The hardest reasoning and long-running agentic coding |
| GPT-6 Sol | $2.00 / $10.00 | $0.80 / $4.00 | Save approximately 60% | Complex coding and agentic workflows (OpenAI's positioning) |
| GPT-6 Luna | $0.10 / $0.50 | $0.04 / $0.20 | Save approximately 60% | The cheapest tier: focused, high-volume tasks such as classification, extraction, and routing |
| GPT-5.6 Sol | $5.00 / $30.00 Current promotional price $4.00 / $20.00 (at least until November 21, 2026) | $2.00 / $12.00 | Save approximately 60% approximately 40% cheaper than the promotional price | GPT-5.6 flagship for heavy reasoning and coding |
| GPT-5.6 Terra | $2.00 / $12.00 | $0.70 / $4.20 | Save approximately 65% | GPT-5.6 workhorse for everyday RAG, support, and rewriting |
| GPT-5.5 | $5.00 / $30.00 | $2.00 / $12.00 | Save approximately 60% | Previous-generation flagship for existing codebases |
GPT-6 Astra is OpenAI's latest flagship, announced on September 3, 2026. Its output rate on Kunavo is about 1.7 times that of GPT-5.6 Sol, so reserve it for problems cheaper models can't handle instead of using it as the default. GPT-5.5 and GPT-5.6 Sol have the same per-token price (cache writes are calculated differently; see below). One more thing to know: when called through Kunavo, the GPT series currently accepts only text input, so use Claude for tasks that require image input.
GPT-6 Sol and GPT-6 Luna are two models OpenAI released on September 22, 2026. OpenAI says both build on GPT-6 Astra's technical advances; each has a 1,050,000-token context window and a 128,000-token maximum output. OpenAI positions GPT-6 Sol for complex coding and agentic workflows; on Kunavo, it costs $0.80 / $4.00. GPT-6 Luna is described by OpenAI as its most efficient model for focused, high-volume tasks; on Kunavo, it costs $0.04 / $0.20 and is the cheapest GPT in this table. New projects don't have to start with GPT-5.6 Sol: GPT-6 Sol on Kunavo costs $0.80 / $4.00, with both input and output rates lower than GPT-5.6 Sol's $2.00 / $12.00.
What does one request actually cost?
The rate table is less useful than the cost of an actual request. Everything below is calculated from the Kunavo rates in the table above:
| Use case | Tokens (input / output) | GPT-6 Luna | GPT-5.6 Terra | GPT-6 Sol | GPT-5.6 Sol | GPT-6 Astra |
|---|---|---|---|---|---|---|
| One short conversation turn | 1,000 / 300 | $0.0001 | $0.002 | $0.002 | $0.0056 | $0.01 |
| One RAG answer | 6,000 / 500 | $0.00034 | $0.0063 | $0.0068 | $0.018 | $0.034 |
| One hard problem | 20,000 / 3,000 | $0.0014 | $0.027 | $0.028 | $0.076 | $0.140 |
| 100,000-item batch | 500 / 20 × 100,000 | $2.40 | $43.40 | $48.00 | $124.00 | $240.00 |
GPT-6 Sol or GPT-5.6 Terra?GPT-6 Sol's output rate is lower than Terra's, but its input rate is higher. That means Terra is cheaper for requests with lots of input and little output—such as the RAG response and 100,000-item batch above. The higher the share of output, the more economical GPT-6 Sol becomes.
Two things can make the actual amount higher than the table shows. First, reasoning tokens: reasoning models generate thinking tokens before replying. These don't appear in the reply, but are billed at the output rate, so the difference grows with the difficulty of the task. Second, retries: failed requests aren't billed, but if a request succeeds and you ask again because the format is wrong, each request costs money. Check usage in the response to see the actual amount, then calculate your own figures:
# OpenAI API 的費用,用實際 token 數算。
# Kunavo 費率,每 1M token USD:(input, output)
RATES = {
"gpt-6-astra": (4.00, 20.00),
"gpt-6-sol": (0.80, 4.00),
"gpt-6-luna": (0.04, 0.20),
"gpt-5-6-sol": (2.00, 12.00),
"gpt-5-6-terra": (0.70, 4.20),
}
def cost(model, inp, out):
i, o = RATES[model]
# reasoning token 不會出現在回覆裡,但按「輸出」費率計費
return inp / 1e6 * i + out / 1e6 * o
print(cost("gpt-5-6-terra", 6_000, 500)) # RAG 回答一次
print(cost("gpt-6-astra", 20_000, 3_000)) # 難題一次
print(cost("gpt-6-luna", 500, 20) * 100_000) # 10 萬筆批次Long prompts and caching: two rules that change the rate
Requests exceeding 272K tokens. GPT-5.5, GPT-5.6 Sol and Terra, and GPT-6 Astra, Sol, and Luna—in other words, every model in the table—are billed at 2× input / 1.5× output for the entire request when its prompt exceeds 272K tokens. This is OpenAI's published long-context rule, and Kunavo applies it the same way, so the discount relative to the official price stays the same on either side of the threshold. For long documents that can be split, try to keep each part below 272K.
Caching. Cache reads cost 10% of the input rate; cache writes cost GPT-6 Astra, GPT-6 Sol, GPT-6 Luna, GPT-5.6 Sol, and GPT-5.6 Terra: 1.25 times the input rate and GPT-5.5 at the standard input rate. For workflows that send the same long system prompt every time, this is the biggest lever. See the caching documentation for how it works.
Switching from existing OpenAI code
Kunavo is an OpenAI-compatible endpoint, so you don't need to change SDKs: update base_url and the key, and both chat.completions and responses continue to work as before.
from openai import OpenAI
client = OpenAI(
api_key="sk-kn-...",
base_url="https://api.kunavo.com/v1", # 只改這一行
)
resp = client.chat.completions.create(
model="gpt-5-6-terra",
messages=[{"role": "user", "content": "你好"}],
)
print(resp.choices[0].message.content)Full steps are in the quickstart; endpoint specifications are in the Chat Completions documentation (English).
To be clear: when it makes sense to go directly to OpenAI
The gateway uses shared capacity, with no dedicated quota or contractually guaranteed SLA. For a scale that requires guaranteed quota, a specified rate limit, or an SLA, contract directly with OpenAI. If you only chat in ChatGPT and don't write code, you also don't need the API—as calculated above, a monthly plan is usually a better deal for heavy chat use.
Paying from Taiwan
Kunavo balances use Stripe and support cards including JCB (Visa, Mastercard, American Express, JCB, UnionPay), Apple Pay, and Google Pay. It is prepaid, with a minimum top-up of $10; balances do not expire, and failed requests are not charged. The more you top up at once, the more bonus you receive: $100 credits $110, $1,000 credits $1,200, and $5,000 credits $6,250.
Taiwan has no local payment methods—JKoPay and LINE Pay aren't on the supported list. For team use, each key can have a monthly spending limit (returns 402 when reached) and an IP allowlist configured in /app/keys; costs can be split by person. See the billing documentation for details.
How to lower your GPT API costs
- Choose the model for the task size. For simple, high-volume tasks such as classification, extraction, and routing, try GPT-6 Luna first. Use Terra or GPT-6 Sol for everyday work, and move up to Astra only for tasks they can't handle—the table above shows that the same task can cost up to 100 times as much.
- Keep output short. Use
max_tokensand output format constraints to limit length; lower reasoning effort for tasks where you can. - Use caching effectively. Put the fixed system prompt first and avoid changing its content between calls; cache reads cost only 10% of the input rate.
- Keep prompts below 272K tokens. Above that, the whole request is billed at 2× input / 1.5× output.
To estimate a full month's costs, use the OpenAI API pricing calculator (English). If rate limits, rather than cost, are the issue, see OpenAI API rate limits (English). For the same kind of comparison with other providers, see Claude costs and subscriptions. If you use these GPT models with Codex, see Codex costs for how subscriptions and API keys are billed.
Frequently asked questions
How much does the ChatGPT API cost?
The ChatGPT API is the OpenAI API. It's billed per token, has no monthly fee, and is billed separately from a ChatGPT subscription. October 5, 2026, OpenAI's official price per 1M tokens is GPT-6 Astra input $10.00 / output $50.00, GPT-6 Sol $2.00 / $10.00, GPT-6 Luna $0.10 / $0.50, GPT-5.6 Sol $5.00 / $30.00 (OpenAI currently offers it at the promotional price of $4.00 / $20.00, which the official pricing page states will last at least until November 21, 2026), GPT-5.6 Terra $2.00 / $12.00; the same models on Kunavo cost $4.00 / $20.00, $0.80 / $4.00, $0.04 / $0.20, $2.00 / $12.00, and $0.70 / $4.20.
How does OpenAI API pricing work?
Pricing is per token: what you send (system prompts, conversations, files) counts as input, and what the model generates counts as output. Each has a rate per 1M tokens, and output costs more. Reasoning tokens generated by reasoning models before they respond don't appear in the reply, but are billed at the output rate. There's no monthly fee; no calls means $0. Failed requests aren't billed, but if a request succeeds and you ask again, each request costs money.
Does ChatGPT Plus include API access?
No. ChatGPT plans (Free, Go, Plus, Pro, etc.) are monthly subscriptions for chatgpt.com and the app; they don't include an API key or API credits. The OpenAI API requires a separate key and is billed separately per token. Conversely, paying for the API doesn't raise your ChatGPT usage limits. The two don't offset each other.
How much does GPT-6 cost?
Kunavo currently offers GPT-6 Astra, GPT-6 Sol, and GPT-6 Luna. The flagship GPT-6 Astra, announced by OpenAI on September 3, 2026, is officially priced at $10.00 input / $50.00 output per 1M tokens; Kunavo charges $4.00 / $20.00 (Save approximately 60%). A hard problem (20,000 input / 3,000 output tokens) costs about $0.140. GPT-6 Sol and GPT-6 Luna, released on September 22, cost $0.80 / $4.00 and $0.04 / $0.20 on Kunavo. For all three models, a single request with a prompt exceeding 272K tokens is billed at 2× input / 1.5× output for the entire request.
How much do the GPT-6 Sol and GPT-6 Luna APIs cost?
OpenAI released these two models on September 22, 2026. Both have a 1,050,000-token context window and a 128,000-token maximum output. OpenAI positions GPT-6 Sol for complex coding and agentic workflows. Its official price per 1M tokens is $2.00 input / $10.00 output; on Kunavo it's $0.80 / $4.00 (Save approximately 60%). OpenAI describes GPT-6 Luna as its most efficient model for focused, high-volume tasks. Its official price is $0.10 / $0.50; on Kunavo it's $0.04 / $0.20 (Save approximately 60%). One RAG response (6,000 input / 500 output tokens) costs about $0.00034.
Is there a free allowance for the OpenAI API?
OpenAI doesn't offer a free tier for the GPT API; billing per token begins with the first call. Kunavo isn't free either, but has no monthly fee: the minimum prepaid balance is $10, usage is deducted as you go, the balance never expires, and months with no calls cost $0.
How much does one GPT API call cost?
It depends on the model and token count. For one RAG response on Kunavo (6,000 input / 500 output tokens), the costs are: GPT-6 Luna approximately $0.00034, GPT-5.6 Terra approximately $0.0063, GPT-6 Sol approximately $0.0068, GPT-5.6 Sol approximately $0.018, and GPT-6 Astra approximately $0.034. These amounts exclude reasoning tokens; use the output token count in the response's usage field for the actual figure.
How do I pay for OpenAI API usage in Taiwan?
Kunavo uses Stripe prepaid payments: Visa, Mastercard, American Express, JCB, UnionPay, Apple Pay, and Google Pay are supported; the minimum top-up is $10, balances do not expire, and top-ups of $100 or more receive bonuses. Taiwan has no local payment channel—JKOPay and LINE Pay are not available, so an international card is the only option.
What needs to change in my existing OpenAI SDK code?
Change only two things: set base_url to https://api.kunavo.com/v1 and replace the key with a Kunavo key that starts with sk-kn-. Both chat.completions and responses calls work as before. Set model to an ID such as gpt-6-astra or gpt-5-6-terra; the same key can also call Claude.